Beery VMI Report Write-Up: Structure, Sample Language & Common Errors

The Beery VMI (Beery-Buktenica Developmental Test of Visual-Motor Integration, Sixth Edition) is a norm-referenced battery for ages 2 to 100: a core copying test plus optional Visual Perception and Motor Coordination supplemental tests whose contrast localizes a weakness. OTs, school psychologists, and evaluators use it in handwriting and developmental referrals. Reporting only one score discards the design. This page covers how to write up Beery VMI results, with a fictional sample.

Free to use and share. No signup required.
Already have session bullets or a transcript? Generate a structured draft with BastionGPT — you review and sign it.
Who writes it

Occupational therapists, school psychologists, and psychologists; Pearson qualification level B; scoring is by hand against the manual's criteria

Audience

IEP and Section 504 teams, school OT services, psychoeducational and neuropsychological batteries, pediatric rehabilitation, payers reviewing evaluations, parents

Typical length

200 to 500 words for the results section · core form 10 to 15 minutes; each supplemental about 5 minutes

Format family

Norm-referenced visual-motor battery (core copying test plus optional Visual Perception and Motor Coordination supplementals; ages 2 to 100)

When it's used

Handwriting and fine-motor referrals, psychoeducational and developmental evaluations, OT evaluations, rehabilitation and neuropsychological batteries across the lifespan

Standards context

Sixth edition (Pearson, 2010; adult norms collected 2006) remains current as of August 2026; no authority mandates it; described here for write-up purposes, no test content reproduced

What is the Beery VMI?

The Beery VMI (Beery-Buktenica Developmental Test of Visual-Motor Integration, Sixth Edition; Beery, Buktenica, and Beery; Pearson, 2010) is the most widely used visual-motor measure in school and pediatric practice, and it is a three-test system, not one score. The core test asks the examinee to copy geometric forms of increasing difficulty, a task that blends visual analysis, planning, and motor production; two optional supplemental tests use the same forms to pull those threads apart, Visual Perception (matching, with the motor demand stripped out) and Motor Coordination (controlled pencil work within boundaries, with the perceptual analysis reduced). There are four administration forms behind the three tests: the core exists as a Full Form (ages 2 to 100) and a Short Form (commonly used for ages 2 through 7), and when both supplementals are given the publisher fixes the sequence, core first, then Visual Perception, then Motor Coordination. Each test yields its own standard score (mean 100, SD 15) and percentile, scored by hand against developmental criteria on the examinee's first production; there is no combined three-test composite, no automated scoring platform, and no seventh edition announced as of August 2026.

The load-bearing facts for the write-up are the contrast and its limits. The supplementals are what let a low copying score be localized, toward perception, motor execution, or the integration itself, and the strongest evidence in the literature says they carry independent value: in one achievement study the Visual Perception test out-predicted the core for math. But the contrast is a hypothesis generator, not a decomposition: across studies the two supplementals explain only about 22 to 36 percent of core-score variance, a normal core can coexist with a weak supplemental, and no authority publishes a universal discrepancy cutoff, so pattern language stays at "suggests," never "proves." A core-only administration is legitimate screening when its limits are stated; a core-only report claiming an "integration deficit" is not. Results usually land in a psychoeducational report beside cognitive measures like the WISC-V and achievement measures like the WIAT-4, whose writing-fluency tasks approach the same referral question from the academic side.

Who uses the Beery VMI and when

School-based occupational therapists are the heaviest users, typically on handwriting and fine-motor referrals feeding IEP and Section 504 processes, where federal law requires a variety of tools and prohibits any single measure from deciding eligibility, so the Beery contributes evidence and never a verdict. School psychologists and psychologists fold it into psychoeducational and developmental batteries from age 2 (it is one of very few instruments normed that young); neuropsychologists use it across the lifespan in rehabilitation and dementia work, where the write-up should note that the adult norms were collected in 2006 and carried forward from the fifth edition. Pediatric evaluations pair it with sensory questionnaires like the Sensory Profile 2 (a caregiver-report instrument measuring a different construct) and with the NEPSY-II's design-copying and visuomotor subtests inside fuller neuropsychological batteries. The report's readers are practical: an IEP team deciding services, a teacher looking for classroom implications, a payer reviewing an OT evaluation billed under the 97165 family (the test rides inside the evaluation, not as a separate charge), or a physician who referred for "messy handwriting" and needs to hear what the instrument can and cannot say about it. The broader evaluation frame lives on the developmental assessment page.

How to structure a Beery VMI results section

No authority prescribes a Beery VMI report format, but the instrument's three-test design dictates the write-up's logic: name exactly what was administered, report every administered test separately, put the behavior beside the scores, interpret the pattern as hypothesis, state the boundaries, and land on function. Each section below carries the pitfall that most often undermines it.

Instrument identification, exactly. Name the edition and everything administered: Beery VMI Sixth Edition, core Full Form or Short Form, plus whichever supplementals were given, with the administration context (individual or group; the supplementals are individually administered) and any departure from standard procedure. When both supplementals were given, they followed the core in the publisher's fixed sequence. Pitfall: "Beery VMI: 82." Three tests and four forms live under one name, and a score without its test and form is as ambiguous as a Wechsler score without its index.

Supplementals administered, or their absence justified. The publisher frames Visual Perception and Motor Coordination as optional follow-up tests, so a core-only screening is legitimate, but say so: state that without the supplementals the result cannot determine whether a weakness is predominantly perceptual, motor, or integrative. When the referral asks why performance is low, administer them; the Visual Perception test has out-predicted the core for academic outcomes in published work. Pitfall: A core-only score doing localization work ("a visual-motor integration deficit") or, the opposite error, supplementals described as universally mandatory. The precise rule is optional-but-sequenced, and the limits travel with the choice.

Scores with the right labels. Report each administered test separately: standard score (mean 100, SD 15) and percentile, with the measurement-error framing the manual supports. Age equivalents stay secondary and descriptive. Scoring is by hand, on the first production (no erase-and-retry, no best-of-attempts), against the manual's developmental criteria; there is no tri-test composite to report. Pitfall: Age equivalents headlining ("functions like a 4-year-old"), an invented composite, or a score from a coached, retried administration presented as standardized.

Observations beside the scores. Place the behavior next to the numbers: hand dominance and switching, grasp and pressure, page stabilization, posture and viewing distance, tremor or line control, pace and speed-accuracy trade-off, impulsive starts versus careful checking, fatigue, frustration, and comprehension of the task. These observations carry the functional interpretation and the validity call. Pitfall: Scores reported bare. A rushed one-minute administration is a behavioral finding on an untimed task, not a scoring artifact, and only the observations make that readable.

Pattern interpretation, at hypothesis strength. Interpret the three-score contrast in plain words: low core with stronger supplementals suggests the combined demands were the difficulty; low core with weakest Motor Coordination supports a graphomotor hypothesis; weakest Visual Perception directs attention to form analysis and a vision check. Anchor any between-test gap to measurement error, and remember the supplementals explain only a minority (roughly 22 to 36 percent) of core variance. Pitfall: A contrast promoted to a diagnosis ("proves a motor deficit"), or a universal discrepancy cutoff applied: no public authority establishes one, and independent retest limits of agreement have spanned roughly 27 points.

Boundaries: what the instrument does not establish. State them: the Beery VMI is not a handwriting assessment (in a developmental coordination disorder sample it did not correlate with handwriting product or process), it does not diagnose dysgraphia or determine eligibility (single measures cannot), and it is not an eye examination, though a weak Visual Perception result warrants excluding a near-vision problem, since induced blur demonstrably lowers all three scores. Pitfall: "Low VMI, consistent with dysgraphia." Handwriting conclusions require direct writing samples, and a low score with unchecked vision may be an optometry finding wearing an OT label.

Functional linkage, goals, and progress plan. Connect the localized hypothesis to observed function: classroom writing samples, tool use, endurance, participation. Write goals on the activity (legibility, sustainable output, efficient grasp, accommodations), never on the score, and plan progress measurement through repeated authentic samples: retest reliability is moderate, no minimal clinically important difference has been established, and direct handwriting gains have occurred without Beery score movement. Pitfall: "Increase the Motor Coordination standard score to 90" as a goal, or re-administration every grading period as a progress measure. The instrument is poorly suited to frequent monitoring, and the publisher's own guidance sets a minimum retest interval.

Blank template (copy and adapt)

BEERY VMI RESULTS SECTION SKELETON
Client: [initials]   Evaluation date: [ ]   Evaluator + credentials: [ ]
Referral question: [why visual-motor assessment was selected]
Instrument: Beery-Buktenica Developmental Test of Visual-Motor Integration,
   Sixth Edition
Tests administered: [core Full Form / core Short Form] [+ Visual Perception]
   [+ Motor Coordination] (sequence per manual when both supplementals given)
   Administration: [individual / group core]   Deviations: [ ]
If core-only: [stated limits: result does not determine whether a weakness
   is perceptual, motor, or integrative; reason supplementals omitted]
Scores (each test separately; hand-scored, first production):
   Core VMI: SS [ ] (mean 100, SD 15), percentile [ ]
   Visual Perception: SS [ ], percentile [ ]
   Motor Coordination: SS [ ], percentile [ ]
   (Age equivalents secondary and descriptive only; no tri-test composite)
Observations (beside the scores): [dominance, grasp, pressure, stabilization,
   posture, tremor or line control, pace, speed-accuracy trade-off,
   checking vs impulsivity, fatigue, comprehension]
Vision status: [corrective lenses worn; recent screening; low Visual
   Perception = exclude near-vision problem before interpreting]
Pattern interpretation (hypothesis strength): [what the contrast suggests;
   gaps anchored to measurement error; supplementals explain only a
   minority of core variance]
Boundaries stated: [not a handwriting assessment; does not diagnose
   dysgraphia or determine eligibility; not an eye examination]
Functional corroboration: [classroom or daily-task evidence: legibility,
   endurance, tool use, participation]
Goals + progress plan: [activity-based goals; progress via repeated
   authentic samples, not score re-administration; retest interval
   respected if a retest is planned]
Evaluator signature / credentials:            Date:

Free to use and share, no signup. The PDF includes a one-page cheat sheet with section-by-section pitfalls and a pre-sign checklist; the DOCX is the blank results-section skeleton, ready to adapt. Neither reproduces test forms, geometric stimuli, scoring criteria, or norms.

Sample Beery VMI write-up (fictional)

Scenario: a kindergarten handwriting referral where the supplemental contrast supports a graphomotor hypothesis, corroborated by classroom samples and kept at hypothesis strength. All details are fictional.

Patient: A.K., 5  ·  Setting: School-based occupational therapy evaluation  ·  Clinician: M. Duarte, OTR/L  ·  Note date: 08/22/2026

Referral and instrument: A.K. was referred by his kindergarten team for effortful, poorly legible printing and avoidance of drawing tasks. As part of a broader occupational therapy evaluation, he was administered the Beery-Buktenica Developmental Test of Visual-Motor Integration, Sixth Edition: the core Full Form followed by the Visual Perception and Motor Coordination supplemental tests in the publisher's sequence, individually, under standard conditions. He wears no corrective lenses and passed the school vision screening this month.

Results: Core Visual-Motor Integration: standard score 82 (12th percentile), below the average range. Visual Perception: standard score 98 (45th percentile), squarely age-expected. Motor Coordination: standard score 76 (5th percentile), his weakest performance. Scores are reported separately by design; the instrument produces no combined composite, and each result carries the measurement-error framing of the manual. Age equivalents are omitted as secondary.

Observations: A.K. showed consistent right-hand use with his left hand stabilizing the page. His grasp tightened visibly as precision demands increased, pencil pressure was heavy enough to indent the page, and he slowed markedly to stay within boundaries, trading speed for control until fatigue loosened both. On the matching task he scanned efficiently and answered promptly, with none of the effort that pencil work produced. He understood every task and persisted well with encouragement.

Interpretation: The pattern, a below-average copying score beside age-expected visual matching and a clearly weaker constrained-pencil-control result, supports a working hypothesis that graphomotor precision and efficiency, more than visual analysis, are constraining A.K.'s visual-motor performance. This is stated as a hypothesis rather than proof: the supplemental tests account for only a minority of core-score variance, and score gaps of this size are interpreted alongside, not instead of, the observed behavior. The Beery VMI is not a handwriting assessment, does not diagnose dysgraphia, and is not an eye examination; his classroom writing samples, which show adequate letter recognition but declining legibility and endurance across longer tasks, are the functional evidence, and they corroborate the hypothesis.

Recommendations: Intervention should target the activity, not the score: sustainable pencil control and writing endurance during authentic classroom tasks, an efficient grasp and page-stabilization strategy, and short-burst writing with planned breaks while stamina builds, with accommodations (reduced copying volume, pencil grip trial) as the team judges useful. Progress will be measured through repeated classroom writing samples rather than Beery re-administration, which is poorly suited to frequent monitoring; any future retest will respect the publisher's minimum interval and be interpreted against measurement error. These findings are one contribution to the team's evaluation: no single measure determines eligibility or services, and the full record, teacher input, and family perspective complete the picture.

This sample is fictional and for educational purposes. It does not describe a real child or record; the scores, observations, dates, and details are invented to show write-up structure and are not clinical guidance. Scores are invented for illustration and correspond to no real child or record, and no test stimuli or norm-table values are reproduced.

↑ Back to the template and downloads

Why this sample works

  • Everything administered is named: edition, core form, both supplementals, and their sequence, so "the VMI score" ambiguity never arises.
  • All three tests are reported separately with the right labels, age equivalents are declined, and no invented composite appears.
  • The observations sit beside the scores and do the functional work: the grasp, pressure, and speed-accuracy trade-off are what make the motor hypothesis readable.
  • The interpretation stays at hypothesis strength with its evidence limits stated, and the boundaries (not handwriting, not dysgraphia, not an eye exam) are explicit, with vision checked before the perceptual score was trusted.
  • The goals target classroom function with progress measured by authentic samples, and the eligibility frame keeps the instrument contributory, never decisive.

Writing these after every session? BastionGPT drafts complete notes from bullets, dictation, or a transcript.

Generate a note from bullets

Documentation and compliance considerations

United States: the legal frame is generic, and that is the point. IDEA's evaluation rules require a variety of assessment tools, prohibit any single measure or assessment as the sole criterion for disability determination or programming, require technically sound instruments, and require assessment in all areas of suspected disability, and no category (specific learning disability, developmental delay, orthopedic impairment, autism) names the Beery VMI; Section 504's evaluation rules likewise require multiple sources and suited tests (LAW). So the report contributes to a team decision and never announces one. In tiered handwriting support, direct repeated writing samples, legibility, speed, endurance, and participation, are the appropriate progress measures, with the Beery as an initial hypothesis-former at most (CONVENTION). Billing follows the container: an OT-administered Beery rides inside the occupational therapy evaluation codes (97165 to 97167, re-evaluation 97168) rather than generating a separate testing charge, while a psychologist folding it into a battery works under the 96130-series evaluation and administration codes, all subject to payer medical-necessity and documentation rules (PAYER POLICY). State OT practice acts add documentation requirements that vary by state and never prescribe the instrument (LAW or PROFESSIONAL REGULATION, state by state).

Canada and Australia put the same instrument inside participation-first frameworks. Canadian school OT runs through provincial arrangements and regulatory colleges whose standards require suitable, evidence-informed assessment, competence, and clear documentation without mandating named tools, and Ontario's special-education guidance treats an OT report as one assessment source in the IEP record; Canadian school-practice models emphasize classroom participation, observation, and tiered support, with individual standardized testing reserved for the questions it actually answers (PROFESSIONAL STANDARD and CONVENTION, provincial). Australia's Disability Standards for Education require consultation and reasonable adjustments rather than named tests, the national data collection (NCCD) wants evidence of assessed need, implemented adjustments, and monitoring from multiple sources, and NDIS evidence turns on functional impact in everyday life, so a Beery profile helps only insofar as the report translates it into participation and support needs; the OT board's competency standards govern the professional side (LAW and POLICY). The cross-jurisdiction conclusion is identical everywhere: no law, payer, or professional body in the three countries makes the Beery VMI mandatory for any eligibility category, plan, or funding decision, and its role is evidentiary and hypothesis-generating.

Instrument facts, rights, and evidence honesty complete the picture. The sixth edition (2010) remains current with no announced successor; its child norms date to roughly 2009 and 2010 (about 1,700 participants aged 2 to 18) while the adult norms (about 1,000 participants) were collected in 2006 and carried forward from the fifth edition, which a 2026 adult report should say plainly. The independent evidence urges restraint in three places: retest reliability was only moderate in an independent study (correlations around .54 to .58, with individual limits of agreement spanning roughly 27 points), no minimal clinically important difference could be established in an autism sample whose functional gains outran score movement, and the marketed "culture-free" framing overstates the case, since a 2025 Chinese preschool study found US norms flagged 1.4 percent of children as weak where local norms flagged 14.5 percent; "nonverbal and low in language demand" is the defensible phrasing, with a stated limitation whenever US norms meet a different cultural or educational background. Vision is the quiet confound: induced blur lowers all three scores, and a weak Visual Perception result warrants excluding a near-vision problem before any perceptual interpretation. Rights are Pearson-standard: record forms and stimuli may not be reproduced, scanned, or rebuilt into templates or EHRs (results and conclusions belong in the record; a narrow output-excerpt permission exists for licensed software, and electronic transfer of results is a permission-managed use), scoring is by hand with no authorized digital scorer, and no free web tool is legitimate. Beery VMI and the Beery-Buktenica Developmental Test of Visual-Motor Integration are trademarks used by Pearson (NCS Pearson, Inc. and affiliates). BastionGPT is not affiliated with, or endorsed by, Pearson. This page reproduces no test items, stimuli, norms, or scoring materials.

↑ Back to the template and downloads

Common Beery VMI write-up errors reviewers flag

The numbers behind these errors are specific. Across published studies the two supplemental tests explained only about 22 to 36 percent of core-score variance; independent test-retest correlations ran .54 to .58 with limits of agreement spanning roughly 27 points; in a developmental coordination disorder sample the instrument did not correlate with handwriting product or process; and a 2025 study found US norms flagged 1.4 percent of Chinese preschoolers against 14.5 percent by local norms. The BastionGPT Clinical Advisory Board sees the same errors most often in Beery VMI documentation reviews:

  • "The VMI score," unidentified. A single number with no edition, no core form (Full or Short), and no statement of which of the three tests produced it. The instrument is a three-test system on four forms, scored separately with no combined composite; name everything administered, and when both supplementals were given, note they followed the core in the publisher's sequence.
  • Localization from a core-only score. "A visual-motor integration deficit" concluded from the copying test alone. Without the supplementals the result cannot say whether the weakness is perceptual, motor, or integrative; core-only screening is legitimate only with that limit stated, and the supplementals earn their five minutes each: in published work Visual Perception out-predicted the core for academic achievement.
  • A contrast promoted to a diagnosis. "Proves a motor deficit" from a Motor Coordination gap, or a universal discrepancy cutoff applied (a circulating 12-point rule has no located authority). The supplementals explain only a minority of core variance, gaps must clear measurement error, and the honest verbs are "suggests" and "raises the hypothesis," corroborated by observation and function.
  • Handwriting and dysgraphia claimed from the score. "Low VMI, consistent with dysgraphia." The instrument was not designed to assess handwriting, correlations with legibility and speed are modest and sample-dependent, and one DCD study found none at all; handwriting conclusions require direct writing samples, instructional history, and the broader record, and no single measure determines eligibility anywhere.
  • Progress monitored by re-administration. Beery scores as OT outcome measures, or goals written on the standard score. Retest reliability is moderate with wide individual limits, no minimal clinically important difference has been established, and direct handwriting gains have occurred without score movement; write goals on the activity, measure progress with authentic samples, and respect the publisher's minimum retest interval.
  • Norms, culture, and vision left unexamined. Adult results on 2006 norms with no comment, "culture-free" repeated from marketing when cross-cultural work contradicts it, age equivalents presented as developmental ages, or a weak Visual Perception score interpreted without excluding a near-vision problem, though induced blur lowers all three scores. Each belongs in the validity and limitations lines.
How BastionGPT helps

BastionGPT is specifically trained, tuned, and clinically tested on psychological and psychoeducational evaluation reports.

  • Give it the facts (tests and forms administered, three standard scores with percentiles, observations, vision status, classroom or functional evidence) and it drafts the results section: everything named, scores separated, the pattern interpreted at hypothesis strength with boundaries stated, and goals written on function, ready for your review.
  • Cross-check a finished report for the gaps reviewers flag: an unidentified "VMI score," localization from a core-only administration, a contrast promoted to a diagnosis, a dysgraphia claim without writing samples, or goals written on the score.
  • Draft the companion pieces: the core-only limitation paragraph, the teacher-facing classroom summary, or the re-evaluation section that pairs authentic writing samples with an interval-respecting retest.

See how clinicians use it day to day on the AI therapy notes page.

Many BastionGPT users report saving more than 90 minutes per day on documentation.

HIPAA-compliant with a signed BAA on every plan. Your data is never used to train models. BastionGPT drafts, you review and sign.

Frequently asked questions

The core Visual-Motor Integration test asks the examinee to copy geometric forms of increasing difficulty; the Visual Perception supplemental presents the same forms as a matching task with the motor demand removed; and the Motor Coordination supplemental requires controlled pencil work within boundaries with the perceptual load reduced. The core comes in two alternative formats, a Full Form (ages 2 to 100) and a Short Form commonly used through age 7, so the system is three score-bearing tests on four forms. Each test is scored by hand against the manual's developmental criteria, on the examinee's first production (no erasing and retrying, no best-of-attempts), and yields its own standard score (mean 100, SD 15), percentile, and secondary age equivalent. There is no combined three-test composite and no authorized automated scoring: the design intent is that the three scores be compared, and the profile, not a sum, is the deliverable.

The publisher's precise position: they are optional follow-up tests, and when both are administered they follow the core in a fixed sequence (core, then Visual Perception, then Motor Coordination). The practical rule for reports is sharper: a core-only administration is legitimate screening, but it cannot support any claim about whether a weakness is perceptual, motor, or integrative, and the write-up must say so, in words like "the supplemental tests were not administered, so this result does not determine the source of the observed weakness." When the referral question is why performance is low, the supplementals are what answer it, they take about five minutes each, and the published evidence gives them independent standing: in one achievement study the Visual Perception test was the strongest predictor of math outcomes while the core copying score dropped out of the model.

By itself, that reproducing visual material was hard, without saying why: the core task blends visual analysis, planning, and motor production, and a low score cannot apportion blame among them. The supplemental contrast generates the hypotheses: stronger matching with weaker constrained pencil control points toward graphomotor execution; weaker matching points toward form analysis and earns a near-vision check first, since experimentally induced blur lowers all three scores; all three low suggests a broader developmental picture rather than three separate deficits; and a low core with both supplementals intact leaves integration, planning, attention, or production strategy on the table. Every one of those readings stays at hypothesis strength: the supplementals explain only about 22 to 36 percent of core variance, gaps must clear measurement error, and corroboration comes from observations and real-task function, never from the numbers alone.

No on both counts, and the evidence is blunter than the folklore. The instrument was not developed or intended to assess handwriting: correlations with legibility and speed are modest and sample-dependent, sensorimotor measures explained no more than about a quarter of handwriting variance in one school-age study, and in a developmental coordination disorder sample the Beery showed no significant correlation with handwriting product or process at all. Dysgraphia and written-expression disorder determinations require direct writing samples, instructional and intervention history, achievement and language evidence, and educational impact, and eligibility law in the US prohibits any single measure from deciding. The defensible role: a low Beery result can justify looking closely at written output, and an intact profile can usefully redirect the inquiry toward instruction, automaticity, language, attention, or endurance; the handwriting conclusion itself always rests on the writing.

Poorly, and the write-up should plan around that. Independent test-retest correlations were only moderate (.54 to .58) with 95 percent limits of agreement spanning roughly 27 points, meaning an individual's score can swing substantially without any true change; a study attempting to establish a minimal clinically important difference in an autism sample could not, with scores flat across eleven months of therapist-documented functional improvement; and handwriting-intervention research has found direct writing outcomes improving while Beery scores stood still. The publisher's own guidance sets a minimum retest interval of about a month, which is an administration floor, not an invitation to monthly monitoring. The defensible plan: write goals on the functional activity, measure progress with repeated authentic samples (legibility, speed, endurance, participation), reserve any retest for a genuine re-evaluation at a respectful interval, and interpret change only beyond documented measurement error.

The norms are aging and the culture-free claim needs demotion. The sixth edition (2010) remains current with no announced successor, its child norms were collected around 2009 and 2010 (about 1,700 participants aged 2 through 18), and its roughly 1,000-person adult norms were collected in 2006 and carried forward from the fifth edition, so an adult evaluated in 2026 is being compared with a two-decade-old sample, which the report should say. On culture: "nonverbal and low in language demand" is defensible; "culture-free" is not, because drawing experience, pencil familiarity, preschool curriculum, and local developmental expectations all shape performance, and a 2025 study of 421 Chinese preschoolers found US norms flagged 1.4 percent of children as weak where locally derived norms flagged 14.5 percent, an order-of-magnitude difference. A clinician using US norms with a child from a different cultural or educational background states that limitation explicitly.

It sells at Pearson qualification level B (appropriate graduate training, licensure, or supervised assessment training), the core copying forms can be group-administered while the supplementals are individual, and scoring is by hand: the digital manual sold through the publisher's platform is view-only, and no authorized automated or free web scorer exists, so any online "Beery calculator" is unauthorized on its face. Rights follow Pearson's standard rules: record forms, stimuli, scoring criteria, and norms may not be photocopied, scanned, retyped, or rebuilt into report templates, EHR forms, or software, and permission processes govern electronic uses, with a narrow allowance for excerpting minimal output conclusions into a written evaluation and permission-managed transfer of results into records systems. What always belongs in the chart is yours: the scores, observations, interpretation, and recommendations, written in your own words.

Yes. Give it the facts (edition, core form, supplementals administered, the three standard scores with percentiles, your observations, vision status, and the classroom or functional evidence) and it drafts the results section: everything named, scores separated with the right labels, the pattern interpreted at hypothesis strength with boundaries stated, and recommendations written on function, ready for your review. It can also cross-check a finished report for an unidentified "VMI score," localization claimed from a core-only administration, a contrast promoted to a diagnosis, a dysgraphia claim without writing samples, or goals written on the score, and it can draft the core-only limitation paragraph or the teacher-facing summary. BastionGPT is HIPAA-compliant with a signed BAA on every plan, and your data is never used to train models.

Primary sources

The instrument facts and compliance claims on this page trace to these sources, last verified August 2026:

  1. Publisher record, accessed August 2026: Pearson, Beery VMI Sixth Edition product page (ages, forms, times, manual scoring, qualification level B), forms and sequence support article (optional supplementals; fixed order), first-production scoring guidance, retesting interval guidance, and permissions and licensing.
  2. Supplemental-contrast evidence: Sortor JM, Kulp MT, 2003, Optometry and Vision Science, the supplementals and academic achievement (Visual Perception out-predicted the core for math); Kulp MT, Sortor JM, 2003, supplementals explain about 36 percent of core variance; Avi-Itzhak T, Obler DR, 2008, about 22 percent in preschoolers, with intact cores over weak supplementals.
  3. Reliability and change: Harvey EM and colleagues, 2017, Optometry and Vision Science, interrater and test-retest study (retest correlations .54 to .58; limits of agreement about 27 points); Ohl AM, Schelly D, 2022, British Journal of Occupational Therapy, no estimable minimal clinically important difference in an ASD sample; Pfeiffer B and colleagues, 2015, handwriting gains without corresponding VMI change.
  4. Handwriting boundary: Prunty M, Barnett AL, Wilmut K, Plumb M, 2016, no VMI-handwriting correlation in developmental coordination disorder; Klein S and colleagues, 2011, sensorimotor measures and handwriting variance; NINDS common data elements, Beery VMI catalog entry (not developed to assess handwriting; norm years).
  5. Culture and vision confounds: Tang X and colleagues, 2025, Frontiers in Pediatrics, Chinese preschool norms study (14.5 versus 1.4 percent flagged; supplementals about 25 percent of core variance); Findlay R and colleagues, 2020, PLOS ONE, induced blur lowers all three scores; Pearson, age-equivalent interpretation problems.
  6. Law and billing: eCFR, IDEA evaluation procedures (34 CFR 300.304) (variety of tools; no single measure); Section 504 evaluation rules; APA Services, testing billing and coding guide (96130-series families); CMS OT evaluation coding context for the 97165 family; NDIS, functional supporting-evidence guidance.

Educational content, not legal or billing advice. Sample notes are fictional. Follow your organization's policies and your board, payer, and jurisdiction requirements.