The WIAT-4 (Wechsler Individual Achievement Test, Fourth Edition) is an individually administered academic achievement battery for ages 4 through 50, measuring reading, mathematics, written expression, and oral language with standard scores set to a mean of 100. Psychologists, school psychologists, and specific-learning-disability evaluation teams use it to quantify academic skills against age or grade norms. This page covers how to write up WIAT-4 results, with a fictional sample and a results-section template.
Psychologists, school psychologists, and other qualified assessors; publisher qualification level B
IEP and Section 504 teams, parents, teachers, accommodations reviewers, physicians
500 to 900 words for the results section · administration time varies by grade and subtests given
Norm-referenced academic achievement battery
Specific learning disability and psychoeducational evaluations, dyslexia and dyscalculia referrals, academic accommodation requests, ages 4 to 50
Published by NCS Pearson (2020); current edition as of July 2026; described here for write-up purposes, no test content reproduced
The Wechsler Individual Achievement Test, Fourth Edition (WIAT-4) is an individually administered, norm-referenced academic achievement battery for ages 4:0 to 50:11, published by NCS Pearson in 2020 as the fourth edition of a line that began with the original WIAT in 1992. It measures learned academic skills, word reading, decoding, reading comprehension and fluency, spelling and written expression, math computation, problem solving and fluency, and oral language, and reports them as composites: Reading, Mathematics, Written Expression, and Oral Language, an overall Total Achievement score (which summarizes the reading, math, and writing subtests), and supplemental composites covering basic reading, decoding, fluency, phonological and orthographic processing, and a Dyslexia Index. The current battery carries 20 subtests, five of them new to this edition and one of the five (Orthographic Choice) available only on Q-interactive, and it yields 32 scores in all, so a report should name which subtests were actually given rather than a generic count. One metric detail matters more than it looks: WIAT-4 subtests and composites are both standard scores with a mean of 100 and a standard deviation of 15, unlike Wechsler cognitive batteries, where subtests use scaled scores with a mean of 10. Scoring runs through Q-global or Q-interactive against age-based or grade-based norms (grade norms in fall, winter, and spring windows), with percentile ranks, confidence intervals, and Growth Scale Values for progress monitoring. Edition status as of July 2026: Pearson's product page calls the WIAT-4 "the most current version of the WIAT" and no fifth edition has been announced; Canada gained a Canadian-normed WIAT-4-CDN in 2024 with a French-Canadian adaptation; Australia and New Zealand have no WIAT-4 adaptation, so clinicians there use the US-normed edition or the older WIAT-III A&NZ.
The load-bearing distinction for write-ups is that achievement is not ability. The WIAT-4 quantifies what a person has learned to do academically; it says nothing by itself about reasoning, and it earns its place in an evaluation next to a cognitive measure such as the WISC-V, whose linkage studies support ability-achievement and pattern-of-strengths-and-weaknesses analyses. The second distinction is interpretive altitude. The technical manual, unusually, contains no factor analyses, a gap both published reviews flag: Dombrowski and Casey (2022) judge the test "well-conceptualized, and generally psychometrically sound" while faulting the missing structural evidence, and Beaujean and Parkin (2022), after running the analyses the manual omits, found support for equal-interval score properties insufficient and advised psychologists to "completely refrain" from composite-score uses that require them. The defensible posture, and the one this page teaches, leads with normative composite interpretation, confidence intervals attached, treats ipsative and discrepancy analytics as stated choices with stated limits, and feeds the eligibility decision, which belongs to a team applying whole-evaluation criteria, rather than announcing it.
School psychologists write up WIAT-4 results inside special-education eligibility evaluations, where its domain structure tracks the eight achievement areas named in the federal specific-learning-disability regulation; clinical and pediatric psychologists use it for dyslexia, dyscalculia, and written-language referrals; and neuropsychologists fold achievement testing into a neuropsychological report when learning questions ride along with attention or memory ones. It also supplies the academic evidence in college-admissions and professional-exam accommodation requests, which for adults is why the norms running to age 50 matter. Instrument choice is mostly a pairing decision: per Pearson's product page the WIAT-4 "links directly to the WISC-V and KABC-2 NU," the KTEA-3 (also Pearson, normed for ages 4:0 to 25:11) carries its own dyslexia index, and the Woodcock-Johnson achievement battery pairs with its own cognitive counterpart for CHC-style cross-battery work; head-to-head validity studies among the three are sparse, so the honest defense of your choice is the referral question and the cognitive test you are linking to. For frequent progress monitoring, curriculum-based measures do that job better; the WIAT-4 is the norm-referenced diagnostic snapshot, and its Growth Scale Values add only limited monitoring capacity between full administrations. Achievement results rarely stand alone: they sit beside cognitive results, behavior ratings such as the BASC-3, intervention-response data, and classroom evidence in the full evaluation report.
No statute, payer, or publisher mandates a results-section format. The sequence below is the convention experienced evaluators converge on because it survives review: it names the data before interpreting it, leads with composites and their confidence intervals, and makes the two genuinely contested moves in achievement reporting, which norm set anchors the comparison and how far below the composite you interpret, explicit stated decisions instead of silent ones. Each section carries the pitfall that most often undermines it.
Measures, edition, norms, and format. Name the instrument and edition, the norm set with its basis (US or Canadian; age-based or grade-based, and for grade norms the season window), the subtests administered, the format (paper or Q-interactive, which differ by four subtests), and the scoring platform. State the age-versus-grade choice as a decision: age norms when the comparison is same-age peers or an age-normed cognitive test, grade norms when the question is grade-level attainment, both for retained or young-for-grade students. Pitfall: scores with no stated norm basis. An 88 against age norms and an 88 against grade norms answer different questions, and ability-achievement analyses require the achievement and cognitive scores to share a reference frame.
Behavioral observations and validity. Describe effort, attention, language and sensory considerations, breaks, and any departure from standardized procedure, then state plainly whether the results are valid estimates of current academic functioning. Pitfall: observations that contradict the interpretation. Noting fatigue and rushed responding, then interpreting timed fluency scores without a caveat, hands a reviewer two incompatible accounts in one report.
Composite summary table. Report each composite with its standard score, confidence interval, percentile rank, and descriptor from one labeled system, and state the metric once: WIAT-4 subtests and composites are standard scores with a mean of 100 and a standard deviation of 15. That sentence prevents the most common cross-battery misreading, because Wechsler cognitive subtests use a mean-10 scaled metric and mixed reports invite readers to apply the wrong ruler. Pitfall: bare standard scores. A score without its confidence interval and percentile presents a measured estimate as an exact point, and achievement decisions downstream (eligibility, accommodations) are threshold decisions where error bands matter.
Domain-by-domain results with error analysis. Move through reading, mathematics, written expression, and oral language, spending the prose on functional meaning for this student, and add what the numbers cannot show: how the student erred, decoding attempts, self-corrections, strategy use, because achievement tests support qualitative skills analysis in a way cognitive batteries rarely do. Pitfall: the score dump. A paragraph per subtest in administration order documents that testing occurred; it does not answer the referral question, and it buries the error-pattern observations that make achievement write-ups useful to teachers.
Composite integrity and the interpretive decision. When subtests inside a composite diverge, say so and interpret at the level the data support; when composites are intact, resist drilling down. The published critiques give this section its spine: the technical manual contains no factor analyses, and the independent evaluation found support for equal-interval score properties insufficient, advising psychologists to refrain from composite-score uses that require them, which is a direct caution about ipsative strength-and-weakness analytics and mechanical discrepancy arithmetic. Lead with normative interpretation, and where you run planned comparisons, present them as supplementary with their base-rate context. Pitfall: an intact-looking composite hiding a split. Reporting Reading as one number when word reading and comprehension sit twenty points apart writes the referral question out of the report.
Ability-achievement integration and the SLD model. Name the framework the operative state uses (discrepancy, intervention response, or pattern of strengths and weaknesses), name the cognitive estimate anchoring any comparison and its norm frame, and remember what federal law actually says: states must not require a severe discrepancy, must permit an intervention-response process, and may permit other research-based procedures (34 CFR 300.307). Pitfall: writing "a severe discrepancy is required" in the report. The regulation prohibits states from requiring it; asserting it as federal law dates the report and misstates the rule reviewers know.
Dyslexia Index and screening boundaries. Report it like any composite, score, confidence interval, percentile, then frame it as what the publisher says it is: a risk screener interpreted inside a comprehensive evaluation. Note the grade-band design where relevant, because the index draws on different subtests before and after grade 4, so index values from different bands describe different skill mixes. Pitfall: "the Dyslexia Index confirms a diagnosis of dyslexia." It flags risk; the diagnosis, where one is made, rests on the whole evaluation, and the word choice is the difference between a defensible report and an exhibit.
Interpretive summary and referral linkage. Answer the referral question directly, state what the achievement data establish and what they do not, and place the WIAT-4 as one source among several: federal law requires that no single measure or assessment serve as the sole criterion for eligibility (34 CFR 300.304(b)(2)), and the eligibility determination carries its own documentation requirements. Pitfall: an eligibility verdict from an achievement score. Identification belongs to the team applying state criteria; the results section supplies calibrated evidence, not the ruling.
Recommendations linkage. Tie each recommendation to a specific finding, keep instructional suggestions at the skill level the data support, and handle grade equivalents deliberately: report standard scores and percentiles, and if a family asks for a grade equivalent, provide it with the explanation that it is an interpolated score, not an instructional placement. Pitfall: recommendations written from grade equivalents. "Reads at a third-grade level, so provide third-grade texts" converts the test's least interpretable score into an instructional decision it cannot support.
MEASURES AND ADMINISTRATION Instrument/edition: WIAT-4 Norms: [US / Canadian] [age / grade (season)] Subtests given: [list; paper battery has 16, Q-interactive 20] Format: [paper / Q-interactive] Scoring: [Q-global / Q-interactive] Norm-choice rationale: [age norms to match cognitive test / grade norms for attainment question / both, with reason] BEHAVIORAL OBSERVATIONS AND VALIDITY [Effort, attention, breaks, departures from standard procedure; statement that results are valid estimates, or the limitation] COMPOSITE SUMMARY (table) [Composite | standard score | 95% CI | percentile | descriptor; one descriptor system; note subtest metric is mean 100, SD 15] DOMAIN-BY-DOMAIN RESULTS (prose + error analysis) Reading / Basic Reading / Fluency: [scores in context + error patterns observed, e.g., decoding attempts, self-corrections] Mathematics: [same] Written Expression: [same] Oral Language: [same] COMPOSITE INTEGRITY NOTES [Within-composite splits named; interpretation kept at the level the data support; planned comparisons labeled supplementary] ABILITY-ACHIEVEMENT INTEGRATION [State SLD model named; cognitive anchor + shared norm frame; no discrepancy-formula-is-required language] DYSLEXIA INDEX (if given) [Score + CI + percentile; framed as risk screening within the comprehensive evaluation, never as a diagnosis] INTERPRETIVE SUMMARY (answer the referral question) [What the data establish and do not; single-measure limits; measurement error acknowledged] RECOMMENDATIONS LINKAGE [Each recommendation tied to a finding; grade equivalents only with context, never as placement decisions] Evaluator signature / credentials: Date:
Free to use and share, no signup. The PDF includes a one-page cheat sheet with section-by-section pitfalls and a pre-sign checklist; the DOCX is the blank results-section skeleton, ready to adapt.
Scenario: a school team refers a 9-year-old in grade 4 after two years of reading intervention with limited response. The achievement profile shows the word-level reading pattern that most often prompts a dyslexia question: low decoding, fluency, and phonological processing against intact oral language and mathematics. This is the achievement results section only, condensed but structurally complete. All details are fictional.
Student: J.T., 9 (grade 4) · Referral: school student-support team, reading concerns after two intervention cycles · Evaluator: L. Okafor, PhD, Licensed Psychologist · Testing date: 07/10/2026 · Report date: 07/16/2026
Measures and administration: The Wechsler Individual Achievement Test, Fourth Edition (WIAT-4) was administered in paper format as part of a broader psychoeducational evaluation and scored on Q-global against the United States age-based norms. Age norms were selected so that achievement and cognitive results in this evaluation share the same reference frame. Subtests were selected for the referral question and sampled reading, mathematics, written expression, and oral language; the Dyslexia Index was also derived. WIAT-4 subtest and composite scores are standard scores with a mean of 100 and a standard deviation of 15, and each composite below is reported with its 90% confidence interval (CI), matching the intervals printed on Q-global score reports, and its percentile rank.
Behavioral observations and validity: J.T. worked cooperatively across two sessions and took one scheduled break. He engaged readily with math and listening tasks. On reading tasks he slowed noticeably, subvocalized while working through unfamiliar words, and self-corrected several times using sentence context. Effort was consistent, administration followed standardized procedures, and the results are considered valid estimates of his current academic skills.
| Composite | Standard score | 90% CI | Percentile |
|---|---|---|---|
| Oral Language | 104 | 99-109 | 61 |
| Mathematics | 101 | 96-106 | 53 |
| Written Expression | 88 | 83-93 | 21 |
| Reading | 82 | 78-86 | 12 |
| Phonological Processing | 81 | 77-86 | 10 |
| Reading Fluency | 79 | 74-84 | 8 |
| Basic Reading | 78 | 74-82 | 7 |
| Decoding | 76 | 72-80 | 5 |
| Dyslexia Index | 80 | 76-85 | 9 |
| Total Achievement | 88 | 85-91 | 21 |
Reading: J.T.'s word-level reading skills are well below age expectations. His Decoding composite of 76 (CI 72-80, 5th percentile) and Basic Reading composite of 78 (7th percentile) reflect consistent difficulty reading real words and pronounceable nonwords, and his Phonological Processing score of 81 (10th percentile) indicates the underlying sound-manipulation skills are similarly weak. The Reading composite of 82 (12th percentile) masks a split worth naming: word reading (subtest standard score 76, 5th percentile) sits well below reading comprehension (90, 25th percentile), which his strong oral language appears to support when passages are short. Reading Fluency of 79 (8th percentile) shows the downstream cost in rate. In error terms, J.T. attempted unfamiliar words by naming the first sound and predicting from word length, confused vowel teams, and read more accurately without time pressure; his self-corrections relied on sentence meaning rather than decoding.
Mathematics and written expression: Mathematics of 101 (CI 96-106, 53rd percentile) is age-typical across computation and problem solving, and no error pattern emerged beyond occasional misread word problems that he solved correctly when reread aloud. Written Expression of 88 (21st percentile) is a relative weakness driven by spelling (subtest standard score 82), while his sentence-level composition (95) carried age-appropriate ideas; his spelling errors mirrored his decoding errors, phonetically plausible but orthographically incorrect.
Oral language: Oral Language of 104 (CI 99-109, 61st percentile) is solidly age-typical. J.T. understood spoken passages and expressed ideas verbally at a level that contrasts sharply with his printed-word skills, the contrast his teachers describe in class.
Dyslexia Index and integration: The Dyslexia Index of 80 (CI 76-85, 9th percentile) is a screening indicator, and this result flags elevated risk consistent with the pattern above; it is not by itself a diagnosis. The Total Achievement score of 88, which summarizes the reading, mathematics, and writing subtests (oral language is not part of it), is reported for completeness but averages across domains that differ sharply, so the domain composites carry the interpretation. Within this evaluation's pattern-of-strengths-and-weaknesses framework, these achievement results supply the academic-weakness evidence; they are read against the cognitive results reported separately, on the same age-norm frame, and against intervention-response and classroom data. Baseline Growth Scale Values were recorded for the reading subtests; per the publisher's guidance, change will be interpreted only after at least three months and across at least three data points.
Summary: Achievement testing shows an age-typical oral-language and mathematics foundation against consistently low word-level reading skills: decoding at the 5th percentile, basic reading at the 7th, reading fluency at the 8th, with spelling weakness on the same phonological line and a Dyslexia Index in the elevated-risk range. This pattern is consistent with the referral concern about word-level reading. It does not by itself establish a specific learning disability: that determination belongs to the eligibility team integrating cognitive results, intervention response, observation, and classroom data, and no single measure may serve as the sole criterion. All scores are estimates that carry measurement error and describe current performance.
Linkage to recommendations: The decoding and phonological findings support recommendation 1 (explicit, systematic instruction in word-level decoding and spelling, delivered with cumulative practice). The fluency finding supports recommendation 2 (a trial of extended time and audio access for content-area text, with the team monitoring whether comprehension holds as passages lengthen). The intact oral-language finding supports recommendation 3 (preserve access to grade-level ideas through discussion and read-aloud while word-level skills are remediated). Progress should be monitored with curriculum-based measures between administrations, with WIAT-4 re-administration reserved for the re-evaluation cycle.
This sample is fictional and for educational purposes. It does not describe a real student or record, and the scores are invented for illustration and correspond to no real child or record.
Writing these after every session? BastionGPT drafts complete notes from bullets, dictation, or a transcript.
Generate a note from bulletsWrite the results section knowing which decision framework will read it, and label the strength of each requirement honestly. In US special education, federal law is instrument-neutral: no regulation names the WIAT-4, and evaluators may "not use any single measure or assessment as the sole criterion" for eligibility (34 CFR 300.304(b)(2)). For specific learning disability the state "must not require the use of a severe discrepancy" between ability and achievement, must permit an intervention-response process, and may permit other research-based procedures (34 CFR 300.307), so which model your write-up serves is state policy, and the same WIAT-4 data serve any of them. The eight achievement areas the regulation lists (34 CFR 300.309(a)(1)) read like a WIAT-4 table of contents, and the technical manual itself says the Basic Reading composite "closely aligns with the definition of basic reading skills" in IDEA. State practice genuinely differs: a comprehensive review of state criteria found only 67% of states allowed the discrepancy approach and 20% prohibited it outright as of the 2013 policy year (Maki, Floyd, and Roberson), and the strengths-and-weaknesses model, adopted by at least 14 states, was found by a systematic review of the diagnostic-accuracy evidence to identify SLD "at the level of chance" (School Psychology Review), which is exactly why the report should name the framework it feeds instead of implying one is federally required. Two federal clarifications belong in every school evaluator's head: intervention response "cannot be used to delay or deny" a full evaluation (OSEP Memo 11-07), and nothing in IDEA prohibits the words dyslexia, dyscalculia, and dysgraphia in evaluations, eligibility determinations, or IEPs (OSERS letter, October 2015); the SLD definition names dyslexia itself (34 CFR 300.8(c)(10)). Accommodations bodies read achievement testing on their own clocks: College Board policy asks that for learning disorders "the educational evaluation and testing should be no more than five years old" while cognitive testing may be older, an IEP or 504 plan alone does not transfer (students "will still need to request accommodations"), and full documentation review "can take up to seven weeks"; ACT's criteria for learning disorders expect a comprehensive battery including "Results of a complete achievement battery" from a credentialed evaluator. Funding is the quiet boundary: school SLD evaluations are education-funded under IDEA, while clinically indicated testing bills payers under the psychological-testing code family (96130 and 96131 for evaluation services, 96136 through 96139 for administration and scoring) as payer policy, and Medicare's statute "does not extend coverage to screening procedures" (Billing Article A57481), so achievement testing run purely for educational eligibility is generally not a medical claim.
Edition and norms currency is quieter here than on the cognitive side, and worth one honest paragraph. As of July 2026 Pearson's own product page states "The WIAT-4 is the most current version of the WIAT," no fifth edition has been announced anywhere, and any WIAT-5 date you encounter is speculation. The norms have a specific property reports should carry quietly in mind: the standardization sample of 1,832 was collected between October 2018 and February 2020, entirely before pandemic school closures (Beaujean and Parkin, 2022), so a score today compares a student to pre-pandemic peers, which matters when instruction was disrupted. Administration time is not a fixed number: per Pearson it "varies by grade level and number of subtests administered," one more reason the report lists what was given. Jurisdiction determines the norms available. Canada has a Canadian-normed WIAT-4-CDN (normative copyright 2024) plus a French-Canadian adaptation, so Canadian reports should say which norm set scored the protocol, especially for records that span the 2024 transition. Australia and New Zealand have no WIAT-4 adaptation: Pearson's local catalogue lists the WIAT-III A&NZ as its standardised achievement edition (checked July 2026), so Australian clinicians either run the US-normed WIAT-4 and say so as a stated limitation, or use the older local standardisation, and Medicare's Better Access items, which fund treatment of a clinically diagnosed mental disorder, are not a funding pathway for educational assessment by policy. Access is one tier broader than the cognitive scales: the WIAT-4 is publisher qualification level B, which is why school psychologists, specialist teachers with assessment credentials, and allied assessors administer it in settings where a level-C instrument would be out of reach; eligibility to purchase is set by Pearson's local affiliate. For re-testing, use the scores built for it: Growth Scale Values track a student against their own past performance (anchored so the average grade-3 GSV is 500 on every subtest), they compare only within the same subtest, and Pearson's GSV guidance cautions against interpreting change when "fewer than 3 months have passed" between sessions and against reading a trend from fewer than three data points; standard scores answer the different question of standing relative to peers.
Wechsler Individual Achievement Test, WIAT, Q-global, and Q-interactive are trademarks, in the US and other countries, of Pearson plc or its affiliates; the WIAT-4 is published by NCS Pearson, Inc. BastionGPT is not affiliated with, or endorsed by, the publisher. This page reproduces no test items, stimuli, norms, or scoring materials.
There is no payer audit series for achievement write-ups; the accountability literature here is psychometric, and it is current. The WIAT-4 is reviewed in the Buros Center's 22nd Mental Measurements Yearbook (published December 2025), the Journal of Psychoeducational Assessment review judged it "well-conceptualized, and generally psychometrically sound" while faulting the manual's missing factor analyses and item-analysis results, and the independent Journal of Intelligence evaluation ran the analyses the manual omits, recommended limiting composite interpretation to a few reading and mathematics scores, and advised psychologists to "completely refrain" from composite-score uses that require equal-interval values. Those findings converge on the same theme: the danger in WIAT-4 reporting is arithmetic that outruns the measurement. The BastionGPT Clinical Advisory Board sees the same errors most often in WIAT-4 report reviews:
BastionGPT is specifically trained, tuned, and clinically tested on psychological and psychoeducational evaluation reports.
See how clinicians use it day to day on the AI therapy notes page.
Many BastionGPT users report saving more than 90 minutes per day on documentation.
HIPAA-compliant with a signed BAA on every plan. Your data is never used to train models. BastionGPT drafts, you review and sign.
One metric almost everywhere: WIAT-4 subtests and composites are both standard scores with a mean of 100 and a standard deviation of 15, which is worth stating in the report because Wechsler cognitive subtests use a mean-10 scaled metric and mixed batteries invite misreading. Each score carries a percentile rank, a confidence interval (Pearson's sample report prints 90% intervals), and a descriptive category. Grade and age equivalents, stanines, and normal-curve equivalents exist for education audiences but carry the least interpretive weight, and Growth Scale Values sit on a separate absolute scale used only for progress comparison against the same student. Write every score as an estimate: number, interval, percentile together.
The WIAT-4 is current everywhere as of July 2026, and no fifth edition has been announced: Pearson's own product page states "The WIAT-4 is the most current version of the WIAT," so treat any WIAT-5 date you read as speculation. What has changed recently is jurisdictional: Canada gained the Canadian-normed WIAT-4-CDN (normative copyright 2024) plus a French-Canadian adaptation, while Australia and New Zealand still have only the WIAT-III A&NZ as a local standardisation. Name the edition and the norm set in every report and the WIAT-5 question takes care of itself when it eventually arrives.
Neither. Federal law names no instrument, forbids any single measure from being the sole criterion (34 CFR 300.304(b)(2)), and prohibits states from requiring a severe ability-achievement discrepancy for SLD identification (34 CFR 300.307): states must permit an intervention-response process and may permit other research-based methods, and intervention response in turn "cannot be used to delay or deny" a full evaluation (OSEP Memo 11-07). Achievement testing stays standard evidence in nearly every framework because the regulation defines SLD through eight achievement areas the WIAT-4 samples directly. Know your state's model and say which one the report feeds; note that the strengths-and-weaknesses model, though adopted by at least 14 states, fared poorly in a recent diagnostic-accuracy review, one more reason to present scores as evidence rather than verdict.
Twenty subtests, per Pearson's own comparison flyer (the WIAT-III had 16; five subtests are new, and one of the five, Orthographic Choice, runs only on Q-interactive). Counting scores rather than subtests, the instrument yields 32, which is why "how many subtests" answers around the web disagree with each other. How many you give is referral-driven: no authority sets a minimum or a required battery, and a dyslexia referral, a math referral, and a full psychoeducational evaluation each justify different selections. The write-up rule that follows is simple: list exactly which subtests were administered and which composites they produced, because "the WIAT-4 was administered" describes twenty different possible batteries.
No. Pearson positions it as a brief performance-based screener that provides "risk assessment, strength of risk, and interpretive information," taking about 5 minutes at the PK to grade 3 level and under 20 minutes for grades 4 through 12 and adults. Two details matter for write-ups. First, composition changes by grade band (per Pearson, Phonemic Proficiency and Word Reading through grade 3; Word Reading, Orthographic Fluency, and Pseudoword Decoding from grade 4), so index values from different bands describe different skill mixes. Second, a low index flags risk to be examined inside a comprehensive evaluation, and the word "diagnosis" belongs to the whole evaluation, not the screener. If dyslexia is confirmed, federal guidance is explicit that nothing in IDEA prohibits using the term in evaluations, eligibility determinations, or IEPs.
Report them only deliberately, and interpret them almost never. A grade equivalent says a student's raw score matched the median raw score of students at some grade point; it is interpolated, it is not an equal-interval scale, and it says nothing about whether the student can do that grade's curriculum. "Reads at a 3rd-grade level" is the most repeated misreading of achievement testing, and instructional placement decisions built on it inherit the error. The defensible pattern: carry the meaning with standard scores, confidence intervals, and percentiles; when a parent or team asks for the grade equivalent, provide it alongside a one-sentence explanation of what it is and is not. The same logic applies to age equivalents, and it is why this page's sample report and templates lead with composites and percentiles throughout.
Choose by the cognitive test you are pairing with and the referral, because head-to-head validity studies are sparse. Pearson's product page states the WIAT-4 "links directly to the WISC-V and KABC-2 NU," which is the practical draw when a WISC-V anchors the evaluation and you want supported ability-achievement or strengths-and-weaknesses analyses on a shared reference frame. The KTEA-3 (also Pearson, norms for ages 4:0 to 25:11) is the sibling battery with its own Dyslexia Index, and the Woodcock-Johnson achievement tests pair with their own cognitive battery for CHC-style cross-battery work. The WIAT-4's distinctive additions are the phonological, orthographic, and fluency processing measures and automated essay scoring. Whichever you choose, the report should say why in one sentence, and for repeat testing, staying within one battery family keeps Growth Scale Values usable.
No. Test items, stimuli, and record forms are protected test materials: purchase and use are restricted by publisher qualification (the WIAT-4 is level B), and evaluators are ethically obligated to protect test security under APA Ethics Standard 9.11 and their publisher agreements. Item exposure also damages the instrument itself, because norms assume examinees have not rehearsed the tasks, and achievement items leak especially easily into test-prep circulation. Reports should carry scores, error-pattern description, and interpretation, never item content: "he confused vowel teams on unfamiliar words" documents the skill; quoting the words he misread would document the test. Records requests that touch protocols route through your test-security obligations, and what families actually need, the scores explained in plain language with next steps, belongs in the report anyway.
Yes. Paste your composite summary (scores, confidence intervals, percentiles) and it drafts the results-section narrative for your review: organized by academic domain, functional meaning in prose, the norm-set and state-model statements in place, and grade equivalents kept out of the interpretation. It can also cross-check a draft you wrote for score-versus-narrative mismatches, metric confusion, and a Dyslexia Index phrased as a diagnosis, and it can produce a plain-language summary of findings for parents and teachers. BastionGPT is HIPAA-compliant with a signed BAA on every plan, your data is never used to train models, and drafting from scores you paste means no protocol or item content ever needs to leave your records.
The instrument facts and compliance claims on this page trace to these sources, last verified July 2026:
Educational content, not legal or billing advice. Sample notes are fictional. Follow your organization's policies and your board, payer, and jurisdiction requirements.