Woodcock-Johnson Report Write-Up: Structure, Sample Language & Common Errors

The Woodcock-Johnson is an individually administered, norm-referenced battery of cognitive and achievement tests built on Cattell-Horn-Carroll theory, currently sold in two editions: the WJ IV (2014) and the digital WJ V (2025). Psychologists, school psychologists, and educational diagnosticians use it in special-education, learning-disability, gifted, and accommodation evaluations. This page covers how to write up Woodcock-Johnson results, with a fictional sample and a results-section template.

Free to use and share. No signup required.
Already have session bullets or a transcript? Generate a structured draft with BastionGPT — you review and sign it.
Who writes it

Psychologists, school psychologists, and educational diagnosticians; graduate-level qualification required by publisher policy

Audience

IEP and Section 504 teams, parents, physicians, gifted programs, postsecondary disability offices

Typical length

500 to 900 words for the results section · administration about 5 to 10 minutes per test administered

Format family

Norm-referenced cognitive and achievement battery

When it's used

SLD and psychoeducational evaluations, gifted identification, adult accommodation documentation, three-year re-evaluations

Standards context

Published by Riverside Insights; WJ V (2025) and WJ IV (2014) both current; described here for write-up purposes, no test content reproduced

What is the Woodcock-Johnson?

The Woodcock-Johnson is a family of individually administered, norm-referenced tests of cognitive abilities and academic achievement, first published in 1977 and now sold by Riverside Insights in two editions at once: the WJ IV (2014, paper-based, ages 2 to 90+) and the WJ V (released February 2025, fully digital, ages 4 through 90+, normed on a 2020-Census-aligned sample collected from 2022 to 2024). It is the most explicit operationalization of Cattell-Horn-Carroll (CHC) theory among the major batteries, and its distinctive design decision is co-norming: the cognitive and achievement batteries share one normative sample, so ability-achievement comparisons run inside a single reference group instead of across linked but separate tests. Alongside the familiar standard scores (mean 100, SD 15) and percentile ranks, WJ reports can carry two metrics most other batteries lack, the Rasch-based W score that underlies all WJ scoring and the criterion-referenced Relative Proficiency Index (RPI), which is written against 90 and predicts the examinee's success rate on tasks typical age or grade peers perform with 90% success.

The load-bearing fact for write-ups, as of July 2026, is that the WJ V is not a reskinned WJ IV, and both editions are in active use with no US retirement announcement for the older one. Per the publisher's own technical documentation, the WJ V recomputed the General Intellectual Ability composite as eight equally weighted tests, removed Auditory Processing (Ga) from the composite and the cognitive battery entirely, split the WJ IV's Long-Term Retrieval cluster into Long-Term Storage (Gl) and Retrieval Fluency (Gr), rebuilt Fluid Reasoning on new tests (Matrices and Verbal Analogies), and eliminated the standalone Oral Language battery, moving those measures into the achievement battery and a new 15-test Virtual Test Library. Scores from the two editions are therefore not interchangeable: the normative samples sit more than a decade apart and the composite and several clusters are built from different tests, which makes the edition-and-norms statement the single highest-value line in a Woodcock-Johnson report. The battery also sits differently in evaluations than a Wechsler scale does: where a WISC-V is typically paired with a separately normed achievement test such as the WIAT-4, the WJ offers both sides co-normed, and many evaluators mix the two traditions in cross-battery work. A WJ score set is evidence, not a verdict: eligibility decisions belong to teams and agencies applying whole-evaluation criteria, and the results section's job is to feed that decision with calibrated, honestly bounded numbers.

Who uses Woodcock-Johnson write-ups and when

School psychologists and educational diagnosticians write up Woodcock-Johnson results inside special-education eligibility evaluations, where the co-normed cognitive and achievement batteries feed every state's specific-learning-disability method, discrepancy, intervention response, or pattern of strengths and weaknesses. Clinical and pediatric psychologists use it in learning-disability and gifted referrals, and neuropsychologists fold WJ achievement clusters into a neuropsychological report beside their cognitive and memory measures. Because the norms extend past age 90, it also carries adult work that child-normed batteries cannot: adult basic education, vocational rehabilitation, and the documentation behind postsecondary and licensing-exam accommodation requests. Instrument choice follows the referral question. Many evaluators pair a WISC-V with the WJ achievement battery, or run WJ COG and ACH together to keep ability-achievement comparisons inside one norm group; the WIAT-4 and KTEA-3 compete on the achievement side with direct statistical links to the Wechsler scales. For Spanish-English bilingual evaluations, the co-normed Batería IV is the WJ IV's Spanish parallel, and no Batería V exists as of July 2026. Oral-language and phonological measures in the WJ overlap with, but do not replace, the comprehensive language batteries a speech-language pathologist administers for a language-disorder diagnosis; in a psychoeducational workup the WJ versions serve as converging evidence, not the diagnostic instrument.

How to structure a Woodcock-Johnson results section

No statute, payer, or publisher mandates a results-section format. The sequence below is the convention experienced evaluators converge on because it survives review: it anchors the scores to an edition and norm set before interpreting them, presents clusters in a table and meaning in prose, and makes the two genuinely WJ-specific moves, explaining the proficiency metrics and handling the WJ IV to WJ V transition, explicit stated decisions instead of silent ones. Each section carries the pitfall that most often undermines it.

Measures, edition, norms, and format. Name the instrument and edition (WJ IV or WJ V), the norm set, the batteries and clusters administered, and the administration format and scoring platform. Pitfall: writing "the Woodcock-Johnson" with no edition. Both editions are on the market at once, and cluster names that look interchangeable are not: the WJ V recomputed the General Intellectual Ability composite, removed Auditory Processing from it, and split Long-Term Retrieval into two clusters, so an edition-less score set is uninterpretable at re-evaluation.

Behavioral observations and validity of the session. Describe effort, engagement, language considerations, breaks, and anything that departed from standardized procedure, then state plainly whether you consider the results valid estimates of current functioning. For WJ V administrations, note the digital format and any connectivity or device accommodations. Pitfall: boilerplate observations that contradict the interpretation. "Worked quickly and confidently" followed by fluency scores you then attribute to slow, effortful responding gives a reviewer two incompatible reports in one document.

Score metrics statement. Say which metrics you are reporting, standard scores (mean 100, SD 15), percentile ranks, and, if you include them, the W score and Relative Proficiency Index, and give lay readers one plain sentence for each WJ-specific metric. The RPI is criterion-referenced and always written against 90, and Riverside's interpretive bulletin fixes the reading: an RPI of 60/90 means about 60% success on tasks that typical age or grade peers perform with 90% success. Pitfall: pasting RPI columns with no explanation. IEP teams read 55/90 as a fraction, a percentile, or a typo; either explain the metric in one sentence where it first appears or deliberately omit it, and apply that choice consistently.

Cognitive results, composite first, then cluster by cluster. Lead with the general ability composite you are using (GIA, or the Brief Intellectual Ability when a screener was given), then move through the broad CHC clusters, spending the prose on what each result means functionally for this examinee rather than re-reading the table aloud. Pitfall: the score dump. A paragraph per test in administration order documents that testing occurred; it does not answer the referral question, and narrow-cluster interpretation is exactly where the independent factor-analytic literature urges the most restraint.

Achievement results by academic domain. Organize by reading, mathematics, written language, and oral language rather than test by test, and tie each domain to the referral concern. State whether scores were computed with age or grade norms if both are in play in your setting. Pitfall: grade equivalents promoted to instructional levels. A grade equivalent of 3.2 means the examinee's raw score matched the median for that grade on this test's items; it is not a placement recommendation, and reviewers flag reports that treat it as one.

Variation and comparison procedures. If you report intra-ability variations or an ability-achievement comparison, name the procedure and the decision framework it feeds (the state's SLD method, or the specific pattern-of-strengths-and-weaknesses model), and describe base rates in words. Pitfall: the unnamed model. A "significant discrepancy" conclusion with no named method, or score-report comparison tables pasted wholesale, reads as software output rather than clinical reasoning, and PSW methods themselves remain scientifically contested, which makes naming your framework more important, not less.

Prior testing and the edition note. When earlier Woodcock-Johnson results exist, name the edition and norms of both administrations. If the editions differ, say directly that WJ IV and WJ V scores are not directly comparable and describe change qualitatively. Pitfall: cross-edition arithmetic. "Broad math fell six points since 2023" across a WJ IV to WJ V gap treats two different instruments as one ruler, the exact move the test's own authors caution against.

Interpretive summary and referral linkage. Answer the referral question directly, state what the data do and do not establish, and place the WJ explicitly as one source among several: federal special-education law requires that no single measure be the sole criterion for eligibility (34 CFR 300.304(b)(2)). Pitfall: an eligibility verdict from a score table. Identification decisions belong to teams and agencies applying their criteria; the results section supplies calibrated evidence, not the ruling.

Recommendations linkage. Tie each recommendation to a specific finding, and where the WJ's proficiency metrics drove it, say so; the instructional-zone language (what this student finds easy versus difficult) converts directly into intervention planning. Pitfall: recommendations that could follow any profile. If fluency supports, chunked instruction, or extended time do not trace back to named cluster findings, either the recommendations or the results section is incomplete.

Blank template (copy and adapt)

MEASURES AND ADMINISTRATION
Instrument/edition: [WJ V / WJ IV]   Norm set: [named]
Batteries and clusters given: [COG / ACH / VTL selections]
Format: [WJ V digital / WJ IV paper]   Scoring: [platform]
BEHAVIORAL OBSERVATIONS AND VALIDITY
[Effort, engagement, language, breaks, departures from
   standard procedure; statement that results are valid
   estimates, or the stated limitation if not]
SCORE METRICS STATEMENT
[Metrics reported (standard score, percentile, RPI, W);
   one plain-language sentence for RPI if included]
COGNITIVE RESULTS (composite first, cluster by cluster)
GIA or BIA: [score, band, percentile + what it means here]
Broad clusters: [functional meaning per cluster, not a
   test-by-test dump]
ACHIEVEMENT RESULTS (by academic domain)
Reading / Math / Written Language / Oral Language:
[domain results tied to the referral question]
VARIATION AND COMPARISON PROCEDURES
[Named procedure and framework (state method or PSW
   model); base rates described in words]
PRIOR TESTING AND EDITION NOTE
[Prior edition and norms named; cross-edition caution
   stated when editions differ]
INTERPRETIVE SUMMARY (answer the referral question)
[What the data establish and what they do not; single-
   measure limits; measurement error acknowledged]
RECOMMENDATIONS LINKAGE
[Each recommendation tied to a finding; re-evaluation plan]
Evaluator signature / credentials:            Date:

Free to use and share, no signup. The PDF includes a one-page cheat sheet with section-by-section pitfalls and a pre-sign checklist; the DOCX is the blank results-section skeleton, ready to adapt.

Sample Woodcock-Johnson write-up (fictional)

Scenario: a three-year re-evaluation for a student found eligible for special education in mathematics in second grade. The prior evaluation used the WJ IV; this one uses the WJ V, so the report has to handle a cross-edition comparison without treating two different instruments as one ruler. This is the Woodcock-Johnson results section only, condensed but structurally complete. All details are fictional.

Student: J.T., 11 (grade 5)  ·  Referral: three-year re-evaluation, continuing eligibility and math concerns  ·  Evaluator: R. Calloway, EdS, NCSP, School Psychologist  ·  Testing dates: 07/08/2026 and 07/10/2026  ·  Report date: 07/16/2026

Measures and administration: The Woodcock-Johnson V Tests of Cognitive Abilities (WJ V COG) and selected clusters from the Woodcock-Johnson V Tests of Achievement (WJ V ACH) were administered digitally on the publisher's platform, examiner device and student tablet, and scored against the WJ V United States age-based norms. Standard scores have a mean of 100 and a standard deviation of 15 and are reported with a 95% confidence band and percentile rank. The Relative Proficiency Index (RPI) is also reported: it is written against 90, so an RPI of 55/90 predicts about 55% success on tasks that typical age peers perform with 90% success. J.T.'s prior evaluation (March 2023, grade 2) used the WJ IV; the two editions are normed and structured differently, and scores are not directly comparable across them (see Prior testing, below).

Behavioral observations and validity: J.T. engaged readily with the tablet format, followed the recorded instructions without repetition, and sustained effort across both sessions with one scheduled break each. On untimed reasoning tasks he worked deliberately and self-corrected. On timed and math tasks he counted on his fingers, skipped items he judged hard, and twice said "math is not my thing," while continuing to work when encouraged. Administration followed standardized procedures, and the results are considered valid estimates of current functioning.

ClusterStandard score (95% band)PercentileRPI
General Intellectual Ability (GIA)96 (92-100)3988/90
Comprehension-Knowledge (Gc)108 (103-113)7095/90
Fluid Reasoning (Gf)104 (99-109)6193/90
Visual Processing (Gv)102 (96-108)5592/90
Auditory Working Memory Capacity (Gwm)82 (77-87)1261/90
Cognitive Processing Speed (Gs)84 (78-90)1466/90
Brief Reading102 (97-107)5593/90
Brief Writing94 (88-100)3484/90
Brief Math78 (73-83)747/90
Number Concepts80 (74-86)952/90

Cognitive results: J.T.'s General Intellectual Ability score of 96 (39th percentile) is in the average range, and the clusters beneath it split in an interpretable way. Comprehension-Knowledge (108, 70th percentile) is a relative strength: his vocabulary and general knowledge are developing ahead of most age peers, consistent with his strong classroom participation in science discussions. Fluid Reasoning (104) and Visual Processing (102) are age-typical: he identifies conceptual rules and works with visual patterns as well as most 11-year-olds. In contrast, Auditory Working Memory Capacity (82, 12th percentile, RPI 61/90) and Cognitive Processing Speed (84, 14th percentile, RPI 66/90) are limited relative to peers: holding and reworking information he hears, and completing simple tasks quickly, are effortful for him. Functionally, this is a student who understands grade-level ideas but loses the middle steps of multi-step directions and works slowly enough that timed tasks go unfinished.

Achievement results: Reading (Brief Reading 102) and writing (Brief Writing 94) are age-appropriate. Mathematics is the area of concern, and the two math clusters agree: Brief Math of 78 (7th percentile) and Number Concepts of 80 (9th percentile). The proficiency metrics say the same thing in instructional terms: with an RPI of 47/90, grade-level math tasks that typical age peers perform with 90% success are predicted to be at about 47% success for J.T., placing them in his difficult range. Error patterns on scored responses showed accurate single-step computation with breakdowns on regrouping and multi-step problems, consistent with the working memory finding; no test items are reproduced here.

Variations and comparisons: Within cognitive abilities, J.T.'s working memory and processing speed clusters are weaknesses relative to his own average, and his math achievement is below the level his overall ability score would predict; differences of this size occur infrequently among students with similar ability scores. These results feed the eligibility team's specific-learning-disability analysis under the state's identification method, alongside classroom data and intervention response; the pattern here, average reasoning and knowledge with concordant weaknesses in working memory, processing speed, and mathematics, is the shape that analysis looks for. The comparison procedures and their base rates come from the publisher's scoring platform and are described here rather than reproduced.

Prior testing and edition note: J.T.'s March 2023 evaluation reported WJ IV mathematics cluster scores in the low 80s. Those scores were produced by a different edition, normed on a different sample collected more than a decade earlier, and several WJ V clusters are built from different tests than their WJ IV counterparts. This report therefore does not read the current Brief Math of 78 as a decline. Stated qualitatively: mathematics was below age expectation in 2023 on the WJ IV, and it remains below age expectation in 2026 on the WJ V, despite documented intervention.

Summary: Cognitive results show average general ability with above-average knowledge and age-typical reasoning over limited working memory and processing speed. Achievement is age-appropriate in reading and writing and well below age expectation in mathematics, with proficiency metrics placing grade-level math in J.T.'s difficult range. This pattern is consistent with his existing identification in mathematics. It does not by itself decide continuing eligibility: that determination belongs to the team, integrating these results with intervention response and classroom evidence, and no single measure may serve as the sole criterion. All scores are estimates that carry measurement error.

Linkage to recommendations: The working memory finding supports recommendations 1 and 2 (multi-step directions chunked and paired with written steps; worked examples visible during practice). The math fluency and Number Concepts findings support recommendation 3 (explicit instruction and daily brief fluency practice at his instructional level, with progress monitoring). The processing speed finding supports recommendation 4 (a trial of extended time on timed math work, with the team monitoring completion). The edition note supports recommendation 5: at the next re-evaluation, compare within the WJ V rather than across editions.

This sample is fictional and for educational purposes. It does not describe a real student or record, and the scores are invented for illustration and correspond to no real child or record.

↑ Back to the template and downloads

Why this sample works

  • The edition, norm set, batteries given, and administration format are all named, so the scores are interpretable now and stay unambiguous while WJ IV and WJ V circulate side by side.
  • The cross-edition question is handled the way the test authors direct: prior WJ IV scores are named but never subtracted from current WJ V scores, and change is described qualitatively.
  • The RPI is explained in one plain sentence where it first appears, then used to translate scores into instructional terms, instead of appearing as unexplained notation.
  • Every number in the prose matches the table, cluster results are written as functional meaning for this student, and the comparison procedures are named and described rather than pasted from the scoring platform.
  • The summary states what the data do not establish, honors the no-single-measure rule, and every recommendation traces to a specific cluster finding.

Writing these after every session? BastionGPT drafts complete notes from bullets, dictation, or a transcript.

Generate a note from bullets

Documentation and compliance considerations

Write the results section knowing which decision framework will read it, and label the strength of each requirement honestly. In US special education, federal law is instrument-neutral: no regulation names the Woodcock-Johnson, evaluators may "not use any single measure or assessment as the sole criterion" for eligibility (34 CFR 300.304(b)(2)), and for specific learning disability the state "must not require the use of a severe discrepancy" between ability and achievement (34 CFR 300.307). Which method your state runs, discrepancy where still permitted, intervention response, or a pattern-of-strengths-and-weaknesses model, is state policy, and the WJ's co-normed comparison procedures serve any of them; what stays constant is the convention of naming the method and model in the report. Gifted identification has no federal mandate at all, so WJ-based thresholds are district and state policy. Testing time is billed to payers under the psychological-testing code families (96130 and 96131 for evaluation services; 96136 through 96139 for administration and scoring) as payer policy with plan-specific units, while school-based IDEA evaluations are education-funded rather than billed to health insurance, a boundary that routinely sends families seeking accommodation documentation to private evaluators. For those accommodation bodies, currency rules are the gatekeeper: ETS's guidelines, for example, describe learning-disability documentation "completed within the past 5 years and/or when the test taker was at least 16" as potentially helpful, treat re-administration of an adult cognitive measure as unnecessary in an update, and ETS's current disability guidance names the WJ V cognitive battery among the adult measures it recognizes. Verify the receiving agency's documentation policy before you choose the battery and norms.

Edition and norms currency is the Woodcock-Johnson's live defensibility question, and as of July 2026 it has an unusual answer: both editions are current products. Riverside released the WJ V in February 2025 as a digital-only system and still sells the WJ IV with no US retirement statement (the publisher posts one when a product ends; the WJ III has one, the WJ IV does not), telling university trainers that short-term WJ IV use is expected while recommending against new WJ IV kit purchases; its Canadian distributor is more concrete, estimating transition support into 2027 to 2028. No IDEA rule forces an edition switch mid-cycle; re-evaluation timing follows the ordinary three-year cycle and district policy. That makes the choice yours, and the documentation non-negotiable: name the edition, the norm basis, and the platform every time, because the WJ V's norms are 2020-Census-aligned and post-pandemic while the WJ IV's date from more than a decade earlier, and standardization samples age measurably, about 2.31 points per decade in the largest Flynn-effect meta-analysis. Scores from the two editions are separately normed and differently structured, so re-evaluations that span the transition should describe change qualitatively and restart trend lines inside the WJ V. Digital administration deserves one line in the report: the WJ V was standardized on a tablet-based format with an offline mode, and remote administration runs through Presence, named its exclusive remote provider in August 2025, with equivalence evidence that is publisher- and vendor-sponsored, worth labeling as such in high-stakes contexts. Access boundaries are the publisher's: Woodcock-Johnson batteries are restricted to qualified professionals under Riverside's user-qualification policy, and outside the US the picture is jurisdiction-specific. Canadian distribution of the WJ V runs digitally through Nelson with data stored on Canadian servers and no separate Canadian norming announced, and Australia's distributor lists the WJ V as the US adaptation while the Australasian adaptation with Australian norms, as of July 2026, remains a WJ IV product, so name the norm set explicitly when writing outside the US.

Woodcock-Johnson is a registered trademark, and WJ IV and WJ V are trademarks, of Riverside Assessments, LLC; the tests are published by Riverside Insights. BastionGPT is not affiliated with, or endorsed by, the publisher. This page reproduces no test items, stimuli, norms, or scoring materials.

↑ Back to the template and downloads

Common Woodcock-Johnson write-up errors reviewers flag

There is no payer audit series for cognitive and achievement write-ups; the accountability literature here is psychometric, and the Woodcock-Johnson has drawn pointed independent scrutiny. The WJ IV was reviewed in the Buros Center's 20th Mental Measurements Yearbook, where the reviewer wrote that "abundant caution should be exercised in interpretation of lower-order scores." Independent factor-analytic studies (exploratory and confirmatory) concluded that the WJ IV cognitive battery supports about four distinct factors rather than the seven it markets, with general ability dominating, and noted that the technical manual reported no separate factor analyses of the cognitive battery on its own. The WJ V is new enough that its independent evidence base is thin: the first independent review appeared in 2026 and, while judging the psychometrics favorably, found the evidence "less convincing for children under six." The errors below are write-up errors, and they are yours. The BastionGPT Clinical Advisory Board sees the same ones most often in Woodcock-Johnson report reviews:

  • Cross-edition scores compared as change. "Math dropped six points since the last evaluation" across a WJ IV to WJ V gap treats two differently normed, differently structured instruments as one ruler: the composite is computed differently, several clusters are built from different tests, and the normative samples sit more than a decade apart. Name both editions and describe change qualitatively.
  • Edition and norms unnamed mid-transition. "The Woodcock-Johnson" with no edition, norm basis, or platform. With both editions on the market, the GIA computed differently in each, and familiar-looking cluster names covering different tests, an unanchored score set cannot be interpreted at re-evaluation or defended under review.
  • Proficiency metrics pasted without translation. An RPI column appears (55/90, 82/90) with no sentence explaining the metric, and readers parse it as a fraction or a percentile. Either explain it in plain language where it first appears, or deliberately omit it; no authority requires the RPI, so the defensible move is a stated house rule, applied consistently.
  • Grade equivalents promoted to placement. "A grade equivalent of 3.2 means fourth-grade math is inappropriate" misreads a developmental score as an instructional prescription. Grade equivalents describe where a raw score falls on this test's growth curve; standard scores, bands, and proficiency metrics carry the interpretive weight.
  • Narrow clusters over-read. Diagnostic conclusions built on a single narrow cluster or test-level contrast, when the independent factor-structure literature finds most of the reliable variance sits with general ability. Interpret composite-first, use broad clusters cautiously, and keep test-level commentary qualitative and bounded.
How BastionGPT helps

BastionGPT is specifically trained, tuned, and clinically tested on psychological and psychoeducational evaluation reports.

  • Paste your cluster summary (standard scores, bands, percentiles, RPIs) and get a drafted results-section narrative organized composite-first and cluster by cluster, with the edition-and-norms statement and the RPI explained in plain language for your review.
  • Cross-check a finished draft for the gaps reviewers flag: numbers that disagree with the table, unexplained proficiency notation, and cross-edition WJ IV to WJ V comparisons written as change.
  • Translate the results section into a plain-language summary for parents and teachers that keeps scores as estimates and converts proficiency metrics into what feels easy and what feels difficult.

See how clinicians use it day to day on the AI therapy notes page.

Many BastionGPT users report saving more than 90 minutes per day on documentation.

HIPAA-compliant with a signed BAA on every plan. Your data is never used to train models. BastionGPT drafts, you review and sign.

Frequently asked questions

Four score types do the work. Standard scores compare the examinee to age or grade peers, with a mean of 100 and a standard deviation of 15, and each carries a percentile rank and a confidence band. Underneath them sits the W score, the Rasch-based equal-interval metric all WJ scoring is built on; because it tracks absolute ability rather than standing among peers, it is the metric that shows real growth over time even when the standard score holds still. The Relative Proficiency Index (RPI) is criterion-referenced and always written against 90: Riverside's own interpretive bulletin gives the reading, an RPI of 60/90 means about 60% success on tasks that typical peers perform with 90% success. Age and grade equivalents describe where a raw score falls on the test's growth curve and deserve the least interpretive weight. Write standard scores as estimates, and translate the WJ-specific metrics in one plain sentence each.

Either is defensible as of July 2026, and that is the unusual part. Riverside released the WJ V in February 2025 as a fully digital system and still sells the WJ IV with no US retirement statement, while its university training page acknowledges some programs will stay on the WJ IV in the short term and recommends against buying new WJ IV kits; the Canadian distributor estimates transition support into 2027 to 2028. No IDEA rule requires the newest edition the day it ships, and re-evaluation timing follows the ordinary three-year cycle. The honest trade: the WJ V carries current, post-pandemic, 2020-Census-aligned norms and the redesigned battery, while the WJ IV carries a decade of familiarity, the paper format, and norms that are now aging. Whichever you choose, name the edition and norms in the report and expect the choice to get asked about mid-transition.

Not as arithmetic. The reasons are structural, not cosmetic: the WJ V's General Intellectual Ability composite is eight equally weighted tests where the WJ IV's was differentially weighted, Auditory Processing no longer contributes to it, the old Long-Term Retrieval cluster is now two clusters, and Fluid Reasoning is measured by new tests, all against a normative sample collected more than a decade after the WJ IV's. Even the vocabulary moved: the WJ IV reports a Short-Term Working Memory cluster, the WJ V an Auditory Working Memory Capacity cluster. A same-named-looking score is not the same measure, and norms drift on their own, about 2.31 points per decade in the largest Flynn-effect meta-analysis. The defensible pattern is the one the sample on this page models: report the prior edition's results as historical findings, describe change qualitatively (still below age expectation, no longer below age expectation), and restart numeric trend lines inside the WJ V.

The Relative Proficiency Index is the WJ's criterion-referenced metric, always written against 90. It answers a different question than a standard score: not "how does this student rank," but "when peers succeed 90% of the time on a task, how often will this student?" An RPI of 47/90 predicts 47% success on material typical peers handle at 90%, which is why the RPI converts so directly into instructional planning, and why it can reveal a proficiency problem a mid-80s standard score understates. No authority requires you to report it, and plenty of experienced evaluators omit it from the body precisely because unexplained 55/90 notation confuses IEP teams. Both practices are defensible; a silent mix of the two is not. Set a house rule, and when the RPI does appear, spend one sentence translating it in plain language.

No. Federal law names no instrument, requires that no "single measure or assessment" be the sole criterion for eligibility (34 CFR 300.304(b)(2)), and forbids states from requiring an ability-achievement discrepancy for specific learning disability (34 CFR 300.307); states choose among discrepancy, intervention response, and pattern-of-strengths-and-weaknesses frameworks. The Woodcock-Johnson's place in those evaluations is convention and district practice, earned by its co-normed cognitive-achievement design, which lets every comparison run inside one norm group. If your district or agency names a specific battery, that is local policy, and the practical answer is the one evaluators have always given: document what the policy requires, and choose the instrument that answers the referral question.

Both stacks are standard; they solve the linking problem differently. The WISC-V with the WIAT-4 gives you the Wechsler clinical tradition and a publisher-linked ability-achievement pair normed on separate but statistically connected samples. The Woodcock-Johnson gives you one co-normed system for both sides, the most explicit CHC-theory alignment of any major battery, and adult norms that extend past age 90. Convention splits by setting: Wechsler scales dominate clinical and neuropsychological work, while the WJ is a school-evaluation workhorse and the backbone of cross-battery assessment, where evaluators combine WJ clusters with Wechsler indexes deliberately. The referral question, your state's SLD method, and what the prior evaluation used (staying inside one battery family preserves comparability) decide it case by case.

No. Test items, stimuli, and record forms are protected test materials: purchase and use are restricted under Riverside's qualification policy, and psychologists are ethically obligated to maintain test security under APA Ethics Standard 9.11 and their publisher agreements. Item exposure damages the instrument itself, because norms assume examinees have not rehearsed the tasks. Reports should carry scores, bands, percentiles, proficiency metrics, and interpretation, never item content; the sample on this page describes error patterns in category terms for exactly that reason. What parents and teams can get is better anyway: the scores with a plain-language explanation of what each cluster means, what the proficiency metrics predict for classroom work, and what happens next.

The WJ V is digital by design rather than a paper test ported to a screen: it was standardized in the tablet-based format, with the examiner running the session from one device and the examinee responding on a tablet, an offline mode for settings without reliable connectivity, and printed response booklets where written responses are required. That means format equivalence with paper is not the right question inside the WJ V; the norms are the digital administration. Two caveats belong in high-stakes write-ups. First, use the standardized device class the publisher specifies, an examinee tablet with a 10-inch-plus screen, "as that is how the test was standardized" in Riverside's words, rather than improvising hardware. Second, remote administration runs through Presence, named the exclusive remote provider in 2025, and the supporting equivalence research is publisher- and vendor-sponsored; naming the administration arrangement in the report is the convention that keeps the scores defensible.

Yes. Paste your cluster summary (standard scores, bands, percentiles, RPIs) and it drafts the results-section narrative for your review: organized composite-first and cluster by cluster, the edition-and-norms statement up front, proficiency metrics translated into plain language, and prior-edition results handled qualitatively instead of as arithmetic. It can also cross-check a draft you wrote for score-versus-narrative mismatches and cross-edition comparisons, and produce a parent-friendly summary of the findings. BastionGPT is HIPAA-compliant with a signed BAA on every plan, your data is never used to train models, and drafting from scores you paste means no protocol or item content ever needs to leave your records.

Primary sources

The instrument facts and compliance claims on this page trace to these sources, last verified July 2026:

  1. Riverside Insights, WJ V product page: the digital-only platform, the 15-test Virtual Test Library, 2022 to 2024 norming against 2020 US Census demographics, the November 1, 2026 Nonverbal launch, and the publisher's own time-saving and post-pandemic-norms claims.
  2. Riverside Insights, WJ V launch release (February 2025): the WJ V COG and ACH batteries and Virtual Test Library available February 1, 2025, hosted on Riverside Score.
  3. Riverside Insights, WJ V Technical Abstract: ages 4 through 90+, the 5,837-person norming sample, the eight-test equally weighted GIA, the removal of Ga from the GIA and COG battery, the Long-Term Retrieval split into Gl and Gr, the new Matrices and Verbal Analogies tests, oral language measures moved into the ACH battery, and the trademark attribution.
  4. Riverside Insights, WJ V brochure: the cluster list, including Auditory Working Memory Capacity (Gwm) and Cognitive Processing Speed (Gs), and the "Coming November 1st, 2026" Nonverbal line.
  5. Riverside Insights, Assessment Service Bulletin on the W score and RPI (Jaffe): the Rasch-based, equal-interval W scale and the RPI's 90% peer-success reading.
  6. Riverside Insights, User Qualifications Guide and university training page: the graduate-level qualification tiers, the short-term WJ IV acknowledgment, and the recommendation against new WJ IV kit purchases.
  7. Riverside Insights, WJ IV Technical Abstract and WJ IV store page: the 2014 publication, ages 2 to 90+, the 1977 first edition, continued availability, and the five-to-ten-minutes-per-test administration estimate.
  8. Nelson, WJ V Canada page: Canadian-server hosting, the transition-support estimate, and the absence of a separate Canadian norming; PAA, WJ IV Australasian Adaptation and WJ V US Adaptation listings: the Australian norms status.
  9. 34 CFR 300.304 and 34 CFR 300.307: the no-single-measure rule and the severe-discrepancy prohibition.
  10. ETS, guidelines for writing diagnostic reports and disability documentation guidance: the five-year currency language, the adult cognitive re-administration position, and the WJ V named among current adult measures.
  11. Buros Center for Testing, Tests reviewed in the Twentieth Mental Measurements Yearbook: independent review status of the WJ IV, with the reviewer's lower-order-scores caution.
  12. Dombrowski, McGill & Canivez, exploratory factor analysis, Psychological Assessment, and confirmatory follow-up, Archives of Scientific Psychology: the four-factor, g-dominant findings; Foster, Perazzo & Decker, Journal of Psychoeducational Assessment (2026): the first independent WJ V review; Trahan et al., Psychological Bulletin (2014): the Flynn-effect meta-analysis.
  13. PR Newswire, Presence named exclusive remote provider of the WJ V (August 2025): the remote-administration arrangement.

Educational content, not legal or billing advice. Sample notes are fictional. Follow your organization's policies and your board, payer, and jurisdiction requirements.