The WPPSI-IV (Wechsler Preschool and Primary Scale of Intelligence, Fourth Edition) is an individually administered intelligence test for ages 2 years 6 months to 7 years 7 months, published in 2012. Psychologists use it for preschool special-education eligibility, kindergarten and gifted decisions, and the cognitive core of early-childhood evaluations. This page covers how to write up WPPSI-IV results, with a fictional sample and a results-section template.
Psychologists and school psychologists; publisher qualification level C
IEP and preschool special-education teams, early-intervention programs, pediatricians, gifted and admissions committees, parents
400 to 800 words for the results section · administration 30 to 60 minutes by age band
Norm-referenced preschool cognitive battery
Preschool special-education and developmental-delay eligibility, early-intervention to preschool transitions, kindergarten-readiness and gifted referrals, autism and developmental evaluations, re-evaluations
Published by Pearson (2012, current edition); described here for write-up purposes, no test content reproduced
The Wechsler Preschool and Primary Scale of Intelligence, Fourth Edition (WPPSI-IV) is an individually administered, norm-referenced intelligence test for children ages 2 years 6 months through 7 years 7 months, published by Pearson in 2012 as the successor to the WPPSI-III (2002) and standardized on 1,700 children. Its defining design fact is that it is effectively two batteries divided at age 4. The younger band (2:6 to 3:11) yields a Full Scale IQ (FSIQ) and three primary index scores, Verbal Comprehension (VCI), Visual Spatial (VSI), and Working Memory (WMI), in the publisher's stated 30 to 45 minutes of core-subtest time. The older band (4:0 to 7:7) adds Fluid Reasoning (FRI) and Processing Speed (PSI) for five primary indexes, at 45 to 60 minutes. Ancillary indexes serve specific questions: the Vocabulary Acquisition Index (VAI) and Nonverbal Index (NVI) at both bands, the General Ability Index (GAI) at both bands with band-dependent composition, and the Cognitive Proficiency Index (CPI) at 4:0 to 7:7 only. Subtest scaled scores run 1 to 19 with 7 to 12 usually considered average; index scores run 40 to 160 with 90 to 109 considered average, each carrying percentile ranks and confidence intervals. Tasks are deliberately preschool-shaped, "engaging tasks for children as young as 2.5 years old" in Pearson's words, including picture-based working-memory subtests and an ink dauber in place of a pencil on processing-speed tasks. Administration is paper or digital via Q-interactive with scoring on Q-global or by hand, and separate Canadian and Australian and New Zealand standardised editions carry local norms. Edition status as of July 2026: the WPPSI-IV remains current, and no WPPSI-5 has been announced in any market. The contrast is visible on Pearson's own pages, where the WISC-V product page banners "WISC-6 is coming" while the WPPSI-IV page carries no equivalent notice.
The load-bearing question for write-ups is architecture and altitude, because most of what ranks in search gets both wrong. Consumer and test-prep pages present a clean five-index battery for all ages, repeat the WPPSI-III's 7:3 upper age bound instead of 7:7, and describe ancillary indexes without noting that the CPI does not exist below age 4 or that the GAI's subtest composition changes at the band boundary, a detail even a published clinical-trial protocol has garbled. The interpretive altitude question is sharper. In the dual Buros review of the WPPSI-IV in the 19th Mental Measurements Yearbook, one reviewer judged it psychometrically strong while Canivez's review documented that exploratory factor analyses "were not reported" in the technical manual despite four subtests being deleted and five added, that incremental validity of the index scores was not reported, and that several index factors rest on only two subtests. His conclusion for practice: clinicians who follow the manual's interpretation schemes "risk overinterpretation and misinterpretation of WPPSI-IV scores in clinical application", and subtest and profile interpretation methods "should not be used in clinical decision-making until psychometric support for them is provided". The defensible posture, and the one this page teaches, names the age band before any number, interprets the FSIQ and reported indexes cautiously, keeps subtest commentary qualitative, and treats the score set as evidence feeding a team or clinical decision, never the decision itself.
School psychologists write up WPPSI-IV results inside preschool special-education eligibility evaluations, most often around the transition from early intervention to preschool services at age 3, where federal law puts an IEP or IFSP in place no later than the third birthday. Clinical child psychologists use it for developmental-delay questions, kindergarten-readiness and gifted referrals, and as the cognitive component of autism and developmental evaluations, where it sits alongside observation measures and adaptive scales such as the Vineland-3 rather than standing alone. It also appears, by convention rather than clinical necessity, in private-school admissions testing in some markets. Age drives instrument choice at both edges: below 2:6, and for very low-functioning children in the overlap years, developmental measures such as the Bayley-4 carry the referral, while at the 6:0 to 7:7 overlap the choice between the WPPSI-IV and the WISC-V follows the publisher's floor-and-ceiling logic, the WPPSI-IV for suspected below-average ability and the school-age scale for high-ability children. A WPPSI-IV section rarely travels alone: it anchors the cognitive portion of a psychoeducational report or broader psychological evaluation report next to speech-language findings, adaptive data, and the behavioral observations that preschool validity rests on.
No statute, payer, or publisher mandates a results-section format. The sequence below is the convention experienced preschool evaluators converge on because it survives review: it names the age band before any score, gives behavioral observation the weight preschool validity actually rests on, keeps composite interpretation at the altitude the psychometric evidence supports, and ends with the caveats that make an early-childhood score usable. Each section carries the pitfall that most often undermines it.
Measures, edition, age band, and format. Name the instrument and edition, the age band administered (2:6 to 3:11 or 4:0 to 7:7), the norm set (US, Canadian, or Australian and New Zealand), the administration format (paper or Q-interactive), the subtests given, and any substitution or proration. Pitfall: an unnamed age band. The two bands yield different index sets, three primary indexes for the younger band and five for the older, so an unlabeled score table is uninterpretable at re-evaluation, and a five-index template forced onto a 3-year-old announces template reuse to any reviewer who knows the test.
Behavioral observations and the validity statement. Document separation from the caregiver, rapport, attention, activity level, frustration tolerance, response to the play-based materials, and breaks, then close with an explicit validity statement tied to what you observed. Preschool cognitive results are only as good as the engagement that produced them. Pitfall: a boilerplate validity sentence. "Results are considered valid" with no observational basis reads as template text; reviewers look for the linkage between observed behavior and the validity claim.
Composite summary. Report the FSIQ and the band's primary indexes, each with confidence interval, percentile rank, and one named descriptor set applied consistently. Pitfall: unlabeled or mixed descriptor vocabularies. The publisher itself calls the qualitative descriptors suggestions rather than evidence-based categories, so the write-up must say which set it uses, and a 92 labeled Average in the table cannot become "solidly average ability" in the summary.
Index-by-index results. Translate each reported index into functional meaning for this child: what the score says about how the child handled that family of tasks, in language a parent and a team can use. Pitfall: subtest-level storytelling. Individual subtest scores are the least reliable numbers on the test, and published reviews of the WPPSI-IV specifically caution against profile and pattern interpretation; keep subtest commentary qualitative and let composites carry the weight.
Ancillary indexes and the reporting decision. When a Vocabulary Acquisition, Nonverbal, General Ability, or (at 4:0 to 7:7) Cognitive Proficiency Index earns a place, report it alongside the FSIQ with the reason stated. Pitfall: a silent composite swap. An ancillary index appearing without rationale reads as score shopping, and at the younger band the GAI still contains an expressive naming task, so it is not the language workaround there that it is at school age.
Stability and norm-currency caveats. State that scores at this age are estimates of current functioning, less stable and less predictive than school-age scores, name the re-evaluation plan, and note the edition and norm vintage. Pitfall: prediction language. "Will remain delayed" and trajectory claims outrun what preschool cognitive scores can support, and reviewers flag them.
Integration with adaptive and developmental data. Read the cognitive results against adaptive behavior, developmental history, speech-language findings, and observation across settings. Pitfall: cognitive-only reasoning. Eligibility and diagnostic frameworks alike require multiple sources; a results section that lets one composite carry the conclusion invites reversal.
Interpretive summary. Answer the referral question, state what the scores do and do not establish, and place the determination where it belongs. Pitfall: eligibility conclusions written as the examiner's verdict. Category decisions belong to teams applying their jurisdiction's criteria; the results section supplies calibrated evidence, not the ruling.
Recommendations linkage. Tie each recommendation to a specific finding, including the re-evaluation timeline and, near age 6 or 7, the instrument-selection note for the next evaluation. Pitfall: recommendations unmoored from findings. A recommendation list that could sit under any profile tells the reader the data did not drive it.
MEASURES AND ADMINISTRATION Instrument/edition: WPPSI-IV (2012) Age band: [2:6-3:11 / 4:0-7:7] Norms: [US / Canadian / A&NZ] Format: [paper / Q-interactive] Subtests given: [core set; any substitution or proration stated] BEHAVIORAL OBSERVATIONS AND VALIDITY [Separation, rapport, attention, activity level, breaks; validity statement tied to observed behavior] COMPOSITE SUMMARY [FSIQ + the band's primary indexes: score, CI, percentile, one named descriptor set applied consistently] INDEX-BY-INDEX RESULTS [Functional meaning per index; no subtest-level claims] ANCILLARY INDEXES (IF REPORTED) [VAI / NVI / GAI, CPI at 4:0-7:7 only; rationale stated; younger-band GAI includes an expressive naming task] STABILITY AND NORM-CURRENCY CAVEATS [Current estimate, not a fixed trait; re-evaluation plan; edition and norm vintage named] INTEGRATION WITH ADAPTIVE AND DEVELOPMENTAL DATA [Adaptive scores, speech-language findings, history, observation; team decision under jurisdiction criteria] INTERPRETIVE SUMMARY (answer the referral question) [What converges; what the scores do NOT establish] RECOMMENDATIONS LINKAGE [Each recommendation tied to a finding; re-evaluation timing; instrument-selection note near the band boundary] Evaluator signature / credentials: Date:
Free to use and share, no signup. The PDF includes a one-page cheat sheet with section-by-section pitfalls and a pre-sign checklist; the DOCX is the blank results-section skeleton, ready to adapt.
Scenario: a 3-year-old approaching the transition from early intervention to preschool special education, referred with expressive-language concerns. Testing uses the younger age band (2:6 to 3:11), which yields three primary indexes rather than five, the architecture detail most sample reports miss. This is the cognitive results section only, condensed but structurally complete. All details are fictional.
Child: L.T., age 3 years 4 months · Referral: preschool special-education team, early-intervention transition, expressive-language concerns · Evaluator: R. Okafor, PsyD, Licensed Psychologist · Testing date: 07/08/2026 · Report date: 07/14/2026
Measures and administration: The Wechsler Preschool and Primary Scale of Intelligence, Fourth Edition (WPPSI-IV) was administered in paper format using the 2:6 to 3:11 age band and United States norms. At this band the battery yields the Full Scale IQ (FSIQ) and three primary index scores, Verbal Comprehension (VCI), Visual Spatial (VSI), and Working Memory (WMI); given the referral question, the Vocabulary Acquisition Index (VAI), Nonverbal Index (NVI), and General Ability Index (GAI) were also derived. All subtests were administered per standard procedures with no substitution or proration. Composites are reported with 95% confidence intervals (CI), percentile ranks, and the qualitative descriptor set printed in the publisher's score reports, named here so the labels are interpretable.
Behavioral observations and validity: L.T. separated from her mother after a brief warm-up period and settled quickly with the play-based materials. She responded readily to pointing-format tasks, used the ink dauber on the visual scanning task with evident enjoyment, and communicated mostly in single words and gestures, consistent with the referral concern. Attention wandered late in the session; one break restored engagement, and standardized procedures were maintained throughout. Because engagement, not effort in the adult sense, is what preschool validity rests on, the validity statement is tied to these observations: the results below are considered a valid estimate of current cognitive functioning, interpreted in the context of her expressive-language presentation.
| Composite | Standard score | 95% CI | Percentile | Descriptor |
|---|---|---|---|---|
| Verbal Comprehension (VCI) | 78 | 73-86 | 7 | Borderline |
| Visual Spatial (VSI) | 92 | 86-100 | 30 | Average |
| Working Memory (WMI) | 85 | 79-93 | 16 | Low Average |
| Full Scale IQ (FSIQ) | 81 | 77-87 | 10 | Low Average |
| Vocabulary Acquisition (VAI) | 71 | 66-81 | 3 | Borderline |
| Nonverbal (NVI) | 90 | 85-96 | 25 | Average |
| General Ability (GAI) | 80 | 75-87 | 9 | Low Average |
Index results: L.T.'s Visual Spatial score (92, Average, 30th percentile) shows age-typical block construction and puzzle assembly, and her Working Memory score (85, Low Average, 16th percentile) sits at the lower edge of the average range on the picture-based memory tasks. Her Verbal Comprehension score (78, Borderline, 7th percentile) is notably lower and should be read alongside a design fact the publisher's own materials state: no subtest contributing to the FSIQ or the primary indexes at either band requires an expressive response on its floor items. Her limited spoken output therefore did not mechanically invalidate these composites, which she could answer by pointing, but the VCI still samples verbal comprehension and is plausibly constrained by her language development rather than by reasoning alone. The Vocabulary Acquisition Index (71, Borderline, 3rd percentile), which does include an expressive naming task, is the profile's lowest score and is reported as the language-acquisition indicator, converging with the speech-language evaluation rather than substituting for it.
Overall ability and the reporting decision: The FSIQ of 81 (CI 77-87, 10th percentile, Low Average) is reported for completeness, but it blends verbal and nonverbal abilities that differ here by nearly a standard deviation. The Nonverbal Index of 90 (CI 85-96, 25th percentile, Average), derived entirely from the four subtests with no expressive-response demand, is used as the anchor for statements about her nonverbal cognitive ability. The General Ability Index (80) is reported but not used as that anchor, for a compositional reason write-ups often miss: at the 2:6 to 3:11 band the GAI includes the expressive naming task, so it does not sidestep the language concern the way the school-age GAI sidesteps working memory and processing speed. Each choice is stated, both scores remain in the table, and no composite has been swapped in silently.
Stability and integration: Cognitive scores obtained at age 3 are current estimates, not fixed traits: they are less stable and less predictive than school-age scores, and re-evaluation is planned before kindergarten entry. On the adaptive side, the Vineland-3 (reported separately) shows Communication well below age expectations with stronger Daily Living and Socialization skills, a pattern consistent with the cognitive profile above. Eligibility for preschool special education is a team determination under this state's developmental-delay criteria; no single measure, including this one, may serve as the sole basis for that decision.
Summary and linkage to recommendations: Testing shows Average nonverbal ability (NVI 90) alongside Borderline verbal comprehension (VCI 78) and a Borderline vocabulary-acquisition indicator (VAI 71), with the FSIQ (81) blending the two domains. The nonverbal-verbal contrast supports recommendation 1 (continuation and intensification of speech-language services through the transition) and recommendation 2 (language-rich preschool programming that lets her demonstrate learning through visual and hands-on formats). The stability caveat supports recommendation 3 (cognitive re-evaluation before kindergarten, noting she will cross into the 4:0 to 7:7 band, where the index architecture changes). The full profile, with the adaptive and speech-language findings, feeds recommendation 4: the eligibility team should apply the state's developmental-delay criteria to the whole evaluation, not to any single score.
This sample is fictional and for educational purposes. It does not describe a real child or record, and the scores are invented for illustration and correspond to no real child or record.
Writing these after every session? BastionGPT drafts complete notes from bullets, dictation, or a transcript.
Generate a note from bulletsWrite the results section knowing which decision framework will read it, and label each requirement's strength honestly. US federal special-education law is instrument-neutral: no regulation names the WPPSI-IV, and evaluators may "not use any single measure or assessment as the sole criterion" for eligibility (34 CFR 300.304(b)(2)). The eligibility category preschool evaluations most often feed, developmental delay, is a state option, not a federal mandate: states may apply it to ages 3 through 9 or any subset (34 CFR 300.8(b)), a district may not use the category unless its state adopted it (300.111(b)), and the ECTA Center's state-by-state table shows 47 of the 50 states use it with age windows that vary widely, while "CA, IA, PR, and TX do not use the developmental delay category" and "BIE applies the developmental delay category to ages 4-9". So the same profile that qualifies a child in one state routes through a different category next door, and eligibility language in the write-up should track your state's definition, not a generic one. Timing is law: the obligation to make FAPE available begins "no later than the child's third birthday" with an IEP or IFSP in effect by that date (300.101(b), 300.124), which is why so many WPPSI-IV administrations cluster just before age 3 transitions, and re-evaluation "must occur at least once every 3 years" unless parent and agency agree otherwise (300.303(b)(2)). Gifted identification has no federal mandate at all: where thresholds exist they are state or district policy, and Pennsylvania's regulation is the instructive exemplar, naming "an IQ of 130 or higher" while directing that "Determination of gifted ability will not be based on IQ score alone" (22 Pa. Code 16.21). Testing time bills to payers under the psychological-testing code families (96130 and 96131 for evaluation services, 96136 and 96137 for administration and scoring, with 96112 and 96113 covering developmental testing) as payer policy with plan-specific rules.
Edition and norms currency is the WPPSI-IV's quiet defensibility question. As of July 2026 the WPPSI-IV (2012) is the current edition and no WPPSI-5 has been announced anywhere, while Pearson's own WISC-V page banners the coming WISC-6, so a report that names the edition is future-proofing itself for the transition question referral sources will eventually ask; when a new edition does publish, Pearson's own transition guidance is that most practitioners move to the new edition within 8 to 12 months of its release. The norms themselves date to the test's 2012 development cycle, and aging norms have a quantified cost: the largest Flynn-effect meta-analysis estimates scores drift about 2.31 points per decade (2.93 for modern Wechsler-family tests), renorming visibly moves individual children, students near eligibility thresholds "lost an average of 5.6 points when retested on a renormed test" in the classic Kanaya, Scullin, and Ceci analysis, and the drift is not uniform across domains, with Wechsler-family analyses noting the particular sensitivity of nonverbal scales and cautioning against mechanical point adjustments. The write-up move is simple: name the edition and norm vintage, never hand-adjust scores for norm age, and compare like with like at re-evaluation. Format deserves one line too: Pearson's equivalence study for digital administration, Q-interactive Technical Report 14, found all subtests within its 0.2 effect-size criterion, concluding "the WPPSI-IV produces consistent scores regardless of format", and the digital kit still uses physical manipulatives, paper response booklets for processing-speed subtests, and a paper stimulus book for Picture Memory, so "Q-interactive" and "paper" are not shorthand for fully digital versus fully analog; as of July 2026 Pearson also lists no WPPSI-IV-specific telepractice equivalence report alongside its WISC-V and WAIS-family guidance, so any remote administration deserves an explicit documented rationale. Access boundaries are the publisher's: qualification level C in the US, with the Canadian edition (French materials available) and the Australian and New Zealand standardised edition, whose normative sample was stratified for "age, gender, parental education, geographic location, and indigenous status", administered through registered psychologists, and Australian school-adjustment and NDIS processes treating cognitive results as supporting evidence alongside functional documentation by convention. Retest timing is convention as well: practice effects make short-interval re-administration hard to interpret, and the publisher's stability data cover only short retest windows, so state the plan rather than implying a rule.
Wechsler, WPPSI, and Pearson are trademarks, in the US and other countries, of Pearson Education, Inc. or its affiliates; WPPSI-IV materials are copyrighted by NCS Pearson, Inc. BastionGPT is not affiliated with, or endorsed by, the publisher. This page reproduces no test items, stimuli, norms, or scoring materials.
There is no payer audit series for preschool cognitive write-ups; the accountability literature here is psychometric, and it cuts close to everyday practice. The WPPSI-IV drew a split verdict in the Buros Center's 19th Mental Measurements Yearbook: one reviewer found it psychometrically strong, while Canivez documented that the manual reported no exploratory factor analyses despite wholesale subtest turnover, no incremental-validity evidence for the index scores, and interpretation schemes resting on what he called "a shared professional myth of subtest and profile utility". Administration mechanics are the publisher's problem; the errors below are write-up errors, and they are yours. The BastionGPT Clinical Advisory Board sees the same ones most often in WPPSI-IV report reviews:
BastionGPT is specifically trained, tuned, and clinically tested on psychological and psychoeducational evaluation reports.
See how clinicians use it day to day on the AI therapy notes page.
Many BastionGPT users report saving more than 90 minutes per day on documentation.
HIPAA-compliant with a signed BAA on every plan. Your data is never used to train models. BastionGPT drafts, you review and sign.
Two metrics carry everything. Subtest scaled scores run 1 to 19, with 7 to 12 usually considered average; composite scores, the FSIQ and the index scores, run 40 to 160, with 90 to 109 considered average, and each composite carries a percentile rank and a confidence interval. The manual supplies estimated-true-score confidence intervals, and its own text notes "there may be a preference for using the obtained score confidence interval" when the question is the child's functioning at the time of testing, a precision nuance the Buros review walks through. Qualitative labels are looser than most readers assume: the publisher's own FAQ says descriptors are "only suggestions and are not evidence-based; alternate terms may be used as appropriate", and Pearson's printed reports use classic Wechsler-era terms such as Borderline and Extremely Low. The defensible write-up reports score, interval, and percentile together, names the descriptor set it uses, and applies it consistently.
As of July 2026 there is no WPPSI-5: no announcement, pre-order page, or standardization recruitment in any market. The absence is informative because Pearson visibly advertises Wechsler revisions that do exist, the WAIS-5 published in 2024 and the WISC-V product page currently banners "WISC-6 is coming", while the WPPSI-IV page carries no such notice. The WPPSI-IV (2012) is therefore fully current and fully defensible; what ages a report is not the edition but the omissions. Norms from the 2012 development cycle do drift, about 2.31 points per decade in the largest Flynn-effect meta-analysis, which is a reason to name the edition and norm vintage in every report, not to hand-adjust scores. And when a fifth edition eventually arrives, the publisher's own transition guidance for the current edition was that most practitioners move to a new edition within 8 to 12 months of its release.
Because the WPPSI-IV is two batteries divided at age 4, and a child tested at 2:6 to 3:11 gets the younger one by design. That band yields the FSIQ plus Verbal Comprehension, Visual Spatial, and Working Memory; Fluid Reasoning and Processing Speed exist only in the 4:0 to 7:7 band, which yields all five primary indexes. The ancillary architecture follows the same split: the Vocabulary Acquisition and Nonverbal Indexes exist at both bands, the General Ability Index exists at both but with band-dependent subtest composition, and the Cognitive Proficiency Index exists only from 4:0. Two related corrections worth knowing: the younger band's FSIQ is built from five core subtests including the Picture Memory working-memory task, which the publisher notes makes it broader than the WPPSI-III's four-subtest version, and the test's upper bound is 7 years 7 months, not the 7:3 figure many secondary pages copied forward from the WPPSI-III. A three-index preschool report is correct, not incomplete.
No law names it. IDEA is instrument-neutral and bars using "any single measure or assessment as the sole criterion" for eligibility (34 CFR 300.304(b)(2)). The developmental-delay category itself is a state option: states may apply it to ages 3 through 9 or any subset (300.8(b)), and a district may not use it unless the state adopted it (300.111(b)). The ECTA Center's current table shows 47 of the 50 states use a developmental-delay category with widely varying age windows, while "CA, IA, PR, and TX do not use the developmental delay category" and "BIE applies the developmental delay category to ages 4-9". What is fixed federally is timing: FAPE must be available "no later than the child's third birthday" with an IEP or IFSP in effect by that date (300.101(b)). So the WPPSI-IV earns its place by convention, as one appropriate cognitive measure inside a multi-source evaluation, and your write-up's eligibility language should track your state's definition.
The 6:0 to 7:7 overlap is deliberate, and the publisher's own selection logic is floor versus ceiling: for children with suspected below-average ability, the WPPSI-IV's lower floor differentiates better, while high-ability children need the school-age scale's higher ceiling. Pearson's overlap FAQ, written when the school-age edition was the WISC-IV, also noted a practical difference, the WPPSI-IV reaches an FSIQ in six subtests versus ten, and the trade-offs carry forward to the WISC-V, whose FSIQ uses seven. Developmental fit matters as much as psychometrics at this age: the WPPSI-IV's game-like, dauber-and-pictures format suits developmentally younger or shy children, while a verbally advanced 7-year-old is usually better served, and better challenged, on the WISC-V. Whichever you choose, say why in the report, name the norm set, and keep the choice consistent with the referral question, because at re-evaluation the two tests' different index architectures make cross-battery score comparison a stated limitation, not a trend line.
Choose by referral question. A developmental measure like the Bayley-4 answers "is development proceeding as expected across domains", covers motor, language, and social-emotional ground the WPPSI-IV never touches, and serves low-functioning children through its extended floor; the WPPSI-IV answers "what is this child's cognitive ability relative to peers" with an extended ceiling and the Wechsler common metric. That floor-versus-ceiling logic is the publisher's own, stated in its instrument-selection FAQ for the previous generation (Bayley-III versus WPPSI-III), and worth knowing in that dated form: the Bayley for suspected delay and trajectory questions, the Wechsler when an ability score is required. Publisher concordance data for the current editions show the two land close on average, a mean WPPSI-IV FSIQ of 103.3 against a Bayley-4 cognitive standard-score equivalent of 98.5 in the comparison sample, so crossing between them at the age-3 transition is expected, not alarming. Either way, cognitive scores at this age belong next to adaptive data such as the Vineland-3 and a speech-language evaluation, not in place of them.
Often yes, and the test's design helps more than most write-ups acknowledge. Pearson's own FAQ states that only Vocabulary and Picture Naming require expressive responses on the floor items, and "neither of these subtests are core to the primary index scores or FSIQ", so a child who points reliably can produce interpretable composites. The Nonverbal Index goes further, drawing entirely on subtests with no expressive demand, while the Vocabulary Acquisition Index serves as the language-acquisition indicator rather than a general-ability score; note that at the younger band the General Ability Index still includes an expressive naming task, so it is not the language workaround it becomes at school age. The non-negotiable part is behavioral documentation: separation, rapport, attention, breaks, and response to the play materials, ending in a validity statement tied to what you observed. When standardized administration genuinely breaks down, the defensible moves are to qualify the results as a likely underestimate, or to shift to a developmental measure and say why.
Not on this page, and not legitimately on any public page. WPPSI-IV items, stimuli, and norms are protected test materials sold at qualification level C, and psychologists carry an ethical duty to maintain test security under APA Ethics Standard 9.11. The pressure here is real: a test-prep market exists around preschool admissions and gifted screening, and that is exactly why coached exposure matters clinically, because familiarity with item types distorts the norm-referenced comparison every downstream decision relies on, at an age when scores are already their least stable. Even where gifted thresholds are codified, as in Pennsylvania's regulation naming "an IQ of 130 or higher", the same rule directs that the determination "will not be based on IQ score alone". A report may include scores, percentiles, descriptors, and your own prose about what the tasks measure; it may not include item content, stimuli, scoring rules, or reproduced norm tables.
Yes. BastionGPT is trained and clinically tested on psychological and psychoeducational evaluation reports, the parent documents a WPPSI-IV section lives inside. Paste a score summary (age band, composites, CIs, percentiles, descriptors, and your observation notes) and it drafts the results-section narrative with the band architecture right, the validity statement tied to observed behavior, and any FSIQ-versus-ancillary reporting decision framed with its rationale for your review. It can also check a finished section for the errors reviewers flag, an unnamed age band, numbers that disagree with the table, descriptor drift, prediction language, or a silent composite swap, and produce a plain-language summary for parents and preschool teams. BastionGPT is HIPAA-compliant with a signed BAA on every plan, and your data is never used to train models.
The instrument facts and compliance claims on this page trace to these sources, last verified July 2026:
Educational content, not legal or billing advice. Sample notes are fictional. Follow your organization's policies and your board, payer, and jurisdiction requirements.