KTEA-3 Report Write-Up: Structure, Sample Language & Common Errors

The KTEA-3 (Kaufman Test of Educational Achievement, Third Edition) is an individually administered academic achievement battery for ages 4:0 through 25:11, measuring reading, math, written language, and oral language with standard scores set to a mean of 100. Psychologists and school psychologists use it in specific learning disability evaluations for its norm-referenced error analysis and parallel forms. This page covers how to write up KTEA-3 results, with a fictional sample.

Free to use and share. No signup required.
Already have session bullets or a transcript? Generate a structured draft with BastionGPT — you review and sign it.
Who writes it

Psychologists, school psychologists, and educational diagnosticians; publisher qualification level B

Audience

IEP and Section 504 teams, parents and teachers, dyslexia programs, accommodations reviewers, physicians

Typical length

500 to 900 words for the results section · administration 15 to 85 minutes for the ASB Composite by grade

Format family

Norm-referenced academic achievement battery

When it's used

Specific learning disability and psychoeducational evaluations, dyslexia referrals, re-evaluations with parallel forms, ages 4 to 25

Standards context

Published by NCS Pearson (2014); current edition as of July 2026; described here for write-up purposes, no test content reproduced

What is the KTEA-3?

The Kaufman Test of Educational Achievement, Third Edition (KTEA-3) is an individually administered academic achievement battery authored by Alan S. Kaufman and Nadeen L. Kaufman and published by NCS Pearson in 2014, with Pearson stating as of July 2026 that it remains the most current version of the KTEA. Nineteen subtests build fourteen composites across four academic domains, Reading, Math, Written Language, and Oral Language, with the Academic Skills Battery (ASB) Composite serving as the overall score and supplemental composites (Decoding, Sound-Symbol, Reading Fluency, Reading Understanding, Orthographic Processing, Oral Fluency, Comprehension, Expression, and Academic Fluency) sharpening specific referral questions. Every composite and every subtest reports on the same standard-score metric, a mean of 100 and a standard deviation of 15, with age-based norms for 4:0 through 25:11 and grade-based norms for pre-K through 12, plus percentile ranks, normal curve equivalents, stanines, cautioned age and grade equivalents, and Growth Scale Values for comparing performance across administrations. Two parallel forms (A and B) support retesting, administration runs 15 to 85 minutes for the ASB Composite depending on grade, and the battery runs on paper or Q-interactive with scoring by hand or in Q-global. Two satellites extend it: the KTEA-3 Brief Form (2015), a three-composite short form whose standard scores Pearson states may be used interchangeably with Comprehensive Form scores, and the KTEA-3 Dyslexia Index (2018, 2021), a 12-to-20-minute risk screener scored from a small set of Form A or Form B subtests.

Two distinctions carry most KTEA-3 write-up mistakes. First, the metric: KTEA-3 subtests use standard scores with the same 100/15 scale as the composites, while the Wechsler cognitive batteries report subtest scaled scores with a mean of 10, so a psychoeducational report that mixes a WISC-V and a KTEA-3 must keep the two scales visibly separate or a subtest score of 9 and one of 90 will read as comparable. Second, the signature feature: norm-referenced error analysis, available for 10 of the 19 subtests, classifies what a student got wrong and compares the error count in each category with the average number of errors made by grade-level peers, which is what makes the KTEA-3 the achievement battery clinicians choose when the referral question is instructional. Against its nearest neighbors, the WIAT-4 offers newer norms (2020) and native Wechsler pairing and the Woodcock-Johnson offers a broad CHC cross-battery design, while the KTEA-3 competes on error analysis, true parallel forms, and two rapid automatized naming measures.

Who uses KTEA-3 reports and when

Psychologists, school psychologists, and educational diagnosticians interpret the KTEA-3, which Pearson lists at qualification level B, and their write-ups land in front of IEP and Section 504 teams, parents and teachers, dyslexia programs, postsecondary disability offices, and sometimes physicians and payers. The classic setting is a specific learning disability evaluation: the KTEA-3 supplies the achievement side of a school or clinic psychoeducational report, paired with a cognitive measure such as the WISC-V, for which Pearson's scoring platforms produce achievement and ability comparisons. It earns its seat over neighboring batteries in three situations: when the team needs error analysis translated into instructional targets, when a re-evaluation calls for a genuine parallel form so practice effects stay out of the comparison, and when a dyslexia referral wants rapid automatized naming measures and the Dyslexia Index screener inside the same kit. Age norms reaching 25:11 also make it a common choice for documenting academic skill deficits in college accommodation files. Upstream of it sit curriculum-based measures and universal screeners that flag risk within RTI or MTSS tiers; the KTEA-3 is the diagnostic, norm-referenced battery the evaluation reaches for once eligibility or diagnosis is on the table, and its results section reads differently from a WIAT-4 section mainly in how much instructional narrative the error analysis supports.

How to structure a KTEA-3 results section

No statute, payer, or publisher mandates a KTEA-3 results-section format. The structure below is the convention experienced evaluators converge on because it survives review: it names the form, norms, and mode before any number, organizes results by academic domain instead of printout order, turns error analysis into instructional narrative, and gives each composite exactly the scope it has. Each section carries the pitfall that most often undermines it.

Measures and methods statement. Name the battery and edition (KTEA-3, 2014), the form administered (A or B), the norm set applied (age-based or grade-based, and which grade norms if applicable), the platform (paper, Q-interactive administration, Q-global or hand scoring), and the mode. If any part was administered via telepractice, Pearson's guidance directs the report to state that and briefly describe the method, because the norms were collected in person. Pitfall: writing "the KTEA-3 was administered" with nothing else. Form, norm type, and mode each change what a score means, and age-based and grade-based scores for the same raw performance can differ visibly.

Behavioral observations during testing. Two or three sentences on engagement, attention, fatigue, sensory barriers, and how the examinee handled timed and open-ended tasks, closing with a statement of whether results are considered a valid estimate of current academic skills and any accommodation or deviation that occurred. Pitfall: boilerplate observations that could describe any child. Reviewers use this paragraph to judge whether the scores deserve their interpretation, and error analysis downstream depends on responses actually attempted.

Domain-by-domain results. Report each relevant domain in a stable order (Reading, Math, Written Language, Oral Language): the domain composite first, then its subtests, each with standard score and percentile, then what the pattern within the domain shows. Describe timed against untimed contrasts (fluency subtests against their untimed counterparts) in prose. Pitfall: marching through the score report in printout order, or drifting between metrics. Every KTEA-3 subtest is a 100/15 standard score; if a Wechsler battery appears in the same report, keep the scales visibly separate and the descriptor system consistent.

Error-analysis narrative. This is the KTEA-3's signature and the reason instructional readers keep the report. For the subtests analyzed (available on 10 of the 19), state which error categories were classified as weaknesses relative to the average number of errors grade peers make, then translate the categories into plain language about what the student can and cannot yet do, and connect each to an intervention target. Pitfall: reporting raw error counts without the normative comparison, or pasting Q-global's interpretive and intervention text. Pearson's report license limits excerpting to the minimum text needed for your conclusions, so the narrative must be yours.

Composite handling. Report the ASB Composite as the overall academic summary while naming its scope: it is built from six reading, math, and written-language subtests and does not include the oral-language subtests. Give supplemental composites (Decoding, Sound-Symbol, Reading Fluency, Orthographic Processing, Academic Fluency) their referral-specific work, and report the Dyslexia Index as the risk screener it is. Pitfall: letting an average ASB imply the whole battery was average, or writing the Dyslexia Index as a diagnosis. Pearson states a single score is not sufficient to diagnose dyslexia.

Cognitive linkage and eligibility framing. Where the evaluation pairs the KTEA-3 with a cognitive measure, report the achievement and ability comparison or pattern-of-strengths-and-weaknesses analysis under the identification model your state actually permits, and label it. Federal regulations bar states from requiring a severe ability-achievement discrepancy and require that an RTI-based process be permitted, so the defensible sentence names the model and the state rule rather than treating one model as universal. Pitfall: presenting a discrepancy table as "the" SLD test. Which models are permissible is state policy, and the eligibility decision belongs to the team, not to any one instrument.

Summary and recommendations linkage. Close the section with the profile in two or three sentences (which domains are impaired, which are preserved, and how the error analysis explains the referral concern), then recommendations that trace to findings: intervention targets from the error categories, accommodations tied to documented fluency or transcription weaknesses, and a re-evaluation plan that names the alternate form and Growth Scale Values for measuring change. Pitfall: recommendations that ignore the error analysis, or a retest plan built on an invented waiting period. No authority sets a fixed retest interval; the parallel form exists so retesting is a design decision, not a countdown.

Blank template (copy and adapt)

KTEA-3 RESULTS SECTION SKELETON (adapt; delete guidance in parentheses before signing)

MEASURES AND METHODS
Edition (KTEA-3, 2014) and form (A / B): ____
Norm set (age-based / grade-based, fall or spring) and why: ____
Platform (paper / Q-interactive; Q-global or hand scoring) and mode
(in person / telepractice, with method described): ____

BEHAVIORAL OBSERVATIONS
Engagement, attention, timed-task behavior, accommodations: ____
Validity statement (results a valid estimate? deviations?): ____

DOMAIN-BY-DOMAIN RESULTS (composite first, then subtests; standard scores + percentiles)
Reading (composite, subtests, timed vs untimed pattern): ____
Math (composite, concepts vs computation pattern): ____
Written Language (composite, spelling vs expression pattern): ____
Oral Language (composite, listening vs speaking pattern): ____

ERROR-ANALYSIS NARRATIVE (for analyzed subtests)
Categories flagged as weaknesses vs grade peers: ____
What the student can / cannot yet do, in plain language: ____
Intervention targets tied to each flagged category: ____

COMPOSITE HANDLING
ASB Composite with its scope stated (no oral-language subtests): ____
Supplemental composites used and why (Decoding, Sound-Symbol,
Reading Fluency, Orthographic Processing, Academic Fluency): ____
Dyslexia Index reported as risk screening, not diagnosis: ____

COGNITIVE LINKAGE AND ELIGIBILITY FRAMING
Ability-achievement / PSW analysis with the state model named: ____

SUMMARY AND RECOMMENDATIONS
Profile summary tied to the referral question: ____
Recommendations traced to findings, incl. re-evaluation plan
(alternate form, Growth Scale Values): ____

Examiner signature, credentials, date: ____

Free to use and share, no signup. The PDF includes a one-page cheat sheet with section-by-section pitfalls and a pre-sign checklist; the DOCX is the blank results-section skeleton, ready to adapt.

Sample KTEA-3 write-up (fictional)

Scenario: an 8-year-old third grader referred for an initial special education evaluation after two years of reading intervention with limited response. The KTEA-3 supplied the achievement component of a school psychoeducational evaluation; this is the KTEA-3 section of the resulting report, condensed but structurally complete. All details are fictional.

Student: J.T., age 8:6, grade 3  ·  Referral: persistent reading difficulty despite tiered intervention; initial SLD evaluation  ·  Examiner: D. Okafor, PsyD, School Psychologist  ·  Testing date: 07/10/2026  ·  Report date: 07/16/2026

Measures and methods: Kaufman Test of Educational Achievement, Third Edition (KTEA-3), Form A, administered in person on paper and scored in Q-global. Age-based standard scores (mean 100, SD 15) are reported below; error analysis, which is normed on grade-level peers, was completed for the decoding, spelling, and phonological subtests given the referral question. The KTEA-3 was selected for this evaluation because its error analysis and naming-fluency measures speak directly to the instructional questions raised by two years of limited intervention response. Cognitive results (WISC-V) are reported in the preceding section.

Behavioral observations: J. worked cooperatively through two sessions with a movement break, attempted every task presented, and stayed engaged with encouragement on reading tasks, where he read quietly and self-corrected often. He completed timed tasks without distress but visibly slowed on unfamiliar words, sounding them out letter by letter. Results are considered a valid estimate of his current academic skills; no deviations from standardized administration occurred.

Results: Composite standard scores appear below; subtest scores follow in the domain narrative. Descriptive categories follow the score report's classification labels.

CompositeStandard scorePercentileDescriptive category
Academic Skills Battery (ASB)8414Below average
Reading8110Below average
Decoding798Low
Sound-Symbol776Low
Reading Fluency765Low
Written Language8110Below average
Math9537Average
Oral Language10050Average

Reading. The Reading Composite of 81 (10th percentile) understates how uneven the underlying skills are. Letter & Word Recognition (82, 12th percentile) and Reading Comprehension (85, 16th percentile) sit at the low end of the Below average range, while the decoding machinery beneath them is weaker still: Nonsense Word Decoding 78 (7th percentile), Phonological Processing 80 (9th percentile), Word Recognition Fluency 76 (5th percentile), and Decoding Fluency 77 (6th percentile). The timed and untimed contrast matters: J. reads some real words accurately given unlimited time, but speed collapses when fluency is required, and pseudowords remove the sight-vocabulary support entirely. Rapid naming was low average to below average (Letter Naming Facility 84, 14th percentile; Object Naming Facility 88, 21st percentile), slower for letters than for objects.

Error analysis (reading and spelling). On Nonsense Word Decoding, J.'s errors in vowel patterns, including short and long vowel forms and vowel teams, were classified as weaknesses, meaning he made more errors in those categories than the average third grader in the norm sample, while single-consonant errors fell in the average range. Spelling errors concentrated in the same place: plausible consonant frames with the vowel wrong (for example, spelling attempts preserving first and last sounds but not the medial vowel). Phonological Processing errors clustered on segmenting and blending items. In plain terms, J. hears and produces beginning and ending sounds reliably but cannot yet map vowels to print or pull apart the middle of words, and that single weakness explains most of what the classroom sees.

Written language and math. The Written Language Composite of 81 (10th percentile) reflects Spelling at 79 (8th percentile) against stronger Written Expression at 86 (18th percentile): sentence-level ideas outrun transcription. Math is a relative strength (composite 95, 37th percentile; Math Concepts & Applications 96, Math Computation 94), with no error categories flagged.

Oral language. The Oral Language Composite of 100 (50th percentile), with Listening Comprehension at 102 (55th percentile) and Oral Expression at 98 (45th percentile), is squarely average. The ASB Composite of 84 (14th percentile) summarizes the six reading, math, and written-language subtests and does not include these oral-language scores, so the overall number should not be read as contradicting the preserved listening and speaking skills.

Summary for the team. Achievement testing shows a specific pattern: intact oral language and average math against weak decoding, spelling, and reading fluency, with error analysis locating the breakdown in vowel-pattern knowledge and phonemic segmentation. These achievement data feed the eligibility analysis in the Data Integration section under the state's criteria; no single score determines eligibility, and that decision belongs to the team.

Recommendations linked to findings. (1) Structured literacy intervention targeting vowel patterns and phoneme segmentation and blending, the categories error analysis flagged, with decodable text practice to build fluency. (2) Classroom supports tied to the documented fluency weakness: extended time for reading-dependent tasks and no cold oral reading. (3) Spelling instruction aligned to the same vowel-pattern sequence rather than memorized lists. (4) At re-evaluation, administer KTEA-3 Form B and compare Growth Scale Values on the targeted subtests so practice effects and measurement error stay out of the progress judgment.

This sample is fictional and for educational purposes. It does not describe a real student or record, and the scores are invented for illustration and correspond to no real child or record.

↑ Back to the template and downloads

Why this sample works

  • The methods statement is complete. Edition, form, norm type, platform, mode, and the reason this battery fits the referral are all named in one place, so any reviewer can reconstruct exactly what produced the numbers, including which scores are age-normed and which comparisons are grade-normed.
  • Scores and narrative agree on one metric and one descriptor system. Every claim in the prose matches the table, all scores are 100/15 standard scores kept apart from the cognitive battery's scaled scores, and the descriptive categories follow the score report's own labels consistently.
  • Error analysis does instructional work. Flagged categories are reported against the grade-norm comparison, translated into plain language about what the student can and cannot yet do, and carried straight into the recommendations, in the evaluator's own words rather than pasted platform text.
  • Every composite gets exactly its scope. The ASB is summarized with its six-subtest coverage stated so it cannot contradict the intact oral-language findings, the supplemental composites answer the referral question, and the eligibility decision is routed to the team under the state's criteria.
  • Recommendations and the retest plan trace to findings. Intervention targets come from the flagged error categories, accommodations tie to the documented fluency weakness, and re-evaluation names Form B with Growth Scale Values instead of an invented waiting period.

Writing these after every session? BastionGPT drafts complete notes from bullets, dictation, or a transcript.

Generate a note from bullets

Documentation and compliance considerations

Label every eligibility and coverage claim by what it actually is, because the KTEA-3 lives in decision contexts with three different kinds of rules. Federal special education law is LAW: the IDEA regulations state that in identifying a specific learning disability a state "must not require the use of a severe discrepancy between intellectual ability and achievement", must permit a process based on the child's response to scientific, research-based intervention, and may permit other alternative research-based procedures, the clause under which pattern-of-strengths-and-weaknesses models operate, and the same regulations require a comprehensive evaluation in which no single measure or score decides anything. Which of those identification models your state actually permits is state POLICY, as are the dyslexia screening statutes that make the KTEA-3's phonological, naming-fluency, and decoding measures attractive in reading referrals. Payment is its own layer: school evaluations under IDEA are publicly funded rather than billed to health insurance, while clinic and private evaluations bill the psychological testing CPT family (evaluation services 96130 and 96131, administration and scoring 96136 through 96139), codes that APA Services notes appear in CPT Appendix P for audio-video and Appendix T for audio-only telemedicine and that are maintained on the CMS Medicare telehealth list for calendar year 2026, which is PAYER POLICY worth re-checking each year. In Canada there is no federal IDEA analogue: identification frameworks come from provincial ministries and school boards, and much psychoeducational testing is accessed privately or through school-board queues, which is CONVENTION rather than entitlement. In Australia, school-eligibility achievement testing is generally not a Medicare-rebatable service and runs through state education pathways or private assessment. In the UK, SASC guidance lists the KTEA-3 as suitable for higher-education specific-learning-difficulty assessment and remains on its current suitable-tests list, with the plain caution that "The tests are only normed up to age 25".

The battery-specific duties sit in the methods and narrative. The KTEA-3's norms date to 2014 and Pearson states it remains the current edition as of July 2026, while the neighboring WIAT-4 (2020) and the fifth-edition Woodcock-Johnson carry newer standardizations, so a defensible report names the edition and norm vintage rather than assuming "current" is self-evident, states whether age-based or grade-based scores are reported (and that error analysis is normed on the grade sample), and dates any edition claim. The same paragraph should respect two publisher boundaries. Purchase requires Pearson qualification level B, and qualification to buy is not competence for every interpretive use, which remains a licensing and professional-standards matter. And the Q-global narrative is not free text: Pearson's own score report licenses excerpting "limited to the minimum text necessary to accurately describe their significant core conclusions", with no adaptations without written permission, so the results narrative and the error-analysis interpretation must be written in your own words with platform output treated as data. Telepractice administration carries a documentation duty of its own, since Pearson's guidance directs the report to state that testing occurred via telepractice and describe the method, the norms having been collected in person. Retention, access, and records handling follow the parent document: a KTEA-3 section inside a school or clinic psychoeducational report inherits that report's education-records and health-records obligations, including parent access and raw-data handling covered there.

KTEA is a trademark of NCS Pearson, Inc. BastionGPT is not affiliated with, or endorsed by, the publisher. This page reproduces no test items, stimuli, norms, or scoring materials.

Common KTEA-3 write-up errors reviewers flag

The published record gives reviewers reasons to read KTEA-3 sections closely. A study of teacher trainees administering core KTEA-3 subtests found administration and clerical errors common, with Reading Comprehension the most error-prone subtest (Lockwood and colleagues, 2020), and an independent factor analysis in Contemporary School Psychology concluded that "the KTEA-3 is factorally complex", cautioning against over-reading structure into score patterns (Parkin and Frisby, 2019). Pearson's own report license, printed on every Q-global output, limits excerpting to the minimum text necessary for your conclusions. The BastionGPT Clinical Advisory Board sees the same errors most often in KTEA-3 write-up reviews:

  • Metric drift in a mixed battery. KTEA-3 subtests are 100/15 standard scores; Wechsler subtests are 10/3 scaled scores. Reports that run both without labeling the scales, or that apply one battery's descriptor bands to the other's scores, misstate results even when every number is transcribed correctly. Name the metric once and keep one descriptor system throughout.
  • Error analysis missing or mishandled. The instructional payload of the KTEA-3 is the norm-referenced error analysis, yet reviewers see sections that skip it, report raw error counts without the comparison to the average number of errors grade peers make, or paste the platform's interpretive and intervention text wholesale. The license allows minimum-text excerpting only; the translation into what the student can and cannot yet do is the evaluator's job.
  • No form, norm-type, or edition statement. Form A and Form B are different item sets, age-based and grade-based scores answer different questions, and the 2014 norm vintage deserves naming while newer neighbor batteries exist. A score whose yardstick is unstated cannot be audited, and a re-evaluation that silently switched forms or norm sets cannot be compared to baseline.
  • Grade equivalents doing placement work. Pearson lists age and grade equivalents among the scores with published cautions, and reviewers flag sentences that treat a grade equivalent of 1.8 as "reads at a first-grade level" or as an instructional placement. Keep interpretation on standard scores and percentiles, and let grade equivalents illustrate rather than decide.
  • Composites given the wrong scope. The ASB Composite summarizes six reading, math, and written-language subtests and omits oral language, so an average ASB cannot clear the whole battery, and the Dyslexia Index is a risk screener whose own product page states a single score "is not sufficient to diagnose dyslexia". Reviewers also flag subtest-scatter stories built on small differences the battery's factor structure will not support.
How BastionGPT helps

BastionGPT is specifically trained, tuned, and clinically tested on psychological and psychoeducational evaluation reports.

  • Paste a KTEA-3 score summary (form, norm set, composites, subtests, flagged error categories) and BastionGPT drafts a domain-organized results narrative with the methods statement, percentile framing, and error-analysis-to-instruction translation set up for your review.
  • It cross-checks a finished section for the reviewer flags on this page: numbers that disagree with the table, subtest standard scores drifting into scaled-score language, a missing form or norm-type statement, grade equivalents written as placements, or a Dyslexia Index phrased as a diagnosis.
  • It turns finished findings into a plain-language summary for parents and teachers that keeps scores as estimates and explains what the eligibility team does next.

See how clinicians use it day to day on the AI therapy notes page.

Many BastionGPT users report saving more than 90 minutes per day on documentation.

HIPAA-compliant with a signed BAA on every plan. Your data is never used to train models. BastionGPT drafts, you review and sign.

Frequently asked questions

Every KTEA-3 score, subtest and composite alike, is a standard score with a mean of 100 and a standard deviation of 15, so 100 sits at the middle of the distribution for same-age or same-grade peers, 85 falls about one standard deviation below it (roughly the 16th percentile), and 115 about one above (the 84th). Percentile ranks, normal curve equivalents, stanines, cautioned age and grade equivalents, and Growth Scale Values round out the score set. Keep the percentile next to the number, and if you use descriptor words, use one labeling system consistently; Pearson's score reports print descriptive categories such as Average and Below average, and mixing systems across a report is a reviewer flag. The error-analysis output is different in kind: it classifies what the student got wrong and compares each category's error count with the average number of errors made by grade-level peers.

No KTEA-4 exists: Pearson's product page states the KTEA-3 is the most current version of the KTEA, a statement verified in July 2026. The norms date to the 2014 publication, which matters because the neighboring WIAT-4 (2020) and the fifth-edition Woodcock-Johnson carry newer standardizations. That does not make KTEA-3 results indefensible; it makes edition and norm-vintage statements mandatory. Name the battery, year, and norm set in the methods sentence, date any "current edition" claim, and if an evaluation pairs the KTEA-3 with newer-normed instruments, keep the vintages visible rather than letting all the scores read as if they share one reference period.

Both are individually administered, comprehensively normed Pearson achievement batteries at qualification level B, and no authority mandates either, so the choice is convention driven by the referral question. The KTEA-3's case: norm-referenced error analysis on ten subtests that converts directly into instructional targets, true parallel Forms A and B for re-evaluation, two rapid automatized naming measures, and age norms to 25:11. The WIAT-4's case: a 2020 standardization and native pairing with the Wechsler cognitive scales. Teams often keep both on the shelf and pick per referral; what reviewers actually flag is a report that mixes the two batteries' scores without labeling metrics and norm vintages, or that switches batteries between baseline and re-evaluation without saying so.

No instrument is. Federal law sets the frame: the IDEA regulations require a comprehensive evaluation using a variety of assessment tools, bar any single measure or score from being the sole criterion, state that states must not require a severe ability-achievement discrepancy for SLD identification, must permit a response-to-intervention process, and may permit alternative research-based procedures such as pattern-of-strengths-and-weaknesses analysis. Which models are actually permitted, and what a dyslexia evaluation must include, is state policy that varies. The KTEA-3 earns its slot as evidence, not as a requirement: it supplies the normed achievement data those models consume, and in reading referrals its phonological, decoding, and naming-fluency measures map onto what state dyslexia rules commonly ask evaluators to examine. The write-up convention that survives review names the state's model and routes the eligibility decision to the team documented in the psychoeducational report.

Report the set that answers the referral question, and say which one you used. Grade-based scores compare the student with others at the same point in schooling, which fits school eligibility and curriculum questions, and they are the reference frame the error-analysis norms are built on (the grade-norm sample). Age-based scores compare the student with same-age peers regardless of grade, which fits developmental and clinical questions and students markedly old or young for grade, and they are the frame that lines up with age-normed cognitive batteries in ability-achievement comparisons. The two sets can differ visibly for the same performance, especially for retained or accelerated students, so the methods sentence names the choice and the reason, and a re-evaluation uses the same frame as baseline or explains the switch.

Error analysis is the KTEA-3's signature: for 10 of the 19 subtests, the system classifies each error by skill category, using examiner-judged classification on some subtests and automatic item-level classification on others, then compares the student's error count in each category with the average number of errors made by grade peers in the norm sample, labeling the category a strength, average, or weakness. That normative comparison is what separates it from informal error inspection, and the standardization research behind the 2017 special issue of the Journal of Psychoeducational Assessment found the error categories form interpretable structures. Include it whenever the referral is instructional: translated into plain language about what the student can and cannot yet do, it is the part of the report teachers and interventionists actually use. Two cautions: report the normative classification rather than raw counts, and write the interpretation yourself rather than pasting the scoring platform's text, which Pearson's license limits to minimum-text excerpting.

No publisher or federal rule sets a fixed waiting period; the parallel forms exist precisely so retesting is a design decision rather than a countdown. The convention: administer the alternate form at re-evaluation (Form B where Form A was baseline, or the reverse), name in the report which form was given each time, keep the norm frame consistent with baseline, and describe change with Growth Scale Values, the Rasch-based scores Pearson provides for comparing performance across administrations, rather than eyeballing standard-score differences that mix real growth with measurement error and practice effects. Short intervals still deserve a stated rationale, since even alternate forms share format familiarity, and the Brief Form offers a lighter progress-monitoring option between full administrations, with Pearson stating Brief standard scores may be used interchangeably with Comprehensive Form scores.

Not from evaluators, and not from this page. Test items, stimulus books, record forms, and scoring materials are copyrighted, sold under Pearson's qualification requirements, and protected by the test-security obligations in the APA Ethics Code, which directs psychologists to make reasonable efforts to maintain the integrity and security of test materials. Item exposure also invalidates the norms: a student coached on the items no longer has a meaningful score. What can be shared is everything this page contains, the structure of a results section, the score metrics, and a fully fictional sample write-up, and families are entitled to a clear explanation of results and, per the education-records rules, access to the reports written about their child. Evaluators handling records requests for raw test data have their own decision to document, balancing records access against test security and publisher agreements.

Yes. BastionGPT is trained and clinically tested on psychological and psychoeducational evaluation reports, the parent documents KTEA-3 sections live inside. Paste a score summary (form, norm set, composites, subtests, flagged error categories) and it drafts a domain-organized results narrative with the methods statement, percentile framing, and error-analysis-to-instruction translation in place for your review; it can also check a finished section for metric drift, score-narrative mismatches, a missing form or norm statement, grade equivalents doing placement work, or a Dyslexia Index phrased as a diagnosis, and produce a plain-language summary for parents and teachers. BastionGPT is HIPAA-compliant with a signed BAA on every plan, your data is never used to train models, and drafting from a score summary you paste means no item or record-form content ever needs to leave your records.

Primary sources

The instrument facts and compliance claims on this page trace to these sources, last verified July 2026:

  1. NCS Pearson, KTEA-3 product page: 2014 publication, the most-current-version statement, ages 4:0 through 25:11, grades pre-K through 12, the 100/15 standard-score set with Growth Scale Values, qualification level B, 15 to 85 minute ASB completion time, Q-interactive and Q-global options, and the parallel-forms, RAN, and error-analysis positioning quoted on this page.
  2. Pearson, KTEA-3 error analysis: error analysis on 10 of the 19 subtests, the within-item and item-level classification methods, the comparison to the average number of errors made by the reference group, and the grade-norm error-norming (n = 2,600 for item-level subtests, a stratified subset for within-item subtests).
  3. Pearson, KTEA-3 sample score report (PDF): the composite architecture and ASB composition, subtest standard scores, 90% confidence intervals, printed descriptive categories, and the report-excerpting license quoted on this page.
  4. Pearson, Telepractice and the KTEA-3 (PDF): in-person normative data, telepractice as a deviation from standardized administration, and the requirement to state and describe telepractice in the written report.
  5. Pearson, KTEA-3 Brief Form and KTEA-3 Dyslexia Index product pages: the 2015 Brief Form's three composites and interchangeable-scores statement; the Dyslexia Index's 2018 and 2021 releases, ages 5 to 25, 12 to 20 minute completion time, and the not-sufficient-to-diagnose statement quoted on this page.
  6. IDEA regulations, 34 CFR 300.307 (govinfo PDF): the severe-discrepancy prohibition, the RTI permission requirement, and the alternative research-based procedures clause; the pattern-of-strengths-and-weaknesses wording sits in 300.309.
  7. O'Brien, Pan, Courville, Bray, Breaux, Avitia & Choi, Exploratory factor analysis of reading, spelling, and math errors, Journal of Psychoeducational Assessment 35(1-2), 7-23 (2017), and Breaux, Bray, Root & Kaufman, Introduction to Special Issue and to KTEA-3 Error Analysis, JPA 35(1-2), 4-6 (2017): the empirical basis for treating error categories as interpretable.
  8. Lockwood, Sealander, Gross & Lanterman, Teacher trainees' administration and scoring errors on the KTEA-3, Journal of Psychoeducational Assessment 38(5), 551-563 (2020): administration and clerical errors among 42 teacher trainees, Reading Comprehension the most error-prone subtest.
  9. Parkin & Frisby, Exploratory factor analysis of the KTEA-3, Contemporary School Psychology 23(2), 138-151 (2019): the factorial-complexity finding quoted on this page.
  10. Breaux & Lichtenberger, Essentials of KTEA-3 and WIAT-III Assessment (Wiley, 2016): the interpretive tradition behind domain-organized reporting and the per-grade administration-time guidance.
  11. Buros Center, Tests reviewed in the Twentieth Mental Measurements Yearbook (2017): independent review listing for the KTEA-3.
  12. SASC, KTEA-3 guidance (June 2016, PDF) and the June 2025 suitable-tests list (PDF): UK higher-education suitability, US norms, and the age-25 norming ceiling quoted on this page.
  13. APA Services, Telehealth reporting updates for 2026, and CMS, Medicare telehealth services list: the psychological-testing CPT family's Appendix P and T listings and its place on the CY 2026 Medicare telehealth list.

Educational content, not legal or billing advice. Sample notes are fictional. Follow your organization's policies and your board, payer, and jurisdiction requirements.