Historical Context & Motivation
The practice of connecting assessment data to specific literacy difficulties has evolved significantly over the past century. Early reading instruction relied almost exclusively on teacher observation and subjective judgment, offering little in the way of systematic diagnosis. As norm-referenced tests emerged in the mid-twentieth century, educators gained a standardized lens through which to compare individual student performance to age- or grade-level peers. However, these tests alone could not reveal why a student struggled—only that a gap existed. The subsequent rise of informal assessment methods—running records, miscue analyses, writing samples, and curriculum-based measures—provided the granular, qualitative detail that standardized instruments lacked. Together, these two streams of data now form the diagnostic backbone of modern literacy intervention frameworks, including Response to Intervention (RTI) and Multi-Tiered Systems of Support (MTSS).
The central question this lesson addresses is both practical and conceptual: How do educators synthesize quantitative norm-referenced scores with qualitative informal assessment data to pinpoint, classify, and address specific literacy difficulties? Answering this question is essential not only for test preparation but for any professional role that involves diagnosing or remediating reading and writing problems in learners.
Core Principles & Definitions
Before one can relate assessment results to literacy difficulties, it is necessary to understand the fundamental nature of each assessment type and the kinds of evidence they generate. Norm-referenced assessments compare a student's performance to a norming sample—a large, representative group whose scores establish percentile ranks, stanines, grade equivalents, and standard scores. In contrast, informal assessments are criterion-referenced or descriptive in nature, measuring what a student can or cannot do against a defined set of skills or benchmarks rather than against other students. Both data types are indispensable: norm-referenced data tells you where a student stands relative to peers, while informal data tells you what specific skills need attention.
Norm-Referenced Scores Indicate Relative Standing
Informal Assessments Reveal Skill-Level Detail
Triangulation Strengthens Diagnostic Accuracy
Pattern Analysis Links Data to Difficulties
Assessment Drives Instruction, Not Just Classification
Visual Explanation — The Diagnostic Synthesis Model
The following diagram illustrates how norm-referenced and informal assessment data converge in a diagnostic synthesis model. Notice that the process begins with universal screening (a norm-referenced step), moves into targeted informal assessment for students who fall below benchmarks, and culminates in an integrated profile that maps directly onto specific literacy difficulties and corresponding interventions.
The flowchart emphasizes a critical principle: data from both formal and informal sources must be integrated rather than used in isolation. A norm-referenced screening score of the 15th percentile in reading comprehension signals concern, but it does not specify whether the difficulty lies in decoding, vocabulary knowledge, background knowledge, inferencing, or text structure awareness. Informal assessments then provide the diagnostic specificity needed to create an actionable intervention plan. The convergence of both data streams into an integrated profile is what distinguishes competent diagnostic practice from surface-level score reporting.
How It Works — Linking Scores to Literacy Domains
To relate assessment results to literacy difficulties effectively, the diagnostician must understand the statistical framework underlying norm-referenced scores and how those scores map onto qualitative performance levels. Most norm-referenced literacy assessments report standard scores with a mean of 100 and a standard deviation (SD) of 15. Performance is then classified into descriptive bands that correspond to varying degrees of concern.
While these quantitative indices place students along a continuum, they do not by themselves reveal the component literacy skills driving the deficit. That is where informal assessment data enters the diagnostic equation. Consider a student whose Woodcock-Johnson IV Passage Comprehension subtest yields a standard score of 78 (8th percentile). This norm-referenced result signals a comprehension difficulty, but it does not explain the mechanism. An informal reading inventory might then reveal that the student reads text at an independent level with high accuracy but cannot retell key ideas—pointing toward a comprehension-monitoring or inference-making weakness rather than a decoding issue. Alternatively, a running record might show frequent substitutions of visually similar words, suggesting that poor decoding accuracy is undermining comprehension from the bottom up.
Mapping Assessments to the Five Literacy Domains
Effective diagnosis requires mapping both norm-referenced and informal assessment data onto the five pillars of literacy identified by the National Reading Panel: phonemic awareness, phonics (decoding), fluency, vocabulary, and comprehension. The following diagram and table detail which assessment tools typically address each domain and how their results relate to specific difficulties within that domain.
| Literacy Domain | Norm-Referenced Tools | Informal Tools | Typical Difficulty Indicator |
|---|---|---|---|
| Phonemic Awareness | CTOPP-2 (Elision, Blending Words) | Yopp-Singer, PAST, phoneme segmentation tasks | SS < 85; cannot isolate, segment, or blend individual phonemes |
| Phonics / Decoding | WJ-IV Word Attack, WIAT-4 Pseudoword Decoding | Phonics survey, nonsense word fluency, spelling inventories | SS < 85; high error rate on CVC/CVCe patterns; reliance on guessing |
| Fluency | GORT-5, TOWRE-2 | CBM-ORF probes, timed reading, NAEP fluency scale | ORF below 50th percentile; word-by-word reading; flat prosody |
| Vocabulary | PPVT-5, EVT-3, WJ-IV Oral Vocabulary | Cloze tasks, vocabulary matching, word sorts, context-use checks | SS < 85; limited expressive/receptive vocabulary; difficulty with Tier 2 words |
| Comprehension | WJ-IV Passage Comprehension, WIAT-4 Reading Comprehension | QRI-7, retelling rubrics, think-alouds, written responses | SS < 85; poor retelling, inability to infer, weak main-idea identification |
Worked Example — Diagnosing a Third-Grader's Literacy Difficulty
The following scenario demonstrates how a reading specialist integrates norm-referenced and informal data to identify a specific literacy difficulty and plan intervention.
Strengths & Limitations of Each Assessment Type
Neither norm-referenced nor informal assessments are sufficient in isolation. Understanding the strengths and limitations of each type is essential for determining when and how to use them in diagnostic decision-making. The following table organizes these considerations side by side.
| Dimension | Norm-Referenced Assessments | Informal Assessments |
|---|---|---|
| Strengths | Standardized administration ensures consistency; percentile ranks enable comparison to national norms; psychometric properties (reliability, validity) are well-documented; results are recognized for eligibility decisions (e.g., IEPs, 504 plans) | Flexible and responsive to individual students; reveal specific skill-level strengths/weaknesses; can be embedded in daily instruction; provide immediate, actionable data for intervention planning |
| Limitations | Do not specify root cause of difficulty; may lack cultural or linguistic sensitivity; snapshot in time—do not capture day-to-day variability; expensive and time-consuming to administer fully | Lack standardized norms; inter-rater reliability may vary; results may not be accepted for formal eligibility decisions; quality depends heavily on the examiner's expertise |
| Best Used For | Screening, quantifying severity, benchmarking against peers, eligibility determination, pre/post outcome measurement | Diagnosing specific skill deficits, planning instructional interventions, ongoing progress monitoring, confirming or clarifying norm-referenced findings |
| Example Scenario | A GORT-5 Fluency Score of 4 (SS = 70) tells you the student is at the 2nd percentile in oral reading fluency—far below peers—but does not explain whether the cause is poor decoding, limited sight-word automaticity, or an underlying language disorder. | A running record on the same student reveals 85% accuracy with most errors on vowel-team words, no self-corrections, and word-by-word phrasing—pointing to a specific phonics deficit rather than a broader language issue. |
Connection to Advanced Diagnostic Frameworks
The process of relating assessment results to literacy difficulties does not exist in a vacuum—it feeds into broader diagnostic and instructional frameworks that shape how educational professionals serve struggling readers. Two particularly influential frameworks deserve mention: the Simple View of Reading (SVR) and the Multi-Tiered Systems of Support (MTSS) model. The SVR posits that reading comprehension is the product of decoding and linguistic comprehension (R = D × LC), which means that a deficit in either component—or both—can produce a comprehension difficulty. Norm-referenced and informal assessments help diagnosticians determine whether the breakdown resides in decoding, language comprehension, or both, directly informing the type of intervention required. Within the MTSS framework, assessment data drives tier placement and movement: universal screening (Tier 1), targeted diagnostic assessment and intervention (Tier 2), and intensive, individualized evaluation and support (Tier 3).
| Concept in This Lesson | Advanced Framework Connection |
|---|---|
| Norm-referenced screening identifies students below benchmark | MTSS Tier 1 → Tier 2 decision point; SVR helps determine if screening deficit is in D, LC, or both |
| Informal assessments reveal specific skill deficits | SVR component analysis: decoding difficulties → phonics intervention; linguistic comprehension difficulties → vocabulary/language intervention; MTSS Tier 2 intervention design |
| Integrated literacy profile with convergent evidence | Comprehensive evaluation for special education eligibility (Tier 3); differential diagnosis of dyslexia vs. language-based learning disability vs. reading difficulty due to limited English proficiency |
| Progress monitoring with informal CBMs | Data-based individualization (DBI) within MTSS; rate of improvement compared to norm-referenced benchmarks determines whether to intensify, maintain, or fade intervention |
As you advance in your study of literacy assessment, you will encounter increasingly nuanced diagnostic models, including the Dual-Route Cascaded Model of word reading, Scarborough's Reading Rope, and Ehri's phases of word reading development. Each of these theoretical frameworks adds layers of specificity to the diagnostic process, but all depend fundamentally on the ability to relate norm-referenced and informal assessment results to the observable manifestations of literacy difficulty—the very competency this lesson targets.
Practice Problems
Summary
Relating assessment results to literacy difficulties requires the systematic integration of two complementary data streams. Norm-referenced assessments—instruments like the WJ-IV, GORT-5, CTOPP-2, and PPVT-5—provide standardized scores (percentile ranks, standard scores, stanines) that quantify the severity of a student's difficulty relative to peers and satisfy requirements for eligibility decisions. Informal assessments—running records, miscue analyses, phonics inventories, retelling rubrics, and CBM probes—reveal specific skill-level strengths and weaknesses that norm-referenced instruments cannot isolate, enabling precise intervention targeting.
The diagnostic process follows a logical sequence: screen with norm-referenced tools, diagnose with informal measures, triangulate for convergent evidence, and build an integrated literacy profile that maps directly onto the five pillars of literacy (phonemic awareness, phonics, fluency, vocabulary, comprehension). The identical norm-referenced score can reflect different root causes—a principle that underscores why informal data is indispensable. Competent diagnosticians use pattern analysis across multiple data sources to ensure that assessment drives instruction and that every struggling reader receives the specific, evidence-based support they need.