KPEERI • FOUNDATIONAL CONCEPTS

Distinguishing Assessment Types — 1. Distinguish among assessment types Screening Diagnostic Outcome (high stakes testing) Progress-monitoring (formative assessment)

Understanding the four pillars of educational assessment and how each serves a unique decision-making purpose.

Historical Context & Motivation

The practice of formally assessing students has deep roots in educational history, but the systematic classification of assessment into distinct types emerged only in the latter half of the twentieth century. Early assessment in American education was monolithic: a single test at the end of a term determined a student's grade, with little attention to the different purposes that evaluation might serve. As researchers in educational psychology and psychometrics began investigating how assessment data could inform instruction—rather than merely certify achievement—a richer taxonomy began to take shape. The recognition that different educational decisions require different kinds of evidence became the conceptual foundation for distinguishing among screening, diagnostic, outcome, and progress-monitoring assessments.

1967
Scriven's Formative–Summative Distinction
Michael Scriven introduced the terms formative and summative evaluation, establishing that assessment can serve fundamentally different purposes—improving a program while it is active versus judging its merit after completion.
1985
Curriculum-Based Measurement (CBM) Formalized
Stanley Deno and colleagues at the University of Minnesota developed curriculum-based measurement, providing the technical basis for progress-monitoring assessments that track growth over short intervals.
2001
No Child Left Behind Act (NCLB)
Federal legislation mandated annual high-stakes outcome testing in reading and mathematics, dramatically raising the profile—and the consequences—of summative assessment in American schools.
2004
IDEA Reauthorization and RTI
The reauthorization of the Individuals with Disabilities Education Act encouraged Response to Intervention (RTI) models, which depend explicitly on universal screening and progress monitoring to identify students with learning difficulties.
2015
Every Student Succeeds Act (ESSA)
ESSA preserved the requirement for statewide assessments while granting states greater flexibility in designing accountability systems, reinforcing the importance of understanding how each assessment type contributes to the larger evaluation framework.

Against this historical backdrop, a central question emerges: How do educators determine which type of assessment to deploy at each stage of the instructional cycle, and what kinds of inferences does each type legitimately support? Answering this question requires a precise understanding of the purpose, timing, scope, and stakes associated with each of the four major assessment types.

Core Principles & Definitions

Before examining each assessment type in detail, it is essential to ground the discussion in a set of organizing principles. Assessment scholars distinguish among types primarily along four dimensions: purpose (why is this assessment administered?), timing (when in the instructional cycle does it occur?), scope (how broad or narrow is the content sampled?), and stakes (what consequences are attached to the results?). These dimensions interact to define the role each assessment plays within an educational system.

1

Screening

A brief, universal assessment administered to all students to quickly identify those who may be at risk for academic difficulties or who may need enrichment. Screening is efficient and broad rather than deep.
2

Diagnostic

A targeted, in-depth assessment given to individual students already identified as at risk, designed to pinpoint specific skill deficits or cognitive processing weaknesses that explain why a student is struggling.
3

Outcome (High-Stakes Testing)

A summative, standardized assessment administered at the end of an instructional period to determine whether students have met established proficiency standards. Results carry significant consequences for students, teachers, or institutions.
4

Progress Monitoring (Formative Assessment)

Frequent, brief assessments administered at regular intervals throughout instruction to track student growth over time and to inform instructional adjustments. Data drive real-time decision-making.
KEY TAKEAWAY
Think of assessment types as instruments in a physician's office. Screening is like taking a patient's temperature—quick and universal, just enough to flag who needs further attention. Diagnostic assessment is the MRI or blood panel—detailed, targeted, and administered only to those who have been flagged. Outcome assessment is the final pathology report that certifies whether the patient is healthy. And progress monitoring is the series of vital-signs checks during treatment, ensuring the prescribed intervention is actually working.

Visual Explanation — The Assessment Cycle

The assessment cycle begins with universal screening to flag at-risk students, who then receive diagnostic assessment to identify specific needs. Interventions are monitored via progress monitoring, which feeds back into diagnostic refinement. At the end of the instructional period, outcome assessments provide summative accountability data.

The diagram above illustrates the logical sequence in which the four assessment types are typically deployed within a multi-tiered system of supports (MTSS) or Response to Intervention (RTI) framework. Notice the cyclical relationship between diagnostic assessment and progress monitoring: as formative data accumulate, educators revisit their diagnostic hypotheses and adjust interventions accordingly. Outcome assessment, positioned at the terminus of the cycle, serves a fundamentally different function—it evaluates the product of instruction rather than guiding the process.

How Each Assessment Type Works

Screening Assessments

Screening assessments are designed to be administered universally—typically to every student in a grade level or school—at predetermined points during the year (commonly fall, winter, and spring benchmarks). Their defining characteristic is efficiency: they must be short enough to administer to large numbers of students within a practical time frame yet sensitive enough to identify those who are likely to struggle without additional support. Technically, screening instruments are evaluated by their sensitivity (the proportion of truly at-risk students correctly identified) and specificity (the proportion of not-at-risk students correctly identified as such). A high-quality screener aims for sensitivity ≥ 0.90, accepting that some false positives will occur because the cost of missing a struggling student is far greater than the cost of over-identifying.

Diagnostic Assessments

When a screening assessment flags a student, diagnostic assessment follows to determine the nature and source of the difficulty. While screening asks "Is there a problem?" diagnostic assessment asks "What exactly is the problem, and why?" These instruments are typically longer, individually administered, and designed to assess specific sub-skills or cognitive processes. In reading, for example, a diagnostic assessment might separately evaluate phonemic awareness, decoding fluency, vocabulary knowledge, and comprehension strategies to construct a detailed learner profile. The output of diagnostic assessment is an instructional hypothesis—a data-informed explanation of why the student is struggling that directly informs intervention design.

Outcome (High-Stakes) Assessments

Outcome assessments—often synonymous with high-stakes testing—serve an accountability function. They are administered at the end of an instructional period and measure whether students have achieved expected proficiency standards. The results carry significant consequences: student promotion or retention, teacher evaluations, school ratings, and funding allocations may all be tied to outcome data. Because of these high stakes, these assessments must demonstrate strong evidence of validity (the degree to which the test measures what it claims to measure) and reliability (the consistency of scores across administrations). State assessments under ESSA, the SAT, ACT, and Advanced Placement exams are all examples of high-stakes outcome assessments.

Progress Monitoring (Formative Assessment)

Progress monitoring is the most instruction-sensitive of the four types. Administered frequently—weekly, biweekly, or monthly—these brief probes generate a time-series of data points that reveal the student's rate of improvement relative to an aimline (a projected growth trajectory). If a student's slope of improvement falls below the aimline, educators know to intensify or modify the intervention. The key technical requirement is that progress-monitoring measures have alternate forms of equivalent difficulty so that repeated administration does not inflate scores through practice effects. Curriculum-based measurement (CBM) procedures, such as oral reading fluency probes, are paradigmatic examples.

Detailed Classification of Assessment Types

This comparison matrix highlights how each assessment type differs across six critical dimensions. Notice that screening and outcome assessments are both administered to all students, but they differ sharply in timing, stakes, and the instructional decisions they support.

The classification matrix above reveals an important structural pattern: screening and progress monitoring are both low-stakes and relatively brief, but they serve different populations and occur at different frequencies. Screening is administered universally at a few benchmark points, whereas progress monitoring targets only at-risk students but occurs much more frequently. Similarly, diagnostic and outcome assessments both involve comprehensive measurement, but diagnostic is narrow-and-deep while outcome is broad-and-comprehensive. These complementary relationships mean that no single assessment type can substitute for another; each fills a unique niche in the decision-making ecosystem.

⚠️ Common Exam Trap
Test items frequently present a scenario and ask you to identify the assessment type. The most common confusion is between screening and diagnostic. Remember: screening casts a wide net over all students; diagnostic follows up on those who have already been identified. If the scenario says 'all second graders were assessed in September,' that's screening. If it says 'a student flagged as at-risk received a detailed reading battery,' that's diagnostic.

Worked Example — Identifying Assessment Types in a Scenario

Consider the following scenario, which mirrors the type of item you may encounter on the KPEERI examination. A third-grade team has adopted a new reading intervention program. Over the course of the year, they administer several different assessments. Your task is to correctly identify the type of each assessment described.

Classifying Assessments in a School-Based RTI Scenario
1
Step 1 — Identify the Universal BenchmarkIn September, every third grader completes a 5-minute oral reading fluency (ORF) probe. The team uses a cut score of the 25th percentile to flag students who may need additional support. Because this assessment is administered to all students, is brief, and is used to sort students into at-risk and not-at-risk categories, this is a screening assessment.
Assessment Type: Screening
2
Step 2 — Identify the Targeted Deep DiveTwelve students fall below the 25th percentile. Each is then given a comprehensive reading diagnostic battery (e.g., the CTOPP-2 for phonological processing and the QRI for comprehension) to determine the specific subskills that are deficient. This assessment is administered only to flagged students, is detailed, and is used to inform the design of a specific intervention plan.
Assessment Type: Diagnostic
3
Step 3 — Identify the Repeated ProbesThe twelve at-risk students begin receiving a targeted phonics intervention. Every two weeks, each student completes an alternate-form ORF probe. The teacher graphs the data to determine whether each student's rate of growth is on track relative to the aimline. Because this assessment is administered repeatedly at regular intervals to inform instructional adjustments, this is progress monitoring.
Assessment Type: Progress Monitoring (Formative)
4
Step 4 — Identify the Summative EvaluationIn April, all third graders take the state's annual English Language Arts assessment. Results are reported as proficiency levels (Below Basic, Basic, Proficient, Advanced) and are used to evaluate school performance for accountability purposes. This test is summative, standardized, and consequential.
Assessment Type: Outcome (High-Stakes)
5
Step 5 — Verify by Cross-Referencing DimensionsTo confirm each classification, check the four defining dimensions: purpose, timing, scope, and stakes. The September ORF (screening) was universal, brief, and low-stakes. The diagnostic battery was individual, in-depth, and moderate-stakes. The biweekly probes (progress monitoring) were frequent, narrow, and low-stakes. The state exam (outcome) was universal, comprehensive, and high-stakes. Each assessment maps cleanly to exactly one type when all four dimensions are considered simultaneously.
All four classifications confirmed.

Strengths, Limitations, and Common Misapplications

Comparative strengths and limitations of the four assessment types
Assessment TypeStrengthsLimitations
ScreeningEfficient, cost-effective, universal; enables early identification before students fall significantly behind; provides a data-driven entry point for tiered support systems.Produces false positives (students identified as at-risk who are not) and false negatives (at-risk students missed); provides no information about why a student is struggling; a single snapshot may not reflect typical performance.
DiagnosticProvides fine-grained information about specific skill deficits; directly informs intervention design; can illuminate underlying cognitive or processing difficulties.Time-intensive and resource-demanding; typically requires trained specialists to administer and interpret; not feasible for universal administration; may over-identify within certain populations if norms are not representative.
Outcome (High-Stakes)Provides standardized, comparable data across schools, districts, and states; supports systemic accountability; can identify large-scale achievement trends and equity gaps.Occurs too late to inform current instruction; can narrow the curriculum ('teaching to the test'); results may be influenced by test anxiety, linguistic bias, or cultural bias; single-occasion measurement introduces error.
Progress MonitoringProvides real-time feedback on student growth; enables data-based instructional adjustments; allows educators to evaluate intervention effectiveness before committing to long-term placements.Requires fidelity of implementation—inconsistent administration undermines data quality; alternate-form reliability can be difficult to establish; demands teacher time and training for data interpretation.
KEY TAKEAWAY
Each assessment type occupies a unique quadrant in the decision-making landscape; misapplying one type in place of another leads to poor decisions. Using a high-stakes outcome test to diagnose individual deficits is like using satellite imagery to find a crack in your foundation—the resolution is wrong for the task. Conversely, relying solely on progress-monitoring data to evaluate an entire school system would be like judging a hospital's effectiveness based only on one patient's vital signs. The right tool for the right question is the governing principle.

Connections to Advanced Theory & Multi-Tiered Systems

The four assessment types do not operate in isolation; they are woven into broader systemic frameworks. In Multi-Tiered Systems of Support (MTSS) and Response to Intervention (RTI) models, assessment types map directly onto the three tiers. Tier 1 relies on universal screening to ensure all students receive adequate core instruction. When screening identifies at-risk learners, Tier 2 interventions are implemented and monitored through progress-monitoring probes. If a student does not respond adequately, diagnostic assessment informs the more intensive, individualized Tier 3 interventions. Outcome assessments operate at the systems level, evaluating whether the entire MTSS framework is producing acceptable results for the school or district.

From foundational assessment concepts to advanced psychometric applications
Foundational ConceptAdvanced Application
Screening identifies at-risk students using cut scores.Classification accuracy research (Silberglitt & Hintze, 2005) uses ROC curves to optimize cut scores, balancing sensitivity and specificity to minimize both false positives and false negatives.
Progress monitoring tracks growth over time.Growth modeling techniques (hierarchical linear modeling, piecewise regression) formalize the analysis of slope data, enabling more precise comparisons between a student's growth rate and normative expectations.
Diagnostic assessment maps skill profiles.Cognitive diagnostic models (CDMs) in psychometrics represent student knowledge as vectors of latent attributes, enabling probabilistic estimation of which specific subskills a student has or has not mastered.
Outcome assessment evaluates proficiency.Item response theory (IRT) models underpin modern high-stakes tests, placing student ability and item difficulty on a common metric to enable adaptive testing and equating across test forms.

As you advance in your study of educational assessment, you will encounter debates about the boundaries between these categories. For example, some scholars argue that computer-adaptive testing blurs the line between screening and diagnostic assessment because a single adaptive platform can both identify at-risk students and provide detailed sub-skill profiles. Similarly, the growing emphasis on interim assessments—administered periodically throughout the year with moderate breadth—occupies a hybrid space between progress monitoring and outcome assessment. Understanding the classical four-type framework equips you to critically evaluate these emerging hybrid models.

Practice Problems

PROBLEM 1CONCEPTUAL
A school psychologist explains that a particular assessment is designed to "cast a wide net" and identify students who may need additional support. The assessment is given to every student in the school three times a year. Which assessment type is the psychologist describing, and what is the single most important psychometric property for this type of instrument?
PROBLEM 2BASIC CALCULATION
A universal screening instrument is administered to 500 students. Follow-up diagnostic testing reveals that 60 students are truly at risk. The screener correctly identified 54 of these 60 students. It also flagged 40 students who turned out not to be at risk upon further evaluation. Calculate the screener's sensitivity and specificity.
PROBLEM 3INTERMEDIATE
A reading specialist administers weekly oral reading fluency (ORF) probes to a student receiving a Tier 2 intervention. After eight weeks, the student's data show a slope of 1.2 words correct per minute (WCPM) gained per week. The aimline—based on normative expectations—projects a needed slope of 2.0 WCPM per week. Based on this information: (a) What type of assessment is being used? (b) What decision should the reading specialist make, and why?
PROBLEM 4APPLIED
A district superintendent receives the following report: 'Sixty-eight percent of fourth graders scored Proficient or Advanced on the state ELA assessment, up from 62% the previous year. However, only 45% of English learners reached proficiency.' The superintendent wants to (1) identify which specific ELA subskills English learners are struggling with, and (2) implement a system for tracking the effectiveness of a new vocabulary intervention for these students. For each goal, identify the appropriate assessment type, justify your choice, and name one concrete instrument or approach that could be used.
PROBLEM 5CRITICAL THINKING
A school district adopts a new computer-adaptive platform that is administered to all students three times per year. The platform provides (a) an overall risk classification (at-risk / not-at-risk), (b) detailed subskill profiles in reading and math, and (c) a growth percentile comparing each student's growth to national norms. A district official claims that this single platform eliminates the need for separate screening, diagnostic, and progress-monitoring assessments. Critically evaluate this claim. Under what conditions might the claim be justified, and under what conditions would it be problematic?

Summary

Educational assessment is not a monolithic activity; it encompasses four distinct types, each tailored to a specific decision-making purpose. Screening assessments are brief, universal instruments administered at benchmark intervals to identify students who may be at risk—they prioritize sensitivity to minimize missed cases. Diagnostic assessments follow screening, providing targeted, in-depth analysis of specific skill deficits or processing weaknesses to generate an instructional hypothesis that directly informs intervention design. Outcome (high-stakes) assessments are summative, standardized evaluations administered at the end of an instructional period to determine whether students have met proficiency standards; their results carry significant consequences for students, educators, and institutions. Progress monitoring (formative assessment) involves frequent, repeated probes that generate a time-series of growth data, enabling educators to evaluate intervention effectiveness and make real-time instructional adjustments based on the student's slope of improvement relative to an aimline.

The four types are distinguished along dimensions of purpose, timing, scope, and stakes. No single assessment type can substitute for another; each fills a unique niche within the multi-tiered decision-making ecosystem of modern education. Within MTSS and RTI frameworks, these four types work in concert—screening initiates the cycle, diagnostic assessment targets the problem, progress monitoring evaluates the solution, and outcome assessment judges the system. Mastery of this taxonomy is foundational to the KPEERI examination and to effective educational practice.

Varsity Tutors • KPEERI • Distinguishing Assessment Types