2.1 Assessment Typology: Screening, Diagnostic, Progress Monitoring, & Summative
Key Takeaways
- Universal screening is administered three times per year (BOY, MOY, EOY) to all students as a brief, standardized measure to detect early reading risk and dyslexia indicators.
- Diagnostic assessments are administered selectively following screening to pinpoint specific, granular deficits in phonological processing, phonics, decoding, or language comprehension.
- Progress monitoring utilizes frequent, brief Curriculum-Based Measurements (CBM) to evaluate intervention efficacy and track growth slopes against established aimlines.
- Summative assessments measure cumulative mastery of grade-level TEKS at the end of instructional units, semesters, or school years, informing program evaluation rather than daily instruction.
- Valid score interpretation requires distinguishing criterion-referenced mastery standards from norm-referenced comparative percentiles, avoiding misleading grade-equivalent interpretations.
2.1 Assessment Typology: Screening, Diagnostic, Progress Monitoring, & Summative
Assessment is the engine that drives evidence-based reading instruction. In a scientific reading model aligned with the Texas Essential Knowledge and Skills (TEKS) and the Multi-Tiered System of Supports (MTSS) framework, assessment is an ongoing, systematic process. It identifies student vulnerabilities early, evaluates instructional efficacy, and guides targeted interventions before reading difficulties become entrenched.
The Assessment Architecture in Standards-Based Reading
Standards-based reading instruction requires educators to distinguish between four distinct assessment categories based on their operational purpose, administration timeline, and instructional utility. Within a preventative MTSS framework, assessment guides decision-making across three tiers of instruction: Tier 1 universal core instruction for all students, Tier 2 targeted small-group supplemental intervention for students at risk, and Tier 3 intensive, individualized intervention for students with persistent, severe reading deficits.
Universal Screening: Early Identification of Reading Risk
Universal screening serves as an early warning system. Administered to all students three times per year—Beginning of Year (BOY), Middle of Year (MOY), and End of Year (EOY)—screeners are brief, standardized, and highly reliable measures. They identify students at risk of reading failure and flag characteristics of dyslexia, as mandated by the Texas Dyslexia Handbook.
Texas-approved screening tools—such as the Texas Kindergarten Entry Assessment (TX-KEA), TPRI, Tejas LEE (for Spanish-language assessment), and early literacy batteries like DIBELS—measure foundational predictors: phonemic awareness, letter naming, letter-sound correspondence, and decoding automaticity. Because screeners are brief, timed indicators, they identify which students need support, but require follow-up diagnostic testing to determine specific instructional needs.
Diagnostic Assessments: Pinpointing Specific Deficits
When universal screening indicates reading risk, teachers administer diagnostic assessments. While screeners detect that a problem exists, diagnostic assessments pinpoint the exact underlying phonological, orthographic, or linguistic breakdown.
Diagnostic assessments are untimed, criterion-referenced, and comprehensive. They evaluate specific subskills along a developmental continuum:
- Phonemic Awareness Diagnostics: Assessing blending, segmenting, and phoneme manipulation (deletion, substitution) with complex consonant clusters.
- Diagnostic Phonics Surveys: Systematically testing grapheme-phoneme correspondences, including short vowels, consonant digraphs, blends, VCe syllables, vowel teams, and r-controlled vowels.
- Structural Analysis Surveys: Evaluating syllable division rules (VC/CV, V/CV) and morphemic analysis (prefixes, base words, derivational suffixes).
Formative Assessment & Progress Monitoring: Evaluating Growth
Formative assessment is woven into daily instruction through informal checks for understanding—such as white-board responses, exit tickets, and oral blending checks—allowing teachers to adjust daily pacing immediately.
Progress monitoring is the formalized, quantitative component of formative assessment within MTSS. Using Curriculum-Based Measurement (CBM) probes (such as Nonsense Word Fluency or Oral Reading Fluency), educators collect regular performance data:
- Tier 2 Interventions: Monitored bi-weekly.
- Tier 3 Interventions: Monitored weekly.
Teachers plot data points on a progress monitoring chart against an aimline connecting baseline performance to an end-of-year goal. Standard data-decision rules dictate that if three to four consecutive data points fall below the aimline, the teacher must modify the intervention by increasing session frequency, decreasing group size, or intensifying explicit teacher modeling.
Summative Assessments: Measuring Grade-Level TEKS Mastery
Summative assessments measure cumulative learning and standards mastery at the end of an instructional cycle, grading period, or school year. Examples include chapter post-tests, district benchmark exams, and the State of Texas Assessments of Academic Readiness (STAAR). Summative data informs campus accountability and curriculum evaluation rather than weekly lesson adjustments.
Psychometric Foundations: Validity, Reliability, & Cultural Fairness
Accurate instructional decisions depend on sound psychometric properties:
- Validity: The degree to which an assessment measures what it claims to measure. Content validity ensures test items align with TEKS standards; construct validity ensures the test measures the target reading construct without contamination; criterion-related validity evaluates how well scores predict external reading benchmarks.
- Reliability: The consistency of assessment scores across repeated testing (test-retest), across different examiners (inter-rater reliability), or across parallel forms (alternate-form reliability).
- Equity and Cultural Fairness: Assessments must be culturally responsive and linguistically unbiased. For Emergent Bilinguals, teachers must distinguish normal second-language acquisition patterns from true reading disorders, utilizing dual-language assessment tools when appropriate.
Score Interpretation: Criterion-Referenced vs. Norm-Referenced Tests
| Assessment Dimension | Criterion-Referenced Tests | Norm-Referenced Tests |
|---|---|---|
| Primary Goal | Measure mastery of specific skills or state standards (TEKS). | Compare student performance against a national normative sample. |
| Score Formats | Percent correct, mastery cut scores (e.g., 80% on phonics screener). | Percentile ranks, stanines (1–9), normal curve equivalents. |
| Best Use | Planning targeted phonics and comprehension instruction. | Determining program eligibility and broad group comparisons. |
The Danger of Grade-Equivalent (GE) Scores
A Grade-Equivalent score of 4.3 earned by a second grader does not indicate readiness for fourth-grade reading materials. It simply means the student scored as well on that second-grade test as a typical fourth grader in the third month of school would score on the same test. Teachers should use criterion-referenced skill profiles and percentile ranks instead of grade equivalents.
Comparative Analysis of Reading Assessments
| Category | Frequency | Target Population | Core Question Answered | Primary Instructional Decision |
|---|---|---|---|---|
| Screening | 3x/year (BOY, MOY, EOY) | 100% of students | Who is at risk for reading failure? | Assign students to Tier 1 core vs Tier 2 intervention. |
| Diagnostic | Post-screening | Flagged students | What specific subskills are missing? | Plan targeted, explicit skill-based lessons. |
| Progress Monitoring | Weekly or bi-weekly | Tier 2 & Tier 3 students | Is the student responding to intervention? | Adjust intervention intensity or modify strategies. |
| Summative | End of unit or year | All students | Did students master grade-level TEKS? | Assign grades, evaluate curriculum effectiveness. |
Real Classroom Scenario: Data-Driven Decision Making in Grade 2
At the beginning of the school year, Ms. Hernandez administers a universal early literacy screener to her second-grade class. Two students, Mateo and Sofia, score below the 20th percentile on composite decoding. Ms. Hernandez administers follow-up diagnostic phonics surveys to both students.
The diagnostic data reveals distinct needs: Mateo struggles with vowel-consonant-e (VCe) patterns and vowel teams, while Sofia has mastered single-syllable phonics but cannot divide multisyllabic words using syllable division principles. Ms. Hernandez places Mateo into a Tier 2 group targeting long-vowel orthographic patterns and Sofia into a group targeting syllable division. She monitors their progress bi-weekly using CBM oral reading fluency probes. After six weeks, Mateo's scores climb along his aimline, confirming intervention success. Sofia's data points fall below the aimline across three consecutive probes, prompting Ms. Hernandez to reduce Sofia's group size from five to three students and introduce explicit syllable-chunking manipulatives.
What is the primary operational purpose and recommended administration frequency of universal screening assessments within a tiered reading framework?
A first-grade universal screening assessment indicates that a student is performing well below the benchmark cut score on composite early literacy measures. What is the teacher's most appropriate next assessment step before planning instruction?
A second-grade reading specialist has been providing Tier 2 intervention to a small group for six weeks, monitoring progress bi-weekly using Curriculum-Based Measurement (CBM) oral reading fluency probes. When analyzing the graphed data, the specialist observes that a student's last four consecutive data points fall noticeably below the established aimline. According to data-based decision rules, what action should the specialist take?
A teacher evaluates the standardized reading achievement test results of a third-grade student. The report indicates a Grade-Equivalent (GE) score of 5.4 in reading comprehension. How should the teacher correctly interpret this score for the student's parents?