6.1 Types and Purposes of Assessment in ESE
Key Takeaways
- Screening is brief and universal, designed to flag which students may need closer evaluation — it identifies risk, not diagnosis
- Diagnostic assessment is in-depth and follows a positive screening or referral to pinpoint the specific nature and severity of a skill deficit
- Formative assessment (including curriculum-based measurement) happens during instruction to guide immediate next-step teaching decisions; summative assessment happens after instruction to judge overall mastery or outcomes
- Curriculum-based measurement (CBM) uses brief, standardized, repeatable probes tied directly to the curriculum to track growth over time and inform data-based decisions
- Observation and portfolio assessment capture authentic, ongoing evidence of skills — especially useful for students with significant disabilities, young children, and communication or behavior goals that a single test cannot measure
6.1 Types and Purposes of Assessment in ESE
Quick Answer: Assessment in exceptional student education is not one activity but a family of tools, each matched to a different decision. Screening flags risk quickly and broadly. Diagnostic assessment digs deep once risk is flagged. Formative assessment (including curriculum-based measurement) checks progress during instruction so teachers can adjust in real time. Summative assessment judges outcomes after instruction ends. Observation and portfolio assessment add authentic, ongoing evidence that formal tests cannot capture alone. The FTCE ESE (061) exam tests whether you can match a described scenario to the correct assessment purpose — not whether you can name a specific commercial test.
Why Purpose Comes First
Before selecting or interpreting any assessment, an ESE teacher must ask: what decision does this data need to support? A tool that is excellent for one purpose (say, quickly flagging which students in a class may be at risk for a reading difficulty) is a poor choice for a different purpose (say, determining whether a student meets IDEA eligibility criteria for Specific Learning Disability). Exam scenarios are built around this mismatch — they describe a teacher using the wrong tool for the stated purpose, and the correct answer identifies the mismatch or names the tool that actually fits.
Screening: Casting a Wide, Fast Net
Screening assessments are brief, low-cost, and administered to an entire class, grade, or population — not just students already suspected of a disability. The purpose of screening is narrow: identify which students may be at risk and therefore warrant a closer look. Screening is not diagnostic; a student who screens "at risk" does not automatically have a disability, and a student who screens "not at risk" is not automatically cleared of one. Universal reading and behavior screeners administered three times a year (fall, winter, spring) within a multi-tiered system of supports (MTSS) are a common example — they sort an entire grade level into risk tiers in a matter of minutes per student.
Because screening instruments are designed for breadth and speed, they sacrifice some precision. This tradeoff is intentional: a screener that took 45 minutes per student would defeat its own purpose of quickly surveying an entire population. When a question describes universal, brief, whole-class or whole-grade testing used to sort students by risk level, the answer is screening.
Diagnostic Assessment: Going Deep After a Flag
Diagnostic assessment is triggered by a positive screening result, a teacher referral, or a formal request for evaluation, and it is fundamentally different in depth and scope from screening. Where screening asks "is there a risk?", diagnostic assessment asks "exactly what is the nature, severity, and likely cause of this specific difficulty?" Diagnostic tools are individually administered, take considerably longer, and often break a broad skill area into component sub-skills — for example, a diagnostic reading battery might separately measure phonemic awareness, phonics decoding, sight-word recognition, oral reading fluency, and comprehension, rather than producing one overall reading score.
Diagnostic assessment data is what feeds a full and individual evaluation (FIE) for IDEA eligibility determination, because eligibility teams need specific, in-depth information about the nature of a suspected disability — not just a risk flag. A common exam trap presents a screening tool being used (incorrectly) to make an eligibility decision, or a lengthy diagnostic battery being proposed (inefficiently) for a quick universal check; the correct answer recognizes the purpose-tool mismatch.
Formative vs. Summative: When the Assessment Happens Relative to Instruction
The formative vs. summative distinction is about timing and function relative to instruction, not about a specific test format:
| Feature | Formative Assessment | Summative Assessment |
|---|---|---|
| When | During instruction, ongoing | After instruction, at an endpoint |
| Purpose | Guide immediate next-step teaching decisions | Judge overall achievement or outcome |
| Stakes | Low-stakes, frequent | Higher-stakes, infrequent |
| Example | Exit ticket, running record, CBM probe, quick check for understanding | Unit test, end-of-course exam, statewide assessment |
| Who acts on it first | The teacher, immediately, to adjust tomorrow's lesson | Often used for grading, promotion, or program evaluation |
For students with exceptionalities, formative assessment is especially critical because it is the engine of data-based individualization — the ongoing cycle of instruct, measure, analyze, and adjust that underlies effective IEP progress monitoring and MTSS decision-making (explored further in Section 7.3). A single end-of-year summative test tells you whether a year-long goal was ultimately met, but it arrives too late to change instruction for that year. Formative data, gathered weekly or more often, tells you whether this week's instruction is working while there is still time to change course.
Curriculum-Based Measurement (CBM): A Specific, Powerful Formative Tool
Curriculum-based measurement (CBM) deserves special attention because it is one of the most heavily researched and widely used formative tools in special education. CBM probes share several defining features:
- Brief — typically 1 to 5 minutes per probe
- Standardized — the same administration and scoring procedure every time, so results are comparable across occasions
- Repeatable — equivalent alternate forms allow frequent re-testing (weekly, biweekly) without a practice-effect confound
- Tied to the general curriculum — probes sample the actual skills being taught (e.g., oral reading fluency passages, digit-correct math computation probes, correct writing sequences)
- Sensitive to growth — designed to detect small amounts of change over short time periods, unlike many norm-referenced tests that are only sensitive to change over a year or more
Because CBM produces a number (e.g., words read correctly per minute) at frequent, regular intervals, teachers can graph the data and visually inspect the student's trend line against a goal line to judge whether an intervention is working — the technical foundation for the progress-monitoring content covered later in this study guide. CBM is formative by function even though it produces quantitative scores that might, at a glance, resemble a summative test.
Observation: Authentic Evidence in Natural Settings
Observation is the systematic watching and recording of a student's behavior, skills, or social interactions as they naturally occur — in the classroom, on the playground, during transitions, or across other authentic settings. Structured observation can be as simple as an anecdotal running record or as formal as an interval-recording behavior data sheet used to build a functional behavior assessment (FBA, covered in Chapter 12). Observation is indispensable for skills and behaviors that a paper-and-pencil or computer-based test simply cannot capture: social interaction quality, on-task behavior across the school day, communication attempts by a nonverbal student, or a young child's play skills.
For students with significant cognitive, communication, or behavioral disabilities, observation is often the primary, not supplementary, source of assessment data, because many formal, standardized instruments assume response formats (bubbling in an answer, writing a sentence, sitting for 45 minutes) that these students cannot access without significant modification that would invalidate the standardized score anyway.
Portfolio Assessment: Evidence Collected Over Time
Portfolio assessment is the purposeful collection of a student's work samples, data points, photos, or other artifacts gathered over time to document growth, skill mastery, or progress toward goals. Portfolios are especially valuable for:
- Documenting progress for students on alternate achievement standards, where a single standardized test may not adequately reflect meaningful growth
- Capturing skills across multiple contexts and time points rather than a single test session
- Providing concrete, shareable evidence for IEP team discussions with families who may find raw score reports difficult to interpret
- Supporting alternate assessment participation decisions and progress reporting for students with the most significant cognitive disabilities
Unlike a single test score, a well-constructed portfolio shows a trajectory — where the student started, what changed, and under what conditions — which directly supports instructional planning conversations in ways a single number cannot.
Putting the Purposes Together
Think of these assessment types as answering different questions along a single pipeline: Screening asks "who might need a closer look?" Diagnostic assessment asks "what exactly is going on, and how severe is it?" Formative assessment/CBM asks "is what I'm teaching right now working?" Summative assessment asks "did the student ultimately reach the goal or standard?" Observation and portfolio asks "what does this look like in real, authentic contexts, over time?" On exam day, resist the urge to memorize a list of specific commercial test names — instead, practice sorting scenario descriptions into the correct purpose category, since that is the skill actually being tested.
A school administers a brief, three-times-a-year reading measure to every student in second grade to identify which students may be at risk for reading difficulty. What type of assessment is this?
Which feature is a defining characteristic of curriculum-based measurement (CBM) that makes it especially useful for tracking a student's response to an intervention week to week?
A teacher wants to document a nonverbal student's growth in initiating communication attempts across the classroom, cafeteria, and playground over an entire semester. Which assessment approach best fits this purpose?