12.1 Assessment Types: Diagnostic, Formative, Summative & Performance
Key Takeaways
Language assessment functions as an instructional compass that diagnoses baseline competencies, guides ongoing learning, and certifies communicative attainment.
Diagnostic assessment identifies specific linguistic strengths and interlanguage gaps prior to instruction, directly informing course placement and curriculum differentiation.
Formative assessment (Assessment for Learning / AfL) relies on dynamic, low-stakes feedback loops (Black & Wiliam) to adapt real-time teaching and foster learner metacognition.
Summative assessment evaluates cumulative attainment at course conclusion against established benchmarks such as the CEFR, whereas performance assessment evaluates authentic, task-based language use.
Criterion-referenced testing measures individual student mastery against explicit performance standards, contrasting with norm-referenced testing which ranks learners along a peer distribution curve.
12.1 Assessment Types: Diagnostic, Formative, Summative & Performance
Note
Language assessment in English Language Teaching (ELT) serves as an instructional compass. Distinguishing among diagnostic, formative, summative, and performance-based measures empowers educators to align measurement instruments with pedagogical goals: identifying baseline competencies, calibrating instruction, or certifying communicative proficiency.
Assessment Taxonomy and Primary Purposes in ELT
Assessment in ELT encompasses all systematic procedures used to gather information about learner language development and communicative performance. While testing is a formal, periodic event, assessment is continuous. A comprehensive ELT assessment taxonomy classifies procedures across four core dimensions:
- Administrative Purpose: Diagnostic, formative, or summative.
- Reference Standard: Criterion-referenced (measured against objectives) or norm-referenced (measured against peers).
- Formality: Formal standardized protocols versus informal classroom observations.
- Task Authenticity: Direct performance tasks versus indirect discrete-point measures.
Structuring assessment along these dimensions ensures instruments yield actionable data that foster language acquisition without elevating the affective filter.
Diagnostic Assessment: Baseline Mapping and Placement
Diagnostic assessment occurs prior to or at the start of an instructional sequence. Its primary objective is mapping a learner's linguistic profile, identifying mastered competencies and systemic interlanguage gaps across phonology, lexis, syntax, and discourse.
In language programs, diagnostic assessment operates through two main instruments:
- Language Proficiency Screeners: Broad initial evaluations administered upon institutional entry to determine whether a student requires specialized English language support services.
- Placement Tests: Targeted batteries calibrated to assign learners to appropriate course tiers within a leveled curriculum (e.g., placing an adult learner into High-Beginning versus Low-Intermediate ESL).
Effective diagnostic tools uncover the underlying source of error, distinguishing developmental interlanguage patterns from native language (L1) negative transfer, enabling instructors to differentiate curriculum content and scaffold learning.
Formative Assessment & Assessment for Learning (AfL)
Formative assessment, or Assessment for Learning (AfL), occurs continuously during the learning process. Popularized by Paul Black and Dylan Wiliam (1998), AfL treats assessment not as an evaluative audit, but as an interactive feedback loop that uncovers misunderstandings and dynamically reshapes instruction.
Building on that work, Dylan Wiliam and colleagues (Leahy, Lyon, Thompson, and Wiliam, 2005; Wiliam and Thompson, 2007) summarized five key AfL strategies:
- Clarifying, sharing, and understanding learning intentions and success criteria.
- Engineering classroom discussions, tasks, and activities that elicit evidence of learning.
- Providing descriptive feedback that moves learners forward.
- Activating learners as instructional resources for one another through peer assessment.
- Activating learners as owners of their own learning via self-regulation.
Formative Assessment Quick-Check Toolkit
| Technique | Implementation Mechanism | Target Language Domain | Pedagogical Benefit |
|---|---|---|---|
| Exit Tickets | 1–2 targeted prompts submitted at class conclusion | Morphosyntax, vocabulary | Provides immediate diagnostic data to adjust the subsequent lesson plan |
| Running Records | Systematic notation of oral reading miscues | Decoding, fluency, comprehension | Isolates phonological and syntactic processing breakdowns in authentic reading |
| Traffic Light Rating | Visual cards (green/yellow/red) indicating confidence | Metacognition, readiness | Identifies immediate student groupings for differentiated practice |
| Anecdotal Observation | Structured notes logged during communicative group work | Spoken fluency, pragmatics | Captures spontaneous communicative competence in low-anxiety settings |
| Mini-Whiteboards | Simultaneous display of responses to target prompts | Grammar accuracy, listening | Ensures total active student participation while eliminating public embarrassment |
Comparative Matrix of Primary Assessment Types
| Dimension | Diagnostic Assessment | Formative Assessment (AfL) | Summative Assessment (AoL) | Performance-Based Assessment |
|---|---|---|---|---|
| Primary Purpose | Identify baseline strengths, gaps, and placement | Guide ongoing instruction; close learning gaps | Measure attainment and certify competence | Evaluate real-world communicative task execution |
| Timing | Prior to instruction or at course start | Continuous and embedded throughout instruction | End of instructional unit, term, or program | Embedded within units or serving as capstones |
| Stakes | Low to moderate (placement; no punitive grades) | Zero to low stakes (ungraded or completion marks) | High stakes (determines promotion, grades, certification) | Moderate to high stakes (depends on rubric weighting) |
| Reference Standard | Criterion- or diagnostic-referenced | Criterion-referenced (lesson learning objectives) | Criterion-referenced or norm-referenced | Criterion-referenced (multi-trait analytic rubrics) |
| Feedback Focus | Informs syllabus design and differentiated grouping | Dynamic feed-forward guiding immediate next steps | Static delayed summary of cumulative mastery | Qualitative critique of integrated communicative performance |
Summative Assessment: Measuring Attainment & Certification
Summative assessment, or Assessment of Learning (AoL), evaluates cumulative learning at the end of an instructional cycle. Its primary role is institutional accountability: verifying whether learners met curricular standards, assigning grades, and awarding credentials.
Summative instruments include teacher-constructed final exams and large-scale standardized batteries such as TOEFL, IELTS, and Cambridge English Qualifications. These examinations are frequently calibrated to the Common European Framework of Reference for Languages (CEFR), which scales proficiency from A1 (breakthrough) to C2 (mastery). While summative exams provide cross-institutional comparability, their backward-looking nature offers limited utility for immediate classroom intervention.
Criterion-Referenced versus Norm-Referenced Testing
- Criterion-Referenced Assessment: Evaluates learner performance against predefined, objective criteria or learning standards, independent of peer performance. For example, a rubric determining whether a student writes a coherent formal complaint letter is criterion-referenced; every student demonstrating mastery receives top marks.
- Norm-Referenced Assessment: Ranks learners relative to the statistical distribution of a normative peer group, reporting scores as percentiles or stanines. Designed to maximize variance along a bell curve, norm-referenced tests are common for competitive admissions but counterproductive for measuring curricular mastery.
Formal versus Informal Assessment
- Formal Assessment: Systematic, planned measurement protocols administered under standardized conditions with explicit scoring criteria (e.g., midterm exams, standardized proficiency batteries).
- Informal Assessment: Unobtrusive, incidental observations conducted during routine classroom activities without standardized scoring (e.g., verbal questioning, monitoring group work). Informal assessment minimizes the affective filter, capturing spontaneous language use.
Performance-Based Assessment & Authentic Language Tasks
Performance-based assessment requires learners to construct spoken or written responses through authentic tasks mirroring real-world communication:
- Oral Presentations & Debates: Evaluating spoken fluency, discourse management, and pragmatic register in real time.
- Workplace & Everyday Simulations: Situating communication in contextualized scenarios (e.g., job interviews, customer complaints).
- Language Portfolios: Purposeful collections of student artifacts (drafts, recordings, reflections) documenting longitudinal growth, fostering learner autonomy, and integrating self-evaluation.
Tip
Always share performance rubrics with learners before task execution. Transparent criteria reduce test anxiety, clarify communicative expectations, and guide self-directed practice.
A secondary school ESL teacher administers an extensive language battery during the first week of the academic year. The instrument assesses phonemic discrimination, lexical breadth, and verb tense morphology, not to calculate report card grades, but to identify specific interlanguage fossilizations and group students for targeted intervention. Which assessment type does this instructional practice exemplify?
Summative assessment designed to assign formal credit and measure institutional accountability
Diagnostic assessment designed to uncover baseline competencies and linguistic deficits
High-stakes certification testing designed to award academic credentials
Norm-referenced proficiency screening designed to rank students on a competitive bell curve
An English language institute evaluates adult immigrant learners using an end-of-module writing assessment. To pass, students must write a 250-word formal complaint letter that includes an opening statement of purpose, chronological details of the grievance, polite modal requests for redress, and accurate conventional salutations. Regardless of how peers perform, every student who satisfies these explicit performance descriptors earns a passing certificate. Which measurement framework does this exam utilize?
Norm-referenced testing, because it evaluates student performance relative to the average cohort distribution
Criterion-referenced testing, because student performance is measured against predetermined behavioral standards
Informal assessment, because the testing procedure relies on unstandardized observational notes
Ipsative assessment, because the score is derived solely by comparing the student's current draft to their own prior writing
During an intermediate communicative English lesson on polite requests, a teacher notices several pairs struggling with modal verb inversion. The teacher immediately pauses the pair-work, conducts a brief two-minute clarification mini-lesson with choral repetition, and asks all students to write one revised request on an individual mini-whiteboard before resuming the communicative activity. According to Black and Wiliam's assessment framework, this intervention represents:
A standardized performance capstone measuring long-term second language acquisition
A formative assessment loop that adjusts teaching based on emerging evidence
A norm-referenced placement screener identifying low-performing language candidates
A summative evaluation measuring terminal attainment of English modal grammar
Sections you finish are checked off in the contents.