12.1 Assessment Types: Diagnostic, Formative, Summative & Performance

Key Takeaways

  • Language assessment functions as an instructional compass that diagnoses baseline competencies, guides ongoing learning, and certifies communicative attainment.

  • Diagnostic assessment identifies specific linguistic strengths and interlanguage gaps prior to instruction, directly informing course placement and curriculum differentiation.

  • Formative assessment (Assessment for Learning / AfL) relies on dynamic, low-stakes feedback loops (Black & Wiliam) to adapt real-time teaching and foster learner metacognition.

  • Summative assessment evaluates cumulative attainment at course conclusion against established benchmarks such as the CEFR, whereas performance assessment evaluates authentic, task-based language use.

  • Criterion-referenced testing measures individual student mastery against explicit performance standards, contrasting with norm-referenced testing which ranks learners along a peer distribution curve.

Last updated: October 2026

12.1 Assessment Types: Diagnostic, Formative, Summative & Performance

Note

Language assessment in English Language Teaching (ELT) serves as an instructional compass. Distinguishing among diagnostic, formative, summative, and performance-based measures empowers educators to align measurement instruments with pedagogical goals: identifying baseline competencies, calibrating instruction, or certifying communicative proficiency.

Assessment Taxonomy and Primary Purposes in ELT

Assessment in ELT encompasses all systematic procedures used to gather information about learner language development and communicative performance. While testing is a formal, periodic event, assessment is continuous. A comprehensive ELT assessment taxonomy classifies procedures across four core dimensions:

  1. Administrative Purpose: Diagnostic, formative, or summative.
  2. Reference Standard: Criterion-referenced (measured against objectives) or norm-referenced (measured against peers).
  3. Formality: Formal standardized protocols versus informal classroom observations.
  4. Task Authenticity: Direct performance tasks versus indirect discrete-point measures.

Structuring assessment along these dimensions ensures instruments yield actionable data that foster language acquisition without elevating the affective filter.

Diagnostic Assessment: Baseline Mapping and Placement

Diagnostic assessment occurs prior to or at the start of an instructional sequence. Its primary objective is mapping a learner's linguistic profile, identifying mastered competencies and systemic interlanguage gaps across phonology, lexis, syntax, and discourse.

In language programs, diagnostic assessment operates through two main instruments:

  • Language Proficiency Screeners: Broad initial evaluations administered upon institutional entry to determine whether a student requires specialized English language support services.
  • Placement Tests: Targeted batteries calibrated to assign learners to appropriate course tiers within a leveled curriculum (e.g., placing an adult learner into High-Beginning versus Low-Intermediate ESL).

Effective diagnostic tools uncover the underlying source of error, distinguishing developmental interlanguage patterns from native language (L1) negative transfer, enabling instructors to differentiate curriculum content and scaffold learning.

Formative Assessment & Assessment for Learning (AfL)

Formative assessment, or Assessment for Learning (AfL), occurs continuously during the learning process. Popularized by Paul Black and Dylan Wiliam (1998), AfL treats assessment not as an evaluative audit, but as an interactive feedback loop that uncovers misunderstandings and dynamically reshapes instruction.

Building on that work, Dylan Wiliam and colleagues (Leahy, Lyon, Thompson, and Wiliam, 2005; Wiliam and Thompson, 2007) summarized five key AfL strategies:

  1. Clarifying, sharing, and understanding learning intentions and success criteria.
  2. Engineering classroom discussions, tasks, and activities that elicit evidence of learning.
  3. Providing descriptive feedback that moves learners forward.
  4. Activating learners as instructional resources for one another through peer assessment.
  5. Activating learners as owners of their own learning via self-regulation.

Formative Assessment Quick-Check Toolkit

TechniqueImplementation MechanismTarget Language DomainPedagogical Benefit
Exit Tickets1–2 targeted prompts submitted at class conclusionMorphosyntax, vocabularyProvides immediate diagnostic data to adjust the subsequent lesson plan
Running RecordsSystematic notation of oral reading miscuesDecoding, fluency, comprehensionIsolates phonological and syntactic processing breakdowns in authentic reading
Traffic Light RatingVisual cards (green/yellow/red) indicating confidenceMetacognition, readinessIdentifies immediate student groupings for differentiated practice
Anecdotal ObservationStructured notes logged during communicative group workSpoken fluency, pragmaticsCaptures spontaneous communicative competence in low-anxiety settings
Mini-WhiteboardsSimultaneous display of responses to target promptsGrammar accuracy, listeningEnsures total active student participation while eliminating public embarrassment

Comparative Matrix of Primary Assessment Types

DimensionDiagnostic AssessmentFormative Assessment (AfL)Summative Assessment (AoL)Performance-Based Assessment
Primary PurposeIdentify baseline strengths, gaps, and placementGuide ongoing instruction; close learning gapsMeasure attainment and certify competenceEvaluate real-world communicative task execution
TimingPrior to instruction or at course startContinuous and embedded throughout instructionEnd of instructional unit, term, or programEmbedded within units or serving as capstones
StakesLow to moderate (placement; no punitive grades)Zero to low stakes (ungraded or completion marks)High stakes (determines promotion, grades, certification)Moderate to high stakes (depends on rubric weighting)
Reference StandardCriterion- or diagnostic-referencedCriterion-referenced (lesson learning objectives)Criterion-referenced or norm-referencedCriterion-referenced (multi-trait analytic rubrics)
Feedback FocusInforms syllabus design and differentiated groupingDynamic feed-forward guiding immediate next stepsStatic delayed summary of cumulative masteryQualitative critique of integrated communicative performance

Summative Assessment: Measuring Attainment & Certification

Summative assessment, or Assessment of Learning (AoL), evaluates cumulative learning at the end of an instructional cycle. Its primary role is institutional accountability: verifying whether learners met curricular standards, assigning grades, and awarding credentials.

Summative instruments include teacher-constructed final exams and large-scale standardized batteries such as TOEFL, IELTS, and Cambridge English Qualifications. These examinations are frequently calibrated to the Common European Framework of Reference for Languages (CEFR), which scales proficiency from A1 (breakthrough) to C2 (mastery). While summative exams provide cross-institutional comparability, their backward-looking nature offers limited utility for immediate classroom intervention.

Criterion-Referenced versus Norm-Referenced Testing

  • Criterion-Referenced Assessment: Evaluates learner performance against predefined, objective criteria or learning standards, independent of peer performance. For example, a rubric determining whether a student writes a coherent formal complaint letter is criterion-referenced; every student demonstrating mastery receives top marks.
  • Norm-Referenced Assessment: Ranks learners relative to the statistical distribution of a normative peer group, reporting scores as percentiles or stanines. Designed to maximize variance along a bell curve, norm-referenced tests are common for competitive admissions but counterproductive for measuring curricular mastery.

Formal versus Informal Assessment

  • Formal Assessment: Systematic, planned measurement protocols administered under standardized conditions with explicit scoring criteria (e.g., midterm exams, standardized proficiency batteries).
  • Informal Assessment: Unobtrusive, incidental observations conducted during routine classroom activities without standardized scoring (e.g., verbal questioning, monitoring group work). Informal assessment minimizes the affective filter, capturing spontaneous language use.

Performance-Based Assessment & Authentic Language Tasks

Performance-based assessment requires learners to construct spoken or written responses through authentic tasks mirroring real-world communication:

  • Oral Presentations & Debates: Evaluating spoken fluency, discourse management, and pragmatic register in real time.
  • Workplace & Everyday Simulations: Situating communication in contextualized scenarios (e.g., job interviews, customer complaints).
  • Language Portfolios: Purposeful collections of student artifacts (drafts, recordings, reflections) documenting longitudinal growth, fostering learner autonomy, and integrating self-evaluation.

Tip

Always share performance rubrics with learners before task execution. Transparent criteria reduce test anxiety, clarify communicative expectations, and guide self-directed practice.

Loading diagram...
The Language Assessment Spectrum in ELT
Test Your Knowledge

A secondary school ESL teacher administers an extensive language battery during the first week of the academic year. The instrument assesses phonemic discrimination, lexical breadth, and verb tense morphology, not to calculate report card grades, but to identify specific interlanguage fossilizations and group students for targeted intervention. Which assessment type does this instructional practice exemplify?

A

Summative assessment designed to assign formal credit and measure institutional accountability

B

Diagnostic assessment designed to uncover baseline competencies and linguistic deficits

C

High-stakes certification testing designed to award academic credentials

D

Norm-referenced proficiency screening designed to rank students on a competitive bell curve

Test Your Knowledge

An English language institute evaluates adult immigrant learners using an end-of-module writing assessment. To pass, students must write a 250-word formal complaint letter that includes an opening statement of purpose, chronological details of the grievance, polite modal requests for redress, and accurate conventional salutations. Regardless of how peers perform, every student who satisfies these explicit performance descriptors earns a passing certificate. Which measurement framework does this exam utilize?

A

Norm-referenced testing, because it evaluates student performance relative to the average cohort distribution

B

Criterion-referenced testing, because student performance is measured against predetermined behavioral standards

C

Informal assessment, because the testing procedure relies on unstandardized observational notes

D

Ipsative assessment, because the score is derived solely by comparing the student's current draft to their own prior writing

Test Your Knowledge

During an intermediate communicative English lesson on polite requests, a teacher notices several pairs struggling with modal verb inversion. The teacher immediately pauses the pair-work, conducts a brief two-minute clarification mini-lesson with choral repetition, and asks all students to write one revised request on an individual mini-whiteboard before resuming the communicative activity. According to Black and Wiliam's assessment framework, this intervention represents:

A

A standardized performance capstone measuring long-term second language acquisition

B

A formative assessment loop that adjusts teaching based on emerging evidence

C

A norm-referenced placement screener identifying low-performing language candidates

D

A summative evaluation measuring terminal attainment of English modal grammar

Sections you finish are checked off in the contents.