1.1 TOEFL iBT Reading Redesign & Section Architecture
Key Takeaways
- The TOEFL iBT Reading section is a two-stage multistage adaptive test: a routing Module 1 followed by a higher- or lower-difficulty Module 2, with ETS listing 50 items and about 30 minutes of base time.
- Test takers encounter three task types: Complete the Words (a 10-blank C-test), Read in Daily Life (short everyday and campus texts), and Read an Academic Passage (~200 words followed by 5 questions).
- Reading is reported on a 1.0–6.0 band scale in 0.5-point increments spanning CEFR A1 through C2, where a band of 5.0 or 5.5 is C1 and only 6.0 is C2.
- ETS's comparison chart maps Reading bands to legacy 0–30 scores as 6.0 = 29–30, 5.5 = 27–28, 5.0 = 24–26, 4.5 = 22–23, and 4.0 = 18–21.
- ETS does not publish the Module 1 routing cut-score or the score ceiling for the lower-difficulty path, so specific thresholds quoted by prep sources are estimates rather than official figures.
TOEFL iBT Reading Redesign & Section Architecture
The TOEFL iBT Reading section assesses your ability to read, comprehend, and analyze academic and everyday English texts in higher-education environments. Rather than presenting a static, linear series of long passages, the test employs a streamlined, multi-stage adaptive design that evaluates reading proficiency across multiple task modalities with high psychometric precision.
Understanding the architectural mechanics of this section—how adaptive routing works, what each task modality measures, and how raw responses translate into reported band scores—is essential for optimizing your preparation and test-day strategy.
1. Section Architecture & Time Breakdown
The Reading section consists of approximately 50 item opportunities delivered across a total testing window of 27 to 30 minutes. The section is divided into two discrete, sequential stages known as Module 1 and Module 2.
| Architectural Dimension | Specification |
|---|---|
| Total Section Duration | 27–30 minutes |
| Number of Modules | 2 sequential stages (Module 1: Routing; Module 2: Adaptive Tier) |
| Total Item Opportunities | ~50 items (spanning discrete blanks, short-passage questions, and full passage items) |
| Task Modalities Tested | 3 core formats: Complete the Words, Read in Daily Life, Read an Academic Passage |
| Scoring Scale | 1.0 to 6.0 in 0.5-point increments (with CEFR B1–C2 alignment) |
| Legacy Score Concordance | Mapped to historical 0–30 scaled score range |
| Adaptive Delivery Engine | Multi-Stage Adaptive Testing (MST) at the module level |
+-----------------------------------------------------------------------------+
| TOEFL iBT READING SECTION ARCHITECTURE |
| |
| [MODULE 1: ROUTING STAGE] (~20 Minutes - the longer module) |
| - Complete the Words (C-Test passages) |
| - Read in Daily Life (Campus notices, emails, schedules) |
| - Read an Academic Passage (Introductory academic text) |
| - Balanced cross-section of foundational to advanced difficulty |
| | |
| v |
| [PSYCHOMETRIC ROUTING CUT-SCORE] |
| (Item Response Theory Ability θ) |
| | |
| +------------------+------------------+ |
| | | |
| v (High Accuracy) v (Lower Accuracy) |
| [MODULE 2: HIGHER DIFFICULTY] [MODULE 2: LOWER DIFFICULTY] |
| (~10 Minutes) (~10 Minutes) |
| - Complex academic discourse - Standard academic & campus texts |
| - Advanced lexical deletions - Core syntactic structures |
| - Top of scale reachable - Upper bands unreachable |
+-----------------------------------------------------------------------------+
Unlike traditional linear tests where every candidate answers identical questions, the multi-stage format personalizes the difficulty of Module 2 based on your demonstrated competence in Module 1. This adaptive design delivers an accurate measurement of your reading proficiency in less total testing time.
2. Multi-Stage Adaptive Testing (MST) Mechanics
Multi-Stage Adaptive Testing (MST) differs from item-level computerized adaptive testing (such as the computer-adaptive GRE). In item-level adaptive testing, question difficulty changes after every single question, preventing test takers from moving backward to review earlier answers. In stage-level MST, adaptation occurs strictly between modules.
How Module 1 Routing Works
Module 1 functions as a universal routing module. Every candidate receives a carefully calibrated set of questions that spans a representative distribution of difficulty levels (from introductory to advanced).
- Routing Assessment: As you complete Module 1, the scoring algorithm computes an intermediate ability estimate (denoted psychometrically as $\theta$).
- Threshold Evaluation: When you submit Module 1, the testing engine compares your score against a predetermined cut-score threshold.
- Branching Decision:
- If your accuracy meets or exceeds the threshold, the engine routes you to the higher-difficulty Module 2.
- If your accuracy falls below the threshold, the engine routes you to the lower-difficulty Module 2.
[!IMPORTANT] The Routing Ceiling Effect: Reaching the higher-difficulty Module 2 is what makes the top of the scale attainable: the hardest items carry the information the scoring model needs to place you at Band 5.5 or 6.0 (CEFR C1–C2). If Module 1 goes badly and you are routed to the lower-difficulty Module 2, your reachable score is limited even if you then answer everything correctly, because the easier items cannot demonstrate advanced ability.
ETS does not publish the routing cut-score or the exact ceiling attached to each path, so treat any specific number you see quoted online as an estimate rather than an official figure. Experienced TOEFL instructors who have sat the redesigned test report needing roughly 60% accuracy in Module 1 to reach the harder Module 2. The practical lesson does not depend on the exact threshold: accuracy in Module 1 matters more than accuracy anywhere else in the section.
+-----------------------------------------------------------------------------+
| ROUTING BRANCHES & SCORE CEILINGS |
| |
| Module 1 Performance: |
| High Accuracy (≥ Cut-Score) ---> Routes to HIGHER-DIFFICULTY MODULE 2 |
| - Top of the scale reachable |
| - Bands 5.0-6.0 = legacy 24-30 |
| - CEFR Alignment: C1 to C2 |
| |
| Lower Accuracy (< Cut-Score) ---> Routes to LOWER-DIFFICULTY MODULE 2 |
| - Upper bands no longer reachable |
| - Bands 1.0-4.5 = legacy 0-23 |
| - CEFR Alignment: A1 to B2 |
+-----------------------------------------------------------------------------+
Within each individual module, you retain full navigational freedom: you can skip questions, return to earlier items, change selections, and flag questions for review before finalizing your module submission.
3. The Three Core Task Modalities
The Reading section evaluates your reading ability across three distinct task modalities that mirror the linguistic demands of North American university life.
+-----------------------------------------------------------------------------+
| THE THREE TASK MODALITIES |
| |
| +---------------------------------------------------------------------+ |
| | 1. COMPLETE THE WORDS (C-Test Modality) | |
| | - Lexical, morphological, and syntactic micro-decoding | |
| | - Short paragraph with deleted word halves (10 blanks = 10 items) | |
| | - Tests contextual vocabulary, inflectional affixes, and collocations| |
| +---------------------------------------------------------------------+ |
| | |
| v |
| +---------------------------------------------------------------------+ |
| | 2. READ IN DAILY LIFE (Campus Pragmatics Modality) | |
| | - Everyday institutional, administrative, and social campus texts | |
| | - Syllabi, emails, bulletin notices, policy guides, schedules | |
| | - Tests factual retrieval, policy conditions, and pragmatic intent | |
| +---------------------------------------------------------------------+ |
| | |
| v |
| +---------------------------------------------------------------------+ |
| | 3. READ AN ACADEMIC PASSAGE (Academic Literacy Modality) | |
| | - Short academic excerpts (~200 words, 5 questions each) | |
| | - Natural sciences, social sciences, humanities, and arts | |
| | - Tests inference, rhetorical purpose, clausal syntax, and synthesis| |
| +---------------------------------------------------------------------+ |
+-----------------------------------------------------------------------------+
Modality 1: Complete the Words (C-Test)
- Format: A short (~70-word) informational paragraph. The first and last sentences are left intact to establish context; beginning in the second sentence, exactly 10 words have their second half deleted (e.g., "The discov____ of new foss____ provi____ crucial evid____..."). Each deleted word counts as one item, so a single C-test task supplies 10 of the section's ~50 items. The number of missing letters (between 1 and 6) is shown for each word, and you type the letters directly into the blank.
- Construct Measured: Micro-level syntactic processing, morphological knowledge (prefixes, roots, suffixes), part-of-speech awareness, and collocational familiarity. Test takers must type the missing letters directly into input boxes.
- Psychometric Role: Provides an efficient measure of underlying lexical automaticity and reading fluency.
Modality 2: Read in Daily Life
- Format: Short non-academic and pragmatic texts typical of university campus environments. Examples include registrar policy updates, library borrowing terms, student club announcements, residence hall guidelines, academic advising emails, and campus facility schedules.
- Construct Measured: Functional reading comprehension, rapid factual retrieval, conditional interpretation (understanding rules, exceptions, and deadlines), and pragmatic inference (identifying the writer's underlying purpose or implied instructions).
Modality 3: Read an Academic Passage
- Format: A short excerpt (~200 words) followed by 5 questions. ETS draws topics from history, art and music, business and economics, life science, physical science, and social science. Unlike the pre-2026 test, question stems no longer tell you which paragraph to search — you scan the whole passage for each answer.
- Construct Measured: High-level academic discourse comprehension. Questions target main idea (gist-content), factual details, negative factual statements (NOT/EXCEPT), contextual vocabulary, logical inference, author's rhetorical purpose, the relationship between paragraphs (organization), identifying an important idea by clicking a sentence, and — rarely — text insertion. Note that pronoun-reference, sentence-simplification, prose-summary, and table/categorization questions were retired with the January 2026 redesign; a ~200-word passage is too short to support them.
4. Band Scoring Scale (1.0–6.0), CEFR Levels & Legacy Concordance
The TOEFL iBT reports Reading section performance on a 1.0 to 6.0 band score scale in 0.5-point increments, introduced on January 21, 2026. Every half band maps to a CEFR level, and ETS publishes a comparison chart relating each band to the legacy 0–30 Reading scaled score used before January 21, 2026. Two cautions matter. First, ETS presents this chart as a tool for interpreting older scores, not as a conversion formula applied to your new test — the underlying statistics "do not assume equal interval score points," so the number of legacy points inside each half band varies. Second, the legacy 0–120 total is reported only as a single comparable overall figure during ETS's two-year transition (through January 2028); ETS does not report legacy 0–30 section scores alongside your band.
| Band Score | CEFR Level | Performance Level Descriptor | Legacy 0–30 Equivalent | Admissions Standard / Institutional Interpretation |
|---|---|---|---|---|
| 6.0 | C2 | Expert Academic | 29–30 | Full fluency; comprehends highly complex, abstract academic texts with nuanced rhetorical and conceptual structures. |
| 5.5 | C1 | Advanced High | 27–28 | Strong command; synthesizes dense disciplinary arguments, parses embedded clauses, and infers subtle pragmatic tone effortlessly. |
| 5.0 | C1 | Advanced | 24–26 | Meets standard unconditional admission cutoffs for top graduate and undergraduate programs. Handles academic texts with minimal difficulty. |
| 4.5 | B2 | High Intermediate | 22–23 | Competent reader; extracts main ideas and key factual relationships but may struggle with dense syntactic embedding or obscure technical vocabulary. |
| 4.0 | B2 | Intermediate | 18–21 | Satisfies conditional or pathway program admissions; requires occasional scaffolding for complex disciplinary readings. |
| 3.5 | B1 | Modest | 12–17 | Grasps basic campus announcements and explicit facts; experiences difficulty with academic inferences and complex sentences. |
| 3.0 | B1 | Limited | 6–11 | Basic comprehension of everyday notices; limited ability to process extended university-level discourse. |
| 2.5 | A2 | Elementary High | 4–5 | Reads short, predictable everyday material; cannot sustain extended academic discourse. |
| 2.0 | A2 | Elementary | 3 | Understands simple, concrete phrases in routine contexts. |
| 1.5 | A1 | Novice High | 2 | Recognizes familiar names, words, and very basic phrases on simple notices. |
| 1.0 | A1 | Novice | 0–1 | Minimal reading comprehension; restricted to isolated words and basic phrases. |
[!NOTE] Item Response Theory (IRT) Scoring: Raw item counts do not translate linearly into band scores. ETS uses a 3-parameter logistic (3PL) Item Response Theory model that accounts for item difficulty ($b$), item discrimination ($a$), and pseudo-guessing parameters ($c$). Correctly answering a high-difficulty item in the higher-difficulty Module 2 contributes more to your latent ability estimate than answering an introductory-level item in the lower-difficulty Module 2.
5. Strategic Implications for Test Takers
- Prioritize Module 1 Precision Over Speed: Rushing through Module 1 to bank time introduces unforced errors that can trigger the routing penalty and lock you into the lower-difficulty Module 2. Maintain a steady, methodical pace.
- Master All Three Task Modalities: High academic reading comprehension alone cannot compensate for poor performance on C-test word completions. Practice morphological decoding and campus pragmatic texts alongside traditional academic passages.
- Use the Review Screen Strategically: Because in-module navigation is permitted, flag uncertain items and return to them before submitting Module 1. Once you confirm submission of Module 1, your routing path is locked.
How does the Multi-Stage Adaptive Testing (MST) architecture in the TOEFL iBT Reading section impact a test taker's maximum possible band score?
Which of the following correctly pairs each TOEFL iBT Reading task modality with its primary assessment construct?
A student receives a TOEFL iBT Reading band score of 5.5. How does this score translate within the CEFR framework and the legacy 0–30 scaled score concordance?