4.1 The Architecture of 'Cannot Say': Insufficient Evidence vs. Plausible Assumptions
Key Takeaways
- 'Cannot Say' represents a definitive epistemological verdict of insufficient textual evidence, not an admission of textual ambiguity, drafting flaw, or implausibility.
- The plausibility trap is the single greatest cause of candidate error, tempting test-takers to mark statements 'True' simply because they are realistic, reasonable, or empirically probable.
- Logical consistency with a passage is merely a baseline condition for 'Cannot Say', whereas conclusive deductive entailment is strictly mandatory for a verdict of 'True'.
- If an assertion requires any unstated bridging assumption, missing link, or real-world extrapolation, it fails the evidentiary test for 'True' and must be classified as 'Cannot Say'.
- Civil Service Verbal Test candidates must operate as detached textual auditors, evaluating exclusively what is conclusively established within the four corners of the text.
4.1 The Architecture of 'Cannot Say': Insufficient Evidence vs. Plausible Assumptions
Core Principle: 'Cannot Say' is neither an admission of textual ambiguity nor a verdict that a proposition is improbable. In the Civil Service Verbal Test (CSVT), 'Cannot Say' is a precise epistemological classification: the passage, evaluated strictly as an exhaustive, self-contained universe, provides insufficient evidence to conclusively substantiate or definitively refute the assertion.
In public administration, civil servants routinely review submissions, intelligence dossiers, draft statutory instruments, and audit findings. In these high-stakes operational environments, conflating what is empirically proven with what is merely plausible or likely can lead to severe policy failure, misallocation of public expenditure, or legal challenge under judicial review. Consequently, the Civil Service Verbal Test evaluates candidates' capacity for rigorous evidentiary discipline. The test demands that candidates resist the psychological urge to fill in informational gaps, enforcing an uncompromising standard where unproven assertions must be categorized strictly as 'Cannot Say'.
The Epistemological Status of 'Cannot Say'
To master the CSVT, candidates must understand the formal three-valued logic governing every item. Statements are evaluated against an evidentiary continuum comprising three mutually exclusive categories:
- True: The assertion is explicitly confirmed by the passage or constitutes an inescapable, logically necessary consequence of the text's explicit premises ($T \models P$). If the passage is true, it is logically impossible for the statement to be false.
- False: The assertion is directly contradicted by explicit textual statements or is rendered logically or physically impossible by the facts presented ($T \models \neg P$). If the passage is true, it is logically impossible for the statement to be true.
- Cannot Say: The evidentiary record within the passage is underdetermined ($T \not\models P$ and $T \not\models \neg P$). The text does not provide sufficient data to confirm the statement, nor does it contain sufficient grounds to refute it.
Dispelling the Common Misconceptions
Many candidates misinterpret the meaning of 'Cannot Say', regarding it as a residual option reserved for confusing, poorly written, or nonsensical questions. In reality, test developers construct 'Cannot Say' items with extreme precision. Consider the following crucial operational boundaries:
- 'Cannot Say' does NOT mean the passage is poorly drafted: The passage may be an impeccably structured policy excerpt. 'Cannot Say' simply reflects that the author chose not to address the specific sub-claim made in the statement.
- 'Cannot Say' does NOT mean the statement is absurd or false in the real world: A statement may represent an undeniable historical or scientific reality (e.g., "The Treasury manages fiscal policy in the United Kingdom"), but if the text focuses exclusively on local council waste collection without mentioning central government finance, the correct determination is strictly 'Cannot Say'.
- 'Cannot Say' does NOT indicate examiner trickery: It is an active assessment of your ability to recognize the boundaries of provided evidence without inventing unstated bridges.
The Psychological Trap of Plausibility
The primary cognitive obstacle in verbal reasoning is plausibility bias. Human cognition is fundamentally Bayesian and narrative-driven: when processing incomplete information, our brains automatically deploy System 1 heuristics to synthesize a coherent mental model. In everyday life and general workplace communication, reading between the lines and adopting reasonable assumptions is rewarded as intuitive intelligence. On the Civil Service Verbal Test, however, this habit is disastrous.
Candidates routinely mark a statement 'True' because:
- The asserted outcome is a natural, common-sense byproduct of the policy described.
- The statement describes an outcome that would be considered "good public administration."
- The claim aligns with general political, economic, or operational reality.
The Evidentiary Spectrum: From Improbable to Entailed
To neutralize plausibility bias, candidates must visualize where an assertion falls along the Evidentiary Spectrum:
[ Improbable ] ---> [ Possible ] ---> [ Plausible ] ---> [ Highly Probable ] ||||===> [ Conclusively Entailed ]
(Cannot Say) (Cannot Say) (Cannot Say) (Cannot Say) |||| (TRUE)
^^^^
Evidentiary Threshold
Notice that every single category prior to the threshold—whether possible, plausible, or overwhelmingly probable—collapses into Cannot Say. An assertion does not transition to 'True' when it reaches a 95% likelihood; it becomes 'True' if and only if it reaches 100% deductive necessity based strictly on the text provided.
Logical Consistency vs. Conclusive Entailment
A critical analytical distinction tested on higher-difficulty CSVT items is the difference between logical consistency and conclusive entailment.
- Logical Consistency (Compatibility): A statement $P$ is logically consistent with a passage $T$ if $T$ and $P$ can both be true simultaneously without generating a contradiction ($T \land P \not\vdash \bot$). That is, nothing in the passage forbids or refutes $P$. Candidates often erroneously assume: "The text doesn't say it didn't happen, and it fits the scenario perfectly, so it must be True." This is a fundamental logical error. Consistency merely establishes that the statement could be true; it does not prove that it is true.
- Conclusive Entailment (Deductive Necessity): A statement $P$ is conclusively entailed by passage $T$ if and only if the truth of $T$ guarantees the truth of $P$ in all circumstances ($T \implies P$). There is zero possibility for the statement to be false if the passage premises hold.
Golden Rule: Logical consistency is merely the entry ticket for 'Cannot Say'. Only deductive entailment justifies 'True'.
Worked Example 1: Operational Casework & Backlog Management
Passage
Under the HM Courts & Tribunals Service (HMCTS) North East civil listing pilot, six designated county courts introduced an automated scheduling algorithm for small claims hearings during the 2025–2026 judicial calendar. Under the pilot protocol, the algorithm clustered hearings based on case category and estimated judicial time, reducing the average interval between listing allocation and the initial hearing date from 14 weeks to 8 weeks. Judicial sitting hours and scheduled hearing durations remained unchanged throughout the pilot period. The pilot evaluation recorded an 82% judicial satisfaction rating regarding daily court timetable predictability.
Statement
Litigants participating in small claims hearings across the six designated North East county courts experienced faster overall case resolutions from initial claim submission to final judicial determination during the pilot period.
Step-by-Step Logic Dissection
- Propositional Deconstruction:
- Subject: Litigants participating in small claims hearings across the six designated North East county courts.
- Assertion: Experienced faster overall case resolutions from initial claim submission to final judicial determination.
- Context: During the pilot period.
- Textual Evidence Audit:
- The text confirms that the average interval between listing allocation and the initial hearing date fell from 14 weeks to 8 weeks.
- The text confirms that judicial sitting hours and scheduled hearing durations remained unchanged.
- The text confirms an 82% satisfaction rating regarding timetable predictability.
- Identification of Evidentiary Gaps:
- Does the text disclose the duration between initial claim submission and listing allocation? No. Administrative backlogs prior to listing could have increased.
- Does the text disclose what transpired after the initial hearing date? No. Hearings may have been adjourned, judgments may have been reserved, or multi-session hearings may have occurred.
- Does the text measure total case lifecycle from submission to final judicial determination? No.
- Plausibility vs. Entailment Assessment:
- Plausibility: It is intuitive, reasonable, and highly plausible that shaving six weeks off the scheduling interval would reduce overall time from claim submission to final resolution.
- Deductive Reality: The passage never mentions total lifecycle duration. If pre-listing delays expanded by eight weeks, overall resolution times would have lengthened despite faster scheduling.
- Definitive Determination: Because the text provides partial operational data but omits total resolution timelines, the claim cannot be confirmed or refuted. The required judgement is Cannot Say.
Worked Example 2: Civil Service Energy Efficiency Retrofitting Scheme
Passage
The Department for Energy Security and Net Zero (DESNZ) allocated £18.5 million in capital grants to retrofit commercial heat pump heating systems across 45 regional administrative hubs. An independent technical evaluation conducted after the first full winter of operation revealed that the installations achieved an average 28% reduction in winter building gas consumption compared to the pre-retrofit baseline year. All retrofitted hubs secured five-year fixed-rate maintenance contracts with accredited regional heating engineering consortiums to service the heat pump equipment.
Statement
The installation of heat pump heating systems across the 45 regional administrative hubs resulted in lower overall winter heating operational expenditure compared to the baseline year.
Step-by-Step Logic Dissection
- Propositional Deconstruction:
- Subject: Installation of heat pumps across the 45 regional administrative hubs.
- Assertion: Produced lower overall winter heating operational expenditure compared to the baseline year.
- Textual Evidence Audit:
- The text explicitly verifies a physical resource reduction: an average 28% reduction in winter building gas consumption.
- The text verifies that capital grants funded the installations and that five-year fixed-rate maintenance contracts were secured.
- Identification of Evidentiary Gaps:
- Heat pumps operate on electricity, not gas. What was the net increase in electrical consumption during the winter months?
- What were the unit tariffs for electricity relative to gas during the evaluated period?
- What was the total financial outlay for the newly established five-year maintenance contracts?
- Does the text provide any balance sheet or expenditure data comparing total baseline heating operating costs to post-retrofit operating costs? No.
- The Plausibility Trap:
- Candidates routinely assume that a 28% reduction in fuel consumption automatically translates into monetary savings. While economically plausible, electricity is historically more expensive per kilowatt-hour than natural gas in the UK. If electrical operational costs and maintenance fees exceeded gas savings, total heating operational expenditure might have risen.
- Definitive Determination: The passage establishes physical fuel efficiency but provides zero financial expenditure figures. Because the assertion is unverified by textual data, the required judgement is Cannot Say.
Comparative Reference Table: "Consistent / Plausible" vs. "Conclusively Proven"
The following reference table summarizes the essential criteria that separate plausible candidate assumptions from rigorous textual proof:
| Analytical Dimension | Consistent / Plausible Assertion | Conclusively Proven Assertion ('True') |
|---|---|---|
| Epistemic Relationship | Statement is compatible with the text; no direct contradiction exists ($T \land P$ is possible). | Statement is strictly entailed by explicit premises ($T \implies P$ is inescapable). |
| Evidentiary Threshold | Highly likely, probable, or represents standard operational practice; missing intermediate steps. | Fully verified across all four propositional components: subject, predicate, scope, and modifiers. |
| Candidate Cognitive Trap | Assuming natural consequences (e.g., faster process = cheaper process; reduced fuel = reduced budget). | Doubting textual facts because real-world exceptions are known to exist outside the passage. |
| Required Verification Test | Can you construct an alternative scenario where the passage is true but the statement is false? If yes $\to$ Cannot Say. | Is it logically impossible for the statement to be false while the passage remains true? If yes $\to$ True. |
| Governing CSVT Determination | Cannot Say | True |
Mastering 'Cannot Say' requires candidates to cultivate mental vigilance, rejecting assumptions of probability in favor of uncompromising textual verification.
A Cabinet Office evaluation of a digital specialist recruitment initiative states: 'The Government Digital Service placed 120 software engineers across five central departments on two-year fixed-term contracts to modernize legacy public databases. All 120 recruits completed an intensive six-week induction on public sector cyber security standards before taking up casework. By the end of their first contract year, 94% of the engineers met or exceeded their performance delivery milestones, and departmental supervisors reported marked improvements in database turnaround.' A candidate evaluates the statement: 'The intensive six-week induction on cyber security standards was the direct cause of the high rate of performance milestone achievement among the recruited engineers.' Which evaluation and logical rationale is correct?
A departmental circular regarding licensing casework operations states: 'The Driver and Vehicle Licensing Agency (DVLA) automated its medical licensing triage workflow, categorizing inbound applications into standard renewals and complex medical reviews. Applications categorized as standard renewals were processed through an algorithmic decision-support tool, reducing standard case handling times from 22 working days to 9 working days. All complex medical reviews were routed manually to senior medical caseworkers for individualized appraisal.' What is the correct judgement for the statement: 'Senior medical caseworkers at the DVLA experienced reduced individual workloads following the implementation of the automated medical licensing triage workflow'?