15.3 English Prosody & Suprasegmentals in Sign-to-Voice Interpreting
Key Takeaways
- Role Delineation Task 10A explicitly requires appropriate register, pausing, rhythm, intonation, pitch, and other suprasegmental features in English output, and the Performance Exam blueprint names prosodic production for hearing candidates.
- Producing interpreted English differs from speaking English: it is built from meaning rather than form, carries someone else's stance, and forms hearing decision-makers' entire impression of the Deaf person.
- ASL prominence devices map onto distinct English equivalents — amplitude to stress, topicalization to fronting or clefting, contrastive body shift to contrastive stress, list buoys to explicit enumeration.
- Register deflation is the more damaging mismatch because listeners attribute the simplicity to the Deaf person rather than to the interpretation.
- Interpreter fillers are attributed to the consumer and are usually a symptom of too short a lag; silence while analysis completes costs listeners far less.
NIC Role Delineation Task 10A requires interpreters to "use English proficiently to construct an equivalent message in the target language, including appropriate vocabulary choice, tone, grammar, and syntax, with appropriate use of register, pausing, rhythm, intonation, pitch, and other supra-segmental features." CASLI's Performance Exam blueprint requires hearing candidates to "produce an interpretation that captures prosodic information (e.g., in English: rhythm, volume, pitch, pausing, etc.)… understand and match intent, and incorporate non-verbal cues."
English production is explicitly assessed, and it is the half of the work that many interpreters practise least.
1. Why voicing gets under-trained
Interpreter education tends to concentrate on ASL production, because that is where hearing learners start from zero. English feels like a solved problem — it is the interpreter's first language. But producing interpreted English is a different task from speaking English:
- The source structure is ASL, and English output must be built from meaning rather than tracked from form.
- The interpreter is voicing someone else's message, stance, and register, not their own.
- The output is monologic and continuous, with no interlocutor shaping it.
- In legal, medical, and employment settings, hearing decision-makers form impressions of the Deaf person entirely from the interpreter's voice.
That last point is why voicing quality is an equity issue and not a matter of polish.
2. The suprasegmental inventory
| Feature | What it carries |
|---|---|
| Pitch range and contour | Question versus statement; enthusiasm, doubt, sarcasm, finality |
| Stress placement | Which element is in focus — I didn't say that / I didn't say that |
| Rhythm and rate | Urgency, hesitation, deliberation, narrative pacing |
| Pausing | Unit boundaries, emphasis, thinking time, dramatic weight |
| Volume | Emphasis, emotional intensity, confidentiality |
| Voice quality | Tension, warmth, fatigue, distress |
| Fillers and disfluency | Represent the signer's own hesitation — but interpreter fillers misrepresent it |
3. Mapping ASL prominence to English prominence
ASL and English mark emphasis differently, so voicing is a translation of the prominence system, not a copy:
| ASL device | English equivalent |
|---|---|
| Larger, longer, more forceful sign | Stress, increased volume, lengthened vowel |
| Repetition of a sign | Intensifier, repetition, or a stronger lexical choice |
| Topicalization with raised brows and a hold | Fronting, a cleft construction ("What she wanted was…"), or a pause after the topic |
| Contrastive body shift between two loci | Contrastive stress: "He said yes; she said no" |
| Rhetorical question structure | "The reason is…", "Why? Because…" |
| Sustained non-manual intensity over a span | Sustained vocal intensity over the corresponding clause |
| List buoy structure | Explicit enumeration: "Three things. First… second… third…" |
4. Register matching
Task 10A names register explicitly. A Deaf professional presenting at a conference must be voiced in professional English; a teenager telling a story to a friend must be voiced as a teenager. Two failure modes:
- Register inflation — voicing a casual signer in formal, complete, elevated English. This is the more common error and is usually well-intentioned, but it misrepresents the person and can read as the interpreter improving on them.
- Register deflation — voicing a sophisticated, academically fluent signer in simple, halting English. This is the more damaging error, because hearing listeners attribute the simplicity to the Deaf person. It is a well-documented mechanism by which interpreting quality shapes hearing people's assessments of Deaf professionals' competence.
Neither is neutral. Both are inaccuracies in the same sense that a wrong lexical item is.
5. Disfluency: whose is it?
Distinguish carefully:
- The signer's disfluency — a genuine hesitation, self-correction, or false start — is part of the message and should appear in the English.
- The interpreter's disfluency — "um," "uh," "like," extended vowel fillers, trailing sentences, and restarts caused by your own processing — is noise that the listener attributes to the Deaf person.
The Role Delineation Study names "distracting mannerisms, fillers, anomalies" among the things interpreters must minimize. The most effective remedy is a longer lag: fillers are almost always the sound of an interpreter producing before analysis is complete. Silence while you finish analysing is far less costly than a stream of fillers, and listeners tolerate it easily.
6. Practical voicing technique
- Breathe and project. Under-supported voice reads as uncertainty. Sit or stand so your breath is available.
- Finish sentences. Trailing off mid-clause forces listeners to guess and is one of the strongest markers of an interpreter behind the source.
- Let pauses do work. A pause before a key point creates emphasis that no amount of volume can.
- Use first person consistently, and mark interpreter statements clearly in the third person when you speak as the interpreter — "the interpreter needs a repetition."
- Match speed to content, not to the signer's signing rate. Dense ASL rendered at conversational English speed will lose the listener; English is more word-dense per proposition than ASL is sign-dense.
- Record and review your own voicing. The single most effective practice intervention, and one most interpreters avoid because it is unpleasant.
7. Comprehending English source (Task 10B)
The mirror task requires comprehending English "including appropriate vocabulary choice, tone, grammar, syntax, appropriate use of register, pausing, rhythm, intonation, pitch." Receptive English demands that regularly defeat interpreters:
- Non-native and regional accents in speakers from anywhere.
- Rapid read-aloud text, which is far denser than spontaneous speech.
- Jargon and acronyms in every professional setting.
- Sarcasm and irony, where the literal proposition is the opposite of the intended one — which must be rendered as intended, not literally.
- Idiom and metaphor, which require conceptual interpretation, not lexical substitution.
- Degraded audio from masks, speakerphones, video platforms, and noisy environments.
Preparation, a longer lag, and an early clarification request are the available controls for all six.
A Deaf engineer with a doctorate presents technical findings in sophisticated ASL. The interpreter voices the content accurately but in simple, halting English sentences. What is the consequence?
An interpreter voicing for a Deaf consumer produces frequent fillers — "um," "uh," "like" — that the consumer did not produce. What is the most effective remedy?
A hearing manager says, with heavy sarcasm, "Oh, that's just brilliant." How should this be interpreted into ASL?