End-of-Utterance Timeout
End-of-utterance (EOU) timeout controls how long the model waits in silence after a speaker stops talking before it flushes the transcript as final. Tuning this value lets you balance responsiveness against cutting users off mid-thought.
With endpointing=true (the default), eou_timeout_ms sets the trailing-silence window (600 ms if unset). With endpointing=false, it sets the model’s own end-of-utterance timer (800 ms if unset). See Finalization and Endpointing for every finalization trigger and the recommended parameters per orchestration.
How It Works
When speech pauses, Pulse starts a silence timer. If no additional speech is detected within the eou_timeout_ms window, the current transcript segment is returned with is_final: true.
- Lower values: faster turn detection, but more likely to split natural pauses
- Higher values: more tolerant of pauses, but slower finalization
Enabling EOU Timeout
Add eou_timeout_ms to your WebSocket connection query parameters. The value must be an integer from 100 to 10000. If unset, the window is 600 ms with endpointing=true (the default) and 800 ms with endpointing=false.
How to Tune It
For conversational voice agents, start at 1000 ms. Mid-sentence pauses in phone conversations often last around 800 ms, so shorter windows close the final before the speaker has finished.
- Decrease only for short, one-word answers where speed matters more than sentence completeness
- Increase for meeting or dictation workflows where speakers pause longer
Tuning Guide
Trade-offs
Example
A voice agent needs one final per caller sentence:
A meeting transcription system should wait for natural pauses: