This guide only applies to cascading agents. If you are using speech-to-speech models, this feature does not apply.
Transcription Modes
Optimize for Speed
Uses the latest interim results with a low endpointing setting for downstream processing. Best latency, slightly less accurate on entities like numbers and dates.
Optimize for Accuracy
Uses results with a higher endpointing setting, waiting longer with more context to generate more accurate transcripts. Incurs ~200ms additional latency.