Skip to main content
Call settings control how an agent behaves in the audio stream — when it stops talking, how long it waits before responding, what it does when the line goes quiet. Defaults work well for most agents; reach for these when an agent talks over people, pauses awkwardly, or responds too eagerly. All settings live on the agent’s Call Settings tab, in six sections: In-Call Limits, Silence Handling, Voice & Audio, Interruptions, Response Timing, and Supporting LLM.

In-call limits

In-Call Limits with max call duration and end call on silenceIn-Call Limits with max call duration and end call on silence

Silence handling

Silence Handling with nudge on silence and the silence messageSilence Handling with nudge on silence and the silence message
When the caller goes quiet, the agent checks in rather than sitting in dead air:
  • Nudge User on Silence — how long to wait before the agent says something.
  • Use static silence message — on, the agent speaks the fixed Silence message (for example “Hey are you there…”); off, it generates one from the conversation so it fits naturally.
  • If the caller still doesn’t respond after the nudge, the call ends on the End Call on Silence limit above.

Voice and audio

Voice and Audio with background audio, detection sensitivity, noise cancellation, and voice detectionVoice and Audio with background audio, detection sensitivity, noise cancellation, and voice detection

Interruptions and turn-taking

Interruptions with allow interruptions, min words to trigger, and interruption sensitivityInterruptions with allow interruptions, min words to trigger, and interruption sensitivity

Response timing

Response Timing with fast, medium, slow, and custom response modesResponse Timing with fast, medium, slow, and custom response modes
How long the agent waits after the caller stops speaking before it replies. Pick a Response Mode:

Supporting LLM

Supporting LLM with provider, model, mode, filler instruction, temperature, and max tokensSupporting LLM with provider, model, mode, filler instruction, temperature, and max tokens
For agents where every millisecond of response delay matters, enable a second, smaller model that speaks a brief acknowledgement (“Sure, let me check that…”) while the main model composes the full answer. The result is an agent that never leaves a long thinking gap.
Change one setting at a time and re-test in the web-call playground. Interruption and endpointing settings interact — adjusting both at once makes it hard to tell which change helped.