Skip to main content
These settings make the difference between “obviously a robot” and a conversation people stay in. All of them are per-assistant.

Background audio

Complete silence sounds artificial on the phone. You can layer in:
  • Ambient sound — office, city, forest, crowded room, or keyboard typing, with adjustable volume (or your own uploaded audio).
  • Thinking sound — a subtle sound while the assistant is thinking — during the LLM wait and while the caller waits for a tool — so processing never sounds like a dropped call. It doesn’t play while a tool with Talk While Waiting runs in the background, because the conversation continues during that time.
Hold music is not part of this section: it’s configured directly on the warm transfer tool (built-in track or your own upload), since it only ever plays while a supervisor is being briefed.

Interruption handling

Callers interrupt — good assistants deal with it gracefully.
  • Allow interruptions — on by default; the assistant stops speaking when the caller talks over it.
  • Minimum interruption duration — ignores very short noises so a cough doesn’t cut the assistant off.
  • Adaptive interruptions — backchanneling like “mhm”, “okay”, or “right” is not treated as an interruption, so the assistant keeps talking through natural listener feedback but yields to a real interjection. In compatible realtime modes, select Adaptive under Turn detection to use this behavior.
  • Resume after false interruptions — in supported modes, the assistant resumes its interrupted output when an apparent interruption produces no transcript during the two-second detection window. This control is disabled when interruptions are off.
As a starting point, outbound assistants usually work better when they’re easier to interrupt — a caller who wants to jump in shouldn’t have to talk over a pitch. Reception and front-desk assistants often benefit from a bit more stability, so a stray cough or background voice doesn’t cut the assistant off mid-sentence.

Noise cancellation and voice detection

Under Conversation → Latency, Noise & background voice cancellation is a single on/off toggle. It suppresses environmental noise and competing background voices before speech detection, with the appropriate audio model selected automatically for phone or web calls. Browser echo cancellation is separate and remains active independently of this assistant setting. The Custom voice detection threshold control lives under Conversation → Turn taking. It changes how readily audio is classified as speech; it does not filter the audio. Auto uses 0.50 for pipeline and adaptive voice detection. Server voice detection uses 0.50 for web calls and 0.70 for phone calls. Lower values hear quieter speech but may react to noise; higher values reject more noise but may miss soft voices. A live latency waterfall (STT / LLM TTFT / TTS TTFB / E2E) appears only in the studio web-call test panel — never in the public widget.
Telecom guidance (ITU-T G.114) targets one-way latency under 150 ms for a call to feel natural — a useful benchmark when deciding how aggressively to tune interruption and VAD sensitivity for your use case.
Conversation settings with turn detection, silence and voice detection controls

Assistant Settings → Advanced → Conversation: Turn taking contains Mode, Min. silence (VAD) and Custom voice detection threshold.

Filler phrases & async tools

When a tool call (CRM lookup, availability check) takes seconds, the assistant can bridge the gap. For HTTP API tools, you set this per tool under Voice behavior on the Tools page. It applies to voice conversations:
  • Filler phrase — said when the request starts (“One moment, I’m checking that…”).
  • Talk While Waiting — the request runs in the background, so the conversation continues; the result is woven in once it arrives. If the tool has no filler phrase, a default phrase announces the request.
  • Talk After Action Completed — on by default: the assistant continues once the request finishes. Turn it off to let the action complete silently; the assistant still receives the result and uses it after the caller speaks again.
Without Talk While Waiting, the caller waits until the request finishes, and a thinking sound can fill that pause. Calendar booking tools have their own spoken phrases, which you can customize on the calendar integration in the assistant’s Tools tab. With secondary languages configured, saved tool announcements and filler phrases guide the meaning of a message generated in the caller’s current supported language. The fields remain editable; names, codes and other literal details are preserved. These messages use live generation instead of fixed cached audio. Without secondary languages, the saved wording keeps its existing delivery behavior. Full Duplex may rephrase ordinary announcements, as described under engine modes.

Idle handling

Nobody wants a call that hangs forever in silence:
  • Idle timeout — after this many seconds of silence, the assistant checks in (“Are you still there?”).
  • Idle messages — optional fixed phrases for those check-ins; leave empty for natural LLM-generated ones.
  • Max rounds — after N unanswered check-ins, the assistant says goodbye and hangs up. This also protects your minute balance.
The same idle settings are available through PATCH /api/v1/assistants/{id} and the MCP update_assistant tool. Independently, max call duration caps every call: when reached, the assistant wraps up politely and ends the call.

Voicemail

These settings apply to direct outbound calls — from the dashboard, the API, automations and callbacks. They do not change what happens when a call simply rings out unanswered.
  • Answering machine detection (AMD) — off by default for new assistants. On: before the assistant speaks, the call checks whether a person, a mailbox or a phone menu answered, so people may hear a few seconds of silence before the greeting. Off: the assistant greets as soon as the call is answered, and a mailbox hears the normal greeting. Turn it on where mailboxes are common; keep it off for calls people expect, such as requested callbacks or test calls. With Call Screen handling on, the call still listens for screening services first.
  • Leave a voicemail message — needs answering machine detection. On: at a detected mailbox, the assistant speaks a message, then hangs up. Off: the assistant hangs up immediately, leaving nothing.
  • Voicemail message — the free-text message spoken via TTS when a mailbox is detected. Keep it short and include a callback number. Leave it empty and the greeting / first message is used instead.
The same settings are available through PATCH /api/v1/assistants/{id} and the MCP update_assistant tool. Campaign calls use their own Answering machine detection (AMD), On voicemail and Voicemail message settings in Campaign Settings. These override the assistant for that campaign. Direct calls use Assistant settings → Conversation → Voicemail. An empty message falls back to the greeting. A message counts as left only when playback finishes. An unavailable mailbox cannot receive a message. Automatic IVR navigation enables detection and presses menu digits when workspace Beta Features are enabled. With navigation off, an interactive receptionist remains on the normal conversation path. Call Screen handling takes priority. Turning campaign AMD off disables detection unless automatic IVR navigation or Call Screen handling is enabled.

Ringing

Two independent timeouts control how long a call rings before something happens:
This is separate from the Call transfer tool’s own ringing timeout, which times a single transfer attempt rather than the original call.
Assistant Privacy tab with recording, transcript redaction, caller memory and consent controls

Assistant Settings → Advanced → Privacy: recording and transcript controls are under Recording & PII; the Consent section contains Ask for consent at the start of the call.

For jurisdictions requiring all-party consent (Germany: §201 StGB):
  • Consent announcement — a configurable message played at call start (pre-generated, zero added latency).
  • Consent mode — the caller agrees verbally or by pressing a key (DTMF).
  • On decline — either continue the call, or end it politely.
  • The consent result is stored with the call, audit-proof.
One announcement can cover two separate purposes. Tick them independently in the assistant editor:
  • Recording the call — the recording starts only after consent, never before, and additionally requires Record calls. Asked on every call, because each recording needs its own permission. (In API updates, omitting recording_enabled in a PATCH leaves the setting unchanged, while an explicit boolean updates it.)
  • Remember callers — a granted consent unlocks that caller’s durable memory, so returning callers are recognised without anyone approving them by hand under Audience → Customer memory. Asked once.
A declined consent is never stored as a permanent refusal: the caller is simply asked again on a later call. The memory consent is re-asked only after the retention window under Settings → Memory has lapsed without a call — the expired memory is deleted, so the next call starts fresh.
Each purpose must be named in the announcement. Consent to being recorded is not consent to storing a customer profile — they are different purposes under GDPR. When you tick a purpose, the editor offers matching wording in the assistant’s language; leaving the text empty falls back to a purpose- and language-aware default.If Record calls is on while the announcement does not cover recording, the assistant refuses to record and logs recording_skipped_no_consent — it will never record without a notice.

Recording privacy: redaction and re-transcription

Two more controls sit next to recording, under Settings → Privacy.

Redact PII in transcript

Selected categories of personal data are replaced with [REDACTED:type] before the transcript is stored — data minimisation, not just display masking. It applies to newly finished calls only; transcripts already stored keep their original text. Turning it on starts with four categories selected — email address, phone number, IBAN, and credit card number. The full built-in list is much larger, spanning identity, contact, government-ID, financial, security-credential, and health information, and you can select any combination of it. Beyond the built-in categories, add up to 20 of your own named patterns (regular expressions) for anything specific to your business.

Automatic re-transcription

Independently of the manual Re-transcribe action in History, Re-transcribe recording can run for every recorded call automatically. When it’s on, each finished call’s recording is transcribed again with a higher-accuracy engine once the call ends; the enhanced transcript appears in the call’s detail view next to the live transcript. It requires Record calls to be on.
Automatic re-transcription bills per recorded minute (rounded up) at the workspace’s Re-transcribe rate; current rates are on the Usage page.

Guardrails

Deterministic safety rails on top of the prompt:
  • Manipulation → Prompt injection — blocks attempts to bypass or override system instructions (jailbreak detection in the caller’s speech). Available on every plan, on by default; uncheck to opt out (guardrails.input.jailbreak).
  • Blocked topics — subjects the assistant must refuse; matched output is replaced by your refusal text (Fallbacks & Guardrails).
  • Escalation keywords — if the caller mentions one (e.g. “lawyer”, “emergency”), the platform — not the LLM — triggers an immediate transfer to your escalation number (Fallbacks & Guardrails).
Topic and category filters are enforced in the output path, so they hold even when a clever caller talks the model around its prompt.

Pronunciation & text cleanup

See Models & voices → Speaking style: pronunciation dictionary, markdown/emoji filtering, speaking rate, and output volume.

Troubleshooting

Change one setting at a time and re-test with a realistic call — small, isolated changes are much easier to judge than several at once.