
The assistant editor.
Name and avatar
Give the assistant a name you’ll recognise in the calls list — if you run several lines, name them after the line rather than the persona. Pick an avatar to tell them apart at a glance.Prompt
The system prompt is the assistant’s instructions. Keep it specific about what it should do and how it should sound. You can use variables anywhere in it, so the prompt knows the time, the caller, and anything your systems told it:Voice and models
Browse voices and hear them before choosing. For multilingual calls, Soniox
handles language switching mid-conversation; set language hints for the languages
you expect.
You can also set a fallback model, so a call continues on a second provider if
the first one fails.
What it will cost
As you change models, the editor shows an estimated cost per minute, broken down by the language model, the voice and the transcription. Swap a model and the figure moves before you save anything. Anything running on your own provider key shows as BYOK and isn’t billed by Qall.Voice settings
Pick a voice from the list, or paste a voice id from your provider for a cloned or custom voice. ElevenLabs voices add three more controls:
For phone calls, moderate stability and low style usually sound most natural —
heavy styling reads as theatrical down a phone line, and costs time on every
turn.
Speech-to-text settings
Keyterms are worth the five minutes. If your callers say a brand name the
transcriber keeps mangling, adding it here fixes it more reliably than anything
in the prompt.
Conversation behaviour
These control how the assistant handles the back-and-forth. The defaults work for most phone calls — change them if callers tell you something feels off.Turn-taking
The assistant decides when you’ve finished speaking rather than waiting for a fixed silence. It commits after 0.3 seconds at the earliest and 2.5 seconds at the latest. Lower the minimum for a snappier feel; raise the maximum if your callers pause mid-sentence to think. Two detectors are offered. Audio model is the one to use — it listens to how you’re speaking and decides from that. Text model is kept for older setups; it waits for the transcript before deciding, which adds a noticeable pause on providers that only transcribe once you’ve stopped.When the caller goes quiet
Set idle timeout to the number of seconds of silence after which the assistant checks in. It asks up to three times, then says goodbye and hangs up. Leave idle message empty and the assistant writes its own check-in, in whatever language it has been speaking. Fill it in only if you need exact words — what you write is what’s said, in the language you wrote it.Recording
For the consent gate in full — keypad answers, per-language wording, and what
happens when a caller declines — see
Recording and consent.