
Audio Tab showing language selector with English as primary, Deepgram for Speech-to-Text, ElevenLabs for Text-to-Speech with voice tuning sliders
Languages
Set the languages your agent can understand and speak. Pick a primary language and add secondary languages for multilingual conversations.
Language selector showing English as the primary language, with Dutch and Hindi added as secondary languages
- Primary Language is marked with
(Primary)and is the language your agent uses at the start of every conversation. The main prompt and multilingual settings are tied to this language. - Secondary Languages allow the agent to understand and respond when a caller switches languages mid-call.
- Click + Add Language to add more languages.
- Remove any language by clicking the x next to it.
Changing the Primary Language
Click the crown icon next to any secondary language to make it primary. A tooltip will confirm the action, for example “Make Hindi primary”. This sets the selected language as the default for the main prompt and multilingual settings.
Clicking the crown icon on Hindi shows a tooltip to make it the primary language
Supported Languages
Speech-to-Text
Controls how your agent converts the caller’s spoken words into text before the LLM processes them. For multilingual agents, each language can have its own STT provider and model. Select a language tab to configure its transcription settings independently.
Speech-to-Text configuration with Azure selected as provider and model, and a Keywords field showing Bruce:100
Provider and Model
Choose a transcription provider from the Provider dropdown, then pick the specific model from the Model dropdown.Keywords
Boost recognition accuracy for specific words the transcriber might miss, such as brand names, product names, or technical terms. Enter keywords in the formatword:boost_value (e.g., Bruce:100).
Text-to-Speech
Controls how your agent sounds when speaking to the caller. For multilingual agents, each language can have its own TTS provider, model, and voice. Select a language tab to configure voice settings independently.
Text-to-Speech configuration with Sarvam as provider, Bulbul v2 as model, and Anjura voice selected, along with voice tuning sliders
Provider, Model, and Voice
Select a Provider
Pick a Model
eleven_turbo_v2_5 for low latency).Choose a Voice
Browsing Voices
Click the Voice dropdown to see a searchable list of all available voices. Filter by gender using the All, Male, Female, and Neutral tabs. Each voice shows a play button so you can preview it before selecting.
Voice selector dropdown showing a searchable list of voices with gender filter tabs and play preview buttons
Preview Welcome Message
Click Preview welcome message to hear the selected voice speak your agent’s welcome prompt (configured in the Agent Tab). This lets you test how the voice sounds before going live.Voice Tuning Parameters
Fine-tune your agent’s voice using the sliders below the voice selector. Available parameters may vary by provider.Adding and Cloning Voices
Click the Add Voice + button in the Text-to-Speech section to add a custom voice by ID or clone one from an audio sample.Add a Voice by ID
Use this when you already have a voice ID from your provider’s voice library.
Add Voice dialog with the Add by ID tab selected, ElevenLabs as provider, and a Voice ID input field
Select the Add by ID tab
Choose a Provider
Enter the Voice ID
Click Add voice
Clone a Voice
Create a new voice by uploading an audio recording. Useful for maintaining a consistent brand voice or using a specific person’s voice (with their permission).
Clone Voice dialog with Cartesia as provider, fields for Voice name, Description, Sample language, and a file upload area
Select the Clone Voice tab
Choose a Provider
Enter Voice Details
Select Sample Language
Upload an Audio Sample
Click Clone voice

