
Audio Tab showing language selector with English as primary, Deepgram for Speech-to-Text, ElevenLabs for Text-to-Speech with voice tuning sliders
Languages
The strip at the top of the tab, labelled Settings for, selects which language the voice and transcription settings below apply to. The language marked(Default) is the one your agent opens every conversation in.

Language selector showing English as the default language, with Dutch and Hindi as additional languages
When the Agent Switches Language
For multilingual agents, a summary under the strip recaps the setup — for example “Starts in English, can move to Hindi, Dutch when …” — and ends in a dropdown that controls when a switch is allowed: requested or auto detected (the default) or the caller requested for it. See When the Agent May Switch for what each option means, and How Language Switching Works for exactly what counts as “the caller speaking another language” — why a stray English word or a read-out order number won’t trip a switch.Supported Languages
See the Multilingual Support guide for the full list of languages Bolna can be configured with.Voice
Controls how your agent sounds when speaking to the caller. For multilingual agents, each language can have its own voice provider, model, and voice. Select a language in the strip above to configure its voice settings independently.
Text-to-Speech configuration with Sarvam as provider, Bulbul v2 as model, and Anjura voice selected, along with voice tuning sliders
Provider, Model, and Voice
Select a Provider
Pick a Model
eleven_turbo_v2_5 for low latency).Choose a Voice
Browsing Voices
Click the Voice dropdown to see a searchable list of all available voices. Filter by gender using the All, Male, Female, and Neutral tabs. Each voice shows a play button so you can preview it before selecting.
Voice selector dropdown showing a searchable list of voices with gender filter tabs and play preview buttons
Preview Welcome Message
Click the play button next to the Voice dropdown to hear the selected voice speak your agent’s welcome prompt (configured in the Agent Tab). This lets you test how the voice sounds before going live.Voice Tuning Parameters
Fine-tune your agent’s voice using the sliders under Advanced settings, below the voice selector. Available parameters may vary by provider.Adding and Cloning Voices
Click the Add Voice + button in the Voice section to add a custom voice by ID or clone one from an audio sample.Add a Voice by ID
Use this when you already have a voice ID from your provider’s voice library.
Add Voice dialog with the Add by ID tab selected, ElevenLabs as provider, and a Voice ID input field
Select the Add by ID tab
Choose a Provider
Enter the Voice ID
Click Add voice
Clone a Voice
Create a new voice by uploading an audio recording. Useful for maintaining a consistent brand voice or using a specific person’s voice (with their permission).
Clone Voice dialog with Cartesia as provider, fields for Voice name, Description, Sample language, and a file upload area
Select the Clone Voice tab
Choose a Provider
Enter Voice Details
Select Sample Language
Upload an Audio Sample
Click Clone voice
Supported Languages for Voice Cloning
Both ElevenLabs and Cartesia support cloning in the same languages Bolna supports, plus an additional Indian Multilingual sample-language option.Transcription
Controls how your agent converts the caller’s spoken words into text before the LLM processes them. For multilingual agents, each language can have its own transcription provider and model. Select a language in the strip above to configure its transcription settings independently.
Transcription configuration with Azure selected as provider and model, and a Keywords field showing Bruce:100
Provider and Model
Choose a transcription provider from the Provider dropdown, then pick the specific model from the Model dropdown.Keywords
Boost recognition accuracy for specific words the transcriber might miss, such as brand names, product names, or technical terms. Enter keywords in the formatword:boost_value (e.g., Bruce:100).

