Voice & Language
Voice & Language
Voice Configuration Overview
Language Configuration
Fillers (Natural Speech)
Adding a Language
Basic Configuration
The add_language method configures Text-to-Speech and Speech-to-Text for an agent:
Voice Format
The voice parameter uses the format engine.voice:model where model is optional:
Available TTS Engines
Filler Phrases
Add natural pauses and filler words:
Speech fillers: Used during natural conversation pauses
Function fillers: Used while the AI is executing a function
Multi-Language Support
Use code="multi" for automatic language detection and matching:
The multi code supports: English, Spanish, French, German, Hindi, Russian, Portuguese, Japanese, Italian, and Dutch.
Speech recognition hints do not work when using code="multi". If you need hints for specific terms, use individual language codes instead.
For more control over individual languages with custom fillers:
Pronunciation Rules
Fix pronunciation of specific words:
Set Multiple Pronunciations
Voice Selection Guide
Choosing the right TTS engine and voice significantly impacts caller experience.
Use Case Recommendations
Choosing an Engine
Providers differ in latency, quality, cost, and language coverage. For a comparison of all supported providers and guidance on choosing between them, see Voices and languages — then audition candidate voices with the voice widget on each provider’s page.
Testing and Evaluation Process
Before selecting a voice for production:
- Create test content with domain-specific terms, company names, and typical phrases
- Test multiple candidates from your shortlisted engines
- Evaluate each voice: Pronunciation accuracy, natural pacing, emotional appropriateness, handling of numbers/dates/prices
- Test with real users if possible — internal team members or beta callers
- Measure latency in your deployment environment
Dynamic Voice Selection
Change voice based on context:
Language Codes Reference
Supported language codes: