Skip to main content

⚡ Deepgram Aura: Ultra-Fast TTS

Industry-leading speed with ~75ms latency. Purpose-built for real-time phone calls, live chat, and interactive applications where every millisecond matters.

Quick Setup

1

Get API Key

  1. Visit Deepgram Console and create an account
  2. Navigate to API Keys in the dashboard
  3. Create a new API key with TTS permissions
  4. Copy your API key
2

Configure in Qall

  1. Go to AI ConfigurationTTS tab
  2. Select Deepgram as provider
  3. Paste your API key in the TTS API Key field
3

Select Aura Model

Choose Aura-2 for best quality or Aura for proven stability
Free Credits: Deepgram provides $200 in free credits to get started. Perfect for testing and small-scale deployments.

Available Models

🚀 Aura-2

~75ms latencyNext-generation model with improved qualityBest for: Production phone calls, live applications Status: Latest and recommended

⚖️ Aura

~85ms latencyProven stable model with consistent qualityBest for: Stable production environments Status: Battle-tested, reliable
Recommendation: Use Aura-2 for the latest quality improvements and fastest response times.

Available Voices

Deepgram Aura Voices

Asteria

Warm and expressivePerfect for customer service and supportVoice ID: aura-asteria-en

Luna

Soft and melodicGreat for gentle, calming interactionsVoice ID: aura-luna-en

Stella

Bright and clearExcellent for announcements and alertsVoice ID: aura-stella-en

Athena

Intelligent and clearProfessional and articulateVoice ID: aura-athena-en

Orion

Deep and authoritativePerfect for business and professional useVoice ID: aura-orion-en

Helios

Bright and energeticGreat for upbeat, engaging contentVoice ID: aura-helios-en

Perseus

Strong and heroic (Aura-2 only)Commanding presence for leadership contentVoice ID: aura-2-perseus-en

Apollo

Musical and expressive (Aura-2 only)Rich, versatile voice for varied contentVoice ID: aura-2-apollo-en

Phone Call Optimization

📞 Twilio Integration

Deepgram Aura is specifically optimized for phone systems with built-in Twilio compatibility.

Audio Format Settings

Recommended for Twilio
  • Format: G.711 µ-law
  • Sample Rate: 8kHz
  • Best for: Phone calls, VoIP systems
  • Quality: Optimized for voice clarity over networks
For Phone Calls: Always use µ-law encoding at 8kHz sample rate for optimal compatibility with Twilio and other phone systems.

Performance Metrics

⚡ Latency

~75ms end-to-endFrom text input to first audio chunk3x faster than most competitors

🎯 Accuracy

99.9% uptimeEnterprise-grade reliabilityProduction-ready stability

📊 Throughput

High concurrencyScales automatically with demandNo rate limit bottlenecks

API Integration

Configuration Examples

Optimal Settings for Phone Systems
  • Ultra-low latency for real-time conversation
  • Phone-compatible audio format
  • Clear, professional voice

Best Practices

🚀 Optimization Tips

Maximize Deepgram’s Speed Advantage
Send text in optimal chunks for best latency
  • Ideal chunk size: 20-50 words
  • Avoid: Sending entire paragraphs at once
  • Benefit: Faster time to first audio
Keep WebSocket connections alive for multiple requests
  • Pattern: One connection per conversation
  • Benefit: Eliminates connection overhead
  • Implementation: Reuse WebSocket for entire call session
Implement robust error handling for production

Pricing

💰 Simple, Predictable Pricing

Pay-per-character with volume discounts. No hidden fees or subscription tiers.
Free Credits: New accounts receive 200incredits.Averagephonecallcosts 200 in credits. Average phone call costs ~0.03 in TTS usage.

Language Support

Current Limitation: Deepgram Aura currently supports English only. Multi-language support is planned for future releases.

🇺🇸 English Optimizations

Specialized for English-language applications
  • Native English training data
  • Optimized for American English pronunciation
  • Best-in-class quality for English content
  • Perfect for US-based business applications

Troubleshooting

Problem: Response time slower than expectedSolutions:
  • Verify you’re using Aura-2 model
  • Check network connection quality
  • Ensure µ-law encoding for phone calls
  • Monitor concurrent connection limits
Problem: Distorted or unclear audioSolutions:
  • Use correct encoding (µ-law for phones, linear16 for high quality)
  • Verify sample rate matches your playback system
  • Check API key permissions include TTS
  • Test with shorter text chunks
Problem: WebSocket connection terminates unexpectedlySolutions:
  • Implement connection keepalive pings
  • Add automatic reconnection logic
  • Monitor connection health
  • Use exponential backoff for retries

Migration from Other Providers

Key Differences:
  • 3x faster latency (75ms vs 250ms)
  • English-only vs multilingual
  • WebSocket-only vs REST+WebSocket
  • Different voice ID format
Migration Tips:
  • Map ElevenLabs voices to Deepgram equivalents
  • Update API calls to WebSocket format
  • Adjust audio encoding for your use case

⚡ Ready for Ultra-Fast TTS?

Configure Deepgram Aura in your assistant settings and experience the speed difference in your phone calls!