tools
Deepgram
Deepgram provides APIs for speech-to-text, text-to-speech, voice agents, and audio analysis.

Deepgram is a developer platform for processing speech and building voice applications. Its APIs support real-time and batch transcription, speech synthesis, conversational voice agents, and audio analysis.
Developers, product teams, contact centers, and enterprises use it for transcription, voice interfaces, call analysis, and conversational systems. Usage is billed by audio, characters, tokens, or agent minutes; enterprise deployments and custom models require contacting sales.
Features
- Real-time and pre-recorded speech-to-text APIs
- Text-to-speech with Flux TTS, Aura-2, and Aura-1
- Voice Agent API with turn detection and interruption handling
- Audio Intelligence for summarization and conversation analysis
- Speaker diarization, redaction, keyterm prompting, and smart formatting
- Custom speech-to-text models for proprietary datasets
- Self-hosted and private-cloud deployment options
Use cases
- Transcribe live calls, meetings, podcasts, and recorded audio
- Build conversational voice assistants and customer-service agents
- Generate spoken responses for voice applications
- Analyze conversations for summaries, entities, sentiment, and intent
- Add speech recognition to contact-center and business workflows
Pros
Cons
Pricing
- Starting price
- Free $200 Credit then pay-as-you-go
- Pricing checked
- 2026-09-19
Pay As You Go
Free $200 Credit then pay-as-you-go
- No minimums
- No expiration
- All endpoints in public models
- Community & Discord support
- Standard Uptime
Growth
$4K+ / year
- Pre-paid credits for the year
- Credits redeemed against actual usage
- All endpoints in public models
- Higher concurrency limits
- Community & Discord support
- Standard Uptime
Enterprise
Contact Sales
- Large-volume deployments
- Data or deployment requirements
- Additional support needs