Tools
Deepgram vs Rime Arcana
Deepgram or Rime Arcana? Their plans monthly and yearly, the capabilities their makers state, security and latest updates, side by side — read from the makers' own pages.
In inglese
In short
- Both: Text to speech, Many languages, API
- Only Deepgram states: Speech to text
- Only Rime Arcana states: Voice cloning
Deepgram
Deepgram provides APIs for speech-to-text, text-to-speech, voice agents, and audio analysis.
Plans
Pay As You Go
Free
- Free $200 Credit then pay-as-you-go
- All endpoints in public models
- Speech-to-Text: Up to 50 REST / 150 WSS / 5 Whisper Cloud
- Text-to-Speech: Up to 45 REST + WSS
- Voice Agent: Up to 45 WSS
- Audio Intelligence: Up to 10 REST
Growth
- $4K+ / year
- Save up to 20% with pre-paid credits for the year
- All endpoints in public models
- Speech-to-Text: Up to 50 REST / 225 WSS / 5 Whisper Cloud
- Text-to-Speech: Up to 60 REST + WSS
- Voice Agent: Up to 60 WSS
- Audio Intelligence: Up to 10 REST
Enterprise
Price on request
- For businesses with large volumes, data or deployment requirements, or support needs
Prices checked 2026-09-26 by web search.
Capabilities
- Text to speech — “Text-to-speech built for live conversations.” source
- Speech to text — “Real-time transcription, production accuracy” source
- Many languages — “Nova models support 45+ languages” source
- API — “Deepgram unifies speech-to-text, text-to-speech, and LLM orchestration into a single API, reducing complexity, latency, and cost.” source
Latest updates
- Nova-3 Improved Models for Flemish, German (Switzerland), Lithuanian, and Portuguese
Released improved Nova-3 monolingual models for four existing languages, enhancing transcription quality for batch and streaming workloads.
- Nova-3 Improved Models for Danish, Estonian, Flemish, Italian, Lithuanian, Macedonian, Polish, Urdu, and Vietnamese
Released improved Nova-3 monolingual models for nine existing languages, enhancing transcription quality for batch and streaming workloads.
- Nova-3 Pharma: Speech-to-Text for Pharmaceutical Use Cases (English)
Released nova-3-pharma, a new Nova-3 model purpose-built for pharmaceutical vocabulary, with a focus on accurate drug-name recognition.
Rime Arcana
Rime Arcana is a text-to-speech model for real-time voice apps, generating expressive multilingual speech through Rime’s API.
Plans
Starter
- starting at $0.03 / 1K characters
- $0.05 / 1K characters
- ~800 minutes free (about 800k characters)
- 20 concurrent TTS generations
- Public Slack support
Enterprise
Price on request
- Unlimited concurrent TTS generations
- Unlimited custom TTS voice clones
- SLAs + dedicated support
- Cloud, on-prem, or VPC
- BAA (HIPAA) and SOC 2 reports
Prices checked 2026-09-26 on the maker’s page.
Capabilities
- Text to speech — “Arcana is a multimodal, autoregressive text-to-speech (TTS) model that generates discrete audio tokens from text inputs.” source
- Voice cloning — “Unlimited custom TTS voice clones” source
- Many languages — “Arcana v3's TTS AI supports 10 languages out of the box, with more coming soon.” source
- API — “It’s production-ready and API-accessible from day one.” source
Security
- SOC 2 Type II — “As of May 2025, we are SOC 2 Type 2 compliant .” source
- SOC 2 — “As of May 2025, we are SOC 2 Type 2 compliant .” source
- HIPAA — “As of February 2024, we are HIPAA compliant .” source
- Not trained on your data — “We never use customer text or audio to train our models” source
Latest updates
- Arcana is Now Unlimited
Removal of character limits enables longer speech passages; Arcana accommodates WAV, PCM, and MULAW, adjustable sampling rates, and text normalization.