tools
Rime Arcana
Rime Arcana is a text-to-speech model for real-time voice apps, generating expressive multilingual speech through Rime’s API.

In inglese
Rime Arcana converts text into expressive speech for voice agents, conversational applications, customer support, outbound calling, audiobook narration, and other audio experiences. Arcana v3 supports multilingual code-switching, word-level timestamps, streaming, voice controls, and on-premises deployment.
Developers can use it through Rime’s API, dashboard, or supported voice-agent integrations. Rime’s current pricing page lists general Starter and Enterprise pricing, but does not show a separate Arcana v3 rate; Enterprise deployment, support, and custom voice options cost extra.
Features
- Generates expressive text-to-speech audio
- Supports multilingual code-switching across 10 languages
- Provides word-level timestamps for alignment and captions
- Streams audio through HTTP and WebSockets
- Supports real-time voice agents with low latency
- Runs on cloud, VPC, or on-premises infrastructure
- Supports 100+ concurrent generations per machine
- Offers API access with the arcana model ID
Use cases
- Build real-time customer-support voice agents
- Create multilingual outbound-calling systems
- Generate speech for audiobook narration
- Add expressive voices to conversational applications
- Develop voice interfaces with captions and interruption handling
- Deploy TTS for regulated or on-premises environments
Pros
Cons
Latest updates
- Arcana is Now Unlimited
Removal of character limits enables longer speech passages; Arcana accommodates WAV, PCM, and MULAW, adjustable sampling rates, and text normalization.
Capabilities
- Text to speech — “Arcana is a multimodal, autoregressive text-to-speech (TTS) model that generates discrete audio tokens from text inputs.” source
- Voice cloning — “Unlimited custom TTS voice clones” source
- Many languages — “Arcana v3's TTS AI supports 10 languages out of the box, with more coming soon.” source
- API — “It’s production-ready and API-accessible from day one.” source
Get it
Security
Pricing
- Prices checked
- 2026-09-26
Starter
- starting at $0.03 / 1K characters
- $0.05 / 1K characters
- ~800 minutes free (about 800k characters)
- 20 concurrent TTS generations
- Public Slack support
Enterprise
Price on request
- Unlimited concurrent TTS generations
- Unlimited custom TTS voice clones
- SLAs + dedicated support
- Cloud, on-prem, or VPC
- BAA (HIPAA) and SOC 2 reports