tools
Rime Arcana
Rime Arcana is a text-to-speech model for real-time voice apps, generating expressive multilingual speech through Rime’s API.

Rime Arcana converts text into expressive speech for voice agents, conversational applications, customer support, outbound calling, audiobook narration, and other audio experiences. Arcana v3 supports multilingual code-switching, word-level timestamps, streaming, voice controls, and on-premises deployment.
Developers can use it through Rime’s API, dashboard, or supported voice-agent integrations. Rime’s current pricing page lists general Starter and Enterprise pricing, but does not show a separate Arcana v3 rate; Enterprise deployment, support, and custom voice options cost extra.
Features
- Generates expressive text-to-speech audio
- Supports multilingual code-switching across 10 languages
- Provides word-level timestamps for alignment and captions
- Streams audio through HTTP and WebSockets
- Supports real-time voice agents with low latency
- Runs on cloud, VPC, or on-premises infrastructure
- Supports 100+ concurrent generations per machine
- Offers API access with the arcana model ID
Use cases
- Build real-time customer-support voice agents
- Create multilingual outbound-calling systems
- Generate speech for audiobook narration
- Add expressive voices to conversational applications
- Develop voice interfaces with captions and interruption handling
- Deploy TTS for regulated or on-premises environments
Pros
Cons
Pricing
- Starting price
- $0.03 / 1K characters
- Pricing checked
- 2026-09-19
Starter
$0.03 / 1K characters
- Approximately 800 minutes free
- 20 concurrent TTS generations
- Public Slack support
- No credit card required
Enterprise
Custom
- Unlimited concurrent TTS generations
- Unlimited custom TTS voice clones
- SLAs and dedicated support
- Cloud, on-prem, or VPC deployment
- BAA (HIPAA) and SOC 2 reports