Vai al contenuto
AI.info

tools

Rime Arcana

Rime Arcana is a text-to-speech model for real-time voice apps, generating expressive multilingual speech through Rime’s API.

Rime Arcana

In inglese

Rime Arcana converts text into expressive speech for voice agents, conversational applications, customer support, outbound calling, audiobook narration, and other audio experiences. Arcana v3 supports multilingual code-switching, word-level timestamps, streaming, voice controls, and on-premises deployment.

Developers can use it through Rime’s API, dashboard, or supported voice-agent integrations. Rime’s current pricing page lists general Starter and Enterprise pricing, but does not show a separate Arcana v3 rate; Enterprise deployment, support, and custom voice options cost extra.

Features

  • Generates expressive text-to-speech audio
  • Supports multilingual code-switching across 10 languages
  • Provides word-level timestamps for alignment and captions
  • Streams audio through HTTP and WebSockets
  • Supports real-time voice agents with low latency
  • Runs on cloud, VPC, or on-premises infrastructure
  • Supports 100+ concurrent generations per machine
  • Offers API access with the arcana model ID

Use cases

  • Build real-time customer-support voice agents
  • Create multilingual outbound-calling systems
  • Generate speech for audiobook narration
  • Add expressive voices to conversational applications
  • Develop voice interfaces with captions and interruption handling
  • Deploy TTS for regulated or on-premises environments

Pros

    Cons

      Latest updates

      • Arcana is Now Unlimited

        Removal of character limits enables longer speech passages; Arcana accommodates WAV, PCM, and MULAW, adjustable sampling rates, and text normalization.

      Capabilities

      • Text to speech — “Arcana is a multimodal, autoregressive text-to-speech (TTS) model that generates discrete audio tokens from text inputs.” source
      • Voice cloning — “Unlimited custom TTS voice clones” source
      • Many languages — “Arcana v3's TTS AI supports 10 languages out of the box, with more coming soon.” source
      • API — “It’s production-ready and API-accessible from day one.” source

      Get it

      Security

      • SOC 2 Type II — “As of May 2025, we are SOC 2 Type 2 compliant .” source
      • SOC 2 — “As of May 2025, we are SOC 2 Type 2 compliant .” source
      • HIPAA — “As of February 2024, we are HIPAA compliant .” source
      • Not trained on your data — “We never use customer text or audio to train our models” source

      Pricing

      Prices checked
      2026-09-26

      Starter

      • starting at $0.03 / 1K characters
      • $0.05 / 1K characters
      • ~800 minutes free (about 800k characters)
      • 20 concurrent TTS generations
      • Public Slack support

      Enterprise

      Price on request

      • Unlimited concurrent TTS generations
      • Unlimited custom TTS voice clones
      • SLAs + dedicated support
      • Cloud, on-prem, or VPC
      • BAA (HIPAA) and SOC 2 reports
      Official website