Skip to content
AI.info

Tools

Rime Arcana vs Ursa

Rime Arcana or Ursa? Their plans monthly and yearly, the capabilities their makers state, security and latest updates, side by side — read from the makers' own pages.

In short

  • Both: Text to speech, Many languages, API
  • Only Rime Arcana states: Voice cloning
  • Only Ursa states: Speech to text, Dubbing

Rime Arcana

Rime Arcana is a text-to-speech model for real-time voice apps, generating expressive multilingual speech through Rime’s API.

Plans

Starter

  • starting at $0.03 / 1K characters
  • $0.05 / 1K characters
  • ~800 minutes free (about 800k characters)
  • 20 concurrent TTS generations
  • Public Slack support

Enterprise

Price on request

  • Unlimited concurrent TTS generations
  • Unlimited custom TTS voice clones
  • SLAs + dedicated support
  • Cloud, on-prem, or VPC
  • BAA (HIPAA) and SOC 2 reports

Prices checked 2026-09-26 on the maker’s page.

Capabilities

  • Text to speech — “Arcana is a multimodal, autoregressive text-to-speech (TTS) model that generates discrete audio tokens from text inputs.” source
  • Voice cloning — “Unlimited custom TTS voice clones” source
  • Many languages — “Arcana v3's TTS AI supports 10 languages out of the box, with more coming soon.” source
  • API — “It’s production-ready and API-accessible from day one.” source

Security

  • SOC 2 Type II — “As of May 2025, we are SOC 2 Type 2 compliant .” source
  • SOC 2 — “As of May 2025, we are SOC 2 Type 2 compliant .” source
  • HIPAA — “As of February 2024, we are HIPAA compliant .” source
  • Not trained on your data — “We never use customer text or audio to train our models” source

Latest updates

  • Arcana is Now Unlimited

    Removal of character limits enables longer speech passages; Arcana accommodates WAV, PCM, and MULAW, adjustable sampling rates, and text normalization.

About Rime Arcana

Ursa

Ursa is Speechmatics’ speech-to-text model family for transcribing live and recorded audio across languages, accents, and speakers.

Plans

Free

Free

  • No credit card required
  • $100 credit to get started
  • For developers and early exploration
  • Speech-to-Text: 55+ languages
  • 2 concurrent real-time sessions
  • Text-to-Speech

Pro

  • from $0.129/hr
  • $0.129/hr
  • $0.24/hr
  • $0.40/hr
  • $0.24/hr
  • $0.43/hr
  • $0.16/hr
  • $0.011/1k characters
  • 55+ languages
  • 50 concurrent real-time sessions
  • 10 file jobs per second
  • Multi-region cloud options
  • Low-latency Text-to-Speech
  • Online email support

Enterprise

Price on request

  • All our features, including audio alignment
  • No rate limits
  • Privacy-first deployment options
  • Custom models
  • SaaS or On-premises deployment
  • Prioritized service and support

Prices checked 2026-09-24 on the maker’s page.

Capabilities

  • Text to speech — “Text-to-Speech” source
  • Speech to text — “Real-time speech-to-text is here” source
  • Dubbing — “Our AI model supports 55+ languages for transcription, with 69 pairs supported for AI translation.” source
  • Many languages — “Transcribe 55+ languages with a single model, even when speakers switch between languages naturally throughout a conversation.” source
  • API — “Speech-to-text API built for every language and accent” source
About Ursa