Rime
Enterprise text-to-speech built for high-stakes phone calls, where a mispronounced name loses the customer.
Paid link, we may earn a commission. How this works.
Scored on the same voice-agent rubric as the full platforms, so a building block like this scores low on the axes it does not address. Read its value score against its job.
See how it stacks up · Full rankings →The accuracy specialist. Rime is a voice engine you wire into your own stack (Vapi, LiveKit, Pipecat), pitched at getting names and numbers right on calls you cannot afford to fumble. Per-character pricing from $0.05 per 1,000 characters, with a HIPAA option for healthcare.
About $0.04 to 0.09 for a minute of conversation, once the phone line and the AI are added in.
That's roughly $2.10–5.40 an hour. Plans: $0/mo (Starter).
Pricing
Show the cost breakdown
| What the platform charges to run the agent, before the phone line and the AI usage are added on. | — |
|---|---|
| The step that turns what the caller says out loud into text the AI can read. | — |
| The AI 'brain' that reads what the caller said and works out what to say back. | $0.01 /min |
| The step that turns the AI's written reply back into a spoken voice. | $0.02 /min |
| The phone line itself: the service that connects the call to a real phone number. Usually billed on top of the platform. | $0.01 /min |
| The total you actually pay for one minute of conversation once every piece is added up: the platform, the AI, the voice and the phone line. | $0.04–0.09 /min |
Rime is text-to-speech billed per character, not an all-in platform, so the per-minute figures here are a representative model, not a quoted rate. The pricing page was redesigned between 2026-06-15 and 2026-07-11, and the headline rate was cut again by our 2026-07-30 capture: the Starter card now reads 'STARTING AT $0.03/1K characters', down from $0.05, still with no per-model price lines (previously Mist $0.03, Arcana $0.04, Coda $0.05). The page contradicts itself on that number. The FAQ answer to 'How is Rime priced?' on the same page still says 'Starter begins at $0.05 per 1,000 characters'. We record the card as printed, because it is the primary price display and it is what the dated screenshot shows, and flag the conflict rather than average the two; treat $0.03 as the published rate but provisional until the FAQ copy catches up. The models themselves (Mist v3, Arcana v3, Coda) are still current in the docs, which now advise migrating Arcana traffic to Coda. A minute of agent speech is roughly 900 to 1,000 characters, and an agent speaks for about half of a phone minute, so $0.03/1k works out near $0.014 to 0.015 of TTS per conversation-minute, against $0.023 to 0.025 at the old rate. Wire in your own language model (about $0.01/min) and a phone line such as Twilio (about $0.014/min) and a realistic all-in lands around $0.04 to 0.09; the stored low end of 0.035 survives only if a volume discount takes the rate below the card price. Rime brings no speech-to-text, language model or telephony of its own, so those are not Rime charges. The voice count (300+) and language count (7) are the docs-stated figures (May 2026) and are not separately re-confirmable as single totals; note the pricing FAQ (re-read 2026-07-23, its placeholder copy now resolved) and the 2026-07-15 Series A press release both claim '600+ voices' across '50+ languages', but Rime's own docs still state 7 languages (as of 19 May 2026) and 94 Arcana flagship voices, so we keep the docs-sourced counts until the docs agree and record the larger claim here with attribution. Rime announced a $24M Series A on 2026-07-15, led by M13 with Twilio Ventures, Corazon Capital, Unusual Ventures and Cadenza participating, and named Rafael Valle chief scientist. HIPAA BAA and SOC 2 reports are stated for the Enterprise tier on the pricing page; GDPR is not stated there, so it is left unticked.
Every plan in one place: the monthly fee, what each one includes, and the features it unlocks. Anything beyond a plan's allowance, or on a pay-as-you-go tier, is billed at the per-minute rate above. A blank in the features means the vendor's plan page does not state it for that plan, not that it is unavailable.
| Starter | Enterprise | |
|---|---|---|
| Price | Free | Custom |
| Included | 3,000 minutes | — |
| Plan notes | Pay as you go from $0.05/1k characters (a starting-at rate with volume discounts); 3,000 minutes free to start; 20 concurrent TTS generations; public Slack support. | Custom per-character rates with volume discounts; unlimited concurrency and custom voice clones; SLAs and dedicated support; cloud, on-prem or VPC; BAA (HIPAA) and SOC 2 reports. |
| What each plan unlocks | ||
| Voice cloning | — | Unlimited custom clones |
| Concurrent calls | 20 concurrent TTS generations | Up to unlimited |
| Priority support | Public Slack | SLAs + dedicated support |
- Starter Free3,000 minutes
Pay as you go from $0.05/1k characters (a starting-at rate with volume discounts); 3,000 minutes free to start; 20 concurrent TTS generations; public Slack support.
- Voice cloning
- —
- Concurrent calls
- 20 concurrent TTS generations
- Priority support
- Public Slack
- Enterprise Custom—
Custom per-character rates with volume discounts; unlimited concurrency and custom voice clones; SLAs and dedicated support; cloud, on-prem or VPC; BAA (HIPAA) and SOC 2 reports.
- Voice cloning
- Unlimited custom clones
- Concurrent calls
- Up to unlimited
- Priority support
- SLAs + dedicated support
Each plan bundles a set amount of talk time a month.
Prices in USD as set by the vendor · last checked 2026-07-30 · vendor pricing →
Slide your expected monthly volume to see roughly what Rime would cost.
A rough estimate from Rime's sourced rates, not a quote. Always confirm on the vendor's own pricing page before you commit.
At a glance
- Speech-to-text
- Text-to-speech
- Rime Mist, Rime Arcana, Rime Coda
- Languages
- en, es, fr, pt, de, ja, he
- Integrations
- Vapi, LiveKit, Pipecat, SignalWire, Together AI, Native API / SDK, Rime CLI
Compliance
Our full take
Rime is a voice engine, not a full agent platform. You do not log in and build a phone bot. You take Rime’s text-to-speech and plug it into something else (Vapi, LiveKit, Pipecat) that handles the call. What Rime sells is the voice itself, and specifically a voice that gets the hard words right. Its own line is “built for calls you can’t afford to get wrong,” and that is the honest centre of the pitch: pronunciation accuracy on names, medications, account numbers and addresses, the kind of detail that quietly loses a customer when an agent fumbles it.
The pricing is usage-based and billed per character, so there is no single all-in number to quote, and it got simpler (and dearer at the entry point) in mid-2026. The redesigned pricing page now shows one figure: starting at $0.05 per 1,000 characters, with volume discounts above it. The old per-model rates (Mist at $0.03, Arcana at $0.04) no longer appear anywhere on the page, so the cheap Mist rate is gone as a published price. The models themselves are all still current in the docs: Mist v3 for low latency, Arcana v3 for expressive multilingual work, and Coda, the flagship, which the docs now recommend migrating Arcana traffic to. New accounts start on Starter with 3,000 minutes free and 20 concurrent generations, then pay as they go. Enterprise is a custom per-character rate.
Because every comparison on this site works in per-minute terms, here is the workings. A minute of spoken agent audio is roughly 900 to 1,000 characters. In a real phone call the agent only talks for about half the time, the caller has the other half, so a conversation-minute generates closer to 450 to 500 characters of speech. At $0.05 per 1,000 that is about $0.023 to 0.025 of voice per minute. Add your own language model (call it $0.01 a minute) and a phone line such as Twilio (about $0.014 a minute) and a realistic all-in sits around $0.05 to 0.10 a minute, a touch above where it sat when the $0.03 Mist rate was published. Two things to be clear about: that all-in is a model, not a rate Rime quotes, and the language-model and telephony parts are not Rime charges, they are what you pay other suppliers in the stack Rime sits inside.
One wrinkle worth flagging. “Starting at” is doing real work in that headline: it implies volume discounts the page does not spell out. The page itself has settled since we caught it mid-redesign with placeholder FAQ text on 11 July 2026 (the real FAQ copy was in place at our 23 July re-check), but the advice stands: if you are negotiating, get the current rate card in writing rather than trusting the single published figure, and ask directly whether a cheaper Mist-class rate still exists at volume.
On voices and reach, Rime advertises a library of 300-plus voices, with Arcana v3 carrying 94 flagship voices on its own. The language list is short though: seven languages (English, Spanish, French, Portuguese, German, Japanese and Hebrew, per the docs as of 19 May 2026). A caveat on that: Rime’s pricing FAQ and its July 2026 funding announcement both now claim 600-plus voices across 50-plus languages, figures its own docs do not yet reflect. We keep the docs numbers until the two agree, and if language breadth is your deciding factor, ask Rime for the current list in writing. On the docs’ account, if you need to sound native across twenty markets this is not the engine for that yet. Where it concentrates its effort is English-language realism and getting domain-specific words right, which is a deliberate trade rather than a gap.
Compliance is where Rime is interesting for the regulated buyer, with a caveat. The pricing page states that the Enterprise tier includes a BAA (HIPAA) and SOC 2 reports, alongside on-prem and VPC deployment options. Rime also leans hard into healthcare in its marketing and runs on Oracle Cloud Infrastructure for that market. So the HIPAA claim is sourced and ticked here. SOC 2 is mentioned as “reports” on the Enterprise tier, but we have not yet linked a primary certification letter or stated a Type 1 versus Type 2 level, so both SOC 2 boxes stay unticked until we see the paperwork. GDPR is not mentioned on the pricing page at all, so it is unticked too. If you are in healthcare or finance, treat the HIPAA line as the starting point of a conversation with their sales team, not the finish line, and get the BAA and the SOC 2 report scope in writing before you build.
On the company itself, Rime announced a $24 million Series A on 15 July 2026, led by M13 with Twilio Ventures, Corazon Capital, Unusual Ventures and Cadenza participating, and named Rafael Valle as chief scientist. The release pitches Rime toward an “enterprise-ready speech-to-speech model” and claims a scale of roughly 100 million calls a month, both the vendor’s own framing. Funding is not a product feature, but for a voice engine you wire deep into a call stack, a freshly funded supplier is worth knowing about.
A few practical limits. Rime brings no speech-to-text, no language model and no telephony of its own, so the headline per-character price genuinely is just the voice. That is fine if you are a developer wiring up a stack, and real setup work if you are a non-technical buyer expecting to plug in and go. There is no documented self-serve affiliate or partner programme either, so the link from this page is the plain website, not a tracked one. And the ease-of-use score reflects that Rime is a building block, not a finished product: you need a framework around it.
My read: Rime earns a place on the shortlist when the deciding factor is pronunciation accuracy on high-value, often regulated calls, and you are already building your own agent rather than buying an all-in platform. The per-character pricing is still reasonable at $0.05 per 1,000, though the published $0.03 budget rate is gone and Cartesia now undercuts it on narration cost. The HIPAA path is real, and Mist is built for the low-latency phone use case. Look elsewhere if you need a large multilingual voice library, voice work in languages outside the seven, or a platform that hands you the phone line and the language model in one bill.
The 1 to 10 scores and the latency figure on this page are an editorial preview, our provisional read to get the framework in place, not a measured result. We have not run Rime through our own listening or call tests yet, so there is no Voxrater latency figure here. The pricing, voice and compliance detail is sourced from Rime’s pricing, docs and homepage plus the LiveKit, Pipecat and Vapi integration pages, first captured 2026-05-31 and most recently re-captured 2026-07-23.
Rime compared
Our in-depth pieces that put Rime side by side with the field, with the sourced numbers and a clear pick.
Alternatives to Rime
Other platforms that overlap with Rime on the same kind of work, ranked by how many capabilities they share, then by cheaper all-in cost per minute. Compare any of them side by side on the compare page.
Further reading
Tracking Rime? Get the next test result
We re-test and re-price the platforms we cover. Join the list and the next dated update lands in your inbox.
Newsletter launching soon.
Sources
- Re-captured 2026-07-30 (screenshot in evidence/): the Starter card now reads 'STARTING AT $0.03 / 1K CHARACTERS', down from $0.05, pixel-confirmed on the screenshot. The FAQ on the same page still says 'Starter begins at $0.05 per 1,000 characters', so the page contradicts itself; we record the card as printed and flag the conflict rather than average the two. Enterprise remains custom-priced and still lists BAA (HIPAA) and SOC 2 reports. · captured 2026-07-30
- Re-captured 2026-07-23 (screenshot in evidence/): 'STARTING AT $0.05/1K characters' unchanged; Starter (3,000 free minutes, 20 concurrent, public Slack) and Enterprise (custom, BAA + SOC 2 reports) unchanged; the placeholder FAQ text flagged on 2026-07-11 is resolved, and the live FAQ now claims '600+ voices' with 'real fluency across 50+ languages' (a claim Rime's docs do not yet reflect). · captured 2026-07-23
- Docs re-checked 2026-07-23: the voices page still states 7 supported languages 'as of 19 May 2026' (en, es, fr, pt, de, ja, he; Hebrew Arcana-only) and 94 Arcana v3 flagship voices, so the pricing FAQ's and Series A release's 600+/50+ claims are not yet docs-confirmed; stored voice and language counts stay docs-sourced. · captured 2026-07-23
- BusinessWire press release (2026-07-15): $24M Series A led by M13 with Twilio Ventures, Corazon Capital, Unusual Ventures and Cadenza; Rafael Valle joins as chief scientist; the release positions Rime around 'the world's first enterprise-ready speech-to-speech model', claims Coda ships 'more than 600 voices across 50-plus languages' and a scale of roughly 100M calls a month (vendor claims, not independently verified). · captured 2026-07-23
- Re-captured 2026-07-11 (screenshot in evidence/): the redesigned pricing page shows a single 'STARTING AT $0.05 / 1K characters' rate with no per-model price lines (Mist/Arcana/Coda names no longer appear on the page); Starter still gives 3,000 free minutes; Enterprise custom. Parts of the page carry placeholder FAQ text, so re-check at the next sweep. · captured 2026-07-11
- Docs re-checked 2026-07-11: Mist v3, Arcana v3 and Coda all still current; docs now recommend migrating existing Arcana traffic to Coda; 7 languages unchanged. · captured 2026-07-11
- Rime pricing re-captured 2026-06-15: per-model TTS unchanged (Mist $0.03, Arcana $0.04, Coda $0.05 per 1k chars), Coda the flagship; Starter 3,000 free minutes; Enterprise BAA + SOC 2. · captured 2026-06-15
- Rime pricing page re-captured 2026-06-02 for the quarterly re-verification; pricing reviewed against the live page (screenshot in evidence/). · captured 2026-06-02
- Rime pricing page: Mist $0.03/1k, Arcana $0.04/1k, Coda $0.05/1k chars; Starter 3,000 free minutes, 20 concurrent; Enterprise custom with BAA (HIPAA) and SOC 2 reports, on-prem/VPC · captured 2026-05-31
- New pricing launch (19 Dec 2025): Starter/Growth/Enterprise plans, Mist from $20/million and Arcana from $30/million per-character on Growth · captured 2026-05-31
- Rime voices docs (as of 19 May 2026): 7 languages (en, es, fr, pt, de, ja, he); Arcana v3 has 94 flagship voices; Mist v3/v2, Arcana v3, Coda models · captured 2026-05-31
- Rime homepage: enterprise TTS for healthcare/finance/telecom; named customer outcome stats (Domino's, ConverseNow, Fortune 500 containment) · captured 2026-05-31
- LiveKit Agents integration guide: Rime as a TTS provider (URL updated 2026-07-11; the old /agents/integrations/rime/ path now redirects here) · captured 2026-05-31
- Pipecat integration: RimeTTSService (WebSocket, word-level timing, interruption) and RimeHttpTTSService · captured 2026-05-31
- Vapi docs: Rime listed as a voice/TTS provider · captured 2026-05-31