What is it
Per-Character Pricing is a billing unit where customers are charged per character of text processed — the standard meter for text-to-speech and translation. The vendor bills the text going in, not the audio or translation coming out: a 10,000-character script costs the same whether the voice reads it in six minutes or eight, which makes the unit fully countable before a single second of audio is generated.
The text-to-speech cluster carries the unit most directly. LMNT sells character allowances (200K on the $10 Indie tier, 5.7M on the $199 Premium) with per-1,000-character overage that falls from $0.05 to $0.035 as tiers climb; Hume AI ladders its Octave TTS from a 10K free allowance to 10M characters on the $500 Business tier, with overage descending from $0.15 to $0.05 per 1,000; PlayHT sold roughly 3M characters a year on its Creator plan with $4-per-10,000 overage until Meta’s July 2025 acquisition wound the product down. ElevenLabs and Hedra meter the same unit one abstraction up, through credits — Hedra’s wallet prices speech at exactly 15 credits per 1,000 characters alongside per-second video.
The API end of the cluster quotes the character in bulk. MiniMax prices its Speech 2.8 model at $60 per 1M characters (turbo) or $100 per 1M (HD); xAI’s Grok text-to-speech runs $15 per 1M characters even though its language API bills tokens; and Sarvam AI, India’s sovereign-model vendor, quotes its Bulbul TTS in rupees at ₹15 per 10,000 characters (v2) or ₹30 (v3). Speechmatics sits across both worlds, billing its TTS per character beside per-hour speech-to-text.
Beyond speech, the character is also the translation meter. DeepL bills its translation API per source character — $25.00 per 1,000,000 on the classic API Pro model — even as its consumer apps sell by the seat, and Unbabel’s Widn.ai caps free translations at 1,500 characters each. On the writing side, Rytr gates its free tier at 10,000 characters a month before flat “unlimited” plans take the meter away entirely.
How it works
There are two dominant structures. The subscription-plus-overage shape (bill = tier fee + max(0, characters − included) × rate per 1k) rules the self-serve TTS tier ladders; the flat per-million-character rate shape (bill = characters ÷ 1,000,000 × rate) rules the developer APIs. The levers:
| Lever | What it controls | Example from the corpus |
|---|---|---|
| Included allowance | Effective price floor per character | LMNT: 200K chars at $10 ≈ $0.05/1k; Hume Pro: 1M chars at $70 |
| Overage rate ladder | Reward for committing to a bigger tier | Hume: $0.15 → $0.12 → $0.10 → $0.05 per 1k as tiers rise; LMNT: $0.05 → $0.035 |
| Flat per-million rate | Predictable API pricing without tiers | MiniMax Speech $60–$100/1M; xAI Grok TTS $15/1M; DeepL API $25.00/1M |
| Fixed base fee | Revenue floor on a usage product | DeepL: $5.49/mo (Pro) or $26/mo (Growth) before any characters |
| Credit translation | One wallet across media types | Hedra: 15 credits/1k characters of speech, plus per-second video |
| Dual units | Splitting batch text from live audio | Hume bills Octave TTS per character but EVI voice sessions per minute |
| Hard caps vs overage | Predictability vs elasticity | LMNT Free and Rytr Free stop at the cap; paid tiers open the meter |
Worked example — pricing an audiobook. A 90,000-word manuscript is roughly 500,000 characters. On LMNT Pro ($49, 1.25M included) it fits inside the allowance — effective cost $49. On Hume AI Starter ($3, 30K included) the same job runs ~470K characters of overage at $0.15/1k ≈ $73.50, while on Hume’s $70 Pro tier it fits in the 1M allowance. The unit is identical; the tier choice moves the bill by 25x — which is why the usage-metric guide treats allowance-sizing, not the headline rate, as the real comparison.
Run that same half-million characters through the flat per-million APIs above and the bill collapses to a single multiplication — the arithmetic in the diagram. The flat APIs are far cheaper per character at scale but expose no allowance to shelter behind: every character bills from the first one, and DeepL’s monthly base fee means low-volume API users still pay a premium per effective character. You can model the credit-wallet version of this math with the ElevenLabs pricing calculator.
Companies using this
12 in-corpus companies meter characters, splitting into self-serve TTS ladders, per-million-character developer APIs, translation meters, and a single character-capped free writing tier. The table below sorts by pricing model, billing units, and free-tier availability.
Patterns observed
The defining self-serve pattern is the falling-rate ladder: every multi-tier TTS vendor prices the marginal character cheaper as the subscription grows — Hume’s overage drops 3x from Starter to Business, LMNT’s 30% from Indie to Premium. The tier ladder is a volume-discount curve wearing subscription clothing — a bigger commitment buys a cheaper unit, the same shape covered in the usage-metric guide.
A second pattern separates the self-serve tiers from the flat-rate developer APIs. Where LMNT and Hume ladder the rate, MiniMax, xAI, and DeepL publish one per-million number and let volume do the work — no allowance to size, no tier to jump. The split is a segmentation move: creators and studios want a bundled subscription with a comfort-blanket allowance, while developers want a raw per-unit rate they can multiply in a spreadsheet. Several vendors run both — Speechmatics and Sarvam expose the character as a bare API rate while ElevenLabs wraps it in a subscription — because the same meter serves two buyers who want it packaged differently.
A third pattern is that voice cloning travels free with the meter. LMNT includes unlimited clones on every tier including free, and ElevenLabs ships instant cloning from its $6 Starter, because cloning drives character volume rather than competing with it. And a fourth, quieter pattern: the pure character meter is giving way to credit wallets at the platform end. ElevenLabs denominates everything in credits that map closely to characters; Hedra prices speech, video seconds, and image megapixels through one fungible balance. The character survives as the counting rule inside the wallet even where it disappears from the price card — the mechanics the prepaid-credit guide unpacks in detail.
Counterexamples & variants
Rytr is the cleanest counterexample: it uses the character meter only to define the free tier’s edge (10,000/month, hard cap), then abolishes it — $9/month buys “unlimited generations” with the meter removed entirely. The bet is that for AI writing, predictability sells better than elasticity, the opposite conclusion from the TTS cluster. PlayHT walked into the same tension from the other side: its “Unlimited” $49 plan undercut its own character meter before Meta acquired the company and withdrew self-serve pricing altogether — a reminder that an unmetered flat tier can quietly cannibalize the per-character ladder it sits above.
DeepL is the boldest variant: it runs the character meter and a per-seat meter simultaneously on one translation engine — characters for the developer API, seats for the knowledge-worker apps ($8.74 to $57.49 per user). The wrinkle is the fixed monthly base fee ($5.49 on API Pro, $26 on Growth) that sits under the per-character usage, which makes the API cheap at scale but expensive to start — at low volume that floor can dominate the bill. The character here is a genuine usage meter, but not pure pay-as-you-go — a pattern the invoicing guide flags as the “usage product with a subscription floor.”
Hume AI shows the unit’s boundary inside one product line: batch Octave TTS bills per character, but live EVI voice sessions bill per minute — when latency and turn-taking enter, the input text stops being the cost driver and the clock takes over, the same split that defines media-minute pricing. And Sarvam AI is the multi-meter extreme: it bills LLM inference per token, speech per audio hour, and TTS/translation per 10,000 characters, all in one INR-denominated stack — so the “cost driver” flips by workload. For a voice product the character and audio-hour meters dominate and the tokens are nearly free, which is exactly the blended-cost trap the character meter alone can hide.
What this means for buyers vs vendors
For buyers
Estimate in characters, not words — English runs ~5–6 characters per word, so a “50,000-word” job is a ~300,000-character job — and ask what the meter counts: whitespace, SSML markup, and regenerations typically all bill. Then match the pricing shape to your volume. On self-serve TTS tiers, shop the allowance, not the rate: the included-character ladders are where the real price differences live, and the falling overage curves mean upgrading one tier is often cheaper than paying overage on the current one (the diagram’s Hume case above). At high, steady volume the flat developer APIs win outright — but watch for a fixed API base fee, which taxes low-volume use.
If your workload spans meters, price the dominant unit first. A voice or translation pipeline on Sarvam AI is a character-and-audio-hour bill with almost-free tokens; a live conversational agent on Hume is a per-minute bill where the character meter barely registers. Estimating from the headline character rate alone will mislead you on any product that also charges by the clock. The usage-metric guide works through how to identify which unit actually drives a mixed-meter bill.
Finally, treat the wallet as a character meter in disguise. When ElevenLabs or Hedra quote credits, convert to characters before comparing — the per-1k exchange rate is what lines a credit price up against a per-1k TTS quote. Otherwise a credit price looks cheaper or dearer than it is.
For vendors
The character is the most legible meter in audio and translation AI — buyers can count it in a text editor — so protect that legibility: publish the counting rules, keep one rate per tier, and put the volume discount in the ladder rather than in negotiation, the way LMNT does. Decide deliberately whether you are selling a self-serve subscription (ladder the rate, size the allowance) or a developer API (one flat per-million rate, no allowance games); DeepL shows you can run both, but only if each buyer sees the packaging built for them rather than a compromise between the two.
If you span media types, follow Hedra and keep the character as the speech-side exchange rate inside the credit wallet rather than inventing a new opaque unit — the prepaid-credit mechanics guide covers how to keep that exchange rate visible so buyers still trust the meter. And if part of your product is latency-bound, split the meter explicitly (characters for batch, minutes for live) the way Hume does, rather than averaging the two into a number neither workload can forecast.
Watch the base-fee trade-off. A fixed API floor turns volatile per-character revenue into a predictable stream, but it dominates the bill at low volume and filters out exactly the small developers who might have grown into scale. If you add one, surface an honest “your effective per-character cost at this volume” estimate so it doesn’t arrive as bill shock on the first invoice — the same discipline the invoicing guide recommends for any usage product with a subscription component.
| Company | Product | Pricing model | Billing units | Free tier | Verified |
|---|---|---|---|---|---|
| DeepInfra | Serverless inference cloud — per-token LLM/embedding APIs, per-image and per-minute media models, per-hour on-demand GPU containers, and reserved DeepCluster GPU clusters | No | 2026-07-21 | ||
| DeepL | AI translation, writing, and translation API | Yes | 2026-07-23 | ||
| ElevenLabs | Voice AI platform across ElevenCreative, ElevenAgents, and ElevenAPI | Yes | 2026-06-30 | ||
| Groq | GroqCloud — LPU-based ultra-low-latency inference API for Llama, GPT-OSS, Qwen, Whisper transcription, and Orpheus text-to-speech | Yes | 2026-07-21 | ||
| Hedra | AI video, avatar, image, and audio generation platform (Hedra Studio + API) | Yes | 2026-06-04 | ||
| Hume AI | Empathic Voice Interface (EVI) + Octave TTS + expression-measurement APIs | Yes | 2026-07-23 | ||
| LMNT | Low-latency AI text-to-speech (TTS) API with voice cloning | Yes | 2026-06-04 | ||
| MiniMax | Foundation models, Hailuo video & per-token API | Yes | 2026-07-23 | ||
| PlayHT | Text-to-speech & voice cloning API (PlayAI) | Yes | 2026-06-09 | ||
| RunPod | GPU cloud marketplace — Secure Cloud and Community Cloud Pods, Serverless endpoints, and persistent storage | No | 2026-07-22 | ||
| Rytr | AI writing assistant for short-form marketing copy and content | Yes | 2026-06-07 | ||
| Sarvam AI | Sovereign Indic LLM, speech & translation APIs | Yes | 2026-07-23 | ||
| Speechmatics | Speech-to-text and text-to-speech APIs with per-hour usage pricing | Yes | 2026-07-06 | ||
| Together AI | AI Acceleration Cloud — serverless inference, dedicated endpoints, GPU clusters, Code Sandbox, fine-tuning | Yes | 2026-07-21 | ||
| Unbabel | AI + human (LangOps) translation platform; Widn.ai self-serve AI translation | Yes | 2026-06-08 | ||
| xAI | Grok API and agentic AI stack | Yes | 2026-07-21 |
Explore this theme in the knowledge graph
FAQ
What is per-character pricing?
Per-character pricing is a billing unit where the customer is charged per character of text processed — the input text, not the output audio or translation. It is the standard meter for text-to-speech (ElevenLabs, LMNT, PlayHT, Hume, Speechmatics, MiniMax, xAI's TTS, Sarvam's Bulbul) and for translation (DeepL's API, Unbabel's Widn.ai), and it gates the free tier of AI writing tools like Rytr.
Why do TTS vendors bill characters instead of audio minutes?
Characters are countable before generation — a buyer can paste a script and know the cost exactly, while audio length varies with voice speed and pauses. The trade-off is that characters measure input volume, not synthesis difficulty, so vendors layer model multipliers or credit systems (ElevenLabs credits, Hedra's 15 credits per 1,000 characters) on top.
Which companies use per-character pricing?
Twelve in this corpus: ElevenLabs, Hedra, Hume AI, LMNT, MiniMax, PlayHT, Rytr, Sarvam AI, Speechmatics, Unbabel, xAI, and DeepL. TTS vendors meter characters directly or through credits; DeepL and Unbabel meter translation characters; Rytr uses a character cap as a free-tier gate.
How much does text-to-speech cost per 1,000 characters?
Published self-serve rates in this corpus run from $0.035/1k (LMNT Premium overage) through $0.05–$0.15/1k (Hume's tier ladder) to $0.40/1k (PlayHT's $4 per 10,000 before its wind-down). At the API end, MiniMax Speech is $60–$100 per 1M characters ($0.06–$0.10/1k) and xAI's Grok TTS is $15 per 1M ($0.015/1k).
How does DeepL's per-character translation API price?
DeepL's translation API meters per source character. The classic API Pro model is a $5.49/month base fee plus $25.00 per 1,000,000 characters, with a 500,000-character/month free tier; the newer Growth packaging is $26/month plus $27.50 per 1,000,000 overage. A fixed monthly base fee always sits under the usage, so the API is cheap per character but not cheap to start.
What counts as a character on a TTS bill?
Usually every character submitted — including whitespace, punctuation, and markup — and regenerations bill again. Vendors differ on SSML tags and pause tokens, so the meter can run ahead of the visible script; check the counting rules before estimating from word counts (English averages ~5–6 characters per word).
Related billing units
- Credit-Based BillingA billing unit where customers pre-purchase or are allocated a pool of credits that deplete as they use the product, often at variable rates per feature.
- Token-Based PricingA billing unit common in LLM and AI products, where customers are charged per input and output token processed.
- Per-Seat PricingA billing unit where the vendor charges a fixed fee per named user, regardless of how much each user consumes.
- Per-Resolution PricingA billing unit unique to AI customer-support products, where the vendor charges only when an AI agent resolves a customer issue without escalation.
- Bandwidth-Based PricingA billing unit where customers are charged per gigabyte of data transferred out of the platform.
- Per-Function-Invocation PricingA billing unit where customers are charged per serverless function invocation, often combined with a separate compute-time charge.
- CPU-Hour PricingA billing unit where customers are charged for the CPU time their workloads consume, typically measured in vCPU-seconds or vCPU-hours.
- GB-Hour PricingA billing unit where customers are charged for the memory their workloads consume over time, measured in gigabyte-hours.
- GPU-Hour PricingA billing unit where customers are charged for GPU time consumed, typically measured per-second or per-hour by GPU type.
- Per-API-Call PricingA billing unit where customers are charged per API request, regardless of payload size or processing time.
- Per-GB Storage PricingA billing unit where customers are charged per gigabyte of data stored on the platform per month.
- Media-Minute PricingA billing unit where customers are charged per minute of audio or video processed — used by speech, voice, and video AI vendors.
- Per-Request PricingA billing unit where customers are charged per request served — the generic meter for inference endpoints, search, scraping, and browser infrastructure.
- Per-Event PricingA billing unit where customers are charged per event ingested — the native meter of observability and billing-infrastructure platforms.
- Vector Storage PricingA billing unit where customers are charged for vectors stored or indexed — the storage dimension of vector database pricing.
- Per-Document PricingA billing unit where customers are charged per document processed or generated — common in AI writing, SEO, and document-intelligence tools.
- Per-Page PricingA billing unit where customers are charged per page crawled, parsed, or rendered — the meter for web scraping and document parsing.
- Per-Transaction PricingA billing unit where customers are charged per financial or billing transaction processed — the meter of billing and accounting platforms.
- Active-User PricingA billing unit where customers are charged per monthly or daily active user rather than per provisioned seat.
- Per-Task PricingA billing unit where customers are charged per task an automation or agent executes — Zapier's historical unit, now spreading to AI agents.
- Per-Unit PricingA billing unit used by robotics, hardware AI, and some SaaS companies where the metered object is a physical or abstract 'unit' — a robot deployed, a device sold, or a defined deliverable.
- Workflow Execution PricingA billing unit where each end-to-end workflow or automation run is metered and billed, regardless of the compute steps it contains.
- Per-Message PricingA billing unit where each individual message or reply in a conversation is metered, common in AI chat and voice platforms.
- Per-Invoice PricingA billing unit used by billing infrastructure platforms where each invoice generated or processed is metered as the primary cost driver.
- Per-Action PricingA billing unit where each discrete action taken by an AI agent or automation is metered — common in browser automation and agentic workflow tools.
- Per-Image PricingA billing unit where each AI-generated image is metered, common in image generation APIs and multimodal AI platforms.
- Per-Conversation PricingA billing unit where each complete customer conversation — from first message to resolution — is metered as a single chargeable event.
- Per-Record PricingA billing unit where each data record processed, labeled, or extracted is metered — common in data platforms and web scraping services.
- Per-Word PricingA billing unit common in translation and localization platforms where the metered object is the word count of content processed.
- Per-Video PricingA billing unit where each AI-generated video is metered, common in video generation and synthetic media platforms.
- Milestone-Based PricingA billing unit used in drug discovery and biotech AI where payment is tied to achieving defined research milestones rather than time or compute consumed.
- Per-Outcome PricingA billing unit where payment is triggered by verified outcomes delivered — distinct from outcome-based pricing models, this refers specifically to 'outcomes' as a countable billing unit.
- Per-Datapoint PricingA billing unit where each individual data measurement or signal ingested is metered — common in cloud cost intelligence and ML evaluation platforms.
- Per-Interaction PricingA billing unit where each patient-agent or user-agent interaction is metered, common in healthcare AI and customer engagement platforms.
- Data Licensing PricingA pricing structure where access to proprietary datasets or data assets is licensed separately from the software or services, common in AI training data and clinical data platforms.
- Robot-Hour PricingA billing unit where each hour a robot or autonomous system operates is metered — the robotics equivalent of a GPU-hour.
- Per-Contact PricingA billing unit where each contact or lead in the database is metered, common in AI sales development and outbound automation platforms.
- Per-Mailbox PricingA billing unit where each connected email mailbox or sending account is metered, common in AI outbound sales and email automation platforms.
- Browser-Hour PricingA billing unit where each hour of headless browser compute time is metered, common in web scraping and browser automation platforms.
- Per-Generation PricingA billing unit where each AI-generated creative asset — image, video, or design — is counted as a 'generation' and metered accordingly.
- Per-Ticket PricingA billing unit where each customer support ticket handled by an AI agent is metered — common in AI customer service platforms.
- Per-Log PricingA billing unit where each LLM request log ingested or stored is metered — common in AI observability and evaluation platforms.
- Per-Trace PricingA billing unit where each distributed trace — a complete record of an LLM request chain — is metered, common in AI observability platforms.
- Per-IP PricingA billing unit where each IP address or proxy endpoint allocated is metered — used by web scraping proxy providers.
- Per-Device PricingA billing unit where each hardware device or endpoint connected to the AI platform is metered.
- Per-Case PricingA billing unit used in legal AI platforms where each case or matter processed by the AI is metered.
- Per-Report PricingA billing unit where each AI-generated report or analysis document is metered as a discrete output.