AI Summary
About
Retell AI is a conversational voice-agent platform that lets contact centers and developers deploy AI phone, chat, SMS, and email agents that handle tier-one calls. A one-line integration composes the full real-time voice stack — speech-to-text, a large language model, text-to-speech, and telephony — behind a single API, with pre-built templates, simulation testing, call analytics, and webhooks. Founded in 2023 (Evie Wang, Zexia Zhang, Bing Wu, Todd Li, Weijia Yu) and a Y Combinator W24 company, Retell launched in February 2024 and reached roughly $3M annualized revenue within months. It announced a $4.6M seed led by Alt Capital (with Y Combinator, Carya Venture, and 20+ founders/operators) in September 2024 and has since scaled to about $60M ARR by April 2026 — up ~650% year-over-year — powering 50M+ real-time AI phone calls per month for BPOs and enterprises like Everise, Sunshine Loans, and Matic.
For the most current information, visit Retell AI.
Pricing summary : How Retell AI’s pricing model works
Retell AI is true pay-as-you-go: there is no platform fee, no contract, and no minimum. Voice agents cost $0.07–$0.31 per minute, and the headline figure is unbundled — every minute is itemized into Retell Voice Infra ($0.055/min), the LLM you choose (e.g. GPT 5.1 at $0.04/min standard, GPT 5.5 at $0.16/min, Claude 4.6 Sonnet at $0.08/min), text-to-speech ($0.015/min for most voices, $0.04/min for ElevenLabs), and telephony ($0.015/min via Twilio/Telnyx, or free if you bring your own SIP trunk). Retell’s on-page calculator shows a computed all-in estimate of roughly ~$0.115/min for a default GPT-class agent (see Hidden costs). Chat agents are billed per AI message from $0.002+ (e.g. GPT 5.1 at $0.013/msg). New accounts get $10 in free credits and 20 free concurrent calls to build and launch.
The Enterprise plan is custom-quoted and adds volume pricing, a dedicated stable server, HIPAA/BAA, custom MSA/DPA, custom SSO, role-based access control, a high concurrency cap, and 24/7 omnichannel support with a dedicated portal.
Call-agent pricing has two modes on the pricing page. The Default (cascading) mode is the unbundled stack above. A Speech-to-Speech tab prices realtime speech-to-speech engines as a single per-minute line instead: GPT Realtime 1.5 and GPT Realtime at $0.345/minute, and GPT Realtime mini at $0.07/minute.
What makes this different: Retell exposes the entire cost stack instead of hiding it behind a flat per-minute rate. Buyers pick the LLM, voice, and telephony independently and see each layer’s contribution to the per-minute price — turning model selection into a direct cost lever. Calls are metered to the second with no per-call rounding, though billing continues through silence and hold because the STT engine stays active.
Pricing by product
Retell AI (plans)
| Tier | Price | Included | Key mechanics |
|---|---|---|---|
| Pay-as-you-go | $0.07–$0.31/min voice; from $0.002/AI msg chat | $10 free credits, 20 free concurrent calls, full platform access | No platform fee, contract, or minimum; billed on accumulated minutes each cycle |
| Enterprise | Custom pricing (voice and chat) | Everything in pay-as-you-go plus dedicated stable server, HIPAA/BAA, custom MSA/DPA, custom SSO, RBAC, 24/7 omnichannel support with dedicated portal | Volume pricing, no cap on concurrent calls; sales-led, quoted |
Voice agents — component pricing (Default mode)
| Component | Price | Included | Key mechanics |
|---|---|---|---|
| Retell Voice Infra | $0.055/min | Conversation voice engine on every call | The only Retell-margin layer; everything else is chosen |
| Text-to-speech | $0.015/min (Retell Platform, Minimax, Fish, Cartesia, OpenAI voices); $0.040/min (ElevenLabs voices) | — | Voice choice is a direct cost lever |
| LLM (standard / fast tier) | GPT 5.5 $0.16 / $0.32; GPT 5.4 $0.080 / $0.16; GPT 5.2 $0.056 / $0.112; GPT 5.1 $0.04 / $0.08; GPT 4.1 $0.045 / $0.0675 per minute | Claude 4.6 Sonnet $0.08, Claude 4.5 sonnet $0.08, Claude 4.5 haiku $0.025, Gemini 3.5 Flash $0.081, Gemini 3.0 Flash $0.027, Gemini 2.5 Flash $0.035, Gemini 2.5 Flash Lite $0.006 per minute | GPT 5.5, GPT 5.4, GPT 4.1 and Claude 4.6 Sonnet are flagged “Recommended”; cheapest listed is GPT 5 nano at $0.003/min |
| Telephony | $0.015/min (Twilio/Telnyx, US) | Country picker covers ~19 Twilio/Telnyx routes | ”No Charge for Sip Trunking/Custom Telephony” — bring your own trunk and this line is $0 |
Voice agents — Speech-to-Speech mode
| Engine | Price | Included | Key mechanics |
|---|---|---|---|
| GPT Realtime 1.5 | $0.345/min | Realtime speech-to-speech engine | Replaces the separate LLM + TTS lines in Default mode |
| GPT Realtime | $0.345/min | Realtime speech-to-speech engine | Same rate as GPT Realtime 1.5 |
| GPT Realtime mini | $0.07/min | Realtime speech-to-speech engine | ~5× cheaper realtime option |
Chat agents (per AI message)
| Model | Price | Included | Key mechanics |
|---|---|---|---|
| GPT 5.5 / GPT 5.4 / GPT 4.1 | $0.052 / $0.026 / $0.015 per AI msg | Flagged “Recommended” on the page | Message-metered, not minute-metered |
| Claude 4.6 Sonnet / Claude 4.5 sonnet | $0.03 per AI msg | — | Claude 4.5 haiku $0.007/AI msg |
| GPT 5.2 / GPT 5.1 / GPT 5 | $0.018 / $0.013 / $0.013 per AI msg | — | GPT 5 mini $0.004, GPT 5 nano $0.001 |
| Gemini 3.0 Flash / 2.5 Flash / 2.5 Flash Lite | $0.009 / $0.012 / $0.006 per AI msg | GPT 4.1 mini $0.006, GPT 4.1 nano $0.002 | Cheapest listed message rate is $0.001/AI msg |
| Chat add-on — SMS | $0.01/Msg | — | Same SMS rate as call agents |
Add-ons and monthly subscription items
| Item | Price | Included | Key mechanics |
|---|---|---|---|
| Knowledge Base (per-minute) | +$0.005/minute | — | Charged on top of every minute the KB is attached |
| Batch Call | +$0.005/dial | — | Per dial, not per minute |
| Branded Call | +$0.10/outbound call | — | Per-call, the priciest add-on unit |
| Advanced Denoising / Safety Guardrails | +0.005/min each | — | Stackable per-minute extras |
| PII Removal | +0.01/min | — | Compliance add-on billed per minute |
| AI Quality Assurance | $0.10/min | *First 100 minutes free | Nearly doubles a typical ~$0.115/min minute once the free QA minutes are used |
| SMS | $0.01/Msg | — | Per message |
| Retell Phone Numbers | $2.00/month | *No charge for custom phone numbers | Recurring per number |
| Retell SMS | $20.00/month | *No charge for custom telephony SMS | Recurring |
| Concurrency (active calls) | $8.00/concurrency/month | *Free for first 20 concurrency (active calls) | Add or remove anytime; Enterprise custom concurrency starts at 50+ |
| Knowledge Base (monthly) | $8.00/Knowledge Base/month | *Free for first 10 knowledge bases | Separate from the per-minute KB add-on |
| Verified Phone Number | $10.00/Phone number/month | *“One-time $10 fee” | Page lists both a monthly per-number rate and a one-time $10 fee |
Sales motions across products: self-serve PLG for pay-as-you-go (instant signup, $10 credits, no contract needed to start building), and sales-led for Enterprise (volume pricing, compliance, dedicated support).
Hidden costs : What Retell AI users actually pay
The headline $0.07/min is the floor for the cheapest LLM with no extras. Real bills are shaped by which LLM you pick (a 4x swing from GPT 5.1 to GPT 5.5), the voice (ElevenLabs is $0.04/min vs $0.015/min for platform voices), telephony, per-minute add-ons, and monthly subscription items.
| Line item | Monthly cost (illustrative, 1,000 min) |
|---|---|
| Retell Voice Infra ($0.055/min) | $55 |
| LLM — GPT 5.1 standard ($0.04/min) | $40 |
| TTS — platform voice ($0.015/min) | $15 |
| Telephony — Twilio US ($0.015/min) | $15 |
| Add-ons (e.g. knowledge base +$0.005/min) | $5 |
| Extra concurrency (beyond 20, $8/call/mo) | varies |
| Estimated total (~$0.115/min × 1,000) | ~$115 + add-ons |
Other things to budget for: billing runs during silence and hold because the STT engine stays active; branded outbound calls add $0.10 per call; PII removal is +$0.01/min and AI Quality Assurance is $0.10/min after the first 100 free minutes; extra concurrency above the 20 free is $8/call/month; and Retell phone numbers ($2/mo) and verified numbers ($10/mo) are recurring.
Want to estimate your own Retell AI bill? Use the Retell AI pricing calculator to model your costs based on usage patterns.
Pricing evolution : Retell AI pricing history and changes
Cadence
| Period | Price changes | Product / SKU additions | Notes |
|---|---|---|---|
| 2024 H1 | Launch | Pay-as-you-go voice agents | YC W24; per-minute, no contract |
| 2024 H2 | — | — | $4.6M seed; ~$3M ARR |
| 2025–2026 | Model menu expanded | Chat agents, add-ons, monthly items | $10 credits, 20 free concurrency; ~$60M ARR |
Tracked range: 2024–present. Retell has kept a stable true-pay-as-you-go structure since launch; the menu of LLMs, voices, add-ons, and monthly items has broadened while the unbundled per-minute model stayed constant. (Wayback archive access was unavailable for this capture, so historical snapshot dates are based on company milestones rather than archived pricing pages.)
Notable changes
- 2024-02 — Launches (YC W24) with true pay-as-you-go per-minute voice agents and no platform fee. Reaches ~$3M annualized revenue within months.
- 2024-09 — Announces a $4.6M seed led by Alt Capital; pricing stays transparent and self-serve as BPO/enterprise adoption grows.
- 2025–2026 — Expands the LLM and voice menu (GPT-5 family, Claude 4.x, Gemini), adds chat agents (per AI message), an add-on menu (knowledge base, denoising, PII removal, QA, branded calls), and monthly subscription items. Scales to ~$60M ARR by April 2026.
What’s unique : Retell AI’s distinctive pricing mechanics
1. Fully unbundled per-minute stack. Instead of a single opaque per-minute rate, Retell itemizes every voice minute into Voice Infra, LLM, TTS, and telephony. Buyers choose each layer independently and see its contribution, making model and voice selection a direct, transparent cost lever — unusual in a category that usually quotes one blended number.
2. True pay-as-you-go with no contract. No platform fee, no seat minimum, no annual commitment — a deliberate contrast to voice-AI rivals that require annual contracts “before you write a single line of code.” $10 in free credits and 20 free concurrent calls lower the barrier to launch.
3. Second-level metering with honest edge cases. Calls are tracked to the nearest second with no per-call rounding, and Retell publishes exactly when billing applies (during silence/hold, transfers, voicemail, failed calls). This granular, documented metering builds trust around a notoriously fuzzy unit.
Strengths & weaknesses
| Strengths | Weaknesses |
|---|---|
| Radically transparent, unbundled per-minute pricing | Composite bill is hard to predict without modeling each layer |
| True pay-as-you-go — no contracts, fees, or minimums | No flat-rate or committed-use discount on self-serve |
| $10 credits + 20 free concurrency lower launch friction | Billing during silence/hold can surprise new buyers |
| Wide LLM/voice/telephony menu = cost-tuning flexibility | Many small add-ons and monthly items add up quietly |
| Second-level metering with documented billing rules | Enterprise volume pricing is opaque (sales-led only) |
Billing UX : Retell AI billing controls and transparency
- Billing controls — Fully self-serve: sign up, get $10 in credits, and pay as you scale with no card required to start building. Concurrency and add-ons can be added or removed anytime. Enterprise is invoiced under a custom MSA/DPA.
- Usage visibility — The “Estimate Your Cost” calculator on the pricing page takes monthly call minutes and a chosen LLM, Text To Speech voice, telephony route, and Call Agent Add-ons, then breaks the result into named lines — Cost Per Minute, LLM Cost, Retell Voice Infra, TTS Cost, Telephony Cost, Add-ons, and Total per month (default configuration: $0.115/min, $11.5 for 100 minutes). In-product call analytics and transcripts cover after-the-fact usage, and invoices are available in the Billing tab each cycle.
- Component and mode toggles — The “Detailed Component Pricing” block exposes a Default / Speech-to-Speech toggle for call agents and a per-country telephony picker (~19 Twilio/Telnyx routes), so a buyer can price the exact stack before committing. Concurrency is a self-serve line item — 20 active calls free, $8.00/concurrency/month beyond that, addable or removable anytime.
- Payment options — Card-based self-serve credits for pay-as-you-go; Enterprise is invoiced under custom terms. Telephony can be brought via your own SIP trunk to avoid Retell’s pass-through telephony fee.
Strategic wins : Why Retell AI’s pricing decisions worked
1. Transparency as a wedge against contract-first rivals
By leading with “$0, pay only for what you use” against competitors that demand annual contracts up front, Retell turned pricing into a top-of-funnel advantage. The $10 credits and self-serve launch path converted developers fast — a contributor to ~650% YoY growth. See how AI companies structure pricing.
2. Unbundling the stack to align price with value
Itemizing Voice Infra, LLM, TTS, and telephony lets buyers right-size each layer (cheap LLM for simple flows, premium voice where it matters). This makes the per-minute number feel fair and gives Retell a clean margin on its own $0.055/min infra while passing through model costs. Related: outcome-based pricing trends.
3. Per-minute as the value metric buyers already count
Contact centers already think in agent-minutes and cost-per-minute, so pricing per voice-minute maps directly onto the metric Retell displaces — offshore human agents at $0.30–$0.80/min. Undercutting that by 70–95% makes the ROI obvious. See choosing the right usage metric.
Areas to improve : Gaps in Retell AI’s pricing approach
1. Composite bills are hard to forecast
The flip side of unbundling is that a single voice minute draws from four or five separately priced layers plus add-ons, so estimating a monthly bill requires modeling each one. The calculator helps, but a few representative “recipe” bundles with all-in per-minute prices would reduce buyer effort. See bill shock and cost unpredictability.
2. Billing during silence and hold
Charging for silence and hold time (because STT stays active) is honest and documented, but it can surprise buyers benchmarking against per-spoken-word pricing. Clearer up-front framing of effective cost-per-handled-call would set expectations better.
3. No committed-use discount for self-serve
Self-serve buyers get no volume break — only Enterprise unlocks volume pricing, and that path is sales-led and opaque. A published tiered or committed-use discount would give high-volume PLG accounts a reason to consolidate without a sales call.
Monetization stack & signals : how Retell AI builds & buys its revenue engine
Buys 0 Builds 0 3 signal roles
Retell names no monetization vendor and discloses no in-house build — a second-level usage meter clearly runs behind its $0.055/min Retell Voice Infra line, but who built it is unstated. Hiring is enterprise delivery (forward-deployed/deployment strategist), not deal-desk or a pricing PM — a sales-led overlay on a self-serve core, packaging held deliberately stable.
-
“Each call is tracked to the nearest second... No rounding up per call, no inflated bills. Retell Voice Infra is a $0.055/min line item itemized alongside LLM, TTS and telephony — a usage meter clearly exists, but the pricing page and blog disclose no in-house build and name no vendor (Metronome/Orb/Stripe).”
-
“Self-serve $10 credits and card-based pay-as-you-go billed at the end of each cycle on accumulated minutes; the self-serve billing/payments provider is not disclosed.”
-
The revenue-org investment is enterprise delivery (forward-deployed / deployment strategist), not deal-desk or RevOps — the sales-led motion runs on solution delivery onto the self-serve core, with no quote-to-cash or CPQ build evident.
“You will lead implementations from initial discovery through production launch, helping customers redesign business processes around AI while ensuring successful technical delivery... You'll partner closely with Product, Engineering, Sales, and Customer Success while helping shape the future of how enterprises deploy voice AI.”
-
The revenue-org investment is enterprise delivery (forward-deployed / deployment strategist), not deal-desk or RevOps — the sales-led motion runs on solution delivery onto the self-serve core, with no quote-to-cash or CPQ build evident.
“You will lead implementations from initial discovery through production launch, helping customers redesign business processes around AI while ensuring successful technical delivery... You'll partner closely with Product, Engineering, Sales, and Customer Success while helping shape the future of how enterprises deploy voice AI.”
-
The product org is staffed around agent behavior, onboarding and developer tooling — not around packaging or paywalls. There is no dedicated monetization/pricing PM, consistent with a price model (unbundled pass-through + a thin infra margin) that is deliberately stable, not a roadmap surface.
“Own customer outcomes for core product areas such as voice agent behavior, onboarding, testing, and AI-driven automation. Define strategy, roadmap, and success metrics.”
Signals reviewed · derived from public job posts, product docs
Job postings fill and close over time — once a posting is filled we keep it as a dated citation (the quoted evidence remains); use View open roles for current listings.
Key takeaways
- Unbundling can be a trust feature. Showing exactly what each layer of a voice minute costs turns pricing transparency into a differentiator in a category that usually quotes one blended rate.
- True pay-as-you-go beats contract-first in PLG. Removing contracts, fees, and minimums (plus $10 credits) lowered the barrier enough to drive ~650% YoY growth.
- Price on the metric your buyer already counts. Per-voice-minute maps directly onto the agent-minute economics of the call centers Retell displaces.
- Document the fuzzy edges. Publishing exactly when billing applies (silence, hold, transfers, voicemail, failed calls) defuses disputes around a notoriously ambiguous unit.
- Add-ons and monthly items quietly compound. A long menu of small per-minute and per-month charges can erode the simplicity of a “pay only for what you use” headline.
UBP implications
- Unbundled, pass-through usage pricing aligns vendor margin with buyer choice. Charging a clean margin on your own infra layer while passing through model/voice/telephony costs keeps incentives honest and lets buyers tune spend.
- Pick a value metric the buyer already budgets in. Voice-minutes (vs. tokens) let a contact-center buyer compare directly against human-agent cost-per-minute — the comparison that closes the deal.
- Transparency requires a forecasting aid. When the unit is composite, an interactive calculator (or recipe bundles) is what makes usage-based pricing usable. See usage-based pricing strategy.
Sources
- Retell AI pricing page — live capture (accessed 2026-07-22)
- Retell AI seed announcement (accessed 2026-06-09)
- Y Combinator — Retell AI (W24) (accessed 2026-06-09)
- Sacra — Retell AI at $60M/yr, up 650% YoY (accessed 2026-06-09)
- Y Combinator post — $4.6M seed, $3M ARR since February launch (accessed 2026-06-09)
Bottom line
Retell AI is a YC W24 conversational voice-agent platform that turned radical pricing transparency into a growth engine — true pay-as-you-go voice agents at $0.07–$0.31/min, unbundled into Retell Voice Infra ($0.055/min) plus the LLM, voice, and telephony you choose, with chat agents from $0.002/AI msg, $10 in free credits, and 20 free concurrent calls. No contracts, no platform fee, second-level metering with documented billing rules; Enterprise is custom-quoted with volume pricing, dedicated infrastructure, and compliance. That model helped Retell reach ~$60M ARR by April 2026, up roughly 650% YoY. Browse the pricing blueprint for more fully-researched company profiles.
Want to compare Retell AI against other voice-AI companies like Vapi, Bland, or ElevenLabs? Browse the pricing blueprint.
Pricing timeline : Major events on a vertical axis
Each milestone below corresponds to a public pricing change, product launch, or material adjustment. Major events use a filled marker; minor adjustments use a faded one.
Model menu refresh — Gemini 3.5 Flash added
Routine addition to the selectable voice LLM menu: Gemini 3.5 Flash at $0.081/min (voice). No change to the unbundled per-minute structure, headline $0.07–$0.31/min voice rate, $0.055/min Retell Voice Infra, $0.002+/msg chat, telephony ($0.015/min), add-ons, or monthly items ($2 phone / $20 SMS / $8 concurrency / $8 KB / $10 verified number). All other LLM, TTS, and chat prices unchanged.
Model menu refresh — Claude 4.6, Gemini 3.0 Flash, GPT 5.x reshuffle
Routine refresh of the selectable LLM/voice menu (Claude 4.6 Sonnet $0.08/min, Gemini 3.0 Flash $0.027/min added; GPT 5.x 'Recommended' tiers reshuffled) with no change to the unbundled per-minute structure, headline $0.07–$0.31/min voice rate, $0.055/min Retell Voice Infra, $0.002+/msg chat, telephony ($0.015/min), add-ons, or monthly items ($2 phone / $20 SMS / $8 concurrency / $8 KB / $10 verified number).
Unbundled per-minute stack + chat agents + add-ons
Current structure: voice agents at $0.07–$0.31/min ($0.055/min Retell Voice Infra + LLM + TTS + telephony), chat agents from $0.002/AI msg, $10 free credits and 20 free concurrent calls, an expanded add-on menu (knowledge base, denoising, PII removal, QA, branded calls), monthly subscription items, and a custom-quoted Enterprise plan with volume pricing, dedicated servers, HIPAA/BAA and SSO.
$4.6M seed announced
Announced a $4.6M seed led by Alt Capital (with Y Combinator, Carya Venture, and 20+ founders/operators) while sustaining the unbundled pay-as-you-go pricing. Pricing remained transparent and self-serve as the company scaled into BPO and enterprise contact-center accounts.
Launch — pay-as-you-go voice agents
Retell AI (YC W24) launches in February 2024 with a true pay-as-you-go model: per-minute AI voice agents composed from Retell's voice infrastructure plus a chosen LLM, TTS, and telephony — no contracts, no platform fee. Reached ~$3M annualized revenue within months.
- · Retell AI is a Y Combinator W24 company that was initially rejected by YC, then accepted after sharpening its demo.
- · It went from ~$3M annualized revenue at its September 2024 seed to roughly $60M ARR by April 2026 — up about 650% year-over-year — powering 50M+ real-time AI phone calls a month.
- · Pricing is fully unbundled: a single voice-agent minute is itemized into Retell Voice Infra ($0.055/min), the LLM, TTS, telephony, and add-ons — so buyers see exactly which layer costs what.
Questions & answers
- What is Retell AI's pricing model?
- Retell AI is true pay-as-you-go. Voice agents are billed per minute ($0.07–$0.31/min) as an unbundled stack — $0.055/min Retell Voice Infra plus your chosen LLM, text-to-speech, and telephony. Chat agents are billed per AI message from $0.002+. There are no platform fees, contracts, or minimums; Enterprise is custom-quoted with volume pricing.
- Does Retell AI offer a free tier?
- There is no free plan, but new accounts get $10 in free credits and full platform access with no commitment — enough to build, test, and launch an agent. Every account also includes 20 concurrent calls free, plus 100 free minutes of AI Quality Assurance.
- How much does Retell AI cost per minute?
- A typical voice agent costs about $0.07–$0.31/minute depending on the LLM you select. The bill is composed of $0.055/min Retell Voice Infra, the LLM rate (e.g. GPT 5.1 at $0.04/min, GPT 5.5 at $0.16/min), TTS ($0.015/min for most voices, $0.04/min for ElevenLabs), and telephony ($0.015/min). Retell's calculator shows ~$0.115/min on a GPT-5.1-class setup.
- Is Retell AI pricing usage-based or subscription?
- It is predominantly usage-based — you pay per minute of voice and per message of chat, metered to the nearest second with no per-call rounding. A few optional items are monthly subscriptions (e.g. Retell phone numbers at $2/mo, extra concurrency at $8/call/mo, verified numbers at $10/mo). Enterprise adds custom volume pricing.
- What hidden costs should I budget for with Retell AI?
- Beyond the base per-minute rate, budget for add-ons (knowledge base +$0.005/min, PII removal +$0.01/min, branded call +$0.10/outbound), telephony ($0.015/min), and monthly items like extra concurrency ($8/call/mo) and phone numbers ($2/mo). Billing also runs during silence and hold time because the STT engine stays active.