Retell AI

Retell AI pricing breaks a voice minute into named parts and publishes a rate for each.

Pricing Model:

Pricing Model:

Usage, drawn against prepaid credits

Usage, drawn against prepaid credits

Usage, drawn against prepaid credits

Packaging Model:

Packaging Model:

Freemium, Good / Better / Best (GBB)

Freemium, Good / Better / Best (GBB)

Freemium, Good / Better / Best (GBB)

Credit Model:

Credit Model:

Prepaid, never expire, non-refundable

Prepaid, never expire, non-refundable

Prepaid, never expire, non-refundable

Updated on:

Retell AI pricing: the four parts of a voice minute

Retell AI pricing breaks a voice minute into named parts and publishes a rate for each. You pay $0.055 a minute for Retell's voice engine, plus a rate for the model you picked, plus a rate for the voice, plus telephony. Three of those four you can change or remove. The headline $0.07 to $0.31 range on the pricing page is the part that doesn't survive arithmetic.

Key takeaways

  • Retell charges $0.055 a minute for voice infrastructure and prices the LLM, voice and telephony as separate invoice lines.

  • Bring your own LLM and the LLM line goes to zero. Bring your own SIP trunk and telephony goes to zero.

  • Concurrency costs $8 per concurrent call per month beyond the free 20, and burst calls add $0.10 a minute to the whole call.

  • Prompts over 4,000 tokens scale billed duration: a 60 second call on a 4,800 token prompt bills as 72 seconds.

Retell AI pricing in 2026

Tier

Price

Included

Metered limit

 

Pay as you go

$0 platform fee

Full platform, API, analytics

20 concurrent calls, $10 credit

Enterprise

Custom

SSO, RBAC, custom MSA/DPA/BAA

Concurrency from 50+, volume pricing

Component

Rate



---

---



Retell voice infra

$0.055/min



Voices: Platform, Minimax, Fish, Cartesia, OpenAI, Inworld

$0.015/min



Voices: ElevenLabs

$0.040/min



LLM

$0.0016/min (GPT 5 nano) to $0.64/min (GPT 6 Astra, fast tier)



LLM, recommended models

$0.064/min (GPT 5.6 Terra, Claude 5 Sonnet)



Telephony, Retell Twilio US

$0.015/min



Custom LLM or custom telephony

No charge



What Retell AI actually meters

Connected call seconds, split across four components and totalled by the minute.

Four line items, and a fifth that never appears. Speech to text has no published rate of its own. Retell folds transcription and turn-taking into the $0.055 voice infra line and picks the vendor itself, failing over between Deepgram and Azure mid-call. You can't swap the transcriber or see what it costs.

The other three you control. The LLM line carries a per-minute rate for each of 26 models, several with a fast tier at double the standard rate. Point Retell at your own model and the line reads "No charge". Text to speech is $0.015 a minute across six voice providers and $0.040 for ElevenLabs. Telephony runs $0.015 a minute on a Retell-managed Twilio US number, up to $0.80 for the Philippines, and nothing on your own trunk.

Two adjustments change billed duration rather than the rate. A dynamic opening message carries a 10 second minimum, and prompts above 4,000 tokens scale duration by tokens divided by 4,000, rounded up. Silence bills, because the transcriber keeps listening. Unconnected calls don't.

How credits work

Newer workspaces run on prepaid credits; older ones bill monthly in arrears.

Credit accounts deduct usage from a balance in real time. Credits never expire and are non-refundable once bought, an unusual pairing: most vendors trade one for the other. Balances are scoped per workspace, so a card added in one doesn't carry to the next.

Subscription items sit outside the credit balance and bill to the card at month end: numbers at $2, verified numbers at $10, SMS at $20, knowledge bases beyond the first ten at $8, and concurrency at $8 a slot, charged on purchase and prorated by day since February 2026.

Free allowances: $10 of signup credit once per email address, 20 concurrent calls, ten knowledge bases, the first 100 AI QA minutes, and 30 Conductor messages per user per day.

What happens when you hit the limit

Two limits, two different failures.

Run out of credits and calls stop. Retell blocks new calls at a zero balance until auto recharge fires on a threshold and target you set, so the refill is a fixed top-up, not metered overage. Legacy monthly accounts don't block; they bill in arrears.

Run out of concurrency and it depends on direction. Outbound calls are rejected. Inbound calls wait about 40 seconds, then transfer to a fallback number or end with concurrency_limit_reached. Turn on burst and they go through, capped at the lower of 3x your limit or your limit plus 300, with $0.10 a minute added to the entire call, not just the part above the line.

How Retell AI pricing has changed across all these years

Date

Milestone

Source

 

22 Sep 2026

Inworld joins the $0.015 voice tier. Component rates and the $0.07 to $0.31 range unchanged

Vendor

2 Jul 2026

Free LLM token allowance raised from 3,500 to 4,000

Vendor

10 Feb 2026

TTS and voice engine split into separate invoice lines. Concurrency billed upfront, prorated daily

Vendor

16 May 2025

Token scaling above 3,500 prompt tokens and a 10 second minimum on dynamic openers, from 1 June

Vendor

31 Dec 2024

GPT-4o Realtime cut from $1.50 to $0.50/min. Telephony raised from $0.01 to $0.015/min

Vendor

19 Apr 2024

Pricing made modular: voice engine, LLM and telephony split into separate rates, custom LLM free

Vendor

4 Mar 2024

Discounted enterprise tiered pricing introduced

Vendor

6 Feb 2024

New $0.10/min tier using OpenAI TTS

Vendor

Flexprice’s Take

Retell publishes the best component breakdown in voice AI and then wraps it in a headline range that its own rate card contradicts.

Retell publishes the best component breakdown in voice AI: four line items, a rate for each, and a calculator that adds them up. The structure has held since April 2024.

You can zero two of those lines. Bring your own LLM or SIP trunk and Retell charges nothing for either, and it measures per second with no rounding.

The headline range is where it falls apart. The page says $0.07 to $0.31 a minute. The floor checks out: $0.055 + $0.015 + $0.0016 for GPT 5 nano is $0.0716. The ceiling doesn't. An ElevenLabs voice plus GPT 6 Astra on the fast tier is $0.055 + $0.040 + $0.64, or $0.735 a minute, more than twice the advertised maximum before telephony.

The floor beats Deepgram. That $0.0716 undercuts Deepgram Standard at $0.075 by 4.5%, and $0.055 + $0.015 + $0.064 for GPT 5.6 Terra lands at $0.134, 17.8% below Deepgram Advanced at $0.163.

Best For

Teams that want to control each part of the minute, and already own an LLM or a SIP trunk.

Watch Out For

Budgeting from the $0.31 ceiling, token scaling, and burst surcharges.

Manish Choudhary

CEO & Co-founder, Flexprice

Billing a minute assembled from four metered components?

Flexprice meters multi-part usage in one event stream.

Flexprice’s Take

Retell publishes the best component breakdown in voice AI and then wraps it in a headline range that its own rate card contradicts.

Retell publishes the best component breakdown in voice AI: four line items, a rate for each, and a calculator that adds them up. The structure has held since April 2024.

You can zero two of those lines. Bring your own LLM or SIP trunk and Retell charges nothing for either, and it measures per second with no rounding.

The headline range is where it falls apart. The page says $0.07 to $0.31 a minute. The floor checks out: $0.055 + $0.015 + $0.0016 for GPT 5 nano is $0.0716. The ceiling doesn't. An ElevenLabs voice plus GPT 6 Astra on the fast tier is $0.055 + $0.040 + $0.64, or $0.735 a minute, more than twice the advertised maximum before telephony.

The floor beats Deepgram. That $0.0716 undercuts Deepgram Standard at $0.075 by 4.5%, and $0.055 + $0.015 + $0.064 for GPT 5.6 Terra lands at $0.134, 17.8% below Deepgram Advanced at $0.163.

Best For

Teams that want to control each part of the minute, and already own an LLM or a SIP trunk.

Watch Out For

Budgeting from the $0.31 ceiling, token scaling, and burst surcharges.

Manish Choudhary

CEO & Co-founder, Flexprice

Billing a minute assembled from four metered components?

Flexprice meters multi-part usage in one event stream.

Customer
Sentiment Highlights

"I do wish Retell's pricing was cheaper, though; $6/hr is pretty much the cost of a call center employee in India"

reissbaker, Hacker News, February 2024

"The pricing can also become expensive when scaling or doing many test calls during development."

Verified user in computer software, G2, June 2026

Frequently Asked Questions

Frequently Asked Questions

How much does Retell AI cost per minute?

Does Retell AI charge a platform fee?

How much does Retell AI concurrency cost?

Do Retell AI credits expire?

Launch usage-based billing this week, not next quarter

Launch usage-based billing this week, not next quarter

Get Instant Feedback on Your Pricing | Join the Flexprice Community with 400+ Builders on Slack

Join the Flexprice Community on Slack