Retell vs Vapi vs Bland vs Custom LiveKit (2026): Price per Minute, Latency and Lock-In
- Retell is the fastest route to a working phone agent at $0.125 a minute with its telephony; Vapi lands at about 8 to 14 cents with providers you pick; Bland charges $0.12 to $0.14 plus telephony. A custom LiveKit stack runs at about 2.5 cents a minute in production.
- Choose by monthly minutes, not preference. Under about 20,000 minutes a month a managed platform is the correct commercial answer; at 100,000 minutes the gap is about $10,000 a month.
- The per-minute rate hides the decisions that matter: telephony, concurrency caps, a $2,000 a month HIPAA add-on on Vapi against a free BAA on Retell, and who owns your numbers when you leave.
What is the difference between Retell, Vapi, Bland and a custom LiveKit stack?
Retell, Vapi and Bland are managed platforms: each bundles call transport, speech recognition, a language model and speech synthesis behind a per-minute price. A custom LiveKit stack is an architecture instead. You assemble those parts on LiveKit's open-source framework, pay each provider at list price, and own the pipeline, the on-call and the margin.
The three platforms sell different things. Retell sells the quickest route to a working phone agent, with an itemised price. Vapi sells orchestration: a platform fee, with speech, model and telephony passed through at cost, which suits developers who want to pick each part. Bland sells one flat rate that covers its own models, and sizes its plans by call volume.
I have shipped production voice on both paths. My team integrated Retell in about two days, according to the case study Retell published, and the AI voice interview step it powered cut false-positive assessments from 50% to 15%. On a custom LiveKit stack, I cut voice cost from about 10 cents a minute on a managed platform to about 2.5 cents in production. Both decisions were right at the volume we had when we made them.
| Retell | Vapi | Bland | Custom LiveKit | |
|---|---|---|---|---|
| What you pay for | itemised bundle | platform fee + providers at cost | flat rate | providers at list price |
| Published price per minute | $0.07 to $0.31 | about 8 to 14¢ all in | $0.12 to $0.14 + telephony | about 2.5¢, measured |
| Bring your own LLM | custom LLM socket | ✓ | not listed | ✓ |
| Pick your own speech and voice providers | voice menu | ✓ | ✕ | ✓ |
| Own the turn-taking logic | ✕ | partial | ✕ | ✓ |
| You carry reliability and on-call | ✕ | ✕ | ✕ | ✓ |
| Sensible under 20,000 minutes a month | ✓ | ✓ | ✓ | ✕ |
How much does each option cost per minute in 2026?
On list prices checked 23 September 2026: Retell's default is $0.11 a minute before telephony, Vapi lands between about 8 and 14 cents all in, Bland charges $0.12 to $0.14 plus telephony, and LiveKit's own calculator default is about 4.8 cents. My custom LiveKit stack runs at about 2.5 cents a minute in production.
Retell's pricing page is the fairest baseline because it itemises: $0.055 a minute for voice infrastructure, a per-minute price for each model (for example $0.0064 for GPT 5.6 Luna, $0.064 for GPT 5.6 Terra or Claude 5 Sonnet, $0.16 for GPT 5.5) and $0.015 for most voices, $0.040 for ElevenLabs. Its calculator's default adds up to $0.11, and the headline range it publishes is $0.07 to $0.31.
Vapi charges $0.05 a minute for hosting and passes providers through at cost. On its own listed ranges (transcription $0.0095 to $0.0099, OpenAI models $0.0077 to $0.0452, ElevenLabs voices $0.0146 to $0.0238), the low end of those ranges on Vapi's free telephony is 8.2 cents, and the high end on Twilio outbound is 14.3 cents. Bland sells one rate covering model, transcription and voice: $0.14 on its free Start plan, $0.12 on Build at $299 a month.
LiveKit Cloud charges for the session rather than the stack: 1,000 free agent minutes on Build, then $0.01 a minute beyond the 5,000 included in Ship ($50 a month) or the 50,000 in Scale ($500 a month), with speech and model billed by the providers you choose. Its calculator's default phone agent is $0.0479 a minute. Against Retell's $0.11 default, a measured 2.5 cents is a 4.4x reduction; against LiveKit's own default, 1.9x.
| Option | Published price per minute | Included | Billed on top | Concurrency included |
|---|---|---|---|---|
| Retell | $0.07 to $0.31; $0.11 default | Voice infrastructure, transcription, model, voice | Telephony $0.015/min, numbers $2/month, add-ons such as knowledge base +$0.005/min and PII removal +$0.01/min | 20 calls, then $8 per call a month |
| Vapi | $0.05 platform fee; about 8.2¢ to 14.3¢ all in | Orchestration; Vapi's own telephony and SIP | Transcription, model and voice at cost; Twilio $0.008 in, $0.014 out; HIPAA $2,000/month | 4 calls (10 on $29 Core, 30 on Pro), then $10 per line a month |
| Bland | $0.14 on Start; $0.12 on Build ($299/month) | Model, transcription and voice | Telephony at pass-through; transfer time $0.04 to $0.05/min | 10 calls and 100 a day on Start; 50 calls and 2,000 a day on Build |
| LiveKit Cloud | $0.01 agent session after included minutes; $0.0479 calculator default | Rooms, transport, agent hosting | Speech, model and voice providers; SIP $0.004 or $0.003/min after included minutes; numbers $1/month | 5 (Build), 20 (Ship), up to 600 (Scale) |
| Custom LiveKit, self-hosted | About 2.5¢, measured in production | Everything, at provider list price | Engineering time, servers, on-call | What you provision |
What do the per-minute prices leave out?
Telephony, concurrency, add-ons and plan fees. Retell's default shows telephony at $0.00 until you add its $0.015 a minute; Bland bills telephony separately; Vapi's platform fee excludes the providers it passes through. Concurrency caps and daily call limits then decide which plan you are really on.
Read the billing rules as closely as the rate. Retell bills to the second, charges for silence and hold because transcription keeps listening, does not bill calls that fail to connect, and stops the agent fee after a transfer while telephony keeps running. Add-ons compound: a knowledge base (+$0.005), denoising (+$0.005) and PII removal (+$0.01) add $0.02 a minute, 18% on top of the default.
Plan limits are where a cheap quote turns into a plan change. Vapi includes 4 concurrent calls on usage-only billing and charges $10 a line for more; Retell includes 20 and charges $8 for each call above that; Bland caps Start at 10 concurrent calls and 100 calls a day. Size concurrency to your busiest hour, not your monthly minutes.
The line no platform prints is engineering time on the custom path. Counting two engineer-days a month of maintenance, a custom stack does not beat a vendor until about 20,000 minutes a month; the arithmetic is in voice AI build vs buy. Run your own mix of models, voices and minutes in the voice AI cost calculator.
- Is telephony in the per-minute rate, and what does each number cost a month?
- What is billed during silence, hold, voicemail and after a transfer?
- How many concurrent calls and calls a day are included, and what does each extra line cost?
- Which add-ons will you need: knowledge base, denoising, PII removal, QA?
- Is a BAA available on this plan, and at what price?
- Can you export transcripts, recordings and call data in bulk, by webhook or API?
- Can you bring your own numbers and carrier?the answer decides how hard leaving will be
Which is fastest, and how should you compare latency claims?
Each vendor publishes its own number: Retell says about 600ms, Vapi says under 500ms on average, Bland says 400ms. None of those pages states the method, the region or the percentile, so treat them as claims until you have run the same call script through each option yourself.
All four run the same cascade. Audio arrives, the system decides the caller has finished, speech becomes text, the model starts answering and synthesis starts speaking. On default settings the largest line is usually end-of-turn detection rather than the model, because waiting for silence is how a pipeline decides the caller is done, and that wait is a setting. The stage-by-stage budget is in voice AI latency under 800ms.
Measure the tail, not the average. Put the same 50 recorded calls through each option from the same region and carrier, and log the gap between the caller's last word and the agent's first audio at the 95th percentile. A steady 700ms feels better to a caller than a 400ms median with a long tail, because the tail is where people start talking over the agent.
The platforms tune turn-taking for you, and Retell advertises its own turn-taking model. On LiveKit you own it: the Apache-2.0 agents framework ships a transformer-based turn detector, and endpointing can be tuned per use case. That matters when your callers are unusual (older callers, noisy lines, several languages). For a first agent, a platform default is usually good enough.
How does telephony work on each, and can you keep your own numbers?
All four accept a SIP trunk from your own carrier, and all four also sell telephony: Retell at $0.015 a minute, Vapi on its own free telephony or Twilio, Bland at pass-through cost, LiveKit at $1 a month per number. Buy numbers on your own carrier account from day one, and switching platforms later stays a configuration change.
The rates: Retell charges $0.015 a minute on its numbers, $2 a month per number and $10 for a verified number. Vapi's own telephony and SIP are free, with Twilio at $0.008 a minute inbound and $0.014 outbound, and Telnyx at $0.0055. LiveKit includes one free US number, then $1 a month, and charges $0.004 a minute for third-party SIP on Ship after 5,000 included minutes ($0.003 on Scale). A direct Telnyx trunk is $0.0032 a minute inbound and $0.005 outbound on local US numbers.
The real trade is ownership. A number bought inside a platform has to be ported out when you leave, which is a carrier process with its own lead time. A number on your own Twilio or Telnyx account can point at Retell today and at your LiveKit trunk next year without a caller noticing. Trunk choice and SIP setup are in LiveKit SIP trunking: Twilio vs Telnyx.
Outbound adds rules no platform removes. The FCC's February 2024 ruling treats AI-generated voices as artificial voices under the TCPA, so calls need prior express consent, written consent for telemarketing, and must identify the business. Bland's daily caps (100 calls on Start, 2,000 on Build) are the other outbound limit to check before you plan a campaign.
| Retell | Vapi | Bland | LiveKit Cloud | |
|---|---|---|---|---|
| Numbers sold by the platform | $2 a month | 1 to 10 included by plan | 1 inbound number on Start | 1 free, then $1 a month |
| Minutes on platform numbers | $0.015 | free on Vapi telephony | pass-through | $0.01 inbound after included minutes |
| Bring your own carrier over SIP | ✓ | ✓ | ✓ | ✓ |
| Concurrency included | 20 | 4, 10 or 30 by plan | 10 or 50 by plan | 5, 20 or up to 600 |
| Extra concurrency | $8 per call a month | $10 per line a month | Enterprise | Scale or Enterprise |
| Daily call cap | none listed | none listed | 100 or 2,000 | none listed |
How much can you customise each one, and where is the lock-in?
The prompt is portable everywhere. The rest is not: call flows, turn-taking behaviour, analytics, evaluation history and platform-bought numbers stay with the vendor. Retell and Vapi accept your own language model, Vapi lets you choose each provider, Bland runs its own models, and a LiveKit stack is code in your own repository.
Retell's custom LLM mode opens a WebSocket to your server for each call, while Retell keeps telephony, transcription, turn-taking and synthesis. For most products that is enough control: your logic, its voice pipeline. Vapi goes further, with providers you pick and a custom model endpoint. Bland's flat rate is the opposite trade: its own models, no provider pass-through, and custom voices on Enterprise.
Lock-in is the cost nobody models. Over two years, being able to swap a speech or voice provider the week a better one ships is worth more than the per-minute difference. On LiveKit that swap is a code change behind a flag; on a platform it waits for the vendor's integration list.
On a platform, three habits keep leaving cheap: numbers on your own carrier account, transcripts and recordings exported to your own storage by webhook from day one, and prompts and tool definitions kept in your repository rather than only in a dashboard. Then a move is a migration, not a rewrite, and the playbook is in migrating from Retell or Vapi to LiveKit.
Portable on every option, if they live in your repository and not only in a dashboard.
yours everywhereYour choice on Vapi and LiveKit, a custom LLM socket on Retell, Bland's own models on its flat rate.
rented on BlandA voice menu on Retell, any provider on Vapi and LiveKit, Bland's own on Bland.
swap speed differsThe vendor's model on Retell and Bland, settings on Vapi, your own code on LiveKit.
hardest to moveYours if bought on your own Twilio or Telnyx account; a porting job if bought inside the platform.
decide on day oneExported by webhook to your own storage, or left in the vendor's dashboard.
export from day oneWhich is compliant: HIPAA, SOC 2 and data retention?
All four can support regulated work, at very different prices. Retell offers a BAA you sign yourself at no extra fee; Vapi sells HIPAA as a $2,000 a month add-on; Bland puts its BAA on Enterprise; LiveKit Cloud includes SOC 2 Type II and a BAA from its $500 Scale plan. Self-hosting makes compliance your own work.
The details matter more than the badges. Retell's compliance page lists SOC 2 Type 1 and 2 and per-agent retention from one day to two years. Vapi's HIPAA mode changes how logs, recordings and transcripts are stored, needs the paid add-on and a signed BAA, and cannot run alongside zero data retention. Bland lists SOC 2 Type I and II, PCI DSS and GDPR, with on-premises or VPC deployment on Enterprise.
For a small healthcare pilot the arithmetic is short: Vapi's add-on costs as much as 16,000 minutes on Retell at $0.125. For data that must stay inside your network, the choice narrows to Bland Enterprise on-premises or a self-hosted LiveKit stack, where the BAA chain runs through every provider you call, and a provider that stores ePHI is a business associate even when it cannot view the data.
| Retell | Vapi | Bland | LiveKit Cloud | |
|---|---|---|---|---|
| SOC 2 | Type 1 and 2 | via trust centre | Type I and II | Type II, Scale and up |
| BAA for HIPAA | self-signed, no fee | $2,000 a month add-on | Enterprise | Scale ($500 a month) and up |
| Retention and residency | 1 day to 2 years per agent | 14 to 180 days by plan, zero retention option | data residency on Enterprise | region pinning, Scale and up |
| Runs inside your network | ✕ | ✕ | Enterprise, on-premises or VPC | self-host the open-source stack |
Which should you choose, by use case and volume?
Choose by monthly minutes first, then by constraints. Under about 20,000 minutes a month a managed platform is the correct commercial answer: Retell to ship fastest, Vapi for provider control, Bland for a flat rate at volume. Past that, with engineers who can own a pipeline, a custom LiveKit stack wins on cost and control.
The arithmetic at scale is blunt. At 100,000 minutes a month, Retell's default plus its telephony is $12,500 on pay-as-you-go pricing; the same minutes on a custom stack at 2.5 cents are $2,500. That is $10,000 a month, $120,000 a year, before any enterprise discount a platform would offer at that volume, which you should ask for before deciding.
In between, run both. Put a routing layer in front, send a small share of real calls to your own stack, compare transcripts, latency and transfer rates on identical call types, and move traffic in steps. What decides whether callers notice is turn detection and barge-in, not the model. To price the product around the calls, use the AI product cost estimator.
| Use case | Monthly minutes | Pick | Why |
|---|---|---|---|
| First inbound agent: receptionist, FAQ, booking | Under 20,000 | Retell | Fastest to a working call; $0.125 a minute with telephony; BAA at no extra fee |
| Team that wants to choose each provider | Under 20,000 | Vapi | $0.05 platform fee with providers at cost, and a custom model endpoint |
| Outbound campaigns at a steady volume | 10,000 to 100,000 | Bland Build or Enterprise | One rate covers model, speech and voice; 2,000 calls a day on Build |
| Healthcare pilot with patient data | Under 20,000 | Retell | Self-signed BAA with no add-on fee; Vapi's HIPAA add-on is $2,000 a month |
| Data must stay in your network | Any | Bland Enterprise or self-hosted LiveKit | The two options here that run inside your own infrastructure |
| Steady traffic, engineers to own it | 20,000 to 100,000 | Hybrid, then LiveKit | Route a share of calls to your own stack and compare on identical calls |
| High volume | Over 100,000 | Custom LiveKit | 2.5¢ against 12.5¢ is about $10,000 a month at 100,000 minutes |
Engineering and on-call cost more than the margin you would save. Retell for speed, Vapi for provider control, Bland for a flat rate.
Route a share of live calls to your own stack and compare quality on identical calls before committing.
About $10,000 a month separates 12.5 cents from 2.5 cents at this volume, and provider-swap speed becomes an advantage.
How do you move from a platform to LiveKit without dropping calls?
Not with a cutover. Instrument the platform you have, rebuild to parity, put a routing layer in front of both stacks and shift traffic in steps, with a rollback trigger at each one. Keep the platform account live for about 60 days after the new stack carries all of the traffic.
Most of the parity work is platform features you forgot you depended on: post-call analysis, campaign dialling, knowledge base, denoising, PII redaction, branded caller ID and warm transfer. Each is a line on the platform invoice and a small project on your own stack, so list them from the invoice before you estimate the move.
If you would rather have the LiveKit stack built and instrumented for you, voice AI development is built at $0: the work is split into checkpoints with acceptance criteria agreed before it starts, and each one is invoiced only after you have seen it and accepted it.
Frequently asked questions
→Is Retell or Vapi cheaper?
It depends on the providers you pick. Retell's default is $0.11 a minute plus $0.015 for its telephony, $0.125 in total. Vapi charges $0.05 for its platform and passes providers through at cost, which lands between about 8.2 and 14.3 cents on its own listed ranges. With cheap models Vapi is cheaper; with premium voices it is not.
→How much does Bland AI cost per minute?
Bland lists $0.14 a minute on its free Start plan and $0.12 on Build, which costs $299 a month. The rate covers the model, transcription and voice; telephony is billed separately at pass-through, and transfer time costs $0.04 to $0.05 a minute. Start is capped at 100 calls a day and Build at 2,000.
→When should I move off a managed voice platform?
At around 20,000 minutes a month, if you have engineers who can own a voice pipeline including on-call. Below that, engineering and maintenance time cost more than the margin you save. At 100,000 minutes the gap is large: about $12,500 a month on Retell's default against $2,500 on a custom stack at 2.5 cents.
→How much cheaper is a custom LiveKit stack?
My custom LiveKit stack runs at about 2.5 cents a minute in production, down from about 10 cents on a managed platform. Against Retell's published $0.11 default that is a 4.4x reduction, and against LiveKit's own calculator default of $0.0479 it is 1.9x. The saving only pays for the engineering above roughly 20,000 minutes a month.
→Which voice AI platform is HIPAA compliant?
All four can be. Retell lists SOC 2 Type 1 and 2 and a BAA you can sign yourself at no extra fee. Vapi sells HIPAA as a $2,000 a month add-on with a signed BAA. Bland offers a BAA on Enterprise, and LiveKit Cloud includes a BAA from its $500 a month Scale plan.
→What is the hardest part of building your own?
Turn detection and barge-in. Waiting on a fixed silence timer produces an agent that interrupts people mid-sentence, and failing to cancel speech already in progress when the caller interrupts produces two voices talking over each other. The model and the transcription are the easier parts; conversation timing is where custom stacks win or lose.
Open the article in your assistant with one click and ask it how this applies to your product.