Vapi charges $0.05 a minute, and that number is true. It buys turn detection, interruption handling, call state, function calling and the plumbing between four services you supply yourself.
It does not buy any of those four services.
Add the published list price of what the $0.05 excludes and the floor lands at $0.0965 a minute, which is 1.93 times the headline. That is not a hidden fee. Vapi's comparison table is headed Excludes Model Provider Costs and its FAQ prices speech, model and synthesis as At cost ($0 if you bring your own API key), which is more honest than most of this category manages.
The unusual part is what happens when you make the vendor do the arithmetic itself.
What follows: what the $0.05 covers, the computed floor, Vapi's own calculator driven to a total, the six-fold spread hiding inside one product, the two flat fees that dwarf the meter, the concurrency line nobody models, and a correction this site owes on its own Vapi arithmetic. For the head-to-head against the two platforms people shortlist alongside it, that is a separate page.
What does Vapi cost?
$0.05 a minute advertised, $0.0965 computed all-in. Vapi's own calculator, left on its defaults, returns $0.1103.
| Figure | Value | What it is |
|---|---|---|
| Advertised | $0.05/min | Orchestration only, stated as such on the page |
| Computed all-in floor | $0.0965/min | $0.05 plus the list price of all four excluded components [computed] |
| Vapi's own calculator, defaults | $0.1103/min | $110.30 for 1,000 minutes, itemised by the vendor |
| Cheapest selectable stack | ~$0.055/min | Every menu on its cheapest option |
| Dearest selectable stack | $0.32+/min | Before the model is counted |
There is no platform fee and no seat licence, and evaluation is genuinely free on trial credit. The commitment you make is architectural rather than financial.
What the $0.05 excludes, and why that is defensible
All four components. Speech recognition, the language model, synthesis and telephony are each somebody else's invoice.
| Component | In the $0.05? | List rate this site uses |
|---|---|---|
| Speech to text | No | $0.0077/min |
| Language model | No | $0.0108/min |
| Text to speech | No | $0.0140/min of call |
| Telephony | No | $0.0140/min |
A platform that bundled all four would be taking price risk on four vendors' roadmaps at once, and would have to charge for that risk. Vapi declines to. The trade is real and it is not a pricing trick: bring your own keys and those lines leave Vapi's invoice entirely, which is exactly what the architecture is for.
They leave the invoice. They do not leave the budget. That single distinction is what most comparisons of Vapi against a bundled platform get wrong.
Vapi's own calculator returns a higher number than ours
This site drove the vendor's usage calculator to a total on 6 August 2026, with the three providers Vapi itself names in its FAQ plus Twilio for the line.
| Line | Selection | For 1,000 minutes |
|---|---|---|
| Vapi hosting | $0.05/min | $50.00 |
| Transport | Twilio outbound, $0.014/min | $14.00 |
| Speech to text | Deepgram, $0.01/min | $10.00 |
| Language model | GPT-4o | $0.30 |
| Text to speech | ElevenLabs, $0.036/min | $36.00 |
| Total | $110.30 |
Read that twice, because the direction matters. This site exists on the premise that vendors quote low and buyers discover the rest, and here the vendor's own tool quotes higher than the independent model does. The gap is synthesis: Vapi's default puts ElevenLabs at $0.036 a minute where the cost model carries a mid-market $0.014.
None of the pages ranking for this vendor's pricing query appears to have moved the sliders at all.
The 5.9x spread inside one product
Choose the cheapest option on every menu and the stack lands near $0.055 a minute. Choose the dearest and it passes $0.32 before the model is counted.
That is a factor of nearly six, all of it inside what people call Vapi's price.
Which means the sentence "Vapi costs $0.05 a minute" carries almost no information about your bill. The platform rate is fixed and everything that moves is a dropdown. A team that spends a fortnight comparing platforms and four seconds picking a voice has optimised the smaller variable, and the demo voice is reliably the expensive one because it is the one that sells the product.
Worth noticing while you are inside the calculator: Vapi appears in its own synthesis menu at $0.0216 a minute, selling its own voice beside the third parties, on a page that frames component cost as somebody else's pass-through.
Where the money actually goes
| Layer | Per minute | Share of the computed all-in |
|---|---|---|
| Vapi orchestration | $0.05 | 52% |
| Text to speech | $0.014 | 15% |
| Telephony | $0.014 | 15% |
| Language model | $0.0108 | 11% |
| Speech to text | $0.0077 | 8% |
Only one of those lines has real range. Telephony is a carrier rate you cannot argue with, recognition varies by tenths of a cent between vendors, and the model stays around a tenth of the bill however hard you tune it.
Synthesis is the exception, and on a phone call band-limited to 8kHz the difference between the cheap voice and the premium one is frequently inaudible. This site works through that component by component in the bill audit, which owns that argument rather than repeating it here.
The two flat fees that cost more than your minutes
HIPAA is $2,000 a month. Zero Data Retention is another $1,000. Both are flat, both were read on 11 August 2026, and neither appears in any per-minute comparison of this product.
| Minutes a month | Platform spend at $0.05 | HIPAA fee | Effective HIPAA cost per minute |
|---|---|---|---|
| 5,000 | $250 | $2,000 | $0.400 [computed] |
| 20,000 | $1,000 | $2,000 | $0.100 [computed] |
| 40,000 | $2,000 | $2,000 | $0.050 [computed] |
| 100,000 | $5,000 | $2,000 | $0.020 [computed] |
A flat fee is invisible to every per-minute comparison ever written about this category, and it is the line most likely to decide whether the product is affordable for a healthcare pilot.
Concurrency is metered, and it is the line nobody models
Ten concurrent lines are included. Beyond that it is $10 a line a month.
That is cheap until a campaign day needs fifty simultaneous calls, at which point it is $400 a month you did not budget for, arriving on the one morning the volume mattered. This site treats the general pattern in the concurrency page.
What Vapi costs an hour, which is the unit buyers compare
| Basis | Per minute | Per hour of talk time |
|---|---|---|
| Advertised | $0.05 | $3.00 [computed] |
| Computed all-in floor | $0.0965 | $5.79 [computed] |
| Vapi's calculator at defaults | $0.1103 | $6.62 [computed] |
Per-minute pricing hides the comparison people are actually running, which is against a person. Convert before you take a number to anyone who signs things.
This site's own Vapi arithmetic contradicts itself
The general lesson is that a derived number is only as fresh as its inputs, which is why the figure in this site's data layer is computed in code from a dated component table rather than typed into prose. Prose does not recompute itself.
What you take on with the architecture
Four vendor relationships, four sets of rate limits, four status pages and four invoices to reconcile.
When a call fails you own the diagnosis, and the fault is frequently at a seam rather than inside any one component. Interruption handling in particular has no correct default: endpointing thresholds that feel responsive on a quiet office line clip people speaking from a car park. That tuning is where deployments actually spend their first month, not on the integration.
Against a bundled platform at $0.07, Vapi is cheaper per minute at volume and more expensive for the first three months, because the first three months are engineering.
When Vapi is the wrong purchase
When nobody on your team wants to reason about endpointing thresholds or voice model selection. A bundled platform like Retell or Bland will cost you less in total, and the computed floors are close enough that the difference is staffing rather than price.
When you will not actually do the optimisation. The architecture gets cheaper than bundled only if somebody swaps components as better ones ship, and that is work teams intend to do and frequently never schedule.
And when you would rather have one invoice than four.
How to price your own Vapi minute in ten minutes
- Drive the vendor's calculator yourself with your own provider selections and write down the itemised lines rather than the total. The line that surprises you is the one to negotiate.
- Pick the voice before the platform. It is the widest lever in the stack, and it is chosen in a settings panel by someone who was not in the pricing meeting.
- Add the flat fees before the meter. HIPAA and Zero Data Retention are $3,000 a month together and they do not scale with usage.
- Count your peak concurrency, not your average. Ten lines are included and the eleventh is $10 a month.
- Decide honestly whether anyone will own the four vendor relationships. If the answer is nobody, the bundled platforms are the cheaper purchase on published rates.
What this page does not know
measuredPerMin is null for this vendor as it is for all 264 entities on this site. Budget above the $0.0965 floor rather than at it, and treat the vendor's own $0.1103 as the more realistic of the two.What a cost stack is, and which six entities carry one
A cost stack on this site is a small record holding three things: a vendor's advertised per-minute rate, the list of components that rate excludes, and an all-in figure derived in code from those two rather than typed in by hand. Correcting one component rate therefore corrects every page at once, which is how the speech recognition correction of 7 August 2026 moved every computed figure here by $0.0012 a minute in a single commit.
Six of the 264 entities in this directory carry one: Vapi, Retell, Bland, ElevenLabs Agents, Millis and Ultravox. The qualifying condition is an advertised per-minute rate with a verified exclusion set behind it. A vendor that publishes no rate has nothing to add exclusions to, and a vendor that publishes a bundle without itemising what it covers leaves the exclusion set unverified, which is a gap rather than a zero.
The field beside the computed one is measuredPerMin, and it is null on every entity here. It renders as not measured rather than falling back to the computed figure, which is the whole reason it exists: when measurement eventually starts, the records that have not been measured stay visibly separate instead of the distinction being retrofitted.
Every platform on this site with a computed all-in floor
| Platform | Advertised per minute | Computed all-in floor | Multiple | What the advertised rate excludes | Rate read |
|---|---|---|---|---|---|
| Millis AI | $0.02 | $0.0665 | 3.33x | Speech, model, synthesis and telephony | 2026-08-05 |
| Vapi | $0.05 | $0.0965 | 1.93x | Speech, model, synthesis and telephony | 2026-08-06 |
| Ultravox | $0.05 | $0.064 | 1.28x | Telephony only | 2026-08-05 |
| Retell AI | $0.07 | $0.0948 | 1.35x | Model and telephony | 2026-08-06 |
| ElevenLabs Agents | $0.08 | $0.1048 | 1.31x | Model and telephony | 2026-08-05 |
| Bland AI | $0.11 | $0.124 | 1.13x | Telephony, per the pricing page FAQ | 2026-08-06 |
| Deepgram Voice Agent | $0.056 to 12 Sep 2026, then $0.075 | Not computable | n/a | Not itemised on the page | 2026-08-25 |
| Synthflow | None published | Not computable | n/a | No rate to add exclusions to; $30,000 a year floor | 2026-08-11 |
The cheapest advertised rate carries the largest multiple and the dearest carries the smallest. Millis advertises $0.02 and floors at $0.0665; Bland advertises $0.11 and floors at $0.124. Sorting this category on the headline sorts it almost exactly backwards, which is the single most useful thing a buyer can take from this table.
Vapi and Retell are the pair people actually shortlist, and computed all-in they are a fifth of a cent apart at $0.0965 against $0.0948, because Retell's $0.07 excludes the model just as Vapi's $0.05 does. The bottom two rows are there to show what the method cannot reach: a bundle with no published exclusion list and a vendor with no published rate both produce a blank rather than an estimate.
One caveat on the Bland row, and it is the vendor's rather than ours. Its pricing page FAQ excludes telephony while its docs index states the opposite. If the docs are right, that multiple is 1.00 and the floor equals the headline.