Reviewed by Jonathan West · Updated Aug 22, 2026

ElevenLabs Agents vs Vapi: Which Voice Agent Platform Should You Pick?

A managed, bundled speech stack sold at one published rate versus a bring-your-own-keys orchestration layer you assemble yourself.

Reviewed by Jonathan West · Updated Aug 22, 2026

ElevenLabs Agents and Vapi take opposite approaches to the same problem. ElevenLabs gives you a complete conversational stack at a published per-minute price. Vapi provides the orchestration layer and lets you choose the rest of the stack yourself.

Both can answer phone calls, carry on real conversations, and trigger functions during a call. Both also rely on a language model and a phone carrier. And neither is always the cheaper option.

This page compares their fully loaded per-minute costs, voice quality, model flexibility, telephony, setup effort, and vendor lock-in. Vapi clearly comes out ahead in several areas, and we'll say so when it does. Voice-agent pricing changes frequently, though, so always confirm current rates on each vendor's pricing page before setting a budget.

ElevenLabs Agents vs. Vapi: Side-by-Side

DimensionElevenLabs AgentsVapi
What it isManaged conversational platform with speech, turn-taking and voices bundledModular orchestration layer that wires third-party providers into a call
Published platform rate$0.08 per minute for standard call minutes; $0.16 per minute burst above your concurrency limit$0.05 per minute platform fee; provider costs passed through at cost
What the rate coversSpeech-to-text, text-to-speech, voices and the turn-taking modelOrchestration only; speech-to-text and text-to-speech are billed by your providers
LLM billingBilled separately on top, based on usageAt cost, or $0 on your Vapi bill if you bring your own API key
LLM choiceSupported model list plus a custom LLM endpointAny provider, including custom and self-hosted endpoints
Voice qualityThe strongest single reason to stay; voices are the core productAs good as the voice vendor you plug in, and you pay that vendor directly
Plan fees and concurrencyFree $0 (4 concurrent) to Business $990/mo (40 concurrent); plan fee is prepaid minutesBuild is usage-based with 10 lines included; extra lines $10 per line per month
TelephonyNative Twilio integration plus SIP trunking; carrier billed separatelyTwilio, Vonage, Telnyx or your own SIP trunk; carrier billed separately
ComplianceHIPAA BAA on Enterprise tier only, with Zero Retention Mode requiredSOC 2, HIPAA and PCI on enterprise; HIPAA add-on listed at $2,000/mo
Build effortLowest; one account, one dashboard, one billHigher; you choose, wire and monitor every provider yourself
Best forTeams where voice quality sells the product and speed matters more than unit costEngineering teams with volume, opinions about the stack, and time to tune it
Bottom linePick it for the voices and the short path to a live numberPick it for model freedom, provider-level cost control and no burst penalty

What Is the Difference Between ElevenLabs Agents and Vapi?

The core difference is how much of the voice stack you buy already assembled. ElevenLabs Agents ships speech-to-text, text-to-speech, a voice library and a turn-taking model as one product. Vapi ships the orchestration and expects you to supply the rest.

A voice agent is really four moving parts. It hears the caller, decides what to say, speaks the reply, and knows when the caller has stopped talking. That last part, turn-taking, is where most agents feel broken.

ElevenLabs bundles all four and charges one rate for three of them. Its own documentation describes a fine-tuned speech-to-text model, your chosen language model, a TTS engine, and a proprietary turn-taking model working together.

Vapi treats each part as your decision. You pick the transcription vendor, the model, and the voice vendor, and Vapi glues them into a phone call. That freedom is the whole product.

So the real question is not which platform is better. It is whether your team wants to own those choices or rent them.

  • ElevenLabs Agents: bundled speech stack, one published rate, one dashboard.
  • Vapi: orchestration only, bring your own keys, price every layer yourself.
  • Both still need a language model and a phone carrier billed on top.

Deciding between ElevenLabs Agents and Vapi for your business? We can map both to your workflows, data, and compliance needs.

Book a Consultation

Pricing: What Each One Really Costs per Minute

Neither platform's headline rate is what you pay. ElevenLabs Agents publishes $0.08 per minute for standard call minutes, with the language model and telephony billed separately on top. Vapi publishes a $0.05 per minute platform fee, with speech, model and voice costs passed through at cost.

Those two numbers are not comparable, because they cover different things. The $0.08 rate already includes transcription, the voice and the turn-taking layer. The $0.05 fee includes none of them.

That turns the comparison into one clean question. Can you buy transcription and text-to-speech on Vapi for less than $0.03 per minute combined? If yes, Vapi is cheaper before the model. If no, the bundle wins.

Work a 5,000-minute month. On ElevenLabs Agents the speech side is 5,000 times $0.08, or $400, plus your model and carrier bills. On Vapi the floor is 5,000 times $0.05, or $250, plus transcription, voice, model and carrier.

Here is the detail most comparisons miss. The ElevenLabs plan fee is not a volume discount. Divide any tier by its included minutes and you land on the same $0.08: Starter is $6 for 75 minutes, Pro is $99 for 1,238 minutes, Business is $990 for 12,375 minutes. The subscription is prepaid minutes, not a cheaper rate.

Vapi has the opposite shape. Its $0.05 floor never drops on the Build plan either, but every layer above it is a lever you can pull the same day. Swap a premium voice for a cheaper one and your bill changes immediately.

ElevenLabs Agents: $0.08/min covering speech and turn-taking, plus LLM and telephony. Vapi: $0.05/min orchestration only, plus every other layer at your providers' rates. Confirm both on the vendor pricing pages before budgeting.

Concurrency and the Burst Rate Nobody Budgets For

Concurrency is where the two pricing models diverge hardest, and it catches teams out. ElevenLabs Agents caps concurrent calls by plan tier and charges $0.16 per minute in burst pricing when you exceed that cap. Vapi includes 10 concurrent lines and sells extra lines at $10 per line per month.

The published ElevenLabs tiers run from 4 concurrent calls on Free to 40 on the $990 Business plan. Pro at $99 allows 20.

Think about what that means on a busy Monday. Ten simultaneous callers on a Creator plan is fine. Twelve is not, and the extra two cost double for every minute they stay on the line.

Vapi's model has no doubling. An extra line is a flat monthly fee whether it is busy or idle, which is worse for spiky traffic and better for steady traffic.

This is a real planning difference. Model your peak hour, not your average hour, and price both platforms against that peak.

Our guide to the platform's production limits covers the wider set of production ceilings, and our pricing guide covers plan-tier caps in detail.


Voice Quality: The Row ElevenLabs Wins

Voice quality is the clearest ElevenLabs win, and it is not close. Voices are the company's original product, and the agent platform is built on top of that engine rather than buying it in.

Its documentation lists thousands of voices across dozens of languages available to agents. You are not integrating a voice vendor; you are already inside one.

Vapi can reach the same quality, because you can plug the same voice vendor into it. The difference is that you then pay that vendor directly, on their rate card, on top of the $0.05 platform fee.

So the premium-voice case often favours the bundle. If your shortlist of acceptable voices only contains ElevenLabs voices, buying them inside the ElevenLabs product usually removes a markup layer rather than adding one.

The reverse is also true. If a mid-tier voice is good enough for a booking confirmation, Vapi lets you buy that cheaper voice and keep the savings.

Listen before you decide. Record the same script on both stacks over a real phone line, not through a laptop speaker, because phone codecs flatten a lot of the difference.


LLM Freedom and Model Lock-In

Vapi wins on model freedom, but the gap is smaller than most comparisons claim. ElevenLabs Agents supports a list of approved models and also accepts a custom LLM endpoint, so it is not a closed box. Vapi supports any provider, including custom and self-hosted endpoints, with no approved list at all.

The bigger difference is billing. On Vapi you can bring your own API key, and the model cost then drops to $0 on your Vapi bill because you pay the provider directly.

On ElevenLabs Agents the model is billed separately on top, based on usage. That is still pass-through, but it runs through their meter rather than yours.

Model choice also narrows under compliance. ElevenLabs documentation states that HIPAA mode limits you to approved models from Google, Anthropic, ElevenLabs-hosted options, or a custom LLM integration.

For most teams this row decides less than it looks. A booking agent works fine on a mid-tier model, and the model is rarely the expensive part of a call.

It matters most if you have a fine-tuned model of your own. In that case Vapi is the straightforward choice.


Latency and What Actually Causes Lag

Neither vendor publishes a millisecond figure you can hold them to, so treat any number you read elsewhere as someone else's test, not a spec. ElevenLabs documentation describes a low-latency TTS model without stating a figure. Vapi's latency depends entirely on the providers you choose.

Structure still tells you something. A bundled stack makes fewer hops between vendors, because transcription, voice and turn-taking sit in one system.

An orchestrated stack adds a network hop per provider. A well-chosen set of providers can still be fast, and a badly chosen one will not be.

The failure that actually annoys callers is rarely raw latency. It is the agent talking over someone who paused mid-sentence to check a reference number.

That is turn-taking, and it is a model, not a speed setting. ElevenLabs ships one as part of the platform. On Vapi, interruption handling is configuration you tune.

Test it the ugly way. Call your own agent from a car with the window down and see which one survives.


Telephony: Who Owns the Phone Number

Both platforms connect to real phone lines, and neither includes the carrier bill. ElevenLabs Agents offers a native Twilio integration and SIP trunking to existing telephony. Vapi supports Twilio, Vonage, Telnyx and your own SIP trunk.

Vapi is the more flexible of the two here. More supported carriers means more room to negotiate rates or keep an existing contract.

SIP support on both sides matters more than it sounds. It is the difference between porting your main business number and asking customers to learn a new one.

Ask one question before you sign. Whose account holds the number, yours or the platform's?

If the number lives in your own Twilio account, switching platforms later is a configuration change. If it lives inside the platform, it becomes a porting request with a support queue attached.

That single detail cuts your exit cost more than any pricing row on this page.


Build Effort and Time to a Live Agent

ElevenLabs Agents gets you to a working phone number faster, and for many teams that is the deciding factor. One account covers speech, voices and turn-taking, so there is one signup, one dashboard and one bill.

Vapi asks for more setup before the first call. You need accounts and keys for transcription, the model and the voice, plus a carrier.

That cost is front-loaded, not permanent. Once the stack is wired, changing a provider is a config edit.

Staffing decides this row. A team with one engineer and a launch date should weigh the bundle heavily.

A team with a platform engineer who already runs vendor keys will barely notice Vapi's setup step, and will get more back for it later.

Be honest about who maintains it in month six. Voice agents need prompt fixes, transfer rules and monitoring long after launch day.


Compliance and Regulated Call Data

Both platforms gate their strongest compliance behind their most expensive tier, so regulated teams should price that in early. ElevenLabs documentation states that a BAA is only available for Enterprise tier subscriptions, and that Zero Retention Mode must be engaged before PHI is handled. Vapi lists SOC 2, HIPAA and PCI on enterprise, with a HIPAA add-on listed at $2,000 per month and Zero Data Retention at $1,000 per month.

Read the trade in Zero Retention Mode carefully. ElevenLabs documentation notes that conversation analysis is then restricted to non-sensitive metadata, and that transcripts and audio containing PHI are not retained.

That is the right privacy behaviour and a real operational loss. You cannot review a disputed call recording that was never stored.

Clinics running modest call volume often find both routes expensive at small scale. It is worth pricing Retell AI in the same exercise, since it publishes bundled compliance across plans.

Whatever you pick, get the BAA question answered in writing before you build. Rebuilding a live agent because the contract does not cover PHI is an avoidable cost.

HIPAA on ElevenLabs Agents requires an Enterprise subscription, a signed BAA and Zero Retention Mode. Vapi lists a $2,000/month HIPAA add-on outside enterprise. Verify current terms with each vendor.

Lock-In: What It Costs to Leave Each One

Vapi is materially easier to leave, and that is worth real money over a few years. Because you own the provider accounts, your keys, your voice contract and often your phone numbers move with you.

Leaving ElevenLabs Agents means replacing four things at once. You need new transcription, new voices, new turn-taking behaviour and a new billing relationship.

The voices are the hard part. If callers know your brand voice, a cloned voice you licensed inside one platform is not something you carry out in a suitcase.

Prompts and call flows are the easy part. Those are text, and they port between platforms with light editing.

There is a middle path teams overlook. Build on Vapi but use ElevenLabs voices through it, and you keep the voice quality while owning the orchestration layer.

That setup costs more per minute than either pure route. It buys you the option to change your mind, which is sometimes the cheapest thing you can buy. The wider field of replacements is ranked on our ranked alternatives page.


Which Should You Choose?

Choose ElevenLabs Agents when the voice is part of the product and you need to launch soon. It fits brands where callers judge you in the first three seconds, small teams without a spare engineer, and any project where one bill beats four.

Choose Vapi when you have volume, engineers and opinions. It fits product teams embedding calls in software, anyone running a self-hosted or fine-tuned model, and any operation where saving $0.02 a minute is worth an afternoon of tuning.

Volume flips the answer faster than most teams expect. At a few hundred minutes a month the price gap is noise and build speed wins. At tens of thousands of minutes, provider-level control compounds.

Spiky traffic pushes you toward Vapi. Steady traffic with premium voices pushes you toward the bundle.

If you cannot decide, model your minutes with our voice agent cost calculator, then run both for two weeks on the same call flow. Measure cost per resolved call, not cost per minute, because a cheap agent that fails and transfers to a human is the expensive one.

  • Pick ElevenLabs Agents for voice quality, fastest launch and a single bill.
  • Pick Vapi for model freedom, provider-level cost control and lower exit cost.
  • Pick Vapi if your call volume spikes, to avoid burst pricing on concurrency.

The Verdict

ElevenLabs Agents and Vapi are priced for different buyers, not for the same buyer at different budgets. ElevenLabs publishes $0.08 per minute covering transcription, voices and turn-taking, with the language model and telephony on top. Vapi publishes $0.05 per minute for orchestration alone, with every other layer at your providers' rates.

Vapi wins on model freedom, provider-level cost control, carrier choice and exit cost. It also avoids the $0.16 per minute burst rate that ElevenLabs applies above your plan's concurrency limit, which matters for spiky call volume.

ElevenLabs wins on voice quality, turn-taking behaviour out of the box, and the shortest path from signup to a live phone number. For brands where the voice is part of the pitch, that is often worth the higher floor.

The honest split: if voice quality sells your product and you need to ship this month, take the bundle. If you have engineers, real volume, or a model of your own, take Vapi. Prices in this category move often, so re-check both pricing pages before you sign anything longer than month to month.

Sources & Disclaimer

Researched from primary vendor documentation and public regulator sources. Pricing and availability are accurate as of Aug 22, 2026 and can change — confirm current terms with each vendor before you buy.

Frequently Asked Questions

  • It depends on your speech stack. ElevenLabs Agents publishes $0.08 per minute covering transcription, voices and turn-taking, with the language model and telephony billed on top. Vapi publishes a $0.05 per minute platform fee and passes speech, model and voice costs through at cost. Vapi is cheaper only if your combined transcription and text-to-speech bill comes in under about $0.03 per minute.
  • Yes. Vapi is a bring-your-own-provider platform, so you can plug in a premium voice vendor and pay that vendor directly. You keep the voice quality while owning the orchestration layer. The trade is that you pay the voice bill and the $0.05 platform fee separately, so the total is usually higher than either pure route.
  • Yes, within limits. The platform supports a list of approved models and also accepts a custom LLM endpoint. Vapi is broader, supporting any provider including self-hosted endpoints, and bringing your own API key drops the model cost to $0 on your Vapi bill. Under HIPAA mode, ElevenLabs documentation restricts model choice to approved options.
  • Burst pricing is the $0.16 per minute rate ElevenLabs publishes for call minutes that exceed your plan's concurrency limit. Published concurrency runs from 4 simultaneous calls on the free plan up to 40 on the $990 Business plan. Vapi has no equivalent doubling; it includes 10 lines and charges $10 per extra line per month.
  • Neither vendor publishes a millisecond figure, so any number you see is someone else's test rather than a guarantee. A bundled stack makes fewer hops between vendors, while an orchestrated stack depends on the providers you choose. Test both on a real phone line with background noise, because turn-taking behaviour affects perceived speed more than raw latency does.
  • Only under specific conditions. ElevenLabs documentation states that a BAA is available for Enterprise tier subscriptions only, and that Zero Retention Mode must be engaged before protected health information is handled. In that mode, transcripts and audio containing PHI are not retained and conversation analysis is limited to non-sensitive metadata. Confirm current terms with the vendor before handling patient data.
  • Vapi. Because you hold the provider accounts, API keys and often the phone numbers, moving to another platform is mostly reconfiguration. Leaving ElevenLabs Agents means replacing transcription, voices, turn-taking and billing at once, and a cloned brand voice licensed inside one platform does not transfer. Keeping your phone numbers in your own carrier account cuts exit cost on either platform.

Not sure which voice stack fits your call volume?

Layer3Labs builds custom voice-agent systems on whichever platform actually fits your call flows, budget and compliance needs. We will model your real minutes on both stacks before you commit to either.

Book a Free AI Workflow Audit