Vapi, Retell, ElevenLabs and Twilio publish rates from roughly $0.05 to $0.31 per minute, but each pricing page measures a different thing. What matters more than the rate: only ElevenLabs documents an EU data residency option, and it sits on the enterprise tier. From 2 August 2026 the duty to tell callers they are speaking to AI sits on the company deploying the agent, not on the platform.

Most comparisons of AI voice agent platforms rank the same four or five names on latency and price per minute. Useful, up to a point. They are also written for a US buyer, and they skip the two questions that decide the outcome for a European company: where the recordings of your customers' calls physically sit, and what you have to tell the person on the other end of the line.

That second question stops being theoretical on 2 August 2026. Transparency duties under the EU AI Act start applying, and they land on the company running the agent - not on the vendor selling you the minutes.

So this is a comparison with a different order of priorities. Price is in here, with numbers pulled from each vendor's own pricing page rather than from a competitor's blog post, all checked on 26 July 2026. But it comes after the two questions that are harder to fix later.

One boundary before we start: this article is about choosing between platforms. Whether you need a voice agent at all, and how to roll one out once you have chosen, we answered separately in voice AI for business.

What should you compare when choosing an AI voice agent platform?

Compare five things, in this order: where call data is stored and processed, what the all-in cost per minute actually is, how the agent hands a live call to a human, how it behaves when it fails, and who maintains it after launch. Voice quality sits below all five in 2026, because every serious platform now sounds good enough down a phone line. The order matters more than the list. Four of the five you can change after go-live with a config edit, a new prompt or a renegotiated contract. The first one you cannot, which is why it goes at the top.

Data location is first because it is the only irreversible choice. Pricing you can renegotiate, prompts you can rewrite, voices you can swap in an afternoon. Moving a live phone system to a different provider because your legal team blocked the first one is a migration, and it lands after you have already trained the agent, written the scripts and pointed your numbers at it.

Cost is second because the published rate is not the cost. Every vendor draws the boundary of "per minute" somewhere different. Until you rebuild each quote on the same basis, you are comparing marketing decisions, not prices.

Escalation is third because it is the single feature customers notice. An agent that says "let me put you through to a colleague" and actually does it is a good experience. An agent that loops back to the main menu is the reason people stop calling you.

Failure behaviour is fourth because it is what a demo hides. Every platform sounds excellent answering a question it was built to answer. The difference shows up on the call where the caller mumbles a postcode, changes their mind halfway through and has a television on in the background.

Maintenance is fifth because it is the only line item that never stops. Someone has to update the agent when your prices, opening hours or booking system change. If nobody owns that, the quality of the answers decays quietly and you find out from a customer.

How much do AI voice agent platforms cost in 2026?

Published rates run from roughly $0.05 to $0.31 per minute, but each pricing page measures a different thing, so the comparison only works once you add back the parts each vendor leaves out. Some quote a thin platform fee with speech, language model and telephony billed separately. Others quote an all-in range that already contains those layers. A $0.05 platform fee and a $0.31 all-in rate can end up within a few cents of each other on the same call volume. The figures below come from each vendor's own pricing page, checked on 26 July 2026.

Platform Published price Billed on top
Vapi $0.05/min platform fee Speech-to-text, language model and text-to-speech at cost - $0 if you bring your own API keys. Extra concurrency $10 per line per month above 10. SMS $0.005 per message
Retell AI $0.07-$0.31/min all-in Phone numbers $2/month, concurrency $8/month above the 20 included, knowledge base $8/month, PII removal $0.01/min
ElevenLabs Agents $0.08/min for additional minutes, $0.16 burst above your concurrency limit Language model and telephony billed separately. Plans run from free (15 min) to $990/month (12 375 min)
Twilio ConversationRelay $0.07/min Voice and telephony charged separately on standard Twilio rates

Three things to take from that table.

Bring-your-own-keys changes the maths. Vapi's $0.05 looks like the cheapest entry until you add speech and language model costs, at which point the total lands in the same range as everyone else. It is genuinely cheaper only if you already have negotiated model pricing and someone who wants to manage four vendor relationships instead of one.

The burst clause is the one that surprises people. ElevenLabs doubles the per-minute rate to $0.16 when you exceed your concurrency limit. That is not a penalty for heavy use in general - it is a penalty for heavy use in the same fifteen minutes. If your calls cluster on Monday mornings and the first day back after a holiday, model the peak, not the monthly average.

Normalise before you compare. Rebuild every quote as the same four lines - platform, speech, language model, telephony - on one set of assumptions: your real average call length, your real peak concurrency, a month of your real volume rather than a round number. Add the small recurring items, because phone numbers, extra concurrent lines and features like PII removal are cheap individually and add up. It takes an afternoon and it usually reorders the shortlist.

For a full breakdown of what an AI deployment costs beyond the platform line item, we wrote a separate guide on AI costs for business.

Where does your call data live, and why does it matter in the EU?

Of the four platforms compared here, only ElevenLabs publishes a documented EU data residency option for its agents product, it sits on the enterprise tier, and even there storing data in the EU does not mean every processing step happens in the EU. For a company recording customer calls under GDPR, this is the shortest path to a project your legal team blocks in month three. It is also the one answer you should get in writing before you build anything, because unlike price or voice it cannot be changed later without redoing the work.

What each vendor publishes:

Now the part most comparisons get wrong in the other direction. All four of these platforms are lawfully usable by a European company. Standard Contractual Clauses are a valid transfer mechanism under GDPR, most of the European economy runs on US infrastructure, and "American vendor" is not a compliance finding. Anyone telling you otherwise is selling something.

The point is narrower, and it is about sequence rather than nationality: decide where the data sits before you build, because it is the one variable you cannot change with a config edit. A US-hosted platform chosen deliberately, documented in your processing records and covered by a signed data processing agreement is a defensible position. The same platform chosen by accident, discovered by your client's auditor eighteen months in, is an expensive one. If your customers are hospitals, law firms or public institutions, get the answer during the trial, not after.

We went through this exercise ourselves when building our own voice agent, and it is also why the question of where company data ends up in AI tools deserved its own article.

What does the EU AI Act require from a voice agent?

From 2 August 2026, if a person interacts with your AI system, you have to make sure they know it - and the obligation sits on the company deploying the agent, not on the platform that supplied it. For a phone agent, that means disclosure at the start of the call.

Three specifics worth getting right, because the internet is full of confident wrong versions:

The exemption is narrow. Article 50 does not require disclosure where it is obvious from the circumstances that the person is dealing with an AI. A voice agent designed to sound like a person does not clear that bar. If your agent introduces itself with a human first name and a natural voice, assume you need to say it.

It applies to the deployer. Buying from Retell, Vapi, ElevenLabs or Twilio does not transfer this obligation to them. Their compliance documentation helps you, but the duty to tell the caller is yours. Obligations for general-purpose model providers are a separate part of the regulation and are not your problem as a user.

The penalty tier for Article 50 is real but not the headline one. Breaches of transparency duties fall under the tier of up to 15 million EUR or 3% of worldwide annual turnover. The 35 million EUR figure quoted in most articles applies to prohibited practices such as social scoring, not to a support line that forgot to introduce itself. For SMEs the regulation sets the fine at the lower of the two values rather than the higher.

The practical version is one sentence at the top of the call. Something like "Hi, you're speaking with an automated assistant from [company] - I can help with bookings and pricing, or put you through to a colleague." That sentence costs you nothing and settles the requirement.

Note what this does to the platform decision: almost nothing. None of the four can carry the obligation for you, so no vendor's compliance page is a reason to pick it. What a platform can do is make the disclosure easy to implement and hard to lose - a first-turn message you control, and a script that survives every later edit. Check that in the trial.

Our full breakdown of what the regulation means for an ordinary company is in the EU AI Act guide.

Which platform fits which situation?

Match the platform to who is on the call and to what happens when the agent gets it wrong, not to the benchmark table. There is no winner here and we are not ranking them - each of the four is the right answer for a different constraint, and your job is to work out which constraint binds you first. Four situations cover most of what we see.

You want to test whether voice AI works for you at all. Start where you can get a number answering within a day and cancel without a conversation. ElevenLabs publishes a free tier with 15 minutes, and Retell bills per minute from $0.07 with no seat commitment. Run twenty real calls before you compare anything, because the first twenty will change what you think you are buying.

You already have a technical team and negotiated model pricing. Vapi's bring-your-own-keys model is built for exactly this and will be the cheapest at volume. It also assumes you are comfortable owning the integration when one of the four underlying services changes behaviour, which they do, usually without asking you first.

Your calls are the front door to the business. Reception, bookings, out-of-hours enquiries. Here the deciding features are escalation to a human, reliable handling of accents and interruptions, and clean write-back to your CRM. Pay for the tier with proper support rather than the cheapest minute - at this volume the difference between plans is smaller than the cost of one bad Monday.

You are in a regulated sector or sell to the public sector. Start from the data residency answer and work backwards, in the order set out above. The shortlist may end up being one name, or a European provider outside this comparison, and that is a legitimate outcome rather than a failed evaluation.

What do AI voice agent platform comparisons usually leave out?

The platform is maybe a third of the work of running a voice agent. The rest is the part nobody demos: what the agent says when it does not know, how it hands the call over, and who fixes it in month four. This is also where the money goes, which is why two companies on identical per-minute rates can get completely different results from the same vendor.

Three things we have learned building and running voice agents:

Silence is worse than a wrong answer. A one-second pause feels normal in text and broken on a phone call. Callers fill it - they repeat themselves, talk over the agent, or hang up. Test with real background noise and someone who interrupts mid-sentence, because that is what your customers do and it is not what a scripted demo does.

Escalation needs a destination, not a promise. "I'll pass you to a colleague" is only worth saying if someone picks up. Decide what happens at 22:00, during lunch and when the one person who knows the answer is on holiday. Voicemail with a callback commitment beats a transfer into an empty queue. Whoever picks up also needs to know what the agent can and cannot do, which is a training problem rather than a configuration one - it is why our AI training runs alongside the build rather than after it.

Maintenance is the recurring cost nobody budgets. Change your prices and the agent needs to know. Add a service, same. Switch CRM, rebuild the integration. Deloitte's 2026 State of AI report found only about one in five organisations has a mature governance model for autonomous agents, while 17% have deployed agents and more than 60% expect to within two years. The gap between those numbers is where unmaintained pilots quietly rot.

That pattern is not unique to voice. We wrote about why most deployments stall before production in why AI pilots fail.

How do you pilot a voice agent before committing?

Run a two-week bake-off: two platforms from your shortlist, one call type, one script, the same test calls on both, judged on the things that actually differ between vendors. Anything broader turns into an evaluation project that never ends, and anything narrower just proves that voice agents can talk, which we already know. The output of this exercise is a platform decision, not a business case.

How to run it:

  1. Build the same agent twice. One call type, one script, one escalation rule, deployed on both platforms. Identical inputs are the whole point - if the prompts differ, you are comparing your own writing, not the vendors.
  2. Test what a demo will not show you. Interrupt the agent mid-sentence. Call from a car. Give a half-finished answer. Ask something outside the script and listen to how it fails. Then call at your busiest hour on both platforms and see what concurrency does to response time.
  3. Listen to the recordings yourself. At least thirty per platform, not a dashboard summary of them. Nobody has ever changed their mind about a voice agent by reading a transcript.
  4. Price the winner on your real numbers. Take the normalised four-line cost from the section above and apply your actual volume and peak, not the vendor's example call.

Alongside the technical test, send both vendors the same four questions in writing: where is call data stored, where is it processed, who are the subprocessors, and will you sign a data processing agreement. The answers matter, and so does the response time - a vendor who takes three weeks to answer a standard procurement question during a sales process is showing you what support looks like after you sign.

Then decide on one criterion, agreed before you start. Usually it is calls fully handled without a human, sometimes it is cost at peak. Once the platform is chosen, the rollout is its own project with its own metric, and we set that out step by step in voice AI for business.

Frequently asked questions

What is the best AI voice agent platform in 2026?

There is no single best one. For a fast test, ElevenLabs and Retell get you live quickest. For a technical team with existing model contracts, Vapi is cheapest at volume. For procurement-heavy organisations, Twilio usually clears legal review fastest. The better question is which constraint binds you first: budget, data location or time to launch.

How much does an AI voice agent cost per minute?

Published rates run from about $0.05 to $0.31 per minute depending on what the vendor bundles. Vapi charges $0.05 for the platform with speech and model costs on top, Retell quotes $0.07 to $0.31 all-in, ElevenLabs charges $0.08 for additional minutes, and Twilio ConversationRelay is $0.07 plus telephony. Budget for peak concurrency, not average volume.

Do AI voice agent platforms offer EU data residency?

Rarely, and usually not on standard plans. ElevenLabs offers EU residency as an enterprise feature and notes that processing may still happen outside the chosen region unless you use Zero Retention Mode with the API. Retell states that data is primarily stored and processed in the United States, with Standard Contractual Clauses covering EEA transfers. Ask for this in writing before you build.

Does the EU AI Act require me to tell callers they are talking to AI?

Yes, from 2 August 2026, where a person interacts with an AI system they must be informed unless it is obvious from the circumstances. A voice agent built to sound human does not meet that exemption. The obligation falls on the company deploying the agent, and transparency breaches sit in the tier of up to 15 million EUR or 3% of worldwide annual turnover.

Is a European voice agent platform automatically the safer choice?

No. Standard Contractual Clauses are a lawful transfer mechanism, so a US platform chosen deliberately and covered by a signed data processing agreement is defensible. An EU-hosted vendor with vague subprocessor documentation is not automatically better. Judge the paperwork, not the flag - and get storage, processing and subprocessors confirmed in writing either way.

Do I still need Twilio if I use Vapi, Retell or ElevenLabs?

You need telephony from somewhere, and how it is billed differs. Retell sells phone numbers at $2/month within an all-in per-minute rate. ElevenLabs bills telephony separately from its per-minute agent price, as does Vapi, which lets you bring your own provider keys. Twilio can be the carrier underneath any of them.

Can I switch voice agent platforms later?

Technically yes, and prompts and call flows port over with moderate effort. What does not port cleanly is your phone number configuration, your recording history and any compliance paperwork tied to a specific processor. Treat the data residency decision as the expensive one and the rest as reversible.

Which AI voice agent platform is best for 24/7 lead qualification and appointment booking?

Any of the four can take a call at 3 a.m. The difference is in what happens after the caller says yes. For this use case, judge the platform on five things: a native handoff to a human during business hours, a write-back to your CRM or calendar without a middleware workaround, call recording with consent handling that matches your jurisdiction, the AI disclosure the EU AI Act requires at the start of the call, and the price at your peak hour rather than your average. Vapi and Retell win on integrations, ElevenLabs on voice quality, Twilio when you already run telephony there. If the booking logic has many exceptions - multiple locations, staff with different hours, deposits - the platform stops being the bottleneck and you need a custom agent on top of it.

Want to know which of these fits your call volume?

Tell us how many calls you get, when they cluster and who your customers are. We will tell you which platform we would pick and why - including when the honest answer is that you do not need one yet. We build and run AI agents for a living, including our own voice agent, so we know what the demos leave out.

Let's talk voice agents