Rank #7 of 9 in Voice Agent Platforms
Install
Showcase


Try itExperimental
See what an agent can do with ElevenLabs Agents before you ever sign up. Pick a story: recorded sessions replay real probe-harness transcripts; sandboxed self-drive sessions are designed and gated (docs/TRY-IT.md).
$npx -y @elevenlabs/cli --versionrecorded session — replayed, not liveVerified integrations
Connections to other tracked products — hover a chip for the verbatim evidence quote behind it.
By theme — the product's score on each story themeBy theme
Agent building — building agents — abstractions, tool wiring, control flowAgent buildingevidence →
Building agents — abstractions, tool wiring, control flow
Agenticness — how well agents can access and operate the productAgenticnessevidence →
How well agents can access and operate the product
Automation depth — how much of the product can run unattendedAutomation depthevidence →
How much of the product can run unattended
Compliance trust — stories about compliance trust in this arenaCompliance trustevidence →
Stories about compliance trust in this arena
Deployment scale — stories about deployment scale in this arenaDeployment scaleevidence →
Stories about deployment scale in this arena
Latency turntaking — stories about latency turntaking in this arenaLatency turntakingevidence →
Stories about latency turntaking in this arena
Openness — open source, data portability, and self-hosting storiesOpennessevidence →
Open source, data portability, and self-hosting stories
Pricing plans — plan structure and value — what each tier costs and what it unlocksPricing plansevidence →
Plan structure and value — what each tier costs and what it unlocks
Privacy posture — data-handling and privacy storiesPrivacy postureevidence →
Data-handling and privacy stories
Telephony — stories about telephony in this arenaTelephonyevidence →
Stories about telephony in this arena
Testing analytics — stories about testing analytics in this arenaTesting analyticsevidence →
Stories about testing analytics in this arena
Tools function calling — stories about tools function calling in this arenaTools function callingevidence →
Stories about tools function calling in this arena
Transcription recording — stories about transcription recording in this arenaTranscription recordingevidence →
Stories about transcription recording in this arena
Voices tts — stories about voices tts in this arenaVoices ttsevidence →
Stories about voices tts in this arena
Story verdicts — every judged story with its evidenceStory verdicts
Follow the green: where the map greys out is where ElevenLabs Agents stops today. ✓ full · ~ partial · ! disputed · — none · n/a not applicable.
Agent building — building agents — abstractions, tool wiring, control flowAgent building
Building agents — abstractions, tool wiring, control flow
Agent ops
Build
Agenticness — how well agents can access and operate the productAgenticness
How well agents can access and operate the product
API surface
Drive the product through a documented public API
✓9/10
unlocks → Webhooks · Scoped API keys · Machine-readable spec · Versioning policy · Full data export
Subscribe to events via webhooks
—0/10
Build against official SDKs
~4/10
Issue scoped/least-privilege API credentials for an agent
—0/10
Connect an agent via an official MCP server
✓9/10
Download a machine-readable API spec (OpenAPI or equivalent)
—–
Rely on versioned APIs with a documented deprecation policy
—–
Test against a sandbox environment without touching production data
~4/10
Explore an interactive API reference with runnable examples
—–
Docs for agents
Point an agent at llms.txt or agent-oriented docs
✓9/10
Agentic features
Delegate tasks to a built-in AI assistant inside the product
~5/10
Operate the product with natural-language commands
✓8/10
Plug MCP servers into this product so it can use their tools
✓8/10
Get AI-generated insights and suggestions from my data inside the product
~5/10
Set up automations that run autonomously in the background
~4/10
Automation depth — how much of the product can run unattendedAutomation depth
How much of the product can run unattended
Compliance trust — stories about compliance trust in this arenaCompliance trust
Stories about compliance trust in this arena
Deployment scale — stories about deployment scale in this arenaDeployment scale
Stories about deployment scale in this arena
Latency turntaking — stories about latency turntaking in this arenaLatency turntaking
Stories about latency turntaking in this arena
Openness — open source, data portability, and self-hosting storiesOpenness
Open source, data portability, and self-hosting stories
Pricing plans — plan structure and value — what each tier costs and what it unlocksPricing plans
Plan structure and value — what each tier costs and what it unlocks
Privacy posture — data-handling and privacy storiesPrivacy posture
Data-handling and privacy stories
Telephony — stories about telephony in this arenaTelephony
Stories about telephony in this arena
Call control
Run batch outbound call campaigns with scheduling and throughput controls
—0/10
Provision phone numbers and run both inbound and outbound calls through the platform's API
~6/10
Connect my own carrier or PBX via SIP trunking (or import Twilio/Telnyx numbers) instead of being locked to bundled telephony
~7/10
Testing analytics — stories about testing analytics in this arenaTesting analytics
Stories about testing analytics in this arena
Tools function calling — stories about tools function calling in this arenaTools function calling
Stories about tools function calling in this arena
Transcription recording — stories about transcription recording in this arenaTranscription recording
Stories about transcription recording in this arena
Voices tts — stories about voices tts in this arenaVoices tts
Stories about voices tts in this arena
Sorted by importance (agentic first) (high → low) · 61/61 stories · click a row’s chevron for the rationale and evidence
Connect an agent via an official MCP server G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 3 | full | 9/10 | Tprobed⚿ | |
Drive the product through a documented public API G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 3 | full | 9/10 | Tprobed⚿ | |
Plug MCP servers into this product so it can use their tools G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 3 | full | 8/10 | Cclaimed | |
Delegate tasks to a built-in AI assistant inside the product G Agentic features | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 3 | partial | 5/10 | Tprobed⚿ | |
Point an agent at llms.txt or agent-oriented docs G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | full | 9/10 | Tprobed | |
Use an official CLI G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | full | 9/10 | Tprobed | |
Operate the product with natural-language commands G Agentic features | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | full | 8/10 | Tprobed⚿ | |
Run the product headlessly / in CI for automation G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | partial | 6/10 | Tprobed | |
Get AI-generated insights and suggestions from my data inside the product G Agentic features | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | partial | 5/10 | Cclaimed | |
Build against official SDKs G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | partial | 4/10 | Tprobed | |
Set up automations that run autonomously in the background G Agentic features | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | partial | 4/10 | Cclaimed | |
Issue scoped/least-privilege API credentials for an agent G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | 0/10 | ⚿ | |
Subscribe to events via webhooks G Agent access | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | 0/10 | ||
Download a machine-readable API spec (OpenAPI or equivalent) G Api quality | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | untested | none yet | |
Explore an interactive API reference with runnable examples G Api quality | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | untested | none yet | |
Rely on versioned APIs with a documented deprecation policy G Api quality | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 2 | none | untested | none yet | |
Test against a sandbox environment without touching production data G Api quality | ai-native user | Agenticness — how well agents can access and operate the productAgenticness | 1 | partial | 4/10 | Cclaimed | |
Build a working phone voice agent — prompt, voice, and phone number — and take my first live call within an hour C Build | developer | Agent building — building agents — abstractions, tool wiring, control flowAgent building | 3 | partial | 7/10 | Cclaimed | |
My agent can call external APIs and custom functions mid-conversation and speak the result without awkward dead air C Tools | developer | Tools function calling — stories about tools function calling in this arenaTools function calling | 3 | full | 7/10 | Cclaimed | |
Define rules that trigger actions automatically on events G | ai-native user | Automation depth — how much of the product can run unattendedAutomation depth | 3 | partial | 6/10 | Cclaimed | |
My coding agent can provision a complete voice agent end to end — create the agent, attach a number, and place a call — through the API, CLI, or MCP without touching the dashboard C Agent ops | ai-native user | Agent building — building agents — abstractions, tool wiring, control flowAgent building | 3 | partial | 6/10 | Tprobed⚿ | |
Provision phone numbers and run both inbound and outbound calls through the platform's API C Numbers | developer | Telephony — stories about telephony in this arenaTelephony | 3 | partial | 6/10 | Cclaimed | |
Rely on the agent to handle interruptions (barge-in) gracefully — stopping speech, updating context, and recovering the turn C Turn taking | developer | Latency turntaking — stories about latency turntaking in this arenaLatency turntaking | 3 | partial | 6/10 | Cclaimed | |
See documented end-to-end voice latency numbers or tuning guidance backing the platform's speed claims C Latency | platform-engineer | Latency turntaking — stories about latency turntaking in this arenaLatency turntaking | 3 | partial | 4/10 | Cclaimed | |
Prevent my data from being used to train AI models G | ai-native user | Privacy posture — data-handling and privacy storiesPrivacy posture | 3 | none | 0/10 | ||
Self-host the core product G | ai-native user | Openness — open source, data portability, and self-hosting storiesOpenness | 3 | none | 0/10 | ||
Self-host the voice agent runtime from open-source code on my own infrastructure C Self host | platform-engineer | Deployment scale — stories about deployment scale in this arenaDeployment scale | 3 | none | 0/10 | ⚿ | |
Export all of my data in open formats and leave G | ai-native user | Openness — open source, data portability, and self-hosting storiesOpenness | 3 | none | untested | none yet | |
Ground the agent on my documents with a built-in knowledge base or RAG so it answers from my content C Personalization | developer | Agent building — building agents — abstractions, tool wiring, control flowAgent building | 2 | full | 8/10 | Cclaimed | |
Inject dynamic variables and per-caller context at call time so each conversation is personalized C Personalization | developer | Agent building — building agents — abstractions, tool wiring, control flowAgent building | 2 | full | 8/10 | Cclaimed | |
My voice agent can plug in MCP servers as tool sources so one integration grants it whole toolsets mid-call C Tools | ai-native user | Tools function calling — stories about tools function calling in this arenaTools function calling | 2 | full | 8/10 | Cclaimed | |
Test agents with simulated conversations or evals before putting them on real phone calls C Testing | developer | Testing analytics — stories about testing analytics in this arenaTesting analytics | 2 | full | 8/10 | Cclaimed | |
Connect my own carrier or PBX via SIP trunking (or import Twilio/Telnyx numbers) instead of being locked to bundled telephony C Sip | platform-engineer | Telephony — stories about telephony in this arenaTelephony | 2 | partial | 7/10 | Cclaimed | |
Do everything through the API that I can do in the UI G | ai-native user | Openness — open source, data portability, and self-hosting storiesOpenness | 2 | partial | 7/10 | Tprobed | |
The platform's AI reviews my calls for me — scoring quality, flagging failures, and analyzing resolution automatically C Analytics | ai-native user | Testing analytics — stories about testing analytics in this arenaTesting analytics | 2 | full | 7/10 | Cclaimed | |
Control data retention and deletion G | ai-native user | Privacy posture — data-handling and privacy storiesPrivacy posture | 2 | partial | 6/10 | Cclaimed | |
Design multi-step conversation flows in a visual builder with branching, states, and handoffs without writing code C Build | founder | Agent building — building agents — abstractions, tool wiring, control flowAgent building | 2 | partial | 6/10 | Cclaimed | |
Extract structured data from every call — outcomes, entities, dispositions — delivered via API or webhook after the call C Post call | developer | Tools function calling — stories about tools function calling in this arenaTools function calling | 2 | partial | 6/10 | Cclaimed | |
See call analytics — success rates, durations, outcomes, sentiment — in dashboards without building my own C Analytics | founder | Testing analytics — stories about testing analytics in this arenaTesting analytics | 2 | partial | 6/10 | Cclaimed | |
Choose from a broad voice library or plug in multiple TTS providers to get the voice I want C Voices | developer | Voices tts — stories about voices tts in this arenaVoices tts | 2 | partial | 5/10 | Cclaimed | |
Retrieve full call recordings and transcripts programmatically for every call C Recording | platform-engineer | Transcription recording — stories about transcription recording in this arenaTranscription recording | 2 | partial | 5/10 | Cclaimed | |
Run regulated workloads with HIPAA/BAA support, SOC 2, and data-residency options C Compliance | platform-engineer | Compliance trust — stories about compliance trust in this arenaCompliance trust | 2 | partial | 5/10 | Cclaimed | |
Get accurate real-time transcription with control over the STT provider, language models, or key terms C Transcription | developer | Transcription recording — stories about transcription recording in this arenaTranscription recording | 2 | partial | 4/10 | Cclaimed | |
Meet call-recording consent and disclosure obligations with per-call recording controls and configurable data retention C Compliance | founder | Compliance trust — stories about compliance trust in this arenaCompliance trust | 2 | partial | 4/10 | Cclaimed | |
Run conversations in multiple languages, including detecting and switching language mid-call C Build | developer | Agent building — building agents — abstractions, tool wiring, control flowAgent building | 2 | partial | 4/10 | Cclaimed | |
Choose where my data is stored (region/residency) G | ai-native user | Privacy posture — data-handling and privacy storiesPrivacy posture | 2 | none | 0/10 | ||
Clone a custom brand voice and use it for my agents, with a documented consent process C Voices | founder | Voices tts — stories about voices tts in this arenaVoices tts | 2 | none | 0/10 | ||
Run batch outbound call campaigns with scheduling and throughput controls C Campaigns | founder | Telephony — stories about telephony in this arenaTelephony | 2 | none | 0/10 | ||
Use model-based end-of-turn detection beyond simple VAD silence timeouts so the agent doesn't talk over slow speakers C Turn taking | developer | Latency turntaking — stories about latency turntaking in this arenaLatency turntaking | 2 | none | 0/10 | ||
Escalate a live call to a human with warm or blind transfer, passing context along C Call control | developer | Telephony — stories about telephony in this arenaTelephony | 2 | none | untested | none yet | |
Opt out of telemetry and usage tracking G | ai-native user | Privacy posture — data-handling and privacy storiesPrivacy posture | 2 | none | untested | none yet | |
Perform bulk operations across many items at once G | ai-native user | Automation depth — how much of the product can run unattendedAutomation depth | 2 | none | untested | none yet | |
Read the product's source under an open license G | ai-native user | Openness — open source, data portability, and self-hosting storiesOpenness | 2 | none | untested | none yet | |
Schedule recurring jobs or workflows G | ai-native user | Automation depth — how much of the product can run unattendedAutomation depth | 2 | none | untested | none yet | |
See documented concurrency limits and scale to many simultaneous calls without manual capacity begging C Scale | platform-engineer | Deployment scale — stories about deployment scale in this arenaDeployment scale | 2 | none | untested | none yet | |
See published per-minute or usage pricing and estimate cost per call before committing G Pricing | founder | Pricing plans — plan structure and value — what each tier costs and what it unlocksPricing plans | 2 | none | untested | none yet | |
The platform's own AI helps me author agents — generating or improving prompts, flows, and test cases from a description C Agent ops | ai-native user | Agent building — building agents — abstractions, tool wiring, control flowAgent building | 1 | partial | 5/10 | Tprobed⚿ | |
Monitor live calls in production and get alerts when agents misbehave or error rates spike C Monitoring | platform-engineer | Testing analytics — stories about testing analytics in this arenaTesting analytics | 1 | partial | 4/10 | Cclaimed | |
Enable noise suppression or audio filtering so the agent stays coherent on noisy real-world calls C Turn taking | developer | Latency turntaking — stories about latency turntaking in this arenaLatency turntaking | 1 | none | 0/10 | ||
My agent can send DTMF keypresses, navigate IVR menus, and detect or leave voicemail C Call control | developer | Telephony — stories about telephony in this arenaTelephony | 1 | none | 0/10 | ||
Version, review, and roll back my automations G | ai-native user | Automation depth — how much of the product can run unattendedAutomation depth | 1 | none | untested | none yet |
Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 49 stories with headroom
What would move ElevenLabs Agents’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.
Openness — open source, data portability, and self-hosting storiesExport all of my data in open formats and leave
nonemoves PA Scoreimpact 30
Missing: any documented export API/CLI command, supported open export formats (e.g., JSON/CSV), and confirmation of full data portability/account closure process.
Openness — open source, data portability, and self-hosting storiesSelf-host the core product
nonemoves PA Scoreimpact 30
Missing: any open-source repo or self-hosted deployment package, docs describing running the core service on one's own infrastructure, independent confirmation of self-hosting.
Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models
nonemoves PA Scoreimpact 30
Missing: explicit training-data opt-out policy or setting, terms-of-service language on model training use, any statement distinguishing enterprise vs free-tier data usage for training.
Deployment scale — stories about deployment scale in this arenaSelf-host the voice agent runtime from open-source code on my own infrastructure
nonemoves PA Scoreimpact 30
ElevenLabs Agents is entirely a managed/hosted service — the CLI and MCP server are clients/interfaces to ElevenLabs' cloud infrastructure, not open-source runtime code that can be deployed on a platform-engineer's own servers.
Agenticness — how well agents can access and operate the productIssue scoped/least-privilege API credentials for an agent
nonemoves agent-readyimpact 30
Missing: any docs on API key permission scopes, workspace role-based tokens, or restricted-credential issuance for agents.
Agenticness — how well agents can access and operate the productSubscribe to events via webhooks
nonemoves agent-readyimpact 30
The evidence describes 'webhook tools' that let an agent make outbound calls to external endpoints during a conversation (docs-14, docs-29, docs-31), which is the opposite of subscribing to platform-emitted events via webhooks.
Agenticness — how well agents can access and operate the productExplore an interactive API reference with runnable examples
nonemoves API qualityimpact 30
The evidence pack contains extensive markdown documentation for ElevenLabs Agents (quickstart, customization, tools, etc.) but nothing describes an interactive API reference page with runnable/'try it' examples — no mention of a Swagger/OpenAPI explorer, live code sandbox, or embedded runnable snippets.
Agenticness — how well agents can access and operate the productDownload a machine-readable API spec (OpenAPI or equivalent)
nonemoves API qualityimpact 30
The evidence pack documents the API, CLI, dashboard, and hosted MCP server for ElevenLabs Agents, but nowhere mentions a downloadable OpenAPI/Swagger spec or machine-readable schema for the API.
Showing the top 8 of 49 — every none/partial verdict in the story verdicts table is headroom.
Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.
Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map4 surfaces · 38 covered stories
Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.
docs38 stories
- My coding agent can provision a complete voice agent end to end — create the agent, attach a number, and place a call — through the API, CLI, or MCP without touching the dashboard
- The platform's own AI helps me author agents — generating or improving prompts, flows, and test cases from a description
- Build a working phone voice agent — prompt, voice, and phone number — and take my first live call within an hour
- Run conversations in multiple languages, including detecting and switching language mid-call
- Design multi-step conversation flows in a visual builder with branching, states, and handoffs without writing code
- Inject dynamic variables and per-caller context at call time so each conversation is personalized
- Ground the agent on my documents with a built-in knowledge base or RAG so it answers from my content
- Point an agent at llms.txt or agent-oriented docs
- Run the product headlessly / in CI for automation
- Plug MCP servers into this product so it can use their tools
- Connect an agent via an official MCP server
- Use an official CLI
- Drive the product through a documented public API
- Build against official SDKs
- Get AI-generated insights and suggestions from my data inside the product
- Set up automations that run autonomously in the background
- Delegate tasks to a built-in AI assistant inside the product
- Operate the product with natural-language commands
- Test against a sandbox environment without touching production data
- Define rules that trigger actions automatically on events
- Meet call-recording consent and disclosure obligations with per-call recording controls and configurable data retention
- Run regulated workloads with HIPAA/BAA support, SOC 2, and data-residency options
- See documented end-to-end voice latency numbers or tuning guidance backing the platform's speed claims
- Rely on the agent to handle interruptions (barge-in) gracefully — stopping speech, updating context, and recovering the turn
- Do everything through the API that I can do in the UI
- Control data retention and deletion
- Provision phone numbers and run both inbound and outbound calls through the platform's API
- Connect my own carrier or PBX via SIP trunking (or import Twilio/Telnyx numbers) instead of being locked to bundled telephony
- The platform's AI reviews my calls for me — scoring quality, flagging failures, and analyzing resolution automatically
- See call analytics — success rates, durations, outcomes, sentiment — in dashboards without building my own
- Monitor live calls in production and get alerts when agents misbehave or error rates spike
- Test agents with simulated conversations or evals before putting them on real phone calls
- Extract structured data from every call — outcomes, entities, dispositions — delivered via API or webhook after the call
- My voice agent can plug in MCP servers as tool sources so one integration grants it whole toolsets mid-call
- My agent can call external APIs and custom functions mid-conversation and speak the result without awkward dead air
- Retrieve full call recordings and transcripts programmatically for every call
- Get accurate real-time transcription with control over the STT provider, language models, or key terms
- Choose from a broad voice library or plug in multiple TTS providers to get the voice I want
Package docs6 stories
- My coding agent can provision a complete voice agent end to end — create the agent, attach a number, and place a call — through the API, CLI, or MCP without touching the dashboard
- Run the product headlessly / in CI for automation
- Use an official CLI
- Drive the product through a documented public API
- Build against official SDKs
- Do everything through the API that I can do in the UI
Agents docs5 stories
- Run conversations in multiple languages, including detecting and switching language mid-call
- See documented end-to-end voice latency numbers or tuning guidance backing the platform's speed claims
- Monitor live calls in production and get alerts when agents misbehave or error rates spike
- My agent can call external APIs and custom functions mid-conversation and speak the result without awkward dead air
- Get accurate real-time transcription with control over the STT provider, language models, or key terms
Probe proofs — replayable recordings from the probe harnessProbe proofs
Replayable recordings from our probe harness — see the Prove-It protocol to submit one.
$npx -y @elevenlabs/cli --versionreproduced$ npx -y @elevenlabs/cli --version \|/elevenlabs 1.1.0 \
$curl -si -X POST https://api.elevenlabs.io/v1/mcp -H 'Content-Type: application/json' -d '<jsonrpc initialize>'reproduced$ curl -si -X POST https://api.elevenlabs.io/v1/mcp -H 'Content-Type: application/json' -d '<jsonrpc initialize>'
HTTP/2 401
date: Sat, 05 Sep 2026 01:43:45 GMT
server: uvicorn
www-authenticate: Bearer resource_metadata="https://api.us.elevenlabs.io/.well-known/oauth-protected-resource"
content-length: 60
content-type: application/json
vary: Accept-Language, Accept-Encoding
access-control-allow-origin: *
access-control-allow-headers: *
access-control-allow-methods: POST, PATCH, OPTIONS, DELETE, GET, PUT
access-control-max-age: 600
strict-transport-security: max-age=1800;
x-trace-id: de448e74020f29612dc12866b1dd5e41
x-region: us-central1
via: 1.1 google
alt-svc: h3=":443"; ma=2592000
{"detail":"OAuth bearer [redacted] required for the hosted MCP."}
Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence
4 of 20 testable claims verified · 0 contradicted → integrity 20/100
33 distinct capability claims found in ElevenLabs Agents’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.
4
Verified
16
Unverified
0
Contradicted
18
Undersold
Verified (8)
“Agents can be managed via the dashboard, API, CLI, or hosted MCP server”
Drive the product through a documented public APIfullproof ↗
“Agents can be managed via the dashboard, API, CLI, or hosted MCP server”
“Agents can be managed via the dashboard, API, CLI, or hosted MCP server”
“Official CLI/skill lets an AI coding assistant build and manage voice agents”
“An AI assistant like Claude can create, configure, and manage agents via natural language with nothing to install locally”
“An AI assistant like Claude can create, configure, and manage agents via natural language with nothing to install locally”
Operate the product with natural-language commandsfullproof ↗
“Hosted MCP server lets Claude or any MCP client create and manage agents via natural language”
“Agents can be created via the API or the web dashboard”
Drive the product through a documented public APIfullproof ↗
Unverified (23)
“Visual workflow builder for designing multi-step conversation flows”
Design multi-step conversation flows in a visual builder with branching, states, and handoffs without writing codepartialproof ↗
“Choose from 5,000+ voices across 31 languages with customization options”
Choose from a broad voice library or plug in multiple TTS providers to get the voice I wantpartialproof ↗
“Upload documents and enable RAG so agents give grounded, document-based responses”
Ground the agent on my documents with a built-in knowledge base or RAG so it answers from my contentfullproof ↗
“Agents can call client-side functions and third-party APIs to perform actions”
My agent can call external APIs and custom functions mid-conversation and speak the result without awkward dead airfullproof ↗
“Dynamic variables and overrides allow per-conversation customization”
Inject dynamic variables and per-caller context at call time so each conversation is personalizedfullproof ↗
“HIPAA-eligible service with Business Associate Agreements offered for PHI-handling voice agents”
Run regulated workloads with HIPAA/BAA support, SOC 2, and data-residency optionspartialproof ↗
“Integrates with existing customer phone systems while using ElevenLabs voice AI”
Connect my own carrier or PBX via SIP trunking (or import Twilio/Telnyx numbers) instead of being locked to bundled telephonypartialproof ↗
“Tools can be executed client-side within a web browser or mobile app”
My agent can call external APIs and custom functions mid-conversation and speak the result without awkward dead airfullproof ↗
“Custom JavaScript tools can run sandboxed on ElevenLabs' own infrastructure”
My agent can call external APIs and custom functions mid-conversation and speak the result without awkward dead airfullproof ↗
“Agents can connect to external MCP servers to access and process data from other sources”
Plug MCP servers into this product so it can use their toolsfullproof ↗
“Agents can connect to external MCP servers to access and process data from other sources”
My voice agent can plug in MCP servers as tool sources so one integration grants it whole toolsets mid-callfullproof ↗
“Custom success-evaluation criteria assess conversation quality, goal achievement, and satisfaction”
The platform's AI reviews my calls for me — scoring quality, flagging failures, and analyzing resolution automaticallyfullproof ↗
“Data collection extracts specific structured data points (e.g., contact info) from conversations”
Extract structured data from every call — outcomes, entities, dispositions — delivered via API or webhook after the callpartialproof ↗
“Sentiment analysis reveals user sentiment across completed conversations”
See call analytics — success rates, durations, outcomes, sentiment — in dashboards without building my ownpartialproof ↗
“Conversations can be searched by keyword or semantic meaning”
See call analytics — success rates, durations, outcomes, sentiment — in dashboards without building my ownpartialproof ↗
“Agent testing verifies responses, tool usage, and multi-turn outcomes before deployment”
Test agents with simulated conversations or evals before putting them on real phone callsfullproof ↗
“Real conversations where the agent underperformed can be converted into test cases”
Test agents with simulated conversations or evals before putting them on real phone callsfullproof ↗
“Retention settings control how long transcripts and audio recordings are stored”
“Agents can trigger authenticated actions mid-conversation, like scheduling meetings or initiating order returns”
My agent can call external APIs and custom functions mid-conversation and speak the result without awkward dead airfullproof ↗
“Voice can be customized for pronunciation, speaking speed, and language-specific settings”
Choose from a broad voice library or plug in multiple TTS providers to get the voice I wantpartialproof ↗
“Conversation flow settings configure silence handling, interruptions, and turn-taking behavior”
Rely on the agent to handle interruptions (barge-in) gracefully — stopping speech, updating context, and recovering the turnpartialproof ↗
“Calls can be routed to AI agents without changing existing phone infrastructure”
Connect my own carrier or PBX via SIP trunking (or import Twilio/Telnyx numbers) instead of being locked to bundled telephonypartialproof ↗
“Agents can be configured, deployed, and monitored in 70+ languages with low-latency voice or chat”
Run conversations in multiple languages, including detecting and switching language mid-callpartialproof ↗
Undersold (18)
My coding agent can provision a complete voice agent end to end — create the agent, attach a number, and place a call — through the API, CLI, or MCP without touching the dashboardpartialproof ↗
The platform's own AI helps me author agents — generating or improving prompts, flows, and test cases from a descriptionpartialproof ↗
Build a working phone voice agent — prompt, voice, and phone number — and take my first live call within an hourpartialproof ↗
Point an agent at llms.txt or agent-oriented docsfullproof ↗
Run the product headlessly / in CI for automationpartialproof ↗
Get AI-generated insights and suggestions from my data inside the productpartialproof ↗
Set up automations that run autonomously in the backgroundpartialproof ↗
Delegate tasks to a built-in AI assistant inside the productpartialproof ↗
Test against a sandbox environment without touching production datapartialproof ↗
Define rules that trigger actions automatically on eventspartialproof ↗
Meet call-recording consent and disclosure obligations with per-call recording controls and configurable data retentionpartialproof ↗
See documented end-to-end voice latency numbers or tuning guidance backing the platform's speed claimspartialproof ↗
Do everything through the API that I can do in the UIpartialproof ↗
Provision phone numbers and run both inbound and outbound calls through the platform's APIpartialproof ↗
Monitor live calls in production and get alerts when agents misbehave or error rates spikepartialproof ↗
Retrieve full call recordings and transcripts programmatically for every callpartialproof ↗
Get accurate real-time transcription with control over the STT provider, language models, or key termspartialproof ↗
Claims outside our story set (6)
Real capability claims found in ElevenLabs Agents’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.
“Assistant can be embedded into a website or app for real-time customer support”
source ↗“Custom authentication can be implemented to protect agent access”
source ↗“Agent can switch between different voices mid-conversation for multi-character storytelling or language tutoring”
source ↗“Speech speed can be adjusted between 0.7x and 1.2x”
source ↗“Max conversation duration setting caps how long a call can remain active (default 10 minutes)”
source ↗“Supports choosing from supported LLMs or bringing your own custom model”
source ↗
Pricing signals
- $0.2per minutepay-as-you-goExtra-minute overage rate for TTS usage beyond the Starter plan's included minutes (entry-paid tier's PAYG rate); other tiers show ~$0.36 (Free), ~$0.18 (Creator), ~$0.17 (Pro/Scale/Business).source ↗as of 2026-09-07
- $6per month (entry plan)entry planCheapest paid monthly plan (Starter) shown on the pricing page; includes 30k credits/month.source ↗as of 2026-09-07
Extracted verbatim from the vendor’s own pricing page — hover a figure for the exact quote.
Business model
Credit-based subscription tiers (Free through Business) where agent conversations consume per-minute credits; Enterprise adds custom pricing, BAAs, and data residency.
pricing ↗Score trend
How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.
Try Experimental
Run it in the microterminal →Recorded agent sessions — and a live MCP handshake where the vendor ships one.
Flag
⚑ Flag a verdictThink a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.
For agents
Agent surface uptime MCP 100% · llms.txt 100% (30d, checked every 6h since Sep 8 '26)
