Retell AI
Also known as: Retell, Retell AI
Voice and chat agent platform for phone workflows, pairing a customer-chosen or self-hosted LLM with node-based conversation flows, telephony over SIP or imported numbers, and a testing suite spanning playground, simulation, batch and live A/B.
Retell AI is a platform for building and running AI voice agents that handle business phone calls, built to be AI-native rather than a layer bolted onto legacy call-center scripting. Founded in 2024, it lets companies replace traditional interactive voice menus with agents that answer and place calls, sound human, respond with low latency, and carry on natural, multi-turn conversations across voice, and also chat, SMS, and email.
The platform is designed for fast, largely no-code setup. Teams build agents in a visual drag-and-drop flow builder, defining conversation paths, branching, and transfer logic, while the underlying language model handles the conversational reasoning, so a high-level goal and a few conditions replace pages of brittle scripting. A proprietary turn-taking model judges when the caller has finished speaking so the agent can respond without awkward pauses, and agents can perform warm transfers or escalate to a human when needed.
Retell handles the telephony as well as the conversation. It connects to existing phone numbers and carriers through SIP trunking or providers like Twilio, supports branded and verified caller ID to reduce spam labeling, and runs outbound batch campaigns.
Agents can call tools, fire webhooks, and pull from connected CRMs, knowledge bases, and other systems, including custom Model Context Protocol servers, and they operate in more than thirty languages with automatic detection of the caller's language.
When an outbound call lands in another company's touch-tone menu, the agent can work out which digits to press and press them, and it detects voicemail and responds to it, both shipped as named building blocks in the flow builder.
Because quality matters at call-center scale, the platform includes built-in testing to simulate conversations and A/B test flows, dashboards tracking outcomes, latency, sentiment, and customer satisfaction, and an automated quality-assurance capability that monitors live calls and flags improvements without manual spot-checking. It meets enterprise requirements with SOC 2, HIPAA, and GDPR compliance and single sign-on.
Retell is used for inbound support, appointment scheduling and confirmations, lead qualification, outbound sales follow-ups, and survey calls, across industries like healthcare, financial services, and logistics. Its pitch is speed and simplicity: get production-ready voice agents live in days, automate the bulk of routine tier-one calls, and free human agents for the complex or sensitive ones, all on usage-based pricing.
Vendor details
Canonical URL
https://www.retellai.com
Category
Customer support agent
Company status
independent
Use cases & customers
Target customers
Deployment options
In practice
Your call center drowns in routine tier-one calls and your old phone menu frustrates callers. Retell lets you build a no-code voice agent that answers naturally, handles the common requests, and warm-transfers anything complex to a human.
Missed calls are lost leads after hours. A Retell AI receptionist answers every call around the clock, collects the caller's details, logs them to your CRM, and notifies you, turning a missed call into a captured lead.
You need to confirm hundreds of appointments without staff on the phone all day. Retell runs an outbound batch campaign of calls that confirm or reschedule, in the caller's own language, and writes the results back to your system.
Sources & related URLs
Agentic Index coverage score
12.0 / 14 capabilities · 86%
| Integrations & Tool Calling | Full |
|---|---|
|
An Apps framework holds credentials for each external system, with documented agent functions across CRM (Salesforce, HubSpot, Dynamics 365, Zoho, GoHighLevel), helpdesk (Zendesk), scheduling (Calendly, Cal.com) and file storage (Google Drive, OneDrive), plus custom functions calling any HTTP API mid-call, a code tool running JavaScript, and MCP tools and nodes reaching remote MCP servers during a live conversation. Contact data also syncs both ways with the CRM. Sourcedocs.retellai.com/integrations/overviewread 2026-09-05 |
|
| Workflow Orchestration | Full |
|
Conversation flow agents are node-and-edge graphs with conversation, subagent, function, code, logic split, extract-variable, SMS, transfer and global nodes, transitions gated by either LLM-evaluated prompt conditions or deterministic equations, reusable subflows shared across flows, and Flex Mode compiling the graph into a runtime prompt for varied caller behavior. Multi-prompt state agents and agent-to-agent transfer cover the same ground for prompt-built agents, with versions and environment tags controlling what is live. Sourcedocs.retellai.com/build/conversation-flow/overviewread 2026-09-05 |
|
| Knowledge Grounding & RAG | Full |
|
Knowledge bases are built and maintained by the platform: sources attached as URLs, documents across roughly twenty file formats, or custom text, then chunked, embedded and stored in a vector database at creation time, with retrieval tuned by query-rewrite instructions that condense recent dialogue into a standalone search query. Google Drive and OneDrive sources re-sync on a 24-hour refresh cycle, and knowledge bases are managed independently of any one agent through their own API. Sourcedocs.retellai.com/build/knowledge-baseread 2026-09-05 |
|
| Human Oversight & Guardrails | Full |
|
Guardrails detect prohibited topics in both agent responses and user input and replace flagged content with a safe placeholder before it is spoken, and live monitoring lets a person follow the transcript in real time, listen in silently, take over from the agent mid-call or end the call, with an Update Live Call endpoint to inject context or trigger a response on a running session. Sourcedocs.retellai.com/features/live-monitoringread 2026-09-05 |
|
| Security, Identity & Governance | Full |
|
SOC 2 Type 1 and Type 2 certified, HIPAA and GDPR compliant with a BAA, DPA and a documented GDPR erasure process, alongside customer-facing controls: role-based access control with Admin, Developer, Analyst and Viewer roles, API keys scoped read or edit per area, a webhook signing key with signature verification and IP allowlisting, expiring signed URLs for recordings and logs, per-agent data storage settings and per-agent retention policy. Sourcedocs.retellai.com/general/complianceread 2026-09-05 |
|
| Observability & Auditability | Full |
|
A monitor-call WebSocket streams transcript, tool call and node events from a live session, so the path the agent took through the flow is reconstructable turn by turn, backed by session history with transcripts and CSV export, per-call P50, P90 and P99 latency, disconnection reason codes, call_started, call_ended and call_analyzed webhooks, and AI QA scoring calls on hallucination rate, knowledge base accuracy, tool usage, latency and sentiment against customer-defined cohorts. Sourcedocs.retellai.com/api-references/monitor-call-websocketread 2026-09-05 |
|
| Memory & State Persistence | Full |
|
Contact records keyed by phone number carry custom fields populated by post-call extraction mappings, so the agent recalls preferences, past issues and commitments across every later conversation, with a merged timeline of that contact's calls and chats and a backfill job that re-applies mappings to historical calls. Scope is per contact and the record is separately deletable. It is implemented as contact field mappings rather than a dedicated memory store with its own lifetime, so retention follows the contact record. Sourcedocs.retellai.com/integrations/build-contact-memoryread 2026-09-05 |
|
| Deployment & Data Residency | Not documented |
|
The documentation covers what is stored and for how long through per-agent data storage settings, a per-agent retention policy and expiring signed URLs, but those are privacy controls rather than control over where data lives. No region selection, VPC, single tenant, self hosted or on premises option is documented, and the reliability page describes shared enterprise infrastructure without naming a region. Sourcedocs.retellai.com/accounts/data-retentionread 2026-09-05 |
|
| Prebuilt Agents, Templates & Packs | Full |
|
The quick start has the customer pick a template as the first build step, Agent Handbook ships toggleable preset prompt packs for personality, accuracy and safety, and reusable subflows package parts of a conversation flow as shared components adopted across agents. The template catalog itself sits inside the dashboard rather than on a public page, so its breadth is not published. Sourcedocs.retellai.com/get-started/quick-startread 2026-09-05 |
|
| Triggers & Channel Coverage | Full |
|
Inbound and outbound calling on Retell-managed numbers or imported ones over elastic SIP trunking, with documented setup for Twilio, Telnyx, Vonage, Avaya Aura, Genesys Cloud, Five9 and Amazon Connect, alongside browser web calls, a website widget in chat, voice and callback modes, two-way SMS and MMS, scheduled batch call campaigns, and an inbound call and SMS webhook that picks the agent and passes per-call context at ring time. Sourcedocs.retellai.com/deploy/custom-telephonyread 2026-09-05 |
|
| Model Flexibility & Routing | Full |
|
The customer picks the language model in agent settings and tunes temperature, fast tier, structured output, retries and timeouts, choosing across providers including OpenAI, Anthropic and Google as named in the model deprecation notices, or runs their own model entirely through the custom LLM WebSocket protocol while Retell handles telephony, transcription and turn-taking. Speech recognition and text-to-speech providers are separately selectable. Sourcedocs.retellai.com/build/llm-optionsread 2026-09-05 |
|
| APIs, SDKs & MCP Extensibility | Full |
|
A REST API at api.retellai.com published with an OpenAPI specification and a Postman collection, covering calls, chats, agents, response engines, knowledge bases, contacts, apps, tests and phone numbers; official TypeScript and Python SDKs plus a browser Web SDK; a command line interface; a hosted MCP server at mcp.retellai.com for Cursor, Claude Desktop and Claude Code; signed webhooks; and a documented LLM WebSocket protocol for bringing your own model. A published deprecation feed with dated breaking-change notices sits alongside it. Sourcedocs.retellai.com/api-references/overviewread 2026-09-05 |
|
| Testing, Debugging & Optimization | Full |
|
Five shipped testing surfaces with readable results: an LLM Playground that inspects tool calls and transitions and replays any turn, LLM simulation testing where an AI-simulated user runs a scenario and success criteria grade each run, batch tests over many cases reporting pass rate and per-case results in a testing history, web and phone call testing, and production A/B testing that splits live traffic between agent versions. Test case definitions, batch tests and test runs each have their own API. Sourcedocs.retellai.com/test/test-overviewread 2026-09-05 |
|
| Browser & Computer Use | Not documented |
|
Browser, desktop session and remote computer control are all absent from the documentation. The Press Digit tool and press digit node do let the agent infer and press keypad digits to walk a third party's DTMF input IVR menu on an outbound call, and voicemail handling detects and responds to another system's automated flow; that is real and documented, but it works a telephone keypad rather than a browser, desktop or remote computer. Sourcedocs.retellai.com/build/single-multi-prompt/press-digitread 2026-09-05 |
|
The Agentic Index coverage score grades every vendor Full, Partial or Not documented against the same 14 buyer facing capabilities, from public evidence only. Each capability links to how all vendors in the index score on it. How this evidence is graded
Recent platform changes
Legacy monthly billing at Retell has ended, and every remaining self serve workspace moved to prepaid credits, starting any workspace that had not switched on its own at a $0 balance. Calls and chats now draw down credits and stop at zero unless auto recharge is on, while phone numbers and some add ons still bill monthly and enterprise accounts stay on invoices.
Bears on: Pricing / packaging
View sourceRetell removed legacy list endpoints and pushed users to versioned v2/v3 list APIs with unified pagination.
Bears on: MCP / tool calling / API
View sourcePricing
Pay as you go from $0 ($10 free credits); voice engine $0.07 to $0.08/min plus LLM and telephony billed separately; Enterprise custom
usage (per minute across voice engine, LLM, and telephony components)
Included quota
Pay as you go: no subscription, $10 free credits (roughly 67 to 90 minutes of testing), 20 concurrent calls included. Enterprise: custom concurrency (50+), compliance documentation, dedicated support, and volume discounts.
What is public
The $0.07/min voice engine base rate, component rates, $10 free credits, 20 free concurrent calls, $8 per extra concurrent slot, and a usage calculator are public. Enterprise rates are custom.
Billing mechanics
Retell is modular and metered per minute: the voice engine base is $0.07 to $0.08/min (speech to text, orchestration, text to speech), the LLM you pick bills separately (roughly $0.003/min for light models up to about $0.08/min for top models, or bring your own), and telephony runs about $0.015/min. Failed calls are not billed. Concurrency beyond the free 20 is $8 per call per month.
Cost watchouts
The $0.07/min headline is the voice engine only; LLM and telephony stack on top, pushing real cost to $0.13 to $0.31/min. HIPAA and BAA require Enterprise, premium voices cost more, branded caller ID and international calling add fees, and concurrency slots are a separate monthly line.
Variable cost rationale
Cost is fully per minute and split across separately metered components (voice engine, model, telephony), so identical call volumes produce very different bills depending on the model and voice chosen.
Additional watchouts
Model the full component stack (voice plus model plus telephony) before committing; the advertised base rate understates production cost by roughly 2x to 4x.
Overage / add-ons
Everything is metered per connected minute across components; extra concurrency beyond 20 is $8 per concurrent call per month. No charge for calls that fail to connect.
Sales call required
Mixed (some tiers require a call)
Free / trial
$10 in free credits and 20 concurrent calls to build and test, no card
Lowest paid plan
Pay as you go, voice engine from $0.07/min (plus LLM and telephony)
Commercial notes
Developer first, API and SIP focused voice agent infrastructure with a visual builder. Two tiers: self serve pay as you go and Enterprise. HIPAA and BAA, SSO, and role based access are Enterprise only; Enterprise adds dedicated infrastructure and volume discounts.
Key ambiguities
Effective per minute cost is configuration dependent (model and voice choice) and Enterprise rates are not published, so total spend is hard to forecast without call data.
Related vendors
- Decagon — AI agent platform purpose-built for customer support, with deep…
- 5.Y — Singapore early-stage platform whose GLUCOSE product deploys…
- Ada — Agentic customer experience platform whose AI agents autonomously…
- Aisera — Enterprise agentic AI platform that orchestrates specialized agents…
- ASAPP — Generative AI for enterprise contact centers that plugs into…
- Assembled — AI customer support orchestration platform that unifies autonomous…
Alternatives to Retell AI
The closest documented capability profiles to Retell AI among customer support agents tracked by Agentic Index, ordered by similarity on the same 14 point evidence the rankings use. No vendor pays for placement.
- Gallabox11.5 / 14A lighter documented profile than Retell AI
- Aisera11.0 / 14A lighter documented profile than Retell AI
- Botpress11.0 / 14A lighter documented profile than Retell AI
- Infobip12.0 / 14Adds documented Deployment & Data Residency
- Kore.ai13.0 / 14Adds documented Deployment & Data Residency
- Kustomer11.0 / 14A lighter documented profile than Retell AI
Similarity is computed from each vendor's Agentic Index coverage score evidence, axis by axis, not from the totals. How this evidence is graded