Back to vendors
T

Telnyx

Also known as: Telnyx AI Assistants, telnyx.com, Telnyx Voice AI

Visit site
Entry price$0.05 a minute for the voice engine, including speech to text, text to speech and knowledge retrieval, with LLM tokens and telephony extraFull pricing detail

Telnyx is a communications carrier that runs AI Assistants on its own network and GPUs. Teams configure voice and chat agents in a portal with their choice of model, speech providers, tools, MCP servers and knowledge bases, guide them with conversation workflows, check them with AI tests and release them through versioned canary traffic splits, at $0.05 a minute for the voice engine.

Telnyx sells voice, messaging, numbers, SIP trunking, wireless and AI inference over a network it operates itself. Its AI Assistants product puts hosted voice and chat agents on top of that network, so the phone number, the call, the speech models and the language model can all sit with one supplier.

An assistant is configured in the portal or through the API. A team writes instructions and a greeting with dynamic variables filled from a webhook or SIP headers, picks a language model from Telnyx hosted models such as Kimi, external models such as OpenAI's on its own key, or any OpenAI compatible endpoint, and picks speech to text from Telnyx, Deepgram or Azure and text to speech from Telnyx, AWS, Azure, ElevenLabs or Inworld.

Tools cover webhooks, transfers, SIP Refer, DTMF, hangup, client side actions, MCP servers and integrations with Salesforce, Zendesk and Jira, and knowledge bases take uploaded files or URLs. Conversation workflows turn an assistant into a directed graph of prompt, speak and tool nodes with LLM or variable conditions on each edge, and edges can hand the conversation to another assistant. Memory recalls a caller's earlier conversations and selected insights at the start of a new one.

AI tests check an assistant against named success criteria before release, new versions take a share of traffic by caller rule or percentage with rollback in one step, and insight groups analyze every stored conversation and post the results to a webhook.

The voice engine costs $0.05 a minute and bundles orchestration, speech to text, text to speech and knowledge retrieval, with LLM tokens and telephony billed on top, so a production agent runs about $0.06 a minute by Telnyx's own estimate.

Telnyx holds a SOC 2 Type II report for voice and messaging, ISO 27001:2022 and HIPAA coverage that includes AI inference, and it offers EU data residency for voice AI agents and SAML single sign on for business accounts.

It fits teams that want voice agents with model and speech provider choice, real test runs and gradual rollouts, carried on a carrier's own network. It fits less well for a business that wants a library of ready made agents, or a per action approval step before an assistant acts.

Vendor details

Canonical URL

https://telnyx.com

Category

Voice agent

Subcategory

Hosted voice and chat AI assistants on Telnyx's own carrier network and GPUs, with workflows, tests and canary releases

Funding status

Privately held carrier, Telnyx LLC.

Company status

independent

Use cases & customers

Primary use cases

inbound support calls answered by an AI assistantoutbound reminder and qualification callsmultistep call flows with handoff between specialist assistantsSMS and web chat assistants sharing the same configurationtesting and canary releases of new assistant versions

Target customers

teams building voice AI agents for inbound and outbound callscontact centers moving IVR to AIdevelopers who want model and speech provider choicecompanies already buying numbers and SIP from Telnyx

Deployment options

SaaSEU data residency for voice AI agents

Integrations

Webhook tools reach any HTTP API. MCP servers attach outside tool catalogs, and enterprise integrations connect Salesforce, Zendesk and Jira. Transfer, SIP Refer and DTMF tools work against existing phone systems, client side tools run in the web widget, and integration secrets store keys for providers such as ElevenLabs. Models come from Telnyx's own GPUs, OpenAI and other providers, or any OpenAI compatible endpoint including AWS Bedrock and Azure OpenAI.

In practice

A healthcare scheduling team builds an assistant on a Telnyx hosted model with a knowledge base of clinic policies. A webhook fills in the patient's name at the start of the call, a tool books the slot in the scheduling system, and a transfer tool sends billing questions to the front desk.

A support team splits its assistant into a conversation workflow, where a triage prompt node routes billing, technical and sales callers along edges to three specialist assistants, each with its own tools and instructions.

Before a new greeting and tool set go live, the team runs AI tests against success criteria, then sends ten percent of calls from selected numbers to the new version and rolls back with one click if results slip.

Agentic Index coverage score

11.0 / 14 capabilities · 79%

Integrations & Tool Calling Full

Assistants call outside systems with webhook tools, which can run asynchronously with filler messages while the request completes, attach MCP servers for whole tool catalogs, and connect enterprise integrations for Salesforce, Zendesk and Jira. Call control tools transfer the caller, hand off with SIP Refer, send DTMF tones to another phone system and hang up, client side tools run actions in the web widget, and keys for outside providers are kept as write only integration secrets. Tools live in a shared library that several assistants and workflow nodes can reuse.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/no-code-voice-assistant.mdread 2026-10-11

Workflow Orchestration Full

Conversation workflows store a directed graph on the assistant. Prompt nodes run the model with their own instructions and optional model and voice overrides, speak nodes play a fixed message, and tool nodes run one shared tool and advance.

Edges move between nodes on a natural language condition the model judges, a deterministic comparison against variables, or a default, and a tool node's HTTP status code can steer the next edge.

An edge can also hand the conversation to another assistant, so a triage assistant can route billing, technical and sales callers to specialists that share the conversation's context.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/workflows.mdread 2026-10-11

Knowledge Grounding & RAG Full

Each assistant can use named knowledge bases built from uploaded files or a URL, kept with the account and attached in the assistant's configuration, so one knowledge base can serve several assistants. Retrieval during a call is part of the voice engine's orchestration and included in the per minute price, while the price estimator carries a separate agent knowledge base add on. Telnyx does not say how the content is chunked or indexed.

Sourcetelnyx.com/pricing/conversational-airead 2026-10-11

Human Oversight & Guardrails Partial

Transfer and SIP Refer tools pass a caller to a person or another phone system, with voicemail detection on transfer, and handoff tools move a conversation between assistants. Releases stay under the team's control. A new version is saved with notes, traffic rules send selected callers or a percentage to it, and rollback returns all traffic to the main version at once. No step lets a person approve a tool call or a reply before the assistant acts.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/version-testing-traffic-distribution.mdread 2026-10-11

Security, Identity & Governance Full

Business accounts sign in to the Mission Control Portal through SAML single sign on with Okta, Azure Active Directory, Google, OneLogin, Auth0 or LastPass. Telnyx holds ISO 27001:2022 and ISO 27701, a SOC 2 Type II report under NDA covering voice, messaging, video and wireless, and HIPAA coverage for voice, messaging, fax and AI inference with a BAA available, plus PCI DSS for its own billing. The SOC 2 scope does not name AI assistants, and Telnyx also claims a SOC 2 Type III, a report type the AICPA does not issue.

Sourcetelnyx.com/ai/compliance.mdread 2026-10-11

Observability & Auditability Partial

Every conversation is stored and listable through the conversations API, transcripts show which workflow node produced each assistant message, and the last tool call's HTTP status is kept as a variable. Insight groups run structured and unstructured analysis on each conversation and can post results to a webhook, with Telnyx managed insights alongside custom ones. Version notes record what changed in each release, and Telnyx does not mention an audit trail of who changed an assistant's configuration.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/workflows.mdread 2026-10-11

Memory & State Persistence Partial

Memory brings earlier conversations into a new one. The dynamic variables webhook returns a query in the same form as the List Conversations endpoint, typically the caller's phone number, a time window or custom metadata, plus the IDs of insights to include, and the assistant starts with those conversations and insight results in context, across voice and messaging.

The webhook has one second to answer before the call carries on without memory. What comes back is stored conversation history picked by the customer's query. There is no separate memory store, and Telnyx sets out no retention period or way to delete one person's history.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/memoryread 2026-10-11

Deployment & Data Residency Full

Voice AI agents run in the United States, the primary data center, or in the European Union with EU data residency, where data does not leave the EU. AI inference is also offered in Asia Pacific with regional processing, though voice agents are not offered there, and Canada has neither. The assistant's language model can also point at the customer's own OpenAI compatible endpoint, such as one on AWS Bedrock, Azure OpenAI or a self run vLLM server, which keeps inference inside an environment the customer controls.

Sourcetelnyx.com/ai/compliance.mdread 2026-10-11

Prebuilt Agents / Templates / Packs Partial

New assistants start from a template in the portal, with a blank template as the usual starting point, and Telnyx managed insights come ready to attach to any assistant. A shared tools library lets one set of tools serve many assistants. Telnyx offers no catalog of finished agents for support, sales or scheduling.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/no-code-voice-assistant.mdread 2026-10-11

Triggers & Channel Coverage Full

Calls reach an assistant on an assigned Telnyx number or over SIP, outbound calls start from the portal, the API or the TeXML ai_calls endpoint, and an embeddable widget carries voice and chat on a web page. SMS conversations run through the assistant SMS chat API, with MMS during a call when messaging is on. At the start of each conversation a dynamic variables webhook can load caller details and memory settings, and custom SIP headers or the outbound request can pass variables in.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/no-code-voice-assistant.mdread 2026-10-11

Model Flexibility & Routing Full

The customer picks the model for each assistant on the Agent tab. Options include models hosted on Telnyx GPUs, with Kimi as the default, managed third party models from providers such as OpenAI, Gemini and Groq, with OpenAI models on the customer's own key, and any public OpenAI compatible endpoint such as AWS Bedrock, Azure OpenAI, Baseten, vLLM or SGLang. Prompt nodes in a workflow can override the model step by step, and speech to text and text to speech providers are chosen separately.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/custom-llm.mdread 2026-10-11

APIs / SDKs / MCP Extensibility Full

The API covers assistants end to end, including creating and updating assistants, assistant tests, canary deploys, MCP server records, SMS chat and a chat endpoint in beta, and the TeXML ai_calls endpoint starts outbound calls. Coding assistants can read the whole reference through llms.txt indexes, and the conversations API hands stored conversations to other systems.

Sourcedevelopers.telnyx.com/development/llms/ai-assistants-llms-txtread 2026-10-11

Testing, Debugging & Optimization Full

AI tests give each check a name, a target assistant and success criteria, such as what the greeting must say, and a test run shows progress live and reports whether the assistant met every criterion, with the conversation to review. A new version can be saved with notes and sent a share of traffic by ordered rules on the caller's number, all to one version or split by percentage, with the remainder staying on the main version, and rollback clears every rule at once. Coval can generate test conversations from seed cases.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/version-testing-traffic-distribution.mdread 2026-10-11

Browser / Computer-use Not documented

Assistants act through webhooks, MCP servers, integrations and call control tools, and send keypad tones to other phone systems. No hosted browser, remote desktop or screen control is offered.

Sourcedevelopers.telnyx.com/docs/inference/ai-assistants/tools-library.mdread 2026-10-11

The Agentic Index coverage score grades every vendor Full, Partial or Not documented against the same 14 buyer facing capabilities, from public evidence only. Each capability links to how all vendors in the index score on it. How this evidence is graded

Pricing

$0.05 a minute for the voice engine, including speech to text, text to speech and knowledge retrieval, with LLM tokens and telephony extra

Per minute of voice engine time, rounded up to whole minutes per call, plus LLM tokens, per minute telephony and monthly number rental

Included quota

The $0.05 per minute voice engine includes orchestration, turn taking, interruption handling, tools, speech to text, text to speech and knowledge base retrieval.

What is public

The voice engine rate and what it includes, LLM token pricing references, telephony and number rates, plan minimums and an all in estimate.

Billing mechanics

Pay as you go with a card on file, or commit monthly spend on Growth or Enterprise for discounts.

Cost watchouts

Voice engine time rounds up to 60 second increments per call, so short calls cost more than the per minute figure suggests. Frontier models are billed at provider rates, and telephony, numbers and SIP trunking are separate lines.

Variable cost rationale

Every minute, token and call leg is metered, so spend tracks call volume and length, with the choice of language model moving the LLM line the most.

Additional watchouts

Short calls round up to a full minute of voice engine time, and the language model's token price moves the total more than any other line.

Overage / add-ons

Pure usage billing on pay as you go. Growth requires $2,000 a month of spend for 15 percent off voice, messaging and numbers, and Enterprise starts at $5,000 a month.

Sales call required

Mixed (some tiers require a call)

Free / trial

No free credit, and pay as you go has no minimum

Lowest paid plan

Pay as you go at $0 a month with a card on file, billed at standard rates

Commercial notes

Telnyx runs the calls, speech and hosted models on infrastructure it operates.

Key ambiguities

The price estimator carries an agent knowledge base add on with no price, while the voice engine rate already includes knowledge retrieval.

Missing data

Per token prices for each model, the price of the agent knowledge base add on, and international telephony rates in the AI pricing summary.

Agentic Index verified 2026-10-11

Contact us

Found a vendor we missed? Have feedback on the index? We'd love to hear from you.