Inworld AI Agent Runtime
Also known as: Inworld AI, Inworld Runtime
Realtime agent runtime for graph-based AI systems: provides orchestration nodes, telemetry, memory, experiment support, on-prem TTS, and deployment options for AI product teams. Infrastructure layer, not an end-user assistant.
Inworld AI is a realtime AI model and infrastructure company focused on consumer-scale, interactive applications. Founded in 2021 by a team that had led language-model product work at Google DeepMind and built Dialogflow, it first became known for AI-powered non-player characters in games, working with partners like Xbox, Disney, NVIDIA, and Niantic. It has since broadened into a voice AI and agent infrastructure platform for any real-time application, from companions and tutors to enterprise support agents.
At the center is its realtime voice stack. Inworld's text-to-speech generates expressive, natural speech with low latency, voice cloning from seconds of audio, real-time control over emotion and pace, and inline non-verbal sounds like laughs and sighs, across many languages, with the ability to carry conversational context forward so long interactions stop feeling like disconnected scripts. A companion speech-to-text model adds voice profiling, natural turn detection, word-level timestamps, and speaker labels.
Beyond audio, Inworld provides the orchestration layer that ties an agent together. A Realtime API handles full-duplex speech-to-speech with intelligent turn-taking and mid-session tool calling, while a Router exposes a single interface to hundreds of language models from providers like OpenAI, Anthropic, Google, and Mistral, with fallbacks and dynamic selection and no added markup. The Agent Runtime, a low-latency C++ core, coordinates the model, speech, and tools, including external systems through the Model Context Protocol, and notably is free to use, with customers paying only for the model consumption underneath.
Inworld's pitch is closing the gap between prototype and production. Built-in telemetry and the ability to A/B test models, prompts, and pipelines on live traffic let teams tune for engagement and retention without redeploying, and the runtime is designed to scale from a prototype to millions of concurrent users with minimal code changes. It offers SDKs for environments like Node.js and Unreal Engine, a browser playground, and enterprise options including on-premise deployment, data residency, and zero data retention. Typical uses include companion and social apps, language learning and tutoring, health and wellness coaching, AI characters in games and interactive media, and enterprise voice agents for support, sales, and recruiting.
Vendor details
Canonical URL
https://inworld.ai/runtime
Category
Agent infrastructure
Company status
independent
Use cases & customers
Target customers
Deployment options
In practice
You're building a companion or character app and the voice breaks the spell, flat, robotic, and forgetful between turns. Inworld's realtime text-to-speech adds emotion, pacing, and inline sounds like laughs, and carries context forward across a long conversation.
Your interactive app needs to scale from a prototype to millions of users without re-architecting. Inworld's Agent Runtime orchestrates the model, speech, and tools in a low-latency core and is free, so you pay only for model consumption.
You want to pick the best language model per situation without rewriting integrations. Inworld's Router gives one interface to hundreds of models from providers like OpenAI, Anthropic, and Google, with fallbacks and no added markup.
Sources & related URLs
Agentic Index coverage score
11.0 / 14 capabilities · 79%
| Integrations & Tool CallingOfficial docs 2026-06-08 | Full |
|---|---|
| Workflow OrchestrationGraph runtime orchestration docs 2026-06-08 | Full |
| Knowledge Grounding & RAGOfficial docs 2026-06-08 | Partial |
| Human Oversight & GuardrailsOfficial docs 2026-06-08 | Partial |
| Security, Identity & GovernanceOfficial docs 2026-06-08 | Partial |
| Observability & AuditabilityTelemetry logs traces docs 2026-06-08 | Full |
| Memory & State PersistenceMemory nodes docs 2026-06-08 | Full |
| Deployment & Data ResidencyOn-prem and SaaS docs 2026-06-08 | Full |
| Prebuilt Agents, Templates & PacksTemplates docs 2026-06-08 | Full |
| Triggers & Channel CoverageRuntime channel docs 2026-06-08 | Full |
| Model Flexibility & RoutingOfficial docs 2026-06-08 | Partial |
| APIs, SDKs & MCP ExtensibilityOfficial docs 2026-06-08 | Full |
| Testing, Debugging & OptimizationExperiment and traces docs 2026-06-08 | Full |
| Browser & Computer UseOfficial docs 2026-06-08 | Unable to verify |
The Agentic Index coverage score grades every vendor Full, Partial or Unable to verify against the same 14 buyer facing capabilities, from public evidence only. Each capability links to how all vendors in the index score on it. How this evidence is graded
Recent platform changes
Inworld launched Realtime TTS-2 Flash, a new low-latency text-to-speech model designed for real-time applications. The model achieves a 20-millisecond time to first audio, operating five times faster than the standard TTS-2 model while supporting over 200 languages and instant voice cloning.
Bears on: Agent capability
View sourcePricing
Runtime free · pay per model usage · credit plans to $300/mo Developer
usage
Included quota
Subscription credits equal the plan price each cycle (Developer $300/mo grants $300 in credits) and roll over up to 3 months on active plans. Higher tiers unlock lower per unit model rates (up to ~53% off on Growth), higher rate limits, and more custom voices. LLMs bill at provider passthrough rates.
What is public
Inworld publishes consumption rates and credit plan tiers publicly. The Agent Runtime carries no license fee; all cost is metered model usage. A June 2026 announcement cut prices more than 50% across TTS, STT, and LLM access for most developers.
Billing mechanics
Dollar denominated credits: plan fee converts to monthly credits, usage deducts at tier rates, unused credits roll over up to 3 months, auto reload available. On Demand (no subscription) pays base rates.
Cost watchouts
The free runtime is real but total cost is entirely model consumption, which scales with session minutes: voice heavy consumer apps live and die on per character TTS and per minute STT rates. Downgrades and cancellations forfeit remaining credits at period end. Legacy Founder rates ($5/$10 per 1M chars) closed to new subscribers May 7, 2026.
Variable cost rationale
Pure consumption billing on TTS characters, STT minutes, and LLM tokens; usage growth translates directly into spend.
Additional watchouts
Evaluate the runtime alongside the company's voice first positioning: the orchestration layer is free and production real, but the commercial engine is model consumption, and roadmap emphasis follows it.
Sales call required
Mixed (some tiers require a call)
Free / trial
On Demand free entry (up to 70 min TTS or 400 min STT free); Runtime free at any scale
Lowest paid plan
Credit subscription tiers (plan price converts to monthly usage credits; Developer $300/mo shown as reference tier)
Commercial notes
Independent; founded by the API.AI/Dialogflow team, backed by investors including Microsoft and Disney. Repositioned around realtime voice AI (Realtime TTS-2 launch, June 2026 price cuts) with the C++ Agent Runtime as the orchestration layer; clients include Google, NVIDIA, Meta, Ubisoft, and Xbox.
Key ambiguities
Effective per session cost depends on model mix and tier discounts, so entry price is not comparable to flat subscription vendors. Enterprise and on prem TTS terms are custom.
Missing data
Enterprise/Custom License terms unpublished.
Related vendors
- Acrab — Singapore compute infrastructure company building a full stack…
- AgentOps — Agent observability and reliability platform with broad model and…
- Agno — High-performance agent runtime and framework (formerly Phidata) with…
- AIsa — Unified resource and payment gateway for AI agents that lets them…
- AlphaBitCore — AI control plane that governs how models, agents, tools, and…
- Anchor Browser — Cloud hosted browser infrastructure that lets AI agents operate real…
Alternatives to Inworld AI Agent Runtime
The closest documented capability profiles to Inworld AI Agent Runtime among agent infrastructure platforms tracked by Agentic Index, ordered by similarity on the same 14 point evidence the rankings use. No vendor pays for placement.
- Hatchet9.0 / 14A lighter documented profile than Inworld AI Agent Runtime
- Kestra12.0 / 14Fuller documented coverage on Human Oversight & Guardrails and Security, Identity & Governance
- Pipecat10.0 / 14Fuller documented coverage on Model Flexibility & Routing
- Bernstein10.5 / 14Fuller documented coverage on Security, Identity & Governance and Model Flexibility & Routing
- Inngest10.5 / 14Fuller documented coverage on Human Oversight & Guardrails and Security, Identity & Governance
- LangChain11.5 / 14Fuller documented coverage on Knowledge Grounding & RAG and Human Oversight & Guardrails
Similarity is computed from each vendor's Agentic Index coverage score evidence, axis by axis, not from the totals. How this evidence is graded