Smallest AI
Also known as: Smallest.ai, Atoms, Waves, Lightning, Pulse
Full stack voice AI platform combining Atoms, a no code voice and chat agent builder handling conversations in more than twenty languages, with its own co optimized models: Lightning text to speech at roughly one hundred millisecond latency and Pulse speech to text at sixty four millisecond first transcript.
Smallest AI was founded by Sudarshan Kamath and Akshat Mandloi and has raised eight million dollars in seed funding led by Sierra Ventures, months after a pre seed round led by 3one4 Capital. The company's thesis is precision over scale: many small, efficient models, each tuned for a narrow job and optimized to run together rather than assembled from different vendors. The company says that in its first year it cut text to speech latency from roughly two seconds to one hundred milliseconds, dropped TTS cost from twenty cents to a penny per minute at scale, and moved from prototype to powering millions of enterprise calls each month.
The stack has three layers. Lightning is the text to speech line, with V3 hitting about one hundred millisecond latency across fifteen plus languages including Indic and European coverage, instant voice cloning from ten seconds of audio, and streaming over HTTP, SSE, or WebSocket. Pulse is streaming speech to text with a sixty four millisecond time to first transcript.
Atoms is the agent platform on top: no code Single Prompt and Conversational Flow agents with an Agentic Graph Builder, knowledge bases, webhooks, API calls from within conversations, phone numbers, outbound campaigns and audiences, a web voice and chat widget, post call metrics, conversation logs, versioning, and testing, all documented alongside Node.js and Python SDKs, a WebSocket SDK, mobile integrations, and an MCP section in the developer docs. The full agent pipeline stays under eight hundred milliseconds per turn, and the whole stack is available on AWS Marketplace under consolidated billing.
Smallest AI fits teams building production voice agents for support, sales, and operations who want the model layer and agent layer from one vendor with published usage pricing, and developers who need fast multilingual TTS and STT as raw APIs.
Vendor details
Canonical URL
https://smallest.ai
Category
Voice agent
Subcategory
Full stack voice AI: Atoms agents plus Lightning TTS and Pulse STT
Funding status
Independent, founded by Sudarshan Kamath and Akshat Mandloi. Raised eight million dollars in seed funding led by Sierra Ventures, months after a pre seed led by 3one4 Capital, and reports powering millions of enterprise calls per month within a year of launch.
Company status
independent
Use cases & customers
Primary use cases
Target customers
Deployment options
Integrations
Atoms agents make API calls to outside systems mid call, run a pre call API request, and fire webhooks; developers get REST APIs, SDKs, a WebSocket SDK, mobile SDKs and an Atoms MCP server that creates and configures agents and campaigns. Language models can be Smallest AI's own or any OpenAI compatible endpoint the customer brings, and the full stack is also sold through AWS Marketplace.
In practice
A support team ships a phone agent in an afternoon: four short prompts generate a Single Prompt agent with voice, model, and knowledge base configured, tested over a web call before going live on a real number.
A developer building a voice assistant needs speech that keeps up with conversation. Lightning V3 returns audio in roughly one hundred milliseconds and Pulse transcribes with a sixty four millisecond first token, keeping each turn under eight hundred milliseconds.
An operations team runs outbound campaigns in twenty plus languages, with agents collecting information, completing transactions, and triggering downstream actions through API calls, while post call metrics and conversation logs feed review.
Sources & related URLs
Related / legacy domains
Research sources
Agentic Index coverage score
11.0 / 14 capabilities · 79%
| Integrations & Tool Calling | Full |
|---|---|
|
Agents carry API call tools that make HTTP requests to outside systems mid call, a pre call API request fetches context before the call connects, and the conversation log records the API calls and webhook triggers each call made; tools are added, updated and removed per agent. SourceSmallest AI, docs.smallest.ai Atoms MCP available tools and Conversation logsread 2026-09-28 |
|
| Workflow Orchestration | Full |
|
Multi agent playbooks (SOPs with their own prompt, intent, auth level and tools) sit behind an intent router with a fallback playbook and mid call rerouting, Conversational Flow agents run node based flows, and the conversation log shows the routing path through multi agent setups. SourceSmallest AI, docs.smallest.ai Atoms MCP available tools and conversation logs pagesread 2026-09-28 |
|
| Knowledge Grounding & RAG | Full |
|
Knowledge bases hold uploaded PDFs and scraped web pages, are searched by semantic embeddings and injected into the model's context before the agent answers, and live apart from agents so content can be updated without touching agent configuration. SourceSmallest AI, docs.smallest.ai Knowledge Base overviewread 2026-09-28 |
|
| Human Oversight & Guardrails | Partial |
|
Agents hand calls to a person by cold or warm transfer, with a private whisper to the destination on a warm transfer, when the prompt tells them to, and playbooks carry an auth level. No step where a person approves an agent action before it commits is documented. SourceSmallest AI, docs.smallest.ai Call transfer and MCP available tools pagesread 2026-09-28 |
|
| Security, Identity & Governance | Full |
|
Organizations have Admin and Member roles (only Admins manage the team, settings and billing), SSO is part of the Enterprise plan, and the site states ISO 27001, SOC 2 Type 2, GDPR and HIPAA compliance with a security compliance portal linked in the footer. SourceSmallest AI, docs.smallest.ai Organization members, smallest.ai homepage and pricing pageread 2026-09-28 |
|
| Observability & Auditability | Full |
|
Each call's conversation log shows the timestamped transcript, the tool and function calls, API calls and webhook triggers, per turn latency, the routing path through multi agent setups, the recording, the cost breakdown and an event timeline, exportable as JSON or CSV. SourceSmallest AI, docs.smallest.ai Conversation logsread 2026-09-28 |
|
| Memory & State Persistence | Partial |
|
Session state lasts for the current call only, and the docs tell developers to keep anything that must survive between calls in their own external store, such as Redis keyed by phone number. No platform memory across calls is provided. SourceSmallest AI, docs.smallest.ai State managementread 2026-09-28 |
|
| Deployment & Data Residency | Full |
|
On premises deployment is listed as an option for the voice agent product, and the Enterprise plan offers dedicated infrastructure. The limit: no region list is published. SourceSmallest AI, smallest.ai homepage and pricing pageread 2026-09-28 |
|
| Prebuilt Agents, Templates & Packs | Partial |
|
Agents can start from a template gallery filtered by industry, direction and agent type, and a cookbook repository holds runnable agent templates and examples. The docs do not name the templates or the job each does. SourceSmallest AI, docs.smallest.ai Quick start and Using cookbooksread 2026-09-28 |
|
| Triggers & Channel Coverage | Full |
|
Outbound campaigns start automatically at a scheduled date and time or through the API, with an audience and retry settings, and agents take calls on provisioned phone numbers. SourceSmallest AI, docs.smallest.ai Campaigns page and Atoms MCP available toolsread 2026-09-28 |
|
| Model Flexibility & Routing | Full |
|
Developers set each agent's language model, and bring your own model connects any endpoint that implements the OpenAI chat completions API, including self hosted Ollama, vLLM and LM Studio servers, beside Smallest AI's own Electron model; the customer chooses. SourceSmallest AI, docs.smallest.ai Bring your own modelread 2026-09-28 |
|
| APIs, SDKs & MCP Extensibility | Full |
|
An Atoms MCP server exposes tools that create, configure, duplicate and archive agents, manage their tools and playbooks, and create, start and pause campaigns, with drafts published to go live, and campaigns can also be created over the API; the documentation index lists REST, WebSocket and mobile SDKs beside it. SourceSmallest AI, docs.smallest.ai Atoms MCP available tools and Campaigns pageread 2026-09-28 |
|
| Testing, Debugging & Optimization | Partial |
|
Agents are tried by hand through web calls, telephony calls and chat before going live, agents can be duplicated to make variants, and post call analytics extract a summary and disposition with reasoning per call. No scored test cases, simulations or quality gates are documented. SourceSmallest AI, docs.smallest.ai Evaluations and Call metrics pagesread 2026-09-28 |
|
| Browser & Computer Use | Not documented |
|
Agents act through API call tools, webhooks and telephony; no browser, desktop or computer control is documented. SourceSmallest AI, docs.smallest.ai Atoms pagesread 2026-09-28 |
|
The Agentic Index coverage score grades every vendor Full, Partial or Not documented against the same 14 buyer facing capabilities, from public evidence only. Each capability links to how all vendors in the index score on it. How this evidence is graded
Recent platform changes
Smallest AI added POST /waves/v1/auth/token so servers can issue temporary credentials for browser and mobile speech clients while retaining the API key. Tokens last 30 to 900 seconds, defaulting to 300, and authorize supported speech synthesis, transcription, speech to speech, and voice listing routes. They cannot authorize voice cloning, LLM chat completions, analytics, or further token creation; usage remains billed to the issuing key.
Bears on: Security / enterprise
View sourceSmallest AI shipped new MCP tools for latency summaries and prompt cache hit rates, reworked telephony so SIP trunks are managed as resources, and released SDK 5.5.0.
Bears on: MCP / tool calling / API
View sourceSmallest AI released an integration with Pipecat, the open source framework for voice and multimodal agents, so developers can use Smallest AI's voice stack inside Pipecat.
Bears on: Integrations
View sourcePricing
Pay as you go from $0.05 a minute ($0.09 to $0.21 depending on models), plus $0.01 a minute hosting; $10 in free credits
per agent minute by model choice, plus a per minute hosting fee
What is public
Per minute ranges by model, the hosting fee, the free credits, the included concurrency and the Enterprise inclusions, all on the pricing page; Enterprise rates are by sales.
Billing mechanics
Pay as you go credits: $0.09 to $0.21 a minute depending on the models selected (as low as $0.05), plus $0.01 a minute hosting at a flat rate, with 20 concurrent agents and unlimited agents included and no contract. Enterprise pricing is tailored and adds dedicated infrastructure, a 99.99 percent uptime SLA, forward deployed engineers, priority support and SSO.
Cost watchouts
Telephony carries per minute economics on both agent usage and phone numbers, and long average call durations multiply spend faster than call counts suggest.
Variable cost rationale
Cost tracks conversation minutes: the official FAQ puts Atoms at $0.08 per minute falling two to three times lower with volume, so spend scales directly with call volume and duration. Subscription plans layer monthly credits on top, and the company markets TTS at $0.01 per minute at scale, so the usage axis dominates.
Additional watchouts
Model choice moves the per minute rate by up to four times, and hosting adds $0.01 a minute on top; price the models you will actually run.
Overage / add-ons
No quota; usage is drawn from credits at the per minute rate
Sales call required
Mixed (some tiers require a call)
Free / trial
$10 in free credits on pay as you go; no commitment
Key ambiguities
Which model combinations land at the low and high ends of the $0.05 to $0.21 range is not itemized on the pricing page.
Missing data
Enterprise rates, the per minute rate for each model combination and any volume discount thresholds are not published.
Related vendors
- aiola — Voice agents for field teams on Salesforce: reps speak, and a squad…
- AviaryAI — Outbound AI voice agents for credit unions, banks and insurers,…
- Avoca — AI front office for home services: agents answer every call, text…
- Broccoli — AI voice agents built for the trades, answering every call for…
- Callers — Callers (formerly Voxia): AI agents for inbound and outbound calls,…
- careCycle — careCycle offers AI employees and a softphone platform for Medicare…
Alternatives to Smallest AI
The closest documented capability profiles to Smallest AI among voice agents tracked by Agentic Index, ordered by similarity on the same 14 point evidence the rankings use. No vendor pays for placement.
- Gnani.ai11.5 / 14Fuller documented coverage on Prebuilt Agents, Templates & Packs
- Vapi11.5 / 14Fuller documented coverage on Testing, Debugging & OptimizationSmallest AI vs Vapi →
- Thoughtly10.0 / 14A lighter documented profile than Smallest AI
- Phonely11.5 / 14Fuller documented coverage on Prebuilt Agents, Templates & Packs and Testing, Debugging & Optimization
- Regal10.5 / 14Fuller documented coverage on Testing, Debugging & Optimization
- Synthflow12.5 / 14Fuller documented coverage on Memory & State Persistence and Prebuilt Agents, Templates & Packs
Similarity is computed from each vendor's Agentic Index coverage score evidence, axis by axis, not from the totals. How this evidence is graded