Back to vendors
S

Smallest AI

Also known as: Smallest.ai, Atoms, Waves, Lightning, Pulse

Visit site
Entry pricePay as you go from $0.05 a minute ($0.09 to $0.21 depending on models), plus $0.01 a minute hosting; $10 in free creditsFull pricing detail

Full stack voice AI platform combining Atoms, a no code voice and chat agent builder handling conversations in more than twenty languages, with its own co optimized models: Lightning text to speech at roughly one hundred millisecond latency and Pulse speech to text at sixty four millisecond first transcript.

Smallest AI was founded by Sudarshan Kamath and Akshat Mandloi and has raised eight million dollars in seed funding led by Sierra Ventures, months after a pre seed round led by 3one4 Capital. The company's thesis is precision over scale: many small, efficient models, each tuned for a narrow job and optimized to run together rather than assembled from different vendors. The company says that in its first year it cut text to speech latency from roughly two seconds to one hundred milliseconds, dropped TTS cost from twenty cents to a penny per minute at scale, and moved from prototype to powering millions of enterprise calls each month.

The stack has three layers. Lightning is the text to speech line, with V3 hitting about one hundred millisecond latency across fifteen plus languages including Indic and European coverage, instant voice cloning from ten seconds of audio, and streaming over HTTP, SSE, or WebSocket. Pulse is streaming speech to text with a sixty four millisecond time to first transcript.

Atoms is the agent platform on top: no code Single Prompt and Conversational Flow agents with an Agentic Graph Builder, knowledge bases, webhooks, API calls from within conversations, phone numbers, outbound campaigns and audiences, a web voice and chat widget, post call metrics, conversation logs, versioning, and testing, all documented alongside Node.js and Python SDKs, a WebSocket SDK, mobile integrations, and an MCP section in the developer docs. The full agent pipeline stays under eight hundred milliseconds per turn, and the whole stack is available on AWS Marketplace under consolidated billing.

Smallest AI fits teams building production voice agents for support, sales, and operations who want the model layer and agent layer from one vendor with published usage pricing, and developers who need fast multilingual TTS and STT as raw APIs.

Vendor details

Canonical URL

https://smallest.ai

Category

Voice agent

Subcategory

Full stack voice AI: Atoms agents plus Lightning TTS and Pulse STT

Funding status

Independent, founded by Sudarshan Kamath and Akshat Mandloi. Raised eight million dollars in seed funding led by Sierra Ventures, months after a pre seed led by 3one4 Capital, and reports powering millions of enterprise calls per month within a year of launch.

Company status

independent

Use cases & customers

Primary use cases

Inbound and outbound voice agents for support and salesMultilingual text to speech for real time applicationsStreaming speech to text for voice pipelinesWeb voice and chat widgets on customer sites

Target customers

Teams deploying voice agents for support, sales, and operationsDevelopers building real time voice applicationsContact centers automating high call volumesMultilingual and Indic language market operators

Deployment options

Cloud

Integrations

Atoms agents make API calls to outside systems mid call, run a pre call API request, and fire webhooks; developers get REST APIs, SDKs, a WebSocket SDK, mobile SDKs and an Atoms MCP server that creates and configures agents and campaigns. Language models can be Smallest AI's own or any OpenAI compatible endpoint the customer brings, and the full stack is also sold through AWS Marketplace.

In practice

A support team ships a phone agent in an afternoon: four short prompts generate a Single Prompt agent with voice, model, and knowledge base configured, tested over a web call before going live on a real number.

A developer building a voice assistant needs speech that keeps up with conversation. Lightning V3 returns audio in roughly one hundred milliseconds and Pulse transcribes with a sixty four millisecond first token, keeping each turn under eight hundred milliseconds.

An operations team runs outbound campaigns in twenty plus languages, with agents collecting information, completing transactions, and triggering downstream actions through API calls, while post call metrics and conversation logs feed review.

Agentic Index coverage score

11.0 / 14 capabilities · 79%

Integrations & Tool Calling Full

Agents carry API call tools that make HTTP requests to outside systems mid call, a pre call API request fetches context before the call connects, and the conversation log records the API calls and webhook triggers each call made; tools are added, updated and removed per agent.

SourceSmallest AI, docs.smallest.ai Atoms MCP available tools and Conversation logsread 2026-09-28

Workflow Orchestration Full

Multi agent playbooks (SOPs with their own prompt, intent, auth level and tools) sit behind an intent router with a fallback playbook and mid call rerouting, Conversational Flow agents run node based flows, and the conversation log shows the routing path through multi agent setups.

SourceSmallest AI, docs.smallest.ai Atoms MCP available tools and conversation logs pagesread 2026-09-28

Knowledge Grounding & RAG Full

Knowledge bases hold uploaded PDFs and scraped web pages, are searched by semantic embeddings and injected into the model's context before the agent answers, and live apart from agents so content can be updated without touching agent configuration.

SourceSmallest AI, docs.smallest.ai Knowledge Base overviewread 2026-09-28

Human Oversight & Guardrails Partial

Agents hand calls to a person by cold or warm transfer, with a private whisper to the destination on a warm transfer, when the prompt tells them to, and playbooks carry an auth level. No step where a person approves an agent action before it commits is documented.

SourceSmallest AI, docs.smallest.ai Call transfer and MCP available tools pagesread 2026-09-28

Security, Identity & Governance Full

Organizations have Admin and Member roles (only Admins manage the team, settings and billing), SSO is part of the Enterprise plan, and the site states ISO 27001, SOC 2 Type 2, GDPR and HIPAA compliance with a security compliance portal linked in the footer.

SourceSmallest AI, docs.smallest.ai Organization members, smallest.ai homepage and pricing pageread 2026-09-28

Observability & Auditability Full

Each call's conversation log shows the timestamped transcript, the tool and function calls, API calls and webhook triggers, per turn latency, the routing path through multi agent setups, the recording, the cost breakdown and an event timeline, exportable as JSON or CSV.

SourceSmallest AI, docs.smallest.ai Conversation logsread 2026-09-28

Memory & State Persistence Partial

Session state lasts for the current call only, and the docs tell developers to keep anything that must survive between calls in their own external store, such as Redis keyed by phone number. No platform memory across calls is provided.

SourceSmallest AI, docs.smallest.ai State managementread 2026-09-28

Deployment & Data Residency Full

On premises deployment is listed as an option for the voice agent product, and the Enterprise plan offers dedicated infrastructure. The limit: no region list is published.

SourceSmallest AI, smallest.ai homepage and pricing pageread 2026-09-28

Prebuilt Agents, Templates & Packs Partial

Agents can start from a template gallery filtered by industry, direction and agent type, and a cookbook repository holds runnable agent templates and examples. The docs do not name the templates or the job each does.

SourceSmallest AI, docs.smallest.ai Quick start and Using cookbooksread 2026-09-28

Triggers & Channel Coverage Full

Outbound campaigns start automatically at a scheduled date and time or through the API, with an audience and retry settings, and agents take calls on provisioned phone numbers.

SourceSmallest AI, docs.smallest.ai Campaigns page and Atoms MCP available toolsread 2026-09-28

Model Flexibility & Routing Full

Developers set each agent's language model, and bring your own model connects any endpoint that implements the OpenAI chat completions API, including self hosted Ollama, vLLM and LM Studio servers, beside Smallest AI's own Electron model; the customer chooses.

SourceSmallest AI, docs.smallest.ai Bring your own modelread 2026-09-28

APIs, SDKs & MCP Extensibility Full

An Atoms MCP server exposes tools that create, configure, duplicate and archive agents, manage their tools and playbooks, and create, start and pause campaigns, with drafts published to go live, and campaigns can also be created over the API; the documentation index lists REST, WebSocket and mobile SDKs beside it.

SourceSmallest AI, docs.smallest.ai Atoms MCP available tools and Campaigns pageread 2026-09-28

Testing, Debugging & Optimization Partial

Agents are tried by hand through web calls, telephony calls and chat before going live, agents can be duplicated to make variants, and post call analytics extract a summary and disposition with reasoning per call. No scored test cases, simulations or quality gates are documented.

SourceSmallest AI, docs.smallest.ai Evaluations and Call metrics pagesread 2026-09-28

Browser & Computer Use Not documented

Agents act through API call tools, webhooks and telephony; no browser, desktop or computer control is documented.

SourceSmallest AI, docs.smallest.ai Atoms pagesread 2026-09-28

The Agentic Index coverage score grades every vendor Full, Partial or Not documented against the same 14 buyer facing capabilities, from public evidence only. Each capability links to how all vendors in the index score on it. How this evidence is graded

Recent platform changes

2026-09-22·MCP / tool calling / APIVerified

Smallest AI added POST /waves/v1/auth/token so servers can issue temporary credentials for browser and mobile speech clients while retaining the API key. Tokens last 30 to 900 seconds, defaulting to 300, and authorize supported speech synthesis, transcription, speech to speech, and voice listing routes. They cannot authorize voice cloning, LLM chat completions, analytics, or further token creation; usage remains billed to the issuing key.

Bears on: Security / enterprise

View source
2026-09-14·MCP / tool calling / APIPartially Verified

Smallest AI shipped new MCP tools for latency summaries and prompt cache hit rates, reworked telephony so SIP trunks are managed as resources, and released SDK 5.5.0.

Bears on: MCP / tool calling / API

View source
2026-09-12·IntegrationsVerified

Smallest AI released an integration with Pipecat, the open source framework for voice and multimodal agents, so developers can use Smallest AI's voice stack inside Pipecat.

Bears on: Integrations

View source
View all 5 changes for Smallest AI →Tracked since Jul 2026 · Verified from public vendor sources

Pricing

Pay as you go from $0.05 a minute ($0.09 to $0.21 depending on models), plus $0.01 a minute hosting; $10 in free credits

per agent minute by model choice, plus a per minute hosting fee

Free tier

What is public

Per minute ranges by model, the hosting fee, the free credits, the included concurrency and the Enterprise inclusions, all on the pricing page; Enterprise rates are by sales.

Billing mechanics

Pay as you go credits: $0.09 to $0.21 a minute depending on the models selected (as low as $0.05), plus $0.01 a minute hosting at a flat rate, with 20 concurrent agents and unlimited agents included and no contract. Enterprise pricing is tailored and adds dedicated infrastructure, a 99.99 percent uptime SLA, forward deployed engineers, priority support and SSO.

Cost watchouts

Telephony carries per minute economics on both agent usage and phone numbers, and long average call durations multiply spend faster than call counts suggest.

Variable cost rationale

Cost tracks conversation minutes: the official FAQ puts Atoms at $0.08 per minute falling two to three times lower with volume, so spend scales directly with call volume and duration. Subscription plans layer monthly credits on top, and the company markets TTS at $0.01 per minute at scale, so the usage axis dominates.

Additional watchouts

Model choice moves the per minute rate by up to four times, and hosting adds $0.01 a minute on top; price the models you will actually run.

Overage / add-ons

No quota; usage is drawn from credits at the per minute rate

Sales call required

Mixed (some tiers require a call)

Free / trial

$10 in free credits on pay as you go; no commitment

Key ambiguities

Which model combinations land at the low and high ends of the $0.05 to $0.21 range is not itemized on the pricing page.

Missing data

Enterprise rates, the per minute rate for each model combination and any volume discount thresholds are not published.

Agentic Index verified 2026-09-28

Alternatives to Smallest AI

The closest documented capability profiles to Smallest AI among voice agents tracked by Agentic Index, ordered by similarity on the same 14 point evidence the rankings use. No vendor pays for placement.

  • Gnani.ai11.5 / 14Fuller documented coverage on Prebuilt Agents, Templates & Packs
  • Vapi11.5 / 14Fuller documented coverage on Testing, Debugging & OptimizationSmallest AI vs Vapi →
  • Thoughtly10.0 / 14A lighter documented profile than Smallest AI
  • Phonely11.5 / 14Fuller documented coverage on Prebuilt Agents, Templates & Packs and Testing, Debugging & Optimization
  • Regal10.5 / 14Fuller documented coverage on Testing, Debugging & Optimization
  • Synthflow12.5 / 14Fuller documented coverage on Memory & State Persistence and Prebuilt Agents, Templates & Packs

Similarity is computed from each vendor's Agentic Index coverage score evidence, axis by axis, not from the totals. How this evidence is graded

Head to head

Contact us

Found a vendor we missed? Have feedback on the index? We'd love to hear from you.