Helicone
Also known as: Helicone AI
Open source LLM observability platform and AI gateway that logs, routes, caches, and costs every request through one OpenAI compatible endpoint.
Helicone is an open source platform that pairs LLM observability with an AI gateway. Teams route requests through its OpenAI-compatible gateway, or log them with an SDK, and every request is recorded with its prompt, completion, tokens, cost, latency and errors, grouped into sessions for agents and pipelines, with user analytics, custom properties and HQL, a SQL interface for querying analytics directly.
The AI Gateway reaches more than 100 providers through one API, with automatic provider routing, fallbacks, caching, custom rate limits by request count or cost, and provider keys stored encrypted so customers can bring their own. LLM Security screens user messages with Meta's Prompt Guard and optional Llama Guard and blocks detected threats, currently for OpenAI models, and moderation filters harmful content.
Alerts fire on error thresholds and cost limits, weekly reports go to email or Slack, and teams collect scores and feedback, curate datasets for fine-tuning and evaluation, and version prompts deployed through the gateway. A REST API, webhooks, a Zapier app and a Helicone MCP server make the data scriptable.
Helicone Cloud runs in the EU or US and states SOC 2 compliance, and the platform self-hosts with Docker or Kubernetes, with on-prem on Enterprise. Helicone was acquired by Mintlify on 3 March 2026; its own announcement says the services remain live for the foreseeable future in maintenance mode, with security updates, new models and fixes still shipping. Pricing is a free Hobby tier of 10,000 requests a month, Pro at $79 a month, Team at $799 a month with SOC-2 Type II and HIPAA, and custom Enterprise with SAML SSO and on-prem.
Vendor details
Canonical URL
https://www.helicone.ai
Category
Agent infrastructure
Subcategory
Observability and gateway
Funding status
A Y Combinator Winter 2023 company. The observability platform and its AI Gateway are both open source under Apache 2.0. Reports use by more than 1,000 AI teams and processing of nearly 10 billion requests. Independent.
Company status
acquired
Use cases & customers
Primary use cases
Target customers
Deployment options
Integrations
One line integration by changing the base URL or adding a header, compatible with the OpenAI SDK and working with OpenAI, Anthropic, Google, Azure, AWS Bedrock, LiteLLM, OpenRouter, and Gemini. The open source AI Gateway reaches 100 plus models through one endpoint, and logs export to tools like PostHog.
In practice
You ship an LLM feature with no visibility into cost or failures. You change one base URL to route through Helicone, and every request is logged with its cost, latency, and errors in a queryable dashboard.
You want the cheapest provider per call with automatic failover. Helicone's AI Gateway gives one endpoint to 100 plus models, routes by cost, and fails over to another provider when one is down.
Your AI spend is climbing and finance wants to know which customers drive it. You tag requests with custom properties in Helicone, track per user cost, and set alerts to catch spend spikes early.
Sources & related URLs
Related / legacy domains
Agentic Index coverage score
8.0 / 14 capabilities · 57%
| Integrations & Tool Calling | Partial |
|---|---|
|
Model providers connect through the gateway, and Helicone also works with frameworks such as the Vercel AI SDK, LiteLLM, OpenRouter and CrewAI, sends webhooks out, and offers a Zapier app that runs chat completions through the gateway. These route model traffic and send events; none connects an agent to a real system to take actions. SourceHelicone, docs.helicone.ai/llms.txt (integrations, webhooks, Zapier, web search)read 2026-09-21 |
|
| Workflow Orchestration | Not documented |
|
Per request, the gateway routes and falls back between providers, and prompts are assembled from versioned templates at call time. No workflows that sequence, branch or retry agent steps, or combine deterministic nodes with agent steps, are documented. SourceHelicone, docs.helicone.ai/llms.txt (provider routing, prompt assembly)read 2026-09-21 |
|
| Knowledge Grounding & RAG | Not documented |
|
Model traffic is routed and logged, web search can be added to Anthropic models through the gateway, and datasets hold logged requests for evaluation and fine-tuning. No document ingestion, index, retrieval layer or knowledge API that grounds an agent in company data is documented. SourceHelicone, docs.helicone.ai/llms.txt (web search, datasets)read 2026-09-21 |
|
| Human Oversight & Guardrails | Full |
|
With LLM Security enabled, Helicone screens each user message with Meta's Prompt Guard, and optionally Llama Guard across 14 threat categories, and blocks detected threats with an error response before the request is processed; moderation filters harmful content, and custom rate limits cap requests by count or cost. SourceHelicone, docs.helicone.ai/features/advanced-usage/llm-security and llms.txt (moderations, custom rate limits); docs.helicone.ai/features/advanced-usage/llm-security.mdread 2026-09-21 |
|
| Security, Identity & Governance | Full |
|
The pricing page lists SAML SSO on the Enterprise plan, with identity integration through the customer's own provider and organizations and seats per plan. SOC 2 Type II and HIPAA compliance are listed from the Team plan up, with InfoSec reviews and customized MSAs on Enterprise; the data autonomy reference adds encrypted provider keys, EU or US regions and configurable retention. Roles or permissions inside an organization are not documented. SourceHelicone, helicone.ai/pricing and docs.helicone.ai/references/data-autonomyread 2026-10-01 |
|
| Observability & Auditability | Full |
|
Every request through the gateway or SDK is logged with prompt, completion, tokens, cost, latency and errors, and Helicone groups them into sessions for agents and pipelines, adds user analytics and custom properties, and lets teams query analytics directly with HQL, a SQL interface with row-level security; data export is available, and retention runs 7 days on Hobby, 1 month on Pro, 3 months on Team and indefinitely on Enterprise, with configurable retention there. SourceHelicone, docs.helicone.ai/llms.txt (sessions, HQL, data export) and helicone.ai/pricingread 2026-09-21 |
|
| Memory & State Persistence | Not documented |
|
Logged requests, sessions and prompt versions persist, and the gateway can automatically clear old tool uses and thinking blocks from long-running agent sessions. No session, workflow or long term memory that an agent reads and writes is documented. SourceHelicone, docs.helicone.ai/llms.txt (sessions, context editing)read 2026-09-21 |
|
| Deployment & Data Residency | Full |
|
Customers choose an EU or US region for data residency on Helicone Cloud, the open source platform self-hosts with Docker or Kubernetes and Helm, and the Enterprise plan lists on-prem deployment. SourceHelicone, docs.helicone.ai/references/data-autonomy and llms.txt (self-hosting, Docker, Kubernetes), and helicone.ai/pricing; docs.helicone.ai/references/data-autonomy.mdread 2026-09-21 |
|
| Prebuilt Agents, Templates & Packs | Not documented |
|
Helicone publishes a model registry and pricing data for the models its gateway reaches and cookbooks on building agents. No ready-made agents, templates or packaged workflows a buyer adopts are documented. SourceHelicone, docs.helicone.ai/llms.txt (model registry, cookbooks)read 2026-09-21 |
|
| Triggers & Channel Coverage | Full |
|
Alerts fire when an application crosses error thresholds or cost limits, automated weekly reports deliver usage, cost and performance summaries to email or Slack, and webhooks send request events to the customer's endpoints. Each starts without a person asking, on a threshold or a schedule. SourceHelicone, docs.helicone.ai/llms.txt (alerts, reports, webhooks) and helicone.ai/pricingread 2026-09-21 |
|
| Model Flexibility & Routing | Full |
|
The Helicone AI Gateway gives one OpenAI-compatible API across more than 100 providers with automatic provider routing, fallbacks, caching, custom rate limits by request count or cost, and provider keys stored encrypted in a vault, so customers bring their own keys or use gateway credits. SourceHelicone, docs.helicone.ai/llms.txt (AI Gateway overview, provider routing, error handling and fallback, custom rate limits) and docs.helicone.ai/references/data-autonomyread 2026-09-21 |
|
| APIs, SDKs & MCP Extensibility | Full |
|
Helicone documents a REST API for requests, sessions, scores, evaluations, prompts, webhooks and trace logging, an OpenAI-compatible gateway API, SDK integrations, webhooks, a Zapier app, and a Helicone MCP server that lets MCP-compatible assistants query observability data; the codebase is open source. SourceHelicone, docs.helicone.ai/llms.txt (REST reference, MCP server, webhooks, Zapier)read 2026-09-21 |
|
| Testing, Debugging & Optimization | Partial |
|
Scores and user feedback on requests and sessions come in through the API, and Helicone curates request data into datasets for fine-tuning and evaluation, provides a playground and versioned prompts deployed through the gateway, and tracks score distributions over time. Running an application against fixtures or datasets before production and configurable quality gates are not documented. SourceHelicone, docs.helicone.ai/llms.txt (eval scores, datasets, prompts, evaluations API) and helicone.ai/pricingread 2026-09-21 |
|
| Browser & Computer Use | Not documented |
|
Helicone documents observability, the AI Gateway, prompts, datasets, security features and self-hosting, and no browser, desktop or computer control by an agent is documented. SourceHelicone, docs.helicone.ai/llms.txtread 2026-09-21 |
|
The Agentic Index coverage score grades every vendor Full, Partial or Not documented against the same 14 buyer facing capabilities, from public evidence only. Each capability links to how all vendors in the index score on it. How this evidence is graded
Pricing
From $79/mo · free tier (10k requests/mo) + open source
Subscription tiers by features and seats, with usage based scaling above included request volume; AI Gateway billed as zero markup provider pass through
Included quota
Hobby free 10,000 requests/mo (1GB storage, 1 seat). Pro $79/mo (unlimited seats). Team $799/mo (5 orgs). Enterprise custom. AI Gateway provider usage is pass through at 0% markup.
What is public
Helicone publishes its full tier structure (Hobby, Pro, Team, Enterprise) with request, seat, and feature limits, plus the zero markup gateway model. Enterprise pricing is custom.
Billing mechanics
A feature and seat based subscription on the observability platform, with Pro scaling on request volume. Routing model traffic through the AI Gateway is billed as pass through provider credits at no markup. The open source builds carry no license fee when self hosted.
Cost watchouts
Request volume drives Pro scaling, and gateway provider pass through grows with model usage. Compliance (HIPAA) and SSO sit on higher tiers.
Variable cost rationale
The subscription tiers are predictable, but Pro scales with request volume above the included amount, and if you route model traffic through the AI Gateway you also pay pass through provider costs that grow directly with usage, albeit at zero markup.
Additional watchouts
Two cost layers can stack: the platform subscription and, if you use the gateway, pass through provider spend. HIPAA is gated to the Team tier and above, and SAML SSO and on premises to Enterprise.
Overage / add-ons
Pro scales with usage above the included request volume. AI Gateway usage is billed as pass through provider credits at zero markup, so you pay exactly what providers charge plus the Stripe processing fee. Gateway credits are added at helicone.ai/credits.
Sales call required
Mixed (some tiers require a call)
Free / trial
Free Hobby plan: 10,000 requests a month, 1 GB storage, 1 seat, 7-day retention, no card; 7-day free trial of Pro and Team; open source and self-hostable
Lowest paid plan
Pro $79/mo (unlimited seats, alerts, HQL)
Commercial notes
Open source led, with a free Hobby tier and one-line integration, Pro and Team as self-serve upgrades with free trials, and Enterprise for SSO and on-prem. Helicone is now owned by Mintlify and runs in maintenance mode, per its own announcement.
Key ambiguities
Helicone was acquired by Mintlify on 3 March 2026 and states its services remain live for the foreseeable future in maintenance mode, with security updates, new models and fixes still shipping. Usage-based pricing above the included requests and storage is estimated by a calculator rather than itemized.
Cancellation / refund
Hobby, Pro, and Team are self serve subscriptions with standard cancellation. Enterprise terms are contractual.
Support SLA / resale
Community support on Hobby, standard support on Pro, Slack support on Team, and dedicated support with an SLA on Enterprise.
Missing data
Enterprise pricing is custom. Exact request overage rates on Pro above the included volume are not itemized.
Related vendors
- AgentOps — Agent observability and debugging platform: open source SDKs trace…
- Agno — Python agent framework and AgentOS runtime (formerly Phidata) for…
- AIsa — Resource and payment gateway for AI agents: one key to 110+ models…
- AlphaBitCore — AI control plane for regulated financial firms: one gateway enforces…
- Anchor Browser — Cloud hosted browser infrastructure that lets AI agents operate real…
- Apify — Cloud platform and marketplace of more than 73,000 ready-to-run…
Alternatives to Helicone
The closest documented capability profiles to Helicone among agent infrastructure platforms tracked by Agentic Index, ordered by similarity on the same 14 point evidence the rankings use. No vendor pays for placement.
- Traceloop8.0 / 14Fuller documented coverage on Testing, Debugging & Optimization
- Fiddler AI8.5 / 14Fuller documented coverage on Security, Identity & Governance and Testing, Debugging & Optimization
- F5 AI Guardrails9.0 / 14Adds documented Prebuilt Agents, Templates & Packs
- Freeplay8.0 / 14Fuller documented coverage on Security, Identity & Governance and Testing, Debugging & Optimization
- Galileo9.0 / 14Adds documented Prebuilt Agents, Templates & Packs
- Metorial8.0 / 14Fuller documented coverage on Integrations & Tool Calling and Security, Identity & Governance
Similarity is computed from each vendor's Agentic Index coverage score evidence, axis by axis, not from the totals. How this evidence is graded