Helicone
Helicone is an AI gateway and LLM observability platform that routes, logs, and cost-tracks AI app traffic across 100+ models.
Helicone earns its place when you route across more than one provider and want the gateway and the analytics in one pane. The 0% passthrough markup is the real differentiator — most gateways take a cut, and that compounding fee gets ugly at scale. Hobby's 10,000 requests/month and 7-day retention are tighter than the marketing implies, and Pro at $79/mo plus usage-based overages is where most growing teams land. LangSmith goes deeper on prompt evaluation; OpenRouter competes on raw routing. The March 2026 Mintlify acquisition adds roadmap uncertainty worth pricing in.
Verified 1h ago · liveness 87/100 · cite: rightaichoice.com/tools/helicone
- Fast-growing AI companies routing calls across three or more LLM providers
- Teams that need unified routing, cost tracking, and alerts on production LLM traffic
- Companies migrating off OpenRouter to a gateway with 0% passthrough markup
- Engineering orgs that must enforce rate limits and caching on every model call
- Teams that need a fully self-hosted stack — on-prem is Enterprise-only
- Hobby-tier users hitting the 10,000 requests/month cap or 7-day retention wall
- Projects that prioritize deep prompt evaluation and experiment tooling over gateway control
We scan live Reddit threads, YouTube comments, X posts, G2 reviews and other communities — and hand you an honest verdict in under a minute.
- Honest verdict, not marketing
- Real pros & cons from real users
- Attributed quotes with receipts
3 free scans · no card needed
Skip Helicone if your priority is deep prompt evaluation and experiment scoring rather than gateway routing and cost control, or if you need fully self-hosted deployment without Enterprise pricing.
Pro and Team charge usage-based rates above the 10,000 free requests each month, so a traffic spike lands on your invoice on top of the $79 or $799 flat fee
Hobby is genuinely free and suits a solo developer or prototype under 10,000 requests/month. Pro at $79/mo fits a small team that has outgrown the 7-day retention window and needs unlimited seats plus HQL. Team at $799/mo is the compliance and scale step, and it's priced above lighter rivals like Langfuse while landing below enterprise observability suites. Enterprise is quoted and adds SSO, on-prem, and bulk cloud discounts.
In short
Helicone — Helicone is an AI gateway and LLM observability platform that routes, logs, and cost-tracks AI app traffic across 100+ models. Best for Fast-growing AI companies routing calls across three or more LLM providers, Teams that need unified routing, cost tracking, and alerts on production LLM traffic, Companies migrating off OpenRouter to a gateway with 0% passthrough markup. Free to start; paid plans from $79/mo.
What's new in Helicone
Checked todayAcross the latest 5 updates: 2 feature updates, 1 launch and 2 news mentions.
Helicone acquired by Mintlify after three years and 14.2 trillion tokens
Helicone announced it has been acquired by Mintlify. The announcement cites three years of operation and 14.2 trillion tokens processed through the platform.
MCPs: What, Why, and How
Helicone published a guide to Model Context Protocol covering what it is, why to use it, and how to implement it, tied to Helicone's MCP compatibility.
Claude Sonnet 4 and Sonnet 4.5 now support 1M context window
Sonnet 4 and 4.5 on the Helicone AI Gateway now default to a 1M token context window across Anthropic API, AWS Bedrock, and Google Vertex AI, with no configuration changes needed.
The Helicone AI Gateway, Now On The Cloud
The Helicone AI Gateway moved onto the cloud with passthrough billing from the dashboard, observability built in, and OpenAI-compatible API access to any model.
Control Reasoning Effort in Playground and better feedback on thinking models
The Playground added reasoning effort control with a new minimal option alongside low, medium, and high, plus a visual display of the model's thinking process where available.
Viability Score
How well maintained and how widely used is Helicone? Built from what the vendor actually publishes (docs, changelog, tutorials, integrations, pricing), whether the site is live, and how much real users discuss it. How we calculate this
Last calculated: September 2026
How we score →Key Features
- AI Gateway routing across 100+ LLM providers
- Passthrough billing at 0% markup on provider costs
- Per-request LLM observability logging
- Sessions and users tracking for full journey reconstruction
- Playground with reasoning effort control (minimal/low/medium/high)
- Visual thinking display for reasoning models
- Datasets for versioning prompt data and regression testing
- HQL (Helicone Query Language) SQL-like log analytics
- Prompt Management V2 with typed variables and version rollback
- Caching to reduce latency and provider spend
- Automatic fallbacks between providers
- Rate limiting on gateway traffic
- Alerts and reports on cost and error spikes
- MCP (Model Context Protocol) compatibility
- 1M context window support for Claude Sonnet 4 and 4.5
About Helicone
Helicone does two jobs at once: it sits in front of your model calls as an AI gateway and logs what happens on the other side. You point your OpenAI-compatible requests at it and get multi-provider routing, caching, rate limiting, and automatic fallbacks on the way in, plus per-request logging, session and user tracking, and cost analytics on the way out. Routing covers 100+ providers including OpenAI (GPT-5, GPT-5-Mini, GPT-5-Nano, GPT-5-Chat-Latest), Anthropic (Claude Opus 4.1, Claude Sonnet 4 and 4.5), AWS Bedrock, Google Vertex AI, Azure, OpenRouter, Groq, Fireworks, Together AI, Anyscale, and LiteLLM. Passthrough billing carries 0% markup, so you pay the provider's rate. Observability includes a Playground with reasoning effort control (minimal/low/medium/high) and a visual thinking display for reasoning models, Datasets for versioning prompt data, Prompt Management V2 with typed variables and version rollback, and HQL, a SQL-like query language over your logs. Country-based request filtering arrived in July 2025, and Sonnet 4 and 4.5 now default to a 1M token context window across Anthropic, Bedrock, and Vertex. Helicone reports 14.2 trillion tokens processed over three years. In March 2026, Helicone was acquired by Mintlify — worth weighing if you care about long-term roadmap commitments. Pricing is usage-based on top of flat tiers: Hobby is free, Pro runs $79/mo, Team runs $799/mo, and Enterprise is quoted.
Behind the Verdict
Helicone's pitch is one control plane instead of a stack of half-connected tools, and for engineering teams that already route across multiple providers, that pitch mostly holds up. The gateway side is where the economics live: passthrough billing at 0% markup means you pay the provider's rate and nothing on top, which matters more the larger your token spend gets. Caching and automatic fallbacks cut both latency and spend without you writing provider-specific retry logic, and rate limiting is enforced at the gateway rather than scattered across your services. The observability side is deeper than a simple request log. You get sessions and users tracking, so you can reconstruct a full user journey instead of staring at isolated requests. The Playground supports reasoning effort control — minimal, low, medium, high — with a visual thinking display for reasoning models, which is a genuinely useful debugging aid when a model's answer goes sideways for non-obvious reasons. Datasets let you version prompt data and run regression checks, HQL gives you SQL-like queries over your logs, and Prompt Management V2 supports typed variables inside system prompts, messages, and tool schemas, with version control and rollback that don't require a code deploy. Country-based filtering, added in July 2025, uses the Cloudflare edge server receiving the request, so geographic analytics come for free if you're already on the gateway. Where Helicone trails is deep prompt evaluation and experiment tooling — that's LangSmith and Braintrust territory, and Helicone's own comparison table concedes limited evaluation and experiments next to those products. If your bottleneck is systematic A/B testing of prompts and scoring outputs, you'll likely pair Helicone with something else. Single-provider shops also won't get much from the multi-model routing, though the logging and cost tracking still apply. The scale claims are unusually concrete: 14.2 trillion tokens processed across three years, and 1M context window support for Claude Sonnet 4 and 4.5 across Anthropic, Bedrock, and Vertex. In September 2025 the AI Gateway moved onto the cloud with passthrough billing from the dashboard and an OpenAI-compatible API to any model. The open question is roadmap: Helicone was acquired by Mintlify in March 2026, and while the platform still ships — changelog entries through November 2025 and blog posts into 2026 — buyers making a multi-year commitment should factor in that ownership change. Non-profits, students, open-source projects, and sub-$5M-funded startups under two years old get discounts, so check the discount page before paying list.
Researching Helicone? Get your full AI stack in 60 seconds.
Free, no signup — tell us your goal and get tools matched to your budget & existing stack.
Real-world workflow fit
Concrete scenarios for the personas Helicone actually fits — and what changes day-one when you adopt it.
You point your OpenAI-compatible calls at the Helicone gateway, use the Hobby tier's 10,000 free requests, and watch request logs and cost breakdowns populate in the dashboard.
Outcome: You get provider-agnostic logging and cost visibility without paying anything, and you know exactly when you're approaching the request cap that forces an upgrade.
You route production traffic across OpenAI and Anthropic through the gateway, enable caching and automatic fallbacks, and configure alerts for error and spend spikes on the Pro tier.
Outcome: Failures in one provider route to another automatically, cache hits cut provider spend, and alerts surface cost anomalies before they show up on the monthly invoice.
You move LLM traffic onto the Team or Enterprise tier to cover SOC-2 Type II and HIPAA requirements, then use country-based request filtering and a longer retention window to satisfy audit requests.
Outcome: Compliance review passes with per-request logs and geographic analytics, and your security team gets SAML SSO and configurable retention on Enterprise.
Use Cases
- Monitor all LLM API calls in real time with detailed logs and cost breakdowns
- Route requests across multiple providers to optimize for latency, cost, or reliability
- Debug failed requests and analyze user sessions to improve app performance
- Experiment with different prompts and models in the Playground before deploying
- Set alerts for rate limits, errors, or spending thresholds across your AI stack
- Track GPT-5, GPT-OSS, or Claude Opus 4.1 costs across Fireworks, Groq, and OpenRouter in one dashboard
- Ship prompt changes without code deployments using Prompt Management V2
- Filter and analyze request volume by country of origin
Models Under the Hood
as of 2026-09-22
Limitations
- The free Hobby plan caps at 10,000 requests per month, 10 logs/min ingestion, 1 GB storage, and 7-day data retention.
- Pro ($79/mo) and Team ($799/mo) include usage-based pricing above the free allowance, so requests, storage, and ingestion can push the bill past the flat tier.
- Team raises ingestion to 15,000 logs/min and API access to 1,000 calls/min, while Enterprise is the only tier with configurable retention, SAML SSO, and on-prem deployment.
- Helicone's own comparison table concedes limited evaluation and experiments next to LangSmith and Braintrust.
- The March 2026 acquisition by Mintlify adds long-term roadmap uncertainty.
as of 2026-09-30
Verification history
We have re-verified Helicone 19 times since . Each pass re-reads the vendor's own pages and re-checks every listed field against that evidence; passes where nothing had changed are marked as such.
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-checked, vendor evidence unchanged
Showing the 6 most recent of 19 verification passes.
Free to cite with attribution — this page re-verifies continuously.
12-month cost
Project the real annual outlay, including the implied monthly cost when only an annual tier is published.
Vendor list price only. Add-on usage, seat overages, and contract minimums are surfaced under Hidden costs & gotchas.
Plans compared
For each published Helicone tier: who it actually fits, and what it adds vs. the previous tier. Cross-reference the cost calculator above for projected annual outlay.
Hobby
$0/mo
Ideal for
Solo developer or side project running under 10,000 LLM requests per month who wants logging and cost tracking without paying
What this tier adds
Free entry point: 10,000 requests/month, 1 GB storage, 1 seat, 7-day retention, 10 logs/min ingestion, community support
Pro
$79/mo
Ideal for
Small engineering team that has outgrown 7-day retention and needs unlimited seats, HQL, and alerts on production traffic
What this tier adds
Adds unlimited seats, alerts and reports, HQL query language, 1,000 logs/min ingestion, 1-month retention, and chat and email support over Hobby
Team
$799/mo
Ideal for
Scaling or regulated company that needs SOC-2 Type II and HIPAA coverage plus multiple organizations under one account
What this tier adds
Adds 5 organizations, SOC-2 Type II and HIPAA compliance, a dedicated Slack channel, 15,000 logs/min ingestion, and 3-month retention over Pro
Enterprise
Custom
Ideal for
Large org with procurement, InfoSec, and deployment requirements that needs custom contracts and on-prem or configurable retention
What this tier adds
Adds custom MSA, SAML SSO, on-prem deployment, bulk cloud discounts, configurable retention, and a dedicated support engineer with SLAs over Team
Where the pricing makes sense
The company stage and team size where Helicone's pricing actually pencils out — and where peers do it cheaper.
Hobby is genuinely free and suits a solo developer or prototype under 10,000 requests/month. Pro at $79/mo fits a small team that has outgrown the 7-day retention window and needs unlimited seats plus HQL. Team at $799/mo is the compliance and scale step, and it's priced above lighter rivals like Langfuse while landing below enterprise observability suites. Enterprise is quoted and adds SSO, on-prem, and bulk cloud discounts.
Setup time & first value
How long it actually takes to get something useful out of Helicone — broken out by persona, not the marketing-page minute.
Solo developer: under 30 minutes to first logged request — it's a one-line integration change on your OpenAI-compatible client. Small team on Pro: an afternoon to wire caching, rate limits, and alerts plus invite unlimited seats. Enterprise: days to weeks, mostly procurement and InfoSec review around the custom MSA, SAML SSO, and on-prem deployment.
Switching to or from Helicone
How to bring data in from common predecessors and how to get it back out — written for the switcher, not the buyer.
- →From OpenRouter: point your OpenAI-compatible calls at the Helicone AI Gateway, which serves 100+ models at 0% markup, following the published migration guide
- →From direct provider SDKs: swap the base URL to the Helicone gateway and keep your existing request shapes
- →From LiteLLM: connect the LiteLLM integration so existing proxy routing feeds into Helicone logging and cost analytics
- →From LangSmith: keep Helicone for gateway routing, logging, and cost tracking while you decide what to do with deep evaluation workflows
- ↗To LangSmith: move prompt evaluation and experiment scoring across, accepting weaker multi-provider gateway economics
- ↗To Braintrust: shift experiment and eval workflows, and keep a separate gateway for routing
- ↗To direct provider SDKs: remove the gateway base URL and lose unified cross-provider logging and alerting
- ↗To a self-hosted stack: run your own observability layer instead of Helicone's cloud tiers
Integrations
Resources & Guides
- Resourcehelicone.ai
What Is An Ai Gateway
Helpful link from helicone.ai
- Resourcehelicone.ai
Complete Guide To Helicone Ai Gateway
Helpful link from helicone.ai
- Resourcehelicone.ai
How To Use Helicone In N8n Workflows
Helpful link from helicone.ai
- Resourcehelicone.ai
How To Migrate From Openrouter To Helicone Ai Gateway
Helpful link from helicone.ai
- Resourcehelicone.ai
10 Ways To Use Your Llm Logs
Helpful link from helicone.ai
- Resourcehelicone.ai
How To Use Ai Gateways To Enhance Ai App Reliability
Helpful link from helicone.ai
- Resourcehelicone.ai
What To Do When Openrouter Is Down
Helpful link from helicone.ai
- Resourcehelicone.ai
Top 5 Llm Gateways 2025
Helpful link from helicone.ai
- Resourcehelicone.ai
OpenRouter Alternatives in 2025
We tested every AI gateway on the market. Here are the results for the top 5 AI Gateways on the market.
- Resourcehelicone.ai
The Helicone Ai Gateway Now On The Cloud
Helpful link from helicone.ai
Tutorials & Learning
YouTube returned 6 videos for “Helicone”, and we withheld 6: 6 could not be judged, because “Helicone” is a single word that other videos use for other things. We are showing none, because we could not prove any of them are about Helicone.
Official links
Tools that pair well with Helicone
Common stack mates teams adopt alongside Helicone, with the specific reason each pairing earns its keep.
Confident AI
Enterprise LLM evaluation, observability, and AI red teaming that standardizes quality across every team.
Fiddler AI
Fiddler AI is an enterprise AI control plane for agent observability, guardrails, and governance across the agentic lifecycle.
Weights & Biases
Weights & Biases tracks ML experiments and traces LLM apps so teams can ship AI models faster
Alternatives to Helicone
View allConfident AI
Enterprise LLM evaluation, observability, and AI red teaming that standardizes quality across every team.
Fiddler AI
Fiddler AI is an enterprise AI control plane for agent observability, guardrails, and governance across the agentic lifecycle.
Weights & Biases
Weights & Biases tracks ML experiments and traces LLM apps so teams can ship AI models faster
Frequently Asked Questions
Used Helicone? Help shape our editorial sentiment research.