Gateway

Gateway

Portkey's AI Gateway routes, secures, and observes 3000+ LLMs through one OpenAI-compatible endpoint.

81/100Safe BetFree · from $49/moFreemium

Portkey earns its keep the moment you're calling more than two providers in production. Fallbacks plus semantic caching plus virtual keys solve real incidents, and the 2026 additions — Secret References from AWS/Azure/HashiCorp vaults, endpoint-scoped rate limits, and an Agent Gateway with a Skills Registry — push it from model proxy toward agent control plane. Compare against building on LiteLLM if you want self-hosted and free, or against single-provider consoles if you only ever call one model. If you're on one model with no governance needs, the $49/month Production tier buys you observability you probably won't read.

Verified 4d ago · liveness 81/100 · cite: rightaichoice.com/tools/gateway

Best for
  • Platform teams routing production traffic across multiple LLM providers
  • Developers who need fallbacks, retries, and load balancing without building a proxy
  • Teams cutting spend with semantic caching and provider batch APIs
  • Enterprises requiring SSO, RBAC, SOC 2 Type 2, ISO 27001, GDPR, HIPAA, and VPC hosting
Not ideal for
  • Hobbyists using a single model with no routing or governance needs
  • Non-developers looking for a no-code AI app builder
  • Projects requiring on-device or fully offline inference
Visit Website

IntermediateA developer can point the OpenAI SDK at Portkey's base URL and see traffic in the dashboard in under 15 minutes using the three-line snippets in the docs. Wiring routing rules, fallbacks, and virtual keys is an afternoon for a platform engineer. Enterprise deployments with VPC hosting, SSO, and custom guardrails take weeks because they involve procurement and infrastructure review, not becauseAPI · CLIAPI availableVerified 4d ago
Pricing
Free · from $49/mo
FreemiumFree tier4 plans5 hidden costs
Learning curve
Intermediate
A developer can point the OpenAI SDK at Portkey's base URL and see traffic in the dashboard in under 15 minutes using the three-line snippets in the docs. Wiring routing rules, fallbacks, and virtual keys is an afternoon for a platform engineer. Enterprise deployments with VPC hosting, SSO, and custom guardrails take weeks because they involve procurement and infrastructure review, not because
Runs on
APICLI
API available · 15 integrations
Who it's for
Backend engineer shipping a multi-provider chat featurePlatform team lead running agent workloads in productionCompliance-minded enterprise architect
Live sentiment
Is Gateway actually worth it?

We scan live Reddit threads, YouTube comments, X posts, G2 reviews and other communities — and hand you an honest verdict in under a minute.

  • Honest verdict, not marketing
  • Real pros & cons from real users
  • Attributed quotes with receipts
Run a free scan

3 free scans · no card needed

Skip it if

Skip Portkey if you call a single model, don't need fallbacks, caching, or per-team key governance, or refuse to put a hosted gateway in your request path.

The 30-second take
Biggest gripe

On the free Developer tier, going past 10k recorded logs a month stops new logs from being recorded — you keep serving requests but lose the audit trail until the next month.

Price reality

At $0/month the Developer tier is the cheapest credible way to prototype multi-provider routing; $49/month Production sits in the same band as a small observability tool and undercuts most APM suites. Cheaper only if you self-host the open source gateway. More expensive peers are enterprise LLMOps platforms that bundle governance into seven-figure contracts — Portkey's Enterprise tier keeps that negotiation custom rather than list-priced.

In short

Gateway — Portkey's AI Gateway routes, secures, and observes 3000+ LLMs through one OpenAI-compatible endpoint. Best for Platform teams routing production traffic across multiple LLM providers, Developers who need fallbacks, retries, and load balancing without building a proxy, Teams cutting spend with semantic caching and provider batch APIs. Free to start; paid plans from $49/mo.

What's new in Gateway

Checked 4 days ago

Across the latest 1 update: 1 news mention.

What people actually say about Gateway — is it worth it?

We scanned public community sources for Gateway on Jul 3, 2026 and could not establish that the discussion we found is about this tool rather than something else sharing its name. Our own analysis of that scan says the posts were off-subject. Rather than publish a sentiment score built on the wrong subject, we publish nothing here and re-run the scan.

Viability Score

81/100
Safe Bet

How well maintained and how widely used is Gateway? Built from what the vendor actually publishes (docs, changelog, tutorials, integrations, pricing), whether the site is live, and how much real users discuss it. How we calculate this

Recent activity
90
Traction
100
Site health
95
User sentiment
52
What the vendor publishes
60

Last calculated: October 2026

How we score →

Key Features

  • Unified API across 3000+ LLMs and 1600+ providers
  • Conditional routing on configurable custom rules
  • Automatic fallbacks and failover during provider errors
  • Load balancing across models and providers
  • Canary testing for new models and prompts
  • Automatic retries and configurable request timeouts
  • Simple and semantic caching with unlimited TTL stream-from-cache
  • Smart batching via provider batch APIs
  • Provider-specific fine-tuning through the unified API
  • Multimodal support for vision, audio, and image generation providers
  • Recording of OpenAI real-time API requests with cost tracking
  • Virtual keys in Portkey vault with rotation, revocation, and budgeting
  • Vault-backed Secret References from AWS Secrets Manager, Azure Key Vault, HashiCorp Vault
  • Observability: logs, traces, feedback, custom metadata, filters, alerts
  • Guardrails including LLM, partner, Zscaler AI Guard, Akto, and Bedrock customHost

About Gateway

FreemiumIntermediateAPI availableAPI · CLI

Portkey's AI Gateway sits between your app and every model you call, collapsing 1600+ provider integrations into one unified API. Platform, DevOps, and GenAI teams use it to centralize routing, key management, and observability instead of maintaining a pile of provider SDKs and dashboards. You integrate in roughly three lines of code — Python, NodeJS, REST, or by pointing the OpenAI SDK at Portkey's base URL — and it starts monitoring every LLM request. The gateway is multimodal by design, covering vision, audio, and image-generation providers, and it records OpenAI real-time API requests including cost and guardrail hits. Reliability features are concrete: conditional routing on custom rules, automatic fallbacks, load balancing, canary testing, retries, request timeouts, and both simple and semantic caching with unlimited TTL stream-from-cache. Cost and scale tooling includes smart batching through provider batch APIs, provider-specific fine-tuning via the unified API, and file uploads you can reference in requests. Governance covers virtual keys held in Portkey's vault (rotate, revoke, monitor), plus RBAC, SSO, and enterprise certifications — SOC 2 Type 2, ISO 27001, GDPR, HIPAA — on custom plans. Portkey reports 3000+ GenAI teams and billions of requests processed monthly. Across 2026 the company has pushed into agent governance: a Skills Registry for managing agent skills, an Agent Gateway for agent traffic, MCP governance guidance, and an explainer on why agent vulnerabilities are trust-boundary failures rather than model failures. March 2026 changelog additions include vault-backed Secret References (AWS Secrets Manager, Azure Key Vault, HashiCorp Vault), weekly and endpoint-scoped rate limits, Zscaler AI Guard and Akto guardrails, and new providers DeepInfra, DeepSeek, and Azure AI Foundry rerank. Pricing is straightforward: free forever for prototyping, $49/month for production, custom enterprise for compliance-heavy workloads. Unlike a hand-rolled proxy or a single-provider console, Portkey gives you one control plane for multi-model governance, caching, and spend tracking.

Behind the Verdict

Portkey's pitch is consolidation: instead of wiring 1600+ provider integrations yourself, you call one OpenAI-compatible endpoint and let the gateway handle provider diversity. The reliability layer is where it justifies itself first. Conditional routing, automatic fallbacks during provider errors, load balancing, automatic retries, request timeouts, and canary testing cover the failure modes that otherwise wake you at 3am. Add simple and semantic caching with unlimited TTL stream-from-cache and you have two independent cost levers — fewer upstream calls, and fewer cold ones. The observability layer is the second reason teams stay: logs, traces, feedback, custom metadata, and filters across every LLM call, with cost attribution per use case. The March 2026 changelog tightened governance with vault-backed Secret References pulling from AWS Secrets Manager, Azure Key Vault, and HashiCorp Vault, weekly budget windows, and endpoint-scoped rate limits — the kind of controls auditors ask about. Guardrail coverage now spans LLM and partner guardrails plus Zscaler AI Guard, Akto Agentic Security, and Bedrock Guardrails customHost. The direction of travel in 2026 is agents: Portkey shipped a Skills Registry for agent skills, an Agent Gateway sitting in front of agent traffic, MCP governance guidance, and a blog argument that agent vulnerabilities — MCP tool-description injection, calendar-invite prompt injection, unauthorized tool calls, cost overruns — are trust-boundary failures rather than model failures. That framing is useful, though it also means Portkey is betting its roadmap on agent workloads materializing the way it expects. The honest constraints: the free Developer tier records 10k logs per month with 3-day log retention and 30-day metrics — exceeding the limit stops recording, though your requests are not blocked, so you can silently lose the audit trail you're relying on. Production at $49/month raises that to 100k logs, 30-day log retention, and 90-day metrics, then charges $9 per additional 100k. Advanced security and governance — SSO, custom guardrail hooks, custom retention, granular budgets — sit on the Enterprise tier with custom pricing. And adopting a gateway adds a hop: Portkey benchmarks 20–40ms of added latency versus direct API calls, offset by caching and routing when those apply. Where it fits: platform and platform-adjacent teams running multi-provider production traffic, organizations that need per-team virtual keys with budgets, and compliance-bound enterprises that need VPC hosting and BAAs. Where it doesn't: single-model prototypes, no-code builders, fully offline inference, and teams unwilling to put a hosted gateway in the request path. This tool is an infrastructure control plane, not a thin wrapper over someone else's model.

Researching Gateway? Get your full AI stack in 60 seconds.

Free, no signup — tell us your goal and get tools matched to your budget & existing stack.

Real-world workflow fit

Concrete scenarios for the personas Gateway actually fits — and what changes day-one when you adopt it.

Backend engineer shipping a multi-provider chat feature

Swap the OpenAI SDK's base URL for Portkey's endpoint on a Friday, configure a fallback to Anthropic, and land the change behind the existing code paths.

Outcome: Provider outages no longer page the team; every request shows up in traces with cost attribution without adding a second SDK.

Platform team lead running agent workloads in production

Put the Agent Gateway in front of agent traffic, register reusable agent skills in the Skills Registry, and attach a fleet of per-service virtual keys with weekly budget windows.

Outcome: One budget and routing policy governs every agent, and a runaway tool-calling loop hits a rate limit instead of a surprise invoice.

Compliance-minded enterprise architect

Migrate secrets into Portkey's vault-backed Secret References from AWS Secrets Manager so raw provider keys never sit in application config, then enable guardrails and VPC hosting.

Outcome: Auditors see centralized key custody, guardrail enforcement, and 30-day-plus log retention without each team running its own provider console.

Use Cases

  • Route AI traffic to the cheapest or most reliable model per request
  • Automatically failover to backup providers during outages
  • Cache semantically similar prompts to cut API costs
  • Monitor all LLM calls with logs, traces, and cost breakdowns
  • Enforce guardrails to filter harmful content or validate output
  • Manage API keys centrally with virtual keys and budgets
  • Govern and observe AI agents via the Agent Gateway and Skills Registry
  • Track cost and guardrail hits on OpenAI real-time API requests

Models Under the Hood

GPT-3.5 Turbo

as of 2026-09-24

Limitations

  • The free Developer tier records 10k logs per month with 3-day log retention and 30-day metrics; exceeding the limit stops recording beyond the cap, though requests are not blocked, so you can lose the audit trail you were counting on.
  • Production at $49/month allows 100k logs with 30-day retention and 90-day metrics, then charges $9 per additional 100k.
  • Advanced security and governance — SSO, custom guardrail hooks, custom retention periods, granular budget and rate limits — are reserved for the Enterprise tier at custom pricing.
  • Portkey benchmarks 20–40ms of added latency versus direct API calls, though caching and routing often offset it.

as of 2026-10-04

Verification history

We have re-verified Gateway 8 times since . Each pass re-reads the vendor's own pages and re-checks every listed field against that evidence; passes where nothing had changed are marked as such.

  1. — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
  2. — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
  3. — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
  4. — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
  5. — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
  6. — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it

Showing the 6 most recent of 8 verification passes.

Free to cite with attribution — this page re-verifies continuously.

12-month cost

Project the real annual outlay, including the implied monthly cost when only an annual tier is published.

Annual total
Free
Over 12 months
Effective monthly
—
—

Vendor list price only. Add-on usage, seat overages, and contract minimums are surfaced under Hidden costs & gotchas.

Plans compared

For each published Gateway tier: who it actually fits, and what it adds vs. the previous tier. Cross-reference the cost calculator above for projected annual outlay.

Open Source

$0

Ideal for

Engineering teams that want routing, retries, fallbacks, and load balancing on their own infrastructure and are comfortable running a gateway themselves.

What this tier adds

Self-hosted starting point: you get the universal API and reliability features but no managed dashboard, hosted logs, or support contract.

Developer Free Forever

$0/mo

Ideal for

Solo developers and small teams prototyping a multi-provider feature or evaluating Portkey inside an enterprise POC before committing budget.

What this tier adds

Adds hosted observability and prompt management over the open source gateway, capped at 10k logs per month with 3-day log retention.

Production

$49/mo

Ideal for

Teams ready to run LLM features in production with real traffic, needing caching, RBAC, and a usable retention window but without custom compliance constraints.

What this tier adds

Raises the log cap tenfold to 100k, extends retention to 30 days for logs and 90 days for metrics, and adds semantic caching, unlimited prompt templates, LLM and partner guardrails, and RBAC.

Enterprise

Custom

Ideal for

Organizations with compliance requirements, data residency needs, or high-volume agent and multi-model traffic that outgrows the 100k-log Production cap.

What this tier adds

Adds SSO, custom guardrail hooks, custom retention periods, granular budgets and rate limits, VPC and private cloud hosting, data isolation, and SOC 2 Type 2 / ISO 27001 / GDPR / HIPAA coverage with BAAs.

Hidden costs & gotchas

What the public pricing page doesn't put in bold. Captured from pricing-page footnotes, contract terms, and recurring complaints.

  • On the free Developer tier, going past 10k recorded logs a month stops new logs from being recorded — you keep serving requests but lose the audit trail until the next month.
  • Production's $49/month includes 100k logs; each additional 100k costs $9, which can quietly push a busier app well above the sticker price.
  • The Production plan is explicitly not recommended for organizations needing custom security controls or data residency guarantees — that capability only arrives on Enterprise custom pricing.
  • Semantic caching, unlimited prompt templates, and RBAC are walled off from the free tier, so growing teams hit a jump from $0 to $49/month rather than a middle rung.
  • Log retention on Production is 30 days for logs and 90 days for metrics — anything you need to investigate beyond that window requires Enterprise custom retention.

Where the pricing makes sense

The company stage and team size where Gateway's pricing actually pencils out — and where peers do it cheaper.

At $0/month the Developer tier is the cheapest credible way to prototype multi-provider routing; $49/month Production sits in the same band as a small observability tool and undercuts most APM suites. Cheaper only if you self-host the open source gateway. More expensive peers are enterprise LLMOps platforms that bundle governance into seven-figure contracts — Portkey's Enterprise tier keeps that negotiation custom rather than list-priced.

Setup time & first value

How long it actually takes to get something useful out of Gateway — broken out by persona, not the marketing-page minute.

A developer can point the OpenAI SDK at Portkey's base URL and see traffic in the dashboard in under 15 minutes using the three-line snippets in the docs. Wiring routing rules, fallbacks, and virtual keys is an afternoon for a platform engineer. Enterprise deployments with VPC hosting, SSO, and custom guardrails take weeks because they involve procurement and infrastructure review, not because

Switching to or from Gateway

How to bring data in from common predecessors and how to get it back out — written for the switcher, not the buyer.

Migrating in
  • →From a hand-rolled proxy: replace your routing code with Portkey configs and keep the same request shape via the OpenAI-compatible endpoint.
  • →From LiteLLM: point your base URL at Portkey's gateway and move provider keys into virtual keys, reusing existing model names.
  • →From direct provider SDKs: use the OpenAI SDK plus Portkey headers so application code changes minimally.
  • →From a single-provider console: export existing usage expectations, then rebuild routing rules, fallbacks, and caching in Portkey.
Migrating out
  • ↗To LiteLLM: self-host the open source proxy and re-map Portkey configs to LiteLLM's YAML routing rules.
  • ↗To direct provider SDKs: drop the gateway header and restore per-provider SDKs, accepting the loss of centralized logs and caching.
  • ↗To a bundled observability suite: export Portkey logs to your data lake first, then rebuild cost and trace dashboards natively.

Integrations

OpenAIAnthropicGoogleAzureAWS BedrockGCPHugging FaceCohereMistralStability AIElevenLabsLangChainLlamaIndexZscalerAkto

Resources & Guides

Tutorials & Learning

YouTube returned 6 videos for “Gateway”, and we withheld 6: 6 could not be judged, because “Gateway” is a single word that other videos use for other things. We are showing none, because we could not prove any of them are about Gateway.

Tools that pair well with Gateway

Common stack mates teams adopt alongside Gateway, with the specific reason each pairing earns its keep.

Featured Head-to-Head Comparisons

Alternatives to Gateway

View all
OrcaRouter

OrcaRouter

One OpenAI-compatible endpoint in front of 200+ models, with per-prompt grading that routes each call — and no markup on tokens.

FreemiumTry
OpenRouter Agents

OpenRouter Agents

One OpenAI-compatible API that routes any request across 500+ text, image, video, and audio models from 80+ providers.

FreemiumTry
CometAPI

CometAPI

CometAPI is one OpenAI-compatible API key for 500+ text, image, video and audio models, priced at least 20% below official vendor rates.

FreemiumTry

Frequently Asked Questions

Used Gateway? Help shape our editorial sentiment research.