Portkey Review (2026)
An open-source LLM gateway that routes to 1,600+ models with fallbacks, caching, budgets and guardrails. The gateway self-hosts free under Apache 2.0 - but real logging, traces and analytics require the managed paid tier.
Rating
Starting Price
$49/mo
Free Plan
Yes
SDKs & Frameworks
8
Deployment
3
Best For
Teams that want one gateway to route, cache, budget and guardrail across many LLM providers, and are fine paying for the managed tier once they need real observability
Last Updated:
10 Things You Should Know About Portkey
- 1 The open-source gateway (Apache 2.0, ~12.5k GitHub stars) self-hosts free but gives routing plus a basic dashboard only
- 2 Real logging, traces, analytics and retention require the managed paid tier
- 3 Billing caps logs, not requests - requests keep flowing past the cap, only logs stop recording; overage is $9 per 100k requests
- 4 Founded January 2023 by Rohit Agarwal and Ayush Garg; raised a $3M seed then a $15M Series A led by Elevation Capital
- 5 Portkey claims sub-1ms added latency, but a Kong-run benchmark reported Portkey ~65% higher latency than Kong's gateway - both are labelled claims
Pros & Cons
Pros
- ✓ The open-source gateway is Apache 2.0 and self-hosts free with routing, fallbacks and guardrails
- ✓ Drop-in proxy - point your existing OpenAI SDK base URL at Portkey, no re-instrumentation
- ✓ Strong OpenTelemetry support that enriches traces with cost and token metrics per GenAI conventions
- ✓ One control plane for spend, routing and governance across 1,600+ models
Cons
- ✕ The observability catch - real logging, traces and analytics need the managed paid tier, even though routing is free
- ✕ Log caps, not request caps, with $9 per additional 100k requests
- ✕ Vendor latency and uptime numbers are self-reported, and a Kong benchmark disputes the latency edge
- ✕ No annual discount published
- ✕ Independent Reddit and HN sentiment is thin - lean on GitHub stars and funding for signal
Features
What Portkey actually is
Portkey is an LLM gateway. It sits between your app and your model providers as a drop-in proxy, routing calls to 1,600+ models through one API. On top of routing it adds fallbacks, load balancing, retries, timeouts, caching, budgets, rate limits and 50+ guardrails - plus an observability backend with logs, traces and metrics, and prompt management.
That makes it a different animal from the other tools in this category. The others are eval-first or trace-first. Portkey is a control plane for spend, routing and governance that happens to include observability. If your problem is “we call five providers, costs are exploding, and we need fallbacks and budgets in one place,” Portkey is aimed squarely at you. If your problem is “we need to score whether our agent’s answers are good,” that’s a different tool.
It’s a well-backed one. Portkey was founded in January 2023 by Rohit Agarwal and Ayush Garg, and raised a $3M seed followed by a $15M Series A led by Elevation Capital with Lightspeed. The open-source gateway repo sits at roughly 12,500 GitHub stars.
The distinctive part: the gateway itself
The gateway is the product, and it’s fast and broad. One base URL change routes your existing OpenAI SDK calls through Portkey, auto-enriched with provider config, caching status, retry attempts and prompt versions. No re-instrumentation for the routing path.
Portkey markets it as sub-1ms added latency with a 122kb footprint, handling 10B+ requests a month for 24,000+ orgs. Those are vendor numbers. A benchmark published by Kong, a competing gateway vendor, reported Portkey at about 65% higher latency than Kong’s own gateway under equivalent conditions. Both parties are interested, so both figures are claims - the honest read is that Portkey’s latency is competitive but not independently settled, and you should test on your own path.
The gotcha: observability is gated
This is the thing to understand before you plan a deployment. The open-source gateway gives you routing and a basic dashboard. Real observability lives on the paid tier.
The Apache 2.0 gateway self-hosts free and does the hard networking - routing, retries, fallbacks, load balancing, guardrails. What it does not give you is meaningful logging, traces, analytics or retention. For that you move to managed Production, or you build your own logging stack around the gateway. So a team that self-hosts Portkey expecting a full observability platform gets a proxy and a basic dashboard, and discovers the analytics they wanted are a paid feature.
That’s not dishonest - it’s open-core, and the free part is genuinely useful. But it inverts the usual self-host promise in this category. With Langfuse, self-host gets you the whole product. With Portkey, self-host gets you the pipe, not the dashboard.
Pricing: watch the log meter
| Tier | Price | Logs/mo | Log retention | Overage |
|---|---|---|---|---|
| Open Source | $0 | self-hosted | - | Apache 2.0 |
| Developer | $0 | 10,000 | 3 days | none |
| Production | $49/mo | 100,000 | 30 days | $9/100k requests |
| Enterprise | Custom | 10M+ | configurable | - |
One detail bites. The meter caps logs, not requests. Past the cap, your requests keep routing normally - you just stop recording logs beyond the limit unless you pay the overage, which is $9 per additional 100k requests. That’s friendlier than a hard request cap in one sense (your app never breaks), but it means your observability quietly goes dark at volume unless you’re watching the meter. There’s no annual discount published.
Portkey versus a pure observability tool
If you’re comparing Portkey to Langfuse or Laminar, you’re comparing different jobs. Portkey’s core value is routing, fallbacks, budgets and guardrails across many providers - the gateway. Its observability is a bundled bonus that gets good on the paid tier. A pure observability tool’s core value is the tracing and eval depth, and it doesn’t route your traffic. Plenty of teams run both: Portkey as the gateway, and a dedicated observability backend fed over OpenTelemetry. Portkey’s strong OTel support, enriching traces with cost and token metrics per the GenAI conventions, makes that pairing straightforward.
Should you use it?
Use Portkey if you call multiple LLM providers and want one place to route, cache, budget, rate-limit and guardrail them, and you’re fine paying for the managed tier once you need real logs and analytics.
Don’t use Portkey if you want free, fully self-hosted observability - the OSS build won’t give you that - or your only need is eval depth, where a purpose-built eval tool serves you better.
Bottom line: the best gateway in this set, with a genuinely useful free open-source core. Just map the split before you deploy - routing is free and self-hostable, but the observability most teams want is on the paid tier.
Pricing and features verified against portkey.ai on 23 July 2026. This category ships breaking changes monthly - we re-verify every 30 days.
Pricing Plans
Open Source
$0
- Self-hosted gateway under Apache 2.0
- Routing, retries, fallbacks, load balancing
- Guardrails and a basic dashboard
- Community support
Developer
$0
- Managed cloud, free forever
- 10,000 logs per month
- 3-day log retention, 30-day metric retention
- Prompt management (3 templates), playground
Production
$49/mo
- 100,000 logs per month
- 30-day log retention, 90-day metric retention
- Full gateway - fallbacks, load balancing, retries
- Advanced observability, guardrails, caching, RBAC
- $9 per additional 100k requests
Enterprise
Custom
- 10M+ logs
- Custom guardrail hooks, SSO, granular budgets
- Private and VPC deployment, data-lake exports
- SOC 2 Type 2, GDPR, HIPAA
SDKs & Frameworks
Deployment
Gateway Features
Our Verdict
Portkey is the strongest LLM gateway in this set, and the open-source core is a genuinely good, free, self-hostable proxy for routing, fallbacks and guardrails. The catch is where the value migrates. Self-host gives you routing for free, but the logs, traces and analytics that most teams actually adopt an observability tool for live on the managed tier. Know that split before you plan a self-hosted deployment, or you'll stand up the gateway and then discover the dashboard you wanted is a paid feature.
Similar Tools
TrueFoundry
Enterprises that need an AI gateway deployable on-premise or in their own cloud with governance features, and that want a managed product rather than operating open-source infrastructure.
Vercel AI Gateway
Teams already deploying on Vercel who want provider-agnostic model access at list price with no markup, and who do not need governance features that carry a per-request charge.
Kong AI Gateway
Organisations already standardised on Kong for API management, where bringing LLM traffic under existing governance matters more than LLM-specific features.
Martian
Nobody choosing a gateway today. The RouterBench work remains worth reading if you are evaluating routing approaches generally.
Frequently Asked Questions
Can I self-host Portkey?
The gateway, yes - it's open-source under Apache 2.0 and runs via Docker, giving you routing, retries, fallbacks, load balancing, guardrails and a basic dashboard for free. But that's the catch. Meaningful observability - full logs, traces, analytics and retention - is not in the open-source build. For that you either move to the managed paid tier or stand up your own logging stack. Self-hosting Portkey does not get you self-hosted observability.
Does Portkey support OpenTelemetry?
Yes, and strongly. Portkey ingests telemetry from any OTel-compatible source and automatically enriches OpenTelemetry traces with cost and token metrics following the GenAI Semantic Conventions. It works with any language that supports OpenTelemetry - Python, JavaScript, Java, Go - and with Vercel AI SDK, LlamaIndex and other instrumented frameworks.
How much does Portkey cost?
The self-hosted gateway and the managed Developer tier are both free. Production is $49/mo for 100,000 logs, then $9 per additional 100,000 requests. Note the meter - it caps logs, not requests, so past the cap your requests still route, you just stop recording logs beyond the limit. Enterprise is custom for 10M+ logs plus SSO, VPC deployment and compliance reports.
Is Portkey's gateway really sub-1ms latency?
That's Portkey's own claim - under 1ms added latency and a 122kb footprint. A competing benchmark run by Kong reported Portkey at roughly 65% higher latency than Kong's own gateway under equivalent conditions. Both are interested parties, so treat both numbers as vendor claims and benchmark on your own path before you commit to a latency budget.
What is the gotcha with the open-source version?
Observability is gated. The open-source gateway (about 12.5k GitHub stars) is a real, free, self-hostable proxy for routing and guardrails. But the logs, traces, analytics and retention that most people adopt an observability layer for require the managed Production tier or Enterprise. It's open-core where the free part is the pipe and the paid part is the dashboard.
Related Articles
4 Portkey Alternatives When You Actually Wanted Observability (2026)
Portkey's open-source gateway self-hosts free, but the logs, traces and analytics most teams adopt an observability tool for live on the paid tier. Here are the alternatives, matched to whether you want a gateway or real logging.
July 26, 2026
guidePortkey Pricing Explained (2026) - What You Actually Pay
Portkey's meter caps logs, not requests, so your traffic keeps flowing while your observability quietly goes dark past the limit. Here is how the $49/mo Production tier really works, a worked bill, and cheaper picks for real tracing.
July 26, 2026
comparisonPortkey vs Helicone in 2026 - Why This Comparison Already Has a Winner
Portkey and Helicone both sit in front of your LLM calls as a proxy, but one is a thriving gateway and the other went into maintenance mode in March 2026. Here is the honest head-to-head, plus where Langfuse fits.
July 26, 2026
comparisonPortkey vs Langfuse in 2026 - Gateway or Observability Platform?
Portkey is an LLM gateway that routes to 1,600+ models with fallbacks and budgets - observability is a paid add-on. Langfuse is a full open-source observability platform, self-hostable free. These solve different problems. Here is which you need.
July 26, 2026
best-ofThe Best LLM Guardrails Tools in 2026, by Where They Actually Run
Guardrails split into three jobs - block bad output at runtime, red-team the app before you ship, and catch violations in production monitoring. Three tools, one for each job, and why picking the wrong layer leaves a gap.
July 26, 2026
best-ofThe Best LLM Monitoring Tools in 2026, Ranked for Production Cost and Reliability
Five tools for monitoring LLM apps in production, judged on what a live system actually needs - cost and token visibility, self-host reality, and pricing that does not go dark at volume. One winner, one gateway pick, and one to avoid.
July 26, 2026
best-ofThe Best LLM Observability for OpenAI Apps in 2026, by Use Case
If you call the OpenAI API, four tools cover you cleanly - two open-source tracers, one gateway, and one you should not start on. Judged on OpenAI SDK integration, cost at scale, self-host and the acquisition status that just changed the math.
July 26, 2026
best-ofThe Best Prompt Management Tools in 2026, Ranked for Versioning and Team Workflow
Four tools for managing LLM prompts, judged on what a growing team actually needs - versioning, a playground to iterate, and whether prompts connect to your evals. One free open-source winner, and the expensive one worth its price for LangChain teams.
July 26, 2026
best-ofThe Best LLM Observability Tools in 2026, Ranked and Road-Tested
Eight LLM observability platforms judged on the four things that actually decide the bill and the migration - self-host reality, OpenTelemetry support, pricing at scale, and eval depth. One clear winner, one you should not start on.
July 23, 2026