TrueFoundry Review (2026)
An enterprise AI gateway fronting 250+ models with on-prem deployment and a free developer tier. Worth knowing that it also publishes the competitor pricing guides that rank highly for its rivals' names.
Rating
Starting Price
$499/mo
Free Plan
Yes
SDKs & Frameworks
3
Deployment
4
Best For
Enterprises that need an AI gateway deployable on-premise or in their own cloud with governance features, and that want a managed product rather than operating open-source infrastructure.
Last Updated:
10 Things You Should Know About TrueFoundry
- 1 The Developer plan is permanently free with 50,000 requests per month and 3 users
- 2 Pro is $499 per month and Pro Plus is $2,999 per month, with custom enterprise pricing above
- 3 Enterprise governance, native MCP support, semantic caching and observability are available from the Pro tier
- 4 It is not an open-source gateway - the data plane is proprietary
- 5 Unifies access to 250+ models through a single OpenAI-compatible API
- 6 Deployment options include SaaS, customer cloud and on-premise
- 7 Has raised over $21 million from Intel Capital, Eniac Ventures and Peak XV Partners
- 8 Customers include Cargill, Mavenir, Whatfix, Aviso and Aviva
- 9 Named in Gartner's 2026 Hype Cycle for Platform Engineering across three categories
- 10 Acquired Seldon AI to deepen agentic and model-serving capabilities
Pros & Cons
Pros
- ✓ Genuinely deployable on-premise and in your own cloud, which most managed gateways cannot do
- ✓ Native MCP support is forward-looking and rare in gateways
- ✓ Semantic caching is more sophisticated than exact-match caching and can cut real cost on repetitive workloads
- ✓ A permanently free developer tier at 50,000 requests a month is a proper trial rather than a demo
- ✓ Substantial company - over $21M raised, named in Gartner's 2026 Hype Cycle for Platform Engineering, and it acquired Seldon AI
- ✓ Credible enterprise customers including Cargill, Aviva and Whatfix
Cons
- ✕ Not open source, so unlike LiteLLM you cannot inspect or fork the data plane
- ✕ The jump from free to $499/mo is steep with nothing in between
- ✕ Requires Helm operations for self-managed deployment, which is real infrastructure work
- ✕ Most available material about it is its own marketing, including comparison content about competitors
- ✕ Performance figures are vendor-claimed and not independently verified
Features
The middle position, which is the point
Most AI gateways force a choice.
Managed SaaS - OpenRouter, Cloudflare AI Gateway - means no infrastructure but every prompt passes through a third party. Frequently disqualifying for regulated workloads.
Open source - LiteLLM - runs anywhere but you own the proxy, the database, the monitoring and the on-call.
TrueFoundry deploys as SaaS, into your own cloud, or on-premise, as a commercial product with support.
For a bank or insurer that cannot send prompts outside its network but also cannot staff a team to run open-source middleware, that middle position is the whole value proposition. Few competitors occupy it - Portkey’s open-source gateway self-hosts but gates observability, and Fiddler’s in-VPC deployment starts at Enterprise.
Two capabilities ahead of the field
Semantic caching - caching on meaning rather than exact string match, so a question phrased differently still hits the cache. Most gateways offer exact-match caching only.
Worth measuring before you assume it pays for itself. On repetitive workloads - customer support, internal knowledge assistants, where users ask the same thing a hundred ways - it cuts real cost. On genuinely diverse traffic it saves very little. Check your near-duplicate query rate first.
Customer support is the clearest case, because query repetition there is extreme - a large share of tickets are the same handful of questions in different words. If support is the workload you are building for, the platform choice matters as much as the gateway underneath it, and our sister site AI Customer Service covers that layer specifically.
Native MCP support - the Model Context Protocol is how agents call external tools, and it is becoming the standard integration surface for agentic systems.
A gateway with native MCP support can sit in front of tool calls, not just model calls. That matters because tool calls are where a growing share of agentic risk and cost now lives. This puts TrueFoundry in a small group treating MCP as infrastructure - Prompt Security built an MCP Gateway for the security angle before SentinelOne acquired it. Most gateways still treat the model call as the only thing worth proxying.
Pricing, and a steep step
| Tier | Price | Included |
|---|---|---|
| Developer | $0 | 50,000 requests/mo, 3 users |
| Pro | $499/mo | Governance, MCP, semantic caching, observability |
| Pro Plus | $2,999/mo | Higher limits |
| Enterprise | Custom | SaaS, customer cloud, on-prem |
The free Developer tier at 50,000 requests is a proper trial rather than a demo. But there is nothing between free and $499, and everything interesting - governance, MCP, semantic caching, observability - sits at Pro.
Against Cloudflare AI Gateway at $5/month for a million logs, or LiteLLM self-hosted at infrastructure cost, this is unambiguously positioned at enterprises rather than small teams. That is a legitimate choice, not a criticism, but it should tell you quickly whether you are the customer.
It is not open source - unlike LiteLLM you cannot inspect or fork the data plane - and self-managed deployment requires Helm operations, which is real infrastructure work even though the product is commercial.
Vendor-claimed performance is ~3–4ms latency and 350+ RPS on 1 vCPU. Not independently verified.
A note on their content, not their product
Worth flagging for you as a reader rather than as a buyer.
TrueFoundry publishes detailed pricing and comparison guides for LiteLLM, Portkey, LangChain, LangGraph and others, and those articles rank well for competitor searches.
They are generally reasonable, and we have cited some of the underlying figures ourselves while researching this category. But they are competitor marketing, and the practical consequence is that if you have been researching AI gateway pricing recently, you have probably read several without registering whose site you were on.
This is the same caution we applied to Agenta’s comparison content - where a licence claim about Latitude turned out to be wrong - and to PromptLayer’s assessment of a rival’s continuity risk. It is a reading instruction, not an accusation. It also partly explains why we could not resolve LiteLLM’s enterprise pricing: one of the conflicting figures came from a competitor’s guide.
None of this reflects on the product, which is good.
Company standing
Over $21 million raised from Intel Capital, Eniac Ventures and Peak XV Partners. Customers including Cargill, Mavenir, Whatfix, Aviso and Aviva. Named in Gartner’s 2026 Hype Cycle for Platform Engineering across three categories. And it acquired Seldon AI to deepen agentic and model-serving capabilities.
That last point carries weight. In a market where this site has documented relentless consolidation - Lakera, CalypsoAI, Prompt Security, Protect AI, Robust Intelligence, promptfoo, Traceloop all acquired - being an acquirer rather than a target is a meaningful signal about which side of the consolidation you are on.
Enterprise customers like Cargill and Aviva also imply procurement diligence well beyond anything a review can perform.
Should you use it?
Use TrueFoundry if you need a gateway deployable on-premise or in your own cloud with enterprise governance, and you want a supported product rather than operating open-source infrastructure.
Don’t use it if you are a small team - the step from free to $499 will decide it - or you want an inspectable, forkable data plane.
Bottom line: the strongest option for regulated enterprises that need a gateway inside their own perimeter without owning the code. MCP support and semantic caching are genuinely ahead. Measure your cache hit potential before paying for Pro, and read their competitor comparisons knowing who wrote them.
Pricing tiers, deployment options, funding and customers verified against vendor sources and third-party directories on 3 August 2026. Most available material originates from TrueFoundry’s own marketing, and performance figures are vendor-claimed and untested. This is a researched directory entry - we have not yet instrumented this gateway with our reference application.
Pricing Plans
Developer
$0
- Permanently free
- 50,000 requests per month
- 3 users
Pro
$499/mo
- Enterprise governance features
- Native MCP support
- Semantic caching
- Observability
Pro Plus
$2,999/mo
- Higher limits and expanded capabilities
Enterprise
Custom
- SaaS, customer-cloud or on-prem deployment
- Contact sales
SDKs & Frameworks
Deployment
Eval Methods
Billing Unit
Our Verdict
TrueFoundry is a credible enterprise AI gateway with one capability that genuinely distinguishes it - it deploys on-premise and in your own cloud, not just as SaaS. For regulated organisations that cannot route prompts through a third party but also do not want to own an open-source proxy outright, that middle position is valuable and few competitors occupy it. Native MCP support and semantic caching are both ahead of most of the field, and the free Developer tier at 50,000 requests a month is a real trial. The company looks solid, having raised over $21 million from Intel Capital, Eniac Ventures and Peak XV, appeared in Gartner's 2026 Hype Cycle for Platform Engineering, and acquired Seldon AI. One thing worth knowing as a reader rather than a buyer. TrueFoundry publishes detailed pricing guides for LiteLLM, Portkey, LangChain and others, and those articles rank well for competitor searches. They are reasonable content, but they are competitor marketing, and if you have been researching gateway pricing you have probably read several without noticing whose site you were on.
Similar Tools
Vercel AI Gateway
Teams already deploying on Vercel who want provider-agnostic model access at list price with no markup, and who do not need governance features that carry a per-request charge.
Kong AI Gateway
Organisations already standardised on Kong for API management, where bringing LLM traffic under existing governance matters more than LLM-specific features.
Martian
Nobody choosing a gateway today. The RouterBench work remains worth reading if you are evaluating routing approaches generally.
LiteLLM
Teams with engineering capacity that want provider-agnostic routing with no per-token markup, particularly at volumes where a percentage fee on a managed gateway becomes significant.
Frequently Asked Questions
Why does on-premise deployment matter here?
Because it occupies a gap most gateways leave open. Managed options like OpenRouter and Cloudflare AI Gateway are SaaS only, so every prompt passes through a third party - frequently disqualifying for regulated workloads. Open-source options like LiteLLM run anywhere but require you to own the proxy, the database, the monitoring and the on-call. TrueFoundry sits between the two, offering a commercial product with support that deploys into your own cloud or data centre. For a bank or insurer that cannot send prompts outside its network but also cannot staff an infrastructure team to run open-source middleware, that middle position is the entire value proposition.
What is semantic caching and is it worth paying for?
Caching based on meaning rather than exact string match, so a question phrased slightly differently can still hit the cache. It is more sophisticated than the exact-match caching most gateways offer and can cut real cost on repetitive workloads - customer support and internal knowledge assistants being the obvious cases, where users ask the same thing in a hundred different ways. Whether it justifies the Pro tier depends entirely on how repetitive your traffic is. Measure your near-duplicate query rate before assuming it will pay for itself, because on genuinely diverse workloads semantic caching saves very little.
What is native MCP support and why does it matter?
The Model Context Protocol is how agents call external tools, and it is becoming the standard integration surface for agentic systems. A gateway with native MCP support can sit in front of those tool calls rather than only in front of model calls, which matters because tool calls are where a lot of agentic risk and cost now lives. This puts TrueFoundry in a small group thinking about MCP as infrastructure - Prompt Security built an MCP Gateway for the security angle before SentinelOne acquired it. Most gateways still treat the model call as the only thing worth proxying.
Should I be cautious about their comparison content?
Read it knowing whose it is. TrueFoundry publishes detailed pricing and comparison guides covering LiteLLM, Portkey, LangChain, LangGraph and others, and those articles rank well for competitor searches. They are generally reasonable, and we have cited some of the underlying figures ourselves - but they are competitor marketing. This is the same caution we applied to Agenta's comparison content, where a licence claim about Latitude turned out to be wrong, and to PromptLayer's assessment of a rival's continuity risk. It is not an accusation, it is a reading instruction. If you have been researching AI gateway pricing recently, you have probably read several of these without registering whose site you were on.
What does the pricing jump look like?
Steep. The Developer plan is free with 50,000 requests a month and 3 users, then Pro is $499 a month, then Pro Plus is $2,999. There is nothing between free and $499, so a team outgrowing the developer tier faces a substantial step. Enterprise governance, MCP support, semantic caching and observability all sit at Pro, which means the free tier is genuinely a trial rather than something you can operate on. Compare Cloudflare AI Gateway at $5 a month for a million logs, or LiteLLM self-hosted at infrastructure cost, and TrueFoundry is clearly positioned at enterprises rather than at small teams.
Is the company stable?
It looks solid. TrueFoundry has raised over $21 million from Intel Capital, Eniac Ventures and Peak XV Partners, counts Cargill, Mavenir, Whatfix, Aviso and Aviva among its customers, was named in Gartner's 2026 Hype Cycle for Platform Engineering across three categories, and acquired Seldon AI to deepen its agentic and model-serving capabilities. Being an acquirer rather than a target is a meaningful signal in a market where this site has documented a great deal of consolidation. Enterprise customers like Cargill and Aviva also imply procurement diligence well beyond what a review can perform.