The best LLM gateways in 2026, scored on what they actually do

For most teams the shortlist is three names: LiteLLM if you self-host, OpenRouter if you want one key and 400 models today, and Portkey if governance is the reason you are shopping. The rest of this page is why, and which of the other six beats them in specific situations.

By the Continuum team. We build a workbench that runs Claude Code, Codex, and their peers, so the model rates quoted here are the ones our own cost analytics ship with.

The short version

Nine LLM gateways scored on routing, failover, caching, observability, spend controls, self-hosting, and price, checked August 2026. LiteLLM is the open-source default at 56,700 stars with the widest provider coverage and the highest operations burden. OpenRouter is the fastest path to many models for a 5.5% credit fee. Portkey leads on governance with an MIT gateway plus a hosted platform from $49/mo. Helicone is observability first at $79/mo Pro. Bifrost, Kong, Vercel, Cloudflare, and Continuum each win a narrower case.

What you need to know
  • LiteLLM wins open-source coverage. 56,700 stars, MIT core, an enterprise-licensed directory inside the same repo.
  • OpenRouter wins time to first token. One key, 5.5% on credit purchases, 5% on BYOK above the plan allowance.
  • Portkey wins governance. MIT gateway, hosted Production tier $49/mo for 100,000 logs.
  • Helicone wins if you are really buying dashboards. Free to 10,000 requests, then $79/mo.
  • Vercel charges zero markup, including on BYOK. That is genuinely rare.
  • No gateway is good at all seven jobs. Pick for the one job you are actually buying.

The comparison table

Scored on what the product does today, from vendor documentation and repository metadata read in August 2026. "Good" means it is a reason to choose this product; "basic" means present and shallow; "no" means absent.

Nine LLM gateways, August 2026.
GatewayRoutingFailoverCachingObservabilitySpend controlsSelf-hostEntry price
LiteLLMGoodGoodGoodGoodGoodYes, MIT core$0
OpenRouterGoodGoodBasicBasicBasicNo5.5% on credits
PortkeyGoodGoodGoodGoodGoodYes, MIT gateway$0 OSS / $49 mo
HeliconeBasicBasicGoodGoodGoodYes$0 / $79 mo
BifrostGoodGoodGoodGoodBasicYes, Apache 2.0$0
Kong AI GatewayGoodGoodGoodGoodGoodYes, Apache 2.0$0 OSS / Konnect paid
Vercel AI GatewayGoodGoodBasicGoodGoodNo$0 markup
Cloudflare AI GatewayBasicBasicGoodGoodBasicNo$0 core
ContinuumGoodBasicNoGoodGoodNo$0 BYOK

The three most people should pick from

LiteLLM: the open-source default

The most-adopted open-source gateway by a wide margin, 56,700 GitHub stars in August 2026, and the only one where "does it support provider X" is almost always yes. The core is MIT-licensed; the enterprise/ directory in the same repository is under a separate commercial licence, which is why GitHub reports the repo's licence as unresolved. That split is honest but it does surprise people who assumed the whole thing was MIT.

What you get free: virtual keys, teams, spend tracking, budgets and rate limits, fallbacks, request logging, Prometheus metrics. What Enterprise adds, at a price quoted per gateway capacity rather than per token and never published: SSO and SCIM, OIDC and JWT auth, audit logs, secret-manager integration and key rotation, multi-region control plane, and 24/7 SLAs.

OpenRouter: the fastest path to many models

One account, one key, hundreds of models, and no infrastructure. If a provider errors, OpenRouter falls back to the next provider serving the same model automatically. Routing shortcuts are pleasant: append :nitro to a model slug to sort providers by throughput, :floor to sort by price, and openrouter/auto hands model selection to their classifier.

The cost is a fee and a middleman. Credit purchases carry 5.5% (minimum $0.80) by card and 5% by cryptocurrency, checked against the OpenRouter FAQ in August 2026. Bring-your-own-key usage is fee-free up to a monthly allowance of list-price inference (documented as $25,000/month on pay-as-you-go) and 5% above it. In exchange, every prompt you send crosses a third party that is not the model vendor.

Portkey: the governance pick

Portkey's gateway is MIT-licensed and self-hostable with no paywalled core routing, and the hosted platform adds guardrails, prompt management, and access control. Pricing read from the Portkey pricing page in August 2026: Developer is free with 10,000 recorded logs a month and three prompt templates, Production is $49/mo for 100,000 logs with overage at $9 per additional 100,000 up to three million, Enterprise is custom and quoted at 10 million plus logs with SSO, VPC hosting, and SOC 2 Type 2.

The reason to choose it over LiteLLM is that guardrails and role-based access are first-class rather than bolted on, and the reason not to is that the free tier's 10,000 logs a month is a demo allowance, so the real comparison is $49/mo against your own operations time.

The six that win narrower cases

  • Helicone is an observability product with gateway features, not the other way round. Hobby is free with 10,000 requests a month, 1 GB of storage and 7-day retention; Pro is $79/mo with unlimited seats and one month of retention; Team is $799/mo with five organisations, SOC 2 and HIPAA. Ingestion is rate-limited by tier, from 10 logs per minute on Hobby to 15,000 on Team. Choose it when the thing you cannot see is why a prompt regressed, not when you need routing.
  • Bifrost (Apache 2.0, 7,400 stars) is the performance-first entrant, written in Go and marketed on gateway overhead. The claims are the vendor's own and worth benchmarking rather than believing, but the licence is clean and the code is readable. Smaller community than LiteLLM, so expect to read source.
  • Kong AI Gateway is a set of LLM plugins on Kong Gateway (Apache 2.0, 44,000 stars). Right answer if Kong already fronts your APIs and your security team already knows it. Konnect, the managed control plane, is the paid path and starts around $105/month per gateway with a million requests included, which secondary sources agree on and Kong does not publish in a form you can quote confidently.
  • Vercel AI Gateway charges no markup and no platform fee on tokens, including with your own keys, which is the most aggressive pricing position in the category. The add-ons are metered separately: custom reporting at $0.075 per 1,000 tag writes and $5 per 1,000 report queries, team-wide provider allowlist and team-wide zero data retention at $0.10 per 1,000 requests each. Compelling on Vercel, pointless off it.
  • Cloudflare AI Gateway gives you dashboard analytics, caching, and rate limiting free on any Cloudflare account, with persistent logs capped at 100,000 total on Workers Free and 10 million per gateway on Workers Paid. Unified Billing, where Cloudflare pays the providers for you, carries a 5% fee on purchased credits. Thin on routing, excellent as a free front door.
  • Continuum is the odd one out and belongs here only because people arrive at this page from a coding-agent problem rather than an application one. Its gateway is fused to a workbench that runs Claude Code, Codex, Cursor, Grok and OpenCode, with an auto model router, per-repo cost analytics, and organisation-level model policy and weekly spend caps. BYOK is free forever; hosted inference is $25/mo and up. It is not a general-purpose proxy you would put in front of a production API, and if that is what you need, one of the eight above is the answer.

How to choose in ten minutes

01

Name the one job

Write down the single sentence that made you open this page. "A provider outage took us down." "I cannot attribute spend." "We need an audit log." Gateways are bought for one reason and evaluated on twenty; the twenty are how you end up with the wrong one.

02

Decide hosted or self-hosted before you compare features

This eliminates half the list immediately and it is a constraint, not a preference. If your prompts cannot leave your perimeter, the answer is LiteLLM, Portkey OSS, Bifrost, or Kong, and price comparisons with OpenRouter are irrelevant.

03

Price it against the fee, not the sticker

A $0 self-hosted gateway that costs a quarter of an engineer is more expensive than $49/mo. A 5.5% fee on $200/month of tokens is $11; the same fee on $80,000/month is $4,400 and pays for a lot of Postgres. Run the number for your volume, not for the vendor's example.

04

Test failover deliberately before you trust it

Point a fallback at a deliberately broken key and send real traffic. Most failover configurations are wrong the first time, and you find out during the outage they were meant to survive.

Questions people ask

What is the best LLM gateway?

LiteLLM if you are self-hosting and want the widest provider coverage, OpenRouter if you want many models today with no infrastructure, Portkey if governance and guardrails are why you are shopping. There is no single best because the seven jobs a gateway does are bought by different people for different reasons.

What is the best open source AI gateway?

LiteLLM, on adoption: 56,700 GitHub stars in August 2026, MIT core, the largest provider list. Portkey's gateway (MIT) and Bifrost (Apache 2.0) are the strongest alternatives, and Kong AI Gateway is the right answer if Kong already fronts your APIs.

Is OpenRouter worth the 5.5% fee?

At small volume, easily: 5.5% of $200 a month is $11, and you get one key, hundreds of models, and automatic provider fallback for it. At $50,000 a month it is $2,750 and a self-hosted gateway starts to look cheap. The other half of the answer is not financial: every prompt crosses a third party.

Is Portkey or LiteLLM better?

LiteLLM has broader provider coverage and a larger community; Portkey has better governance, guardrails, and a managed option that does not require you to run Postgres. If you have a platform team, LiteLLM. If you have a compliance requirement and no platform team, Portkey.

Can I use an LLM gateway with Claude Code?

Yes. Claude Code reads ANTHROPIC_BASE_URL, so any gateway that serves an Anthropic-compatible endpoint at that origin works. Gateways serving only an OpenAI-shaped API need a translation layer, which LiteLLM provides. See the OpenAI-compatible API guide for the exact base-URL rules.

Sources

Every figure above was read from these pages on August 2026. Vendors reprice without notice; if you find a stale number, tell us.

  1. OpenRouter FAQ (fees, BYOK allowance)
  2. LiteLLM pricing
  3. Portkey pricing
  4. Helicone pricing
  5. Vercel AI Gateway pricing
  6. Cloudflare AI Gateway pricing
  7. BerriAI/litellm on GitHub 56,700 stars, read August 2026
  8. Portkey-AI/gateway on GitHub MIT, 12,766 stars
Try it

Measure first,
then buy infrastructure.

Continuum turns the logs your coding agents already write into per-repo cost, with nothing in the request path. Free on every platform.

free app · your subscriptions · local-first