AI Gateways & LLM Proxy Infrastructure

Parent: LLM Models and APIs · researched 2026-06-03T22:51:09.481Z· 14 sources · 12 concepts · skill llm-ai-gateways

An AI gateway (LLM gateway / LLM proxy) is the production traffic-and-control

AI Gateways & LLM Proxy Infrastructure

When to use / Skip

The AI-gateway control plane

Caching at the gateway (and the boundary)

Vendor landscape

LiteLLM Proxy (LLM Gateway) — OSS, self-hosted-first

Portkey — OSS gateway + managed/enterprise

Cloudflare AI Gateway — managed, edge, free core

Kong AI Gateway — plugins on Kong Gateway

Helicone — OSS (Rust), observability-first that also proxies

TrueFoundry — enterprise gateway, low-latency

OpenRouter — marketplace/aggregator used as a gateway

Vercel AI Gateway — managed, GA

AWS Bedrock gateway — a *pattern*, not a product

Databricks AI Gateway — evolving rebrand (CONTESTED naming/GA)

Selection / decision guidance

Integration patterns

Anti-patterns & failure modes

2025-2026 frontier

Sources

Children

Frontier under this node: Budget/spend controls & cost attribution/chargeback, Fallback chains, retries & load balancing, Gateway caching (exact-match vs semantic), Governance, RBAC & audit, Guardrail & PII/DLP enforcement at the proxy, MCP/agent/tool-call gateway governance, Observability, logging & tracing hooks, Rate limiting & quotas, Self-hosted vs managed vs edge vs platform-native, Streaming (SSE) pass-through, Unified OpenAI-compatible API surface, Virtual keys & key-vault management (BYOK)

← the whole tree · 3D view· how to read this page