Awesome Infra for AI › LLM Gateways & Proxies

theagentrouter/agent-router

⭐ 2183 Go added to this list on 2026-09-14 repository created 2024-10-21

Agent Router (formerly Envoy AI Gateway) is an open source AI Foundation project that turns Envoy and Envoy Gateway into a control plane for LLM and agent traffic. Application teams get a single OpenAI-compatible API for every model and tool, whether the destination is a hosted provider such as OpenAI, Anthropic, Azure OpenAI, Google Gemini or Bedrock, or a self-hosted inference cluster and MCP servers. Platform teams centralize credentials, routing, quotas, failover and usage attribution in one place, enforced by Envoy's proven proxy data plane rather than a bespoke one. The project introduces a two-tier gateway pattern: a Tier One Gateway acts as the centralized entry point handling authentication, top-level routing and global rate limiting, while Tier Two Gateways provide fine-grained control over self-hosted model access, including endpoint-picker support for inference-serving optimization. It ships a standalone CLI, aigw, that runs a full OpenAI-compatible router locally with one command, as well as Kubernetes-native custom resources (AIGatewayRoute, AIServiceBackend, BackendSecurityPolicy) for production deployments behind Envoy Gateway. The project recently renamed from Envoy AI Gateway to Agent Router under the Agentic AI Foundation, but kept the same code, maintainers, release cadence, Apache 2.0 license, CRDs, CLI name and container image paths, so existing manifests keep working. It targets platform and infrastructure teams that need centralized, policy-enforced routing for both hosted and self-hosted LLM traffic at scale, rather than individual developers wiring up a single provider.

https://github.com/theagentrouter/agent-router

AI gatewayEnvoyKubernetescontrol planeMCPmulti-provider routing

Also in LLM Gateways & Proxies

diegosouzapw/OmniRoute

OmniRoute is a free AI gateway that unifies access to over 170 AI providers, offering token compression, auto-fallback, and aggregating free tiers to provide billions of free tokens monthly.

Mintplex-Labs/anything-llm

AnythingLLM is an all-in-one local-first AI application for chatting with documents, managing AI agents, and integrating with various LLMs and vector databases, offering dynamic model routing and m...

BerriAI/litellm

LiteLLM is an open-source AI Gateway and Python SDK providing a unified interface to over 100 LLM providers, with features like cost tracking, guardrails, load balancing, and observability.

QuantumNous/new-api

new-api is a unified LLM gateway and AI asset management system that enables aggregation, distribution, and cross-conversion of various LLMs into OpenAI, Claude, or Gemini compatible formats, offer...

Kong/kong

Kong Gateway is a cloud-native API and AI gateway offering high performance, extensibility via plugins, and advanced AI traffic capabilities including multi-LLM support, semantic security, and cach...

decolua/9router

9Router is an AI router and token saver that connects various AI coding tools to over 40 AI providers, optimizing usage with auto-fallback, quota tracking, and token compression.

apache/apisix

Apache APISIX is a dynamic, real-time, high-performance API Gateway that can also function as an AI Gateway, providing AI proxying, load balancing for LLMs, and robust security for AI agents.

lidge-jun/opencodex

OpenCodex is a universal local proxy that enables OpenAI Codex, Claude Code, and Grok Build to utilize any LLM provider, offering advanced model routing, account pooling, and API key management.