diegosouzapw/OmniRoute
OmniRoute is a free AI gateway that unifies access to over 170 AI providers, offering token compression, auto-fallback, and aggregating free tiers to provide billions of free tokens monthly.
Awesome Infra for AI › LLM Gateways & Proxies
Weave Router is a drop-in proxy that sits between AI coding tools or applications and multiple LLM providers, selecting the best model for each individual request rather than applying one model to a whole session. It exposes an endpoint compatible with Anthropic Messages, OpenAI Chat Completions and Gemini native APIs, including streaming, tool calls and vision, and can also reach open-source models (DeepSeek, Kimi, GLM, Qwen, Llama, Mistral) through OpenRouter or any OpenAI-compatible endpoint. Routing decisions are made per "action" (a single upstream API request) using a cluster scorer derived from published research on model routing, implemented as a small on-box ONNX embedder rather than a prompt-based heuristic. The project can be used as a hosted service through a one-command npx installer that wires configuration into Claude Code, Codex, opencode or the pi CLI, or self-hosted as a Docker Compose stack consisting of the router, the scorer, a Postgres database for installations, keys and usage, and an optional dashboard UI. Provider API keys are bring-your-own-key by default and stay on the user's machine, encrypted at rest; in self-hosted mode, prompts go directly from the router to the configured provider rather than through Weave's own infrastructure. The router emits OpenTelemetry traces out of the box, viewable in a bundled dashboard or forwarded to external observability backends such as Honeycomb, Datadog or Grafana. It also supports an optional frozen HMM routing policy as a companion sidecar process. The project targets developers and teams using AI coding agents or LLM-backed applications who want automatic, cost-aware model selection and multi-provider access without hardcoding a single model or provider into their tooling.
https://github.com/weave-os/router
OmniRoute is a free AI gateway that unifies access to over 170 AI providers, offering token compression, auto-fallback, and aggregating free tiers to provide billions of free tokens monthly.
AnythingLLM is an all-in-one local-first AI application for chatting with documents, managing AI agents, and integrating with various LLMs and vector databases, offering dynamic model routing and m...
LiteLLM is an open-source AI Gateway and Python SDK providing a unified interface to over 100 LLM providers, with features like cost tracking, guardrails, load balancing, and observability.
new-api is a unified LLM gateway and AI asset management system that enables aggregation, distribution, and cross-conversion of various LLMs into OpenAI, Claude, or Gemini compatible formats, offer...
Kong Gateway is a cloud-native API and AI gateway offering high performance, extensibility via plugins, and advanced AI traffic capabilities including multi-LLM support, semantic security, and cach...
9Router is an AI router and token saver that connects various AI coding tools to over 40 AI providers, optimizing usage with auto-fallback, quota tracking, and token compression.
Apache APISIX is a dynamic, real-time, high-performance API Gateway that can also function as an AI Gateway, providing AI proxying, load balancing for LLMs, and robust security for AI agents.
OpenCodex is a universal local proxy that enables OpenAI Codex, Claude Code, and Grok Build to utilize any LLM provider, offering advanced model routing, account pooling, and API key management.