diegosouzapw/OmniRoute
OmniRoute is a free AI gateway that unifies access to over 170 AI providers, offering token compression, auto-fallback, and aggregating free tiers to provide billions of free tokens monthly.
Awesome Infra for AI › LLM Gateways & Proxies
Nexus LLM Router is designed for AI infrastructure engineers managing multi-model production pipelines, focusing on optimizing quality, latency, and cost. It acts as middleware, abstracting model choice, fallback mechanisms, budgeting, auditing, and routing logic from the application layer. The router classifies prompt complexity, directing simple tasks to cost-effective models and reserving premium models for intricate queries. It incorporates cost-aware routing, budget guardrails, and Prometheus metrics for spend optimization. Nexus addresses quality variations across different prompt domains by applying deterministic policy rules. It ensures application resilience with per-provider circuit breakers and automatic fallback chains, while also tackling latency spikes through latency-aware routing. The router supports A/B testing of models without code changes and provides durable audit records for compliance. Security features include token-bucket rate limiting, session/tenant budget enforcement, and PII redaction. It offers OpenAI API compatibility, normalizing interactions with various LLM providers like OpenAI, Anthropic, Gemini, and Moonshot, and makes routing policy explicit, testable, and observable. The system features a robust router engine with configurable strategies, an async-first adapter pipeline, comprehensive type safety (mypy compliant), and production readiness with Docker and CI/CD.
https://github.com/Francis1998/nexus-llm-router
OmniRoute is a free AI gateway that unifies access to over 170 AI providers, offering token compression, auto-fallback, and aggregating free tiers to provide billions of free tokens monthly.
AnythingLLM is an all-in-one local-first AI application for chatting with documents, managing AI agents, and integrating with various LLMs and vector databases, offering dynamic model routing and m...
LiteLLM is an open-source AI Gateway and Python SDK providing a unified interface to over 100 LLM providers, with features like cost tracking, guardrails, load balancing, and observability.
new-api is a unified LLM gateway and AI asset management system that enables aggregation, distribution, and cross-conversion of various LLMs into OpenAI, Claude, or Gemini compatible formats, offer...
Kong Gateway is a cloud-native API and AI gateway offering high performance, extensibility via plugins, and advanced AI traffic capabilities including multi-LLM support, semantic security, and cach...
9Router is an AI router and token saver that connects various AI coding tools to over 40 AI providers, optimizing usage with auto-fallback, quota tracking, and token compression.
Apache APISIX is a dynamic, real-time, high-performance API Gateway that can also function as an AI Gateway, providing AI proxying, load balancing for LLMs, and robust security for AI agents.
OpenCodex is a universal local proxy that enables OpenAI Codex, Claude Code, and Grok Build to utilize any LLM provider, offering advanced model routing, account pooling, and API key management.