Awesome Infra for AI › LLM Gateways & Proxies

khankamraan2006-crypto/fabric-router-core

⭐ 116 HTML added to this list on 2026-07-06 repository created 2026-06-28

CloudForge AI functions as a multi-node, horizontally scalable Large Language Model (LLM) gateway designed for production use cases requiring high availability, intelligent failover, and granular cost management. It abstracts away the complexities of integrating with diverse LLM providers by offering a unified, policy-driven interface. The platform decouples application logic from specific inference providers, treating LLM capabilities as a dynamic resource pool. Key features include a provider-agnostic routing engine that optimizes requests based on latency, cost, and rate limits; secure OAuth integration; a dynamic provider registry for hot-reloading configurations; and robust cost governance with budget caps. It also offers comprehensive request telemetry for observability and multilingual response normalization to ensure consistent output across different models. CloudForge AI is built with an operational support framework, including health checks and automated failover, to maintain service continuity. Its architecture emphasizes "request normalization before routing," converting client requests into a canonical internal format before dispatching them to the most optimal backend via dedicated provider adapters. This mediation layer ensures adaptability to changing model costs, performance, and new releases, making it ideal for organizations pursuing a multi-cloud AI strategy.

https://github.com/khankamraan2006-crypto/fabric-router-core

LLM gatewayAI proxyinference orchestrationmulti-provider LLMintelligent routingcost governanceAI securityLLM observabilityproduction AIfailoverAPI management for LLMs

Also in LLM Gateways & Proxies

diegosouzapw/OmniRoute

OmniRoute is a free AI gateway that unifies access to over 170 AI providers, offering token compression, auto-fallback, and aggregating free tiers to provide billions of free tokens monthly.

Mintplex-Labs/anything-llm

AnythingLLM is an all-in-one local-first AI application for chatting with documents, managing AI agents, and integrating with various LLMs and vector databases, offering dynamic model routing and m...

BerriAI/litellm

LiteLLM is an open-source AI Gateway and Python SDK providing a unified interface to over 100 LLM providers, with features like cost tracking, guardrails, load balancing, and observability.

QuantumNous/new-api

new-api is a unified LLM gateway and AI asset management system that enables aggregation, distribution, and cross-conversion of various LLMs into OpenAI, Claude, or Gemini compatible formats, offer...

Kong/kong

Kong Gateway is a cloud-native API and AI gateway offering high performance, extensibility via plugins, and advanced AI traffic capabilities including multi-LLM support, semantic security, and cach...

decolua/9router

9Router is an AI router and token saver that connects various AI coding tools to over 40 AI providers, optimizing usage with auto-fallback, quota tracking, and token compression.

apache/apisix

Apache APISIX is a dynamic, real-time, high-performance API Gateway that can also function as an AI Gateway, providing AI proxying, load balancing for LLMs, and robust security for AI agents.

lidge-jun/opencodex

OpenCodex is a universal local proxy that enables OpenAI Codex, Claude Code, and Grok Build to utilize any LLM provider, offering advanced model routing, account pooling, and API key management.