Awesome Infra for AI › LLM Gateways & Proxies

caidaoli/ccLoad

⭐ 418 Go added to this list on 2026-06-14 repository created 2025-09-08

ccLoad is an specialized AI API gateway designed to simplify the operational complexities of managing multiple AI API upstreams like Claude Code, Codex, Gemini, and OpenAI. It acts as a single, stable entry point for clients, abstracting away the intricacies of API key management, rate limits, and provider-specific failures. The gateway features smart routing, which prioritizes channels and uses weighted round-robin for balanced load distribution among upstreams. It ensures high availability through automatic failover mechanisms, which detect and bypass failed keys or channels, and implements exponential cooldowns to prevent hammering unstable services. ccLoad also supports multi-URL scheduling, allowing a single channel to leverage multiple upstream URLs, with selection based on observed latency and health. A key feature is its ability to perform protocol transformations, enabling cross-protocol conversion between Anthropic, OpenAI, Gemini, and Codex APIs, thus allowing a single channel to serve multiple client protocols. For operational visibility, it offers live monitoring with a web dashboard displaying active requests, logs, token usage, time to first byte (TTFB), and costs. It also handles "soft errors" (HTTP 200 responses that contain error content) and provides robust cost control through per-channel daily limits and per-token cost limits. The system includes local token counting for accurate billing estimation, secure authentication, and supports single binary deployment with embedded SQLite.

https://github.com/caidaoli/ccLoad

aiai-gatewayanthropicapi-proxyclaude-apiclaude-codecodexcost-controlfailovergeminigolangllm-proxyload-balancermonitoringopenaiopenai-apiprotocol-transformreverse-proxysmart-routingtoken-counting

Also in LLM Gateways & Proxies

diegosouzapw/OmniRoute

OmniRoute is a free AI gateway that unifies access to over 170 AI providers, offering token compression, auto-fallback, and aggregating free tiers to provide billions of free tokens monthly.

Mintplex-Labs/anything-llm

AnythingLLM is an all-in-one local-first AI application for chatting with documents, managing AI agents, and integrating with various LLMs and vector databases, offering dynamic model routing and m...

BerriAI/litellm

LiteLLM is an open-source AI Gateway and Python SDK providing a unified interface to over 100 LLM providers, with features like cost tracking, guardrails, load balancing, and observability.

QuantumNous/new-api

new-api is a unified LLM gateway and AI asset management system that enables aggregation, distribution, and cross-conversion of various LLMs into OpenAI, Claude, or Gemini compatible formats, offer...

Kong/kong

Kong Gateway is a cloud-native API and AI gateway offering high performance, extensibility via plugins, and advanced AI traffic capabilities including multi-LLM support, semantic security, and cach...

decolua/9router

9Router is an AI router and token saver that connects various AI coding tools to over 40 AI providers, optimizing usage with auto-fallback, quota tracking, and token compression.

apache/apisix

Apache APISIX is a dynamic, real-time, high-performance API Gateway that can also function as an AI Gateway, providing AI proxying, load balancing for LLMs, and robust security for AI agents.

lidge-jun/opencodex

OpenCodex is a universal local proxy that enables OpenAI Codex, Claude Code, and Grok Build to utilize any LLM provider, offering advanced model routing, account pooling, and API key management.