Awesome Infra for AI › LLM Gateways & Proxies

ai-forever/gpt2giga

⭐ 134 Python repository created 2024-12-25

gpt2giga is a FastAPI-based proxy designed to bridge the compatibility gaps between OpenAI/Anthropic APIs and the GigaChat API. Its primary function is to allow applications, SDKs, or agent frameworks built to interact with OpenAI or Anthropic to seamlessly use GigaChat as the backend. The project handles practical incompatibilities such as request format translation, streaming events (SSE), tool schemas, model discovery, authorization, and optional client parameters. It translates OpenAI Chat Completions, Responses, Embeddings, and Anthropic Messages into GigaChat calls. Key features include mapping tools/function calling, structured output, images, and reasoning flags for GigaChat, while safely ignoring optional fields not supported by GigaChat to prevent errors. It also separates client API key authorization from GigaChat credentials and provides a GigaChat model list in an OpenAI-, Anthropic-, or LiteLLM-compatible format. The proxy supports OpenAI-compatible endpoints like `/v1/chat/completions` and `/v1/embeddings`, Anthropic-compatible endpoints such as `/messages`, and LiteLLM-compatible `/model/info`. It explicitly states current limitations, such as not natively supporting OpenAI Files API, Batches API, or full parity for complex features like audio, image generation, fine-tuning, or assistants.

https://github.com/ai-forever/gpt2giga

LLM proxyAPI gatewayOpenAI compatibilityAnthropic compatibilityGigaChatFastAPIinferencemodel servingAPI translation

Also in LLM Gateways & Proxies

diegosouzapw/OmniRoute

OmniRoute is a free AI gateway that unifies access to over 170 AI providers, offering token compression, auto-fallback, and aggregating free tiers to provide billions of free tokens monthly.

Mintplex-Labs/anything-llm

AnythingLLM is an all-in-one local-first AI application for chatting with documents, managing AI agents, and integrating with various LLMs and vector databases, offering dynamic model routing and m...

BerriAI/litellm

LiteLLM is an open-source AI Gateway and Python SDK providing a unified interface to over 100 LLM providers, with features like cost tracking, guardrails, load balancing, and observability.

QuantumNous/new-api

new-api is a unified LLM gateway and AI asset management system that enables aggregation, distribution, and cross-conversion of various LLMs into OpenAI, Claude, or Gemini compatible formats, offer...

Kong/kong

Kong Gateway is a cloud-native API and AI gateway offering high performance, extensibility via plugins, and advanced AI traffic capabilities including multi-LLM support, semantic security, and cach...

decolua/9router

9Router is an AI router and token saver that connects various AI coding tools to over 40 AI providers, optimizing usage with auto-fallback, quota tracking, and token compression.

apache/apisix

Apache APISIX is a dynamic, real-time, high-performance API Gateway that can also function as an AI Gateway, providing AI proxying, load balancing for LLMs, and robust security for AI agents.

lidge-jun/opencodex

OpenCodex is a universal local proxy that enables OpenAI Codex, Claude Code, and Grok Build to utilize any LLM provider, offering advanced model routing, account pooling, and API key management.