Open-source proxies that aggregate and manage requests across OpenAI, Anthropic, and local language models.
TensorZero is an inference gateway and experimentation framework designed to manage the lifecycle of large language models in production environments. It functions as a central proxy that routes requests across multiple artificial intelligence providers while providing the infrastructure necessary to monitor performance, track costs, and ensure service reliability. The platform distinguishes itself by integrating a comprehensive evaluation engine and an observability pipeline directly into the request flow. It enables developers to conduct controlled experiments and A/B tests to compare diffe
TensorZero is an inference gateway that routes requests across multiple AI providers (OpenAI, Anthropic, and likely local models via its extensible proxy) while providing centralized observability, cost tracking, and experimentation, making it a comprehensive fit for a unified LLM API gateway with multi-provider management.
Antigravity-Manager is an artificial intelligence model orchestration platform that functions as a unified gateway for interacting with multiple external service providers. It standardizes heterogeneous vendor data structures into a consistent internal schema, allowing third-party tools to interface with various models through a single, normalized API. The system distinguishes itself through automated infrastructure management, including the lifecycle tracking of service accounts and the secure rotation of authentication credentials. By acting as a middleware layer, it intercepts traffic to p
Antigravity-Manager is an AI model orchestration platform that provides a unified gateway for interacting with multiple external LLM providers, which aligns with your need for a centralized proxy, though explicit support for self-hosted local models and OpenAI-compatible endpoints is not clearly stated in the available description.
gpt4free-ts is a TypeScript-based LLM API proxy and gateway that provides a unified interface for accessing large language models without paid subscriptions or official API keys. It functions as a containerized AI bridge that routes requests to various free third-party providers to retrieve chat completions. The project acts as an OpenAI API wrapper, translating requests and responses into the standard OpenAI chat completions format to ensure compatibility with existing AI tools. It utilizes a provider-based routing system to distribute request loads across available endpoints. The gateway s
xiangsx/gpt4free-ts is an LLM API proxy that routes requests across multiple free third-party providers via an OpenAI-compatible endpoint and includes load balancing, but it does not support local self-hosted models, Anthropic, or built-in caching/key management, so it's the right kind of gateway but narrower than the full set of providers and features you need.
CoAI is an enterprise-grade, self-hostable AI gateway platform that unifies access to over 200 AI models from more than 35 providers through a single OpenAI-compatible API endpoint. It functions as a multi-tenant gateway, routing requests across providers with load balancing, automatic failover, and priority-based routing, while exposing standard OpenAI API endpoints for chat, image generation, model listing, and billing to enable seamless integration with existing tools and clients. The platform distinguishes itself through a comprehensive set of operational capabilities built around the gat
CoAI is a self-hostable AI gateway that unifies access to over 200 models from 35+ providers through a single OpenAI-compatible API, with load balancing, failover, and monitoring—covering most of your requirements, though explicit support for local self-hosted models is not confirmed in the description (likely possible via custom model integration).
Higress is an AI API gateway and cloud-native traffic manager that functions as a Kubernetes ingress controller. It provides a centralized system for routing, securing, and optimizing traffic directed toward large language models, AI agents, and microservice architectures. The project distinguishes itself through deep AI orchestration, including the ability to host and manage Model Context Protocol servers that transform REST APIs into tools for AI agents. It features specialized AI infrastructure for model request proxying, protocol translation across multiple providers, and semantic-based c
Higress is an AI-native API gateway that centrally routes, secures, and optimizes traffic to LLMs from multiple providers and self-hosted models, exactly fitting the need for a unified interface with features like protocol translation, caching, and observability.
LiteLLM is a unified gateway and proxy server designed to centralize access to over one hundred language model providers. It provides a standardized API interface that abstracts vendor-specific schemas, allowing developers to interact with diverse models through a single, consistent format. By acting as a central traffic management layer, it enables organizations to route, secure, and govern model interactions across multiple deployments. The platform distinguishes itself through its policy-driven architecture, which uses configuration-based routing to manage traffic distribution, load balanc
LiteLLM is a unified gateway that centralizes access to over 100 LLM providers including OpenAI and Anthropic, with an OpenAI-compatible API, request routing, load balancing, API key management, and observability—exactly the single interface for multi-provider and local model management you're looking for.
This project is an artificial intelligence gateway that functions as a centralized middleware layer for managing, securing, and observing interactions with language, vision, and audio models. It provides a unified interface that standardizes requests across multiple providers, enabling teams to integrate AI capabilities into their applications through a consistent set of tools and protocols. The gateway distinguishes itself through its comprehensive infrastructure governance and traffic management capabilities. It allows for policy-driven routing, automated failover, and load balancing across
Portkey's Gateway is an LLM API gateway that provides a unified interface to multiple providers including OpenAI and Anthropic, with policy-driven routing, failover, load balancing, and observability, making it a strong fit for centralized management and switching; local self‑hosted models are likely supported through provider‑agnostic endpoints, though not explicitly highlighted in the evidence.
9router is an AI model gateway designed to route requests from AI coding tools to multiple model providers through a single unified API. It provides administration for self-hosted AI proxy deployments, allowing users to manage API keys and model access on local servers or edge networks. The system differentiates itself through multi-provider API normalization, which translates incompatible request and response formats to ensure compatibility across different AI models. It features AI provider failover management to automatically switch between providers or accounts when quotas are exhausted o
9router is an AI model gateway that unifies providers like OpenAI and Anthropic through a single API with routing, failover, and API key management, fitting your need for a centralized LLM gateway, though local model support and caching are not clearly confirmed.
Archgw is a gateway proxy and data plane designed for agentic applications, providing a centralized layer for routing, safety, and orchestration between application logic and multiple large language model providers. It functions as an AI agent orchestrator that automates the execution of agent workflows to remove repetitive plumbing from the core codebase. The system features a provider-agnostic interface layer that standardizes disparate model APIs into a single format and a transparent proxy data plane to intercept traffic. It employs rule-based routing to decouple application logic from sp
Archgw is an LLM API gateway that provides a provider-agnostic interface and rule-based routing to multiple large language model providers, including support for observability and safety, which fits your need for centralized management and switching between services like OpenAI and potentially local models.
Helicone is an AI gateway and observability platform designed to intercept, manage, and monitor interactions with large language models. By acting as a reverse-proxy, it provides a centralized layer for routing requests across multiple AI providers, allowing developers to maintain consistent application logic while gaining deep visibility into model performance, usage, and costs. The platform distinguishes itself through a robust suite of traffic management and prompt engineering tools. It enables policy-driven control, including automatic failover between providers, rate limiting, and edge-b
Helicone is an AI gateway that provides centralized routing, failover, and observability across multiple LLM providers, fitting the gateway category well, though it lacks explicit support for local self-hosted model integration which the search also requires.
Nanobot is an orchestration framework designed for building, deploying, and managing autonomous AI agents. It provides a secure runtime environment that supports persistent memory, multi-step workflow management, and tool integration, allowing agents to maintain context and state across long-running tasks. The platform distinguishes itself through a unified model gateway that normalizes requests across diverse local and remote language models, alongside a multi-channel integration layer that connects agents to various messaging platforms. It enforces security through containerized sandboxing
Nanobot includes a unified model gateway that normalizes requests across diverse local and remote language models (including OpenAI-compatible APIs), fitting the core requirement for an LLM API gateway, though it is built as part of a broader agent orchestration framework rather than a standalone proxy.
This project provides a unified interface for interacting with a wide range of artificial intelligence services, acting as a central orchestration layer for text and image generation. It standardizes access to diverse AI backends, allowing developers to integrate multiple language and vision models through a single, consistent programming interface. By abstracting provider-specific protocols and authentication requirements, the tool simplifies the development of applications that rely on external AI services. The platform distinguishes itself through a resilient request routing architecture d
This repository offers a unified interface for multiple LLM providers and local models, functioning as an AI request router and proxy, but it lacks explicit support for advanced gateway features like response caching, API key management, and observability that the search requires, so it is a narrower but still category-fitting tool.
This project is an artificial intelligence API gateway that centralizes connections to multiple model providers into a single, standardized interface. By acting as a proxy, it translates diverse provider protocols into a format compatible with existing clients, allowing developers to integrate various language models without managing provider-specific software development kits. The gateway distinguishes itself through a robust traffic management layer that includes intelligent request routing, weighted load balancing, and automated failover mechanisms to ensure service availability. It incorp
Uni-API provides a unified OpenAI-compatible API endpoint and supports routing to multiple cloud LLM providers (OpenAI, Anthropic, Gemini, etc.) with load balancing, which meets your need for centralized multi-provider management, but it does not mention integration with self-hosted local models.