For an open source AI gateway for LLM providers, the strongest matches are diegosouzapw/omniroute (OmniRoute is a self-hostable LLM API gateway that provides), coaidev/coai (CoAI is a self-hostable AI gateway and proxy platform) and looplj/axonhub (Axonhub is an open-source AI gateway and multi-model proxy). portkey-ai/gateway and winfunc/deepreasoning round out the shortlist. Each is ranked by relevance to your query, popularity and recent activity.
We curate open-source GitHub repositories matching “open source alternatives to openrouter”. Results are ranked by relevance to your query — pick filters below to narrow, or refine with AI.
OmniRoute is a unified LLM API gateway that connects multiple AI providers to a single endpoint. Its primary purpose is to simplify the integration of various AI models into tools and agents by translating different provider formats into a standardized API. The project distinguishes itself through a multi-strategy request routing system that optimizes for cost, speed, and availability, including automatic model fallbacks and a circuit-breaker resilience model to isolate provider failures. It employs a local-first security posture, using AES-256-GCM encryption to store API keys and conversatio
OmniRoute is a self-hostable LLM API gateway that provides a unified endpoint, multi-provider routing, API key management, and usage tracking with streaming response support.
CoAI is an enterprise-grade, self-hostable AI gateway platform that unifies access to over 200 AI models from more than 35 providers through a single OpenAI-compatible API endpoint. It functions as a multi-tenant gateway, routing requests across providers with load balancing, automatic failover, and priority-based routing, while exposing standard OpenAI API endpoints for chat, image generation, model listing, and billing to enable seamless integration with existing tools and clients. The platform distinguishes itself through a comprehensive set of operational capabilities built around the gat
CoAI is a self-hostable AI gateway and proxy platform that aggregates multiple LLM providers behind a unified OpenAI-compatible API with cost tracking and routing.
Axonhub is an AI gateway and multi-model API proxy that provides a unified interface for routing requests to multiple large language model providers. It functions as a load balancer and translation layer, converting a standardized API format into provider-specific payloads to enable communication with various AI models without provider-specific code. The system manages traffic through rule-based routing and automatic failover to maintain high availability. It differentiates its operations by providing a provider-agnostic interface that decouples client requests from specific model backends us
Axonhub is an open-source AI gateway and multi-model proxy that provides a unified API endpoint and multi-provider routing, though it lacks explicit native support for streaming responses in its listed feature set.
This project is an artificial intelligence gateway that functions as a centralized middleware layer for managing, securing, and observing interactions with language, vision, and audio models. It provides a unified interface that standardizes requests across multiple providers, enabling teams to integrate AI capabilities into their applications through a consistent set of tools and protocols. The gateway distinguishes itself through its comprehensive infrastructure governance and traffic management capabilities. It allows for policy-driven routing, automated failover, and load balancing across
Portkey is a self-hostable AI gateway that aggregates multiple language model providers behind a unified API while delivering key infrastructure capabilities like multi-provider routing, cost and usage tracking, and streaming responses.
DeepReasoning is a self-hosted AI gateway and chat interface that provides an LLM inference API. It functions as a bridge that merges reasoning traces from DeepSeek R1 with the generative capabilities of Claude models to facilitate complex problem solving. The system is delivered as a dockerized application, allowing for deployment on private infrastructure. This architecture enables private LLM inference and secure local management of API keys and authentication tokens on user-controlled hardware. The project covers multi-model orchestration by combining chain-of-thought reasoning and gener
DeepReasoning is a self-hosted AI gateway and chat interface that bridges multiple model providers with a unified API, though it focuses specifically on combining DeepSeek and Claude reasoning flows rather than general multi-provider aggregation.
LiteLLM is a unified gateway and proxy server designed to centralize access to over one hundred language model providers. It provides a standardized API interface that abstracts vendor-specific schemas, allowing developers to interact with diverse models through a single, consistent format. By acting as a central traffic management layer, it enables organizations to route, secure, and govern model interactions across multiple deployments. The platform distinguishes itself through its policy-driven architecture, which uses configuration-based routing to manage traffic distribution, load balanc
LiteLLM is a self-hostable AI gateway and proxy server that provides a unified API, multi-provider routing, API key management, cost tracking, and streaming responses to centralize access to hundreds of large language models.
This project is an AI model API gateway and proxy server designed to provide a unified interface for interacting with diverse artificial intelligence service providers. It functions as a centralized middleware platform that routes, load balances, and translates API requests across multiple models, enabling developers to access text, image, audio, and video generation capabilities through a single, standardized integration. The gateway distinguishes itself through comprehensive administrative and financial controls, including event-driven usage accounting, real-time token consumption tracking,
This project is an open-source AI gateway and proxy server that aggregates multiple large language model providers under a unified API with built-in token usage tracking and self-hosting support.
This project provides a unified interface for interacting with a wide range of artificial intelligence services, acting as a central orchestration layer for text and image generation. It standardizes access to diverse AI backends, allowing developers to integrate multiple language and vision models through a single, consistent programming interface. By abstracting provider-specific protocols and authentication requirements, the tool simplifies the development of applications that rely on external AI services. The platform distinguishes itself through a resilient request routing architecture d
This project serves as a unified interface and reverse-proxy layer for multiple AI services, providing multi-provider routing and abstraction despite lacking native cost tracking and commercial self-hostable proxy management.
Quotio is a local LLM API proxy gateway and credential manager that intercepts and routes requests from command-line tools and integrated development environments to various AI model providers. It serves as a centralized authentication hub, managing API keys and service accounts to provide a unified interface for external AI agents. The project distinguishes itself through a routing engine that implements priority-chain and round-robin load balancing to distribute workloads across multiple accounts. It features automated API key failover, which redirects requests to backup authentication keys
Quotio is a self-hostable AI proxy gateway designed to intercept and route requests across multiple LLM providers with credential management and load balancing, though it is tailored more for local CLI and IDE integrations than broad enterprise usage.
This repository provides a self-hostable load balancer and proxy tailored for Google's Gemini API, offering a unified endpoint and streaming support, though it is narrower in scope than a multi-provider aggregator.
InsForge is a backend-as-a-service platform that provides an integrated suite of tools for managing relational databases, identity provision, object storage, and serverless compute. It functions as an open-source identity provider and a PostgreSQL database manager featuring integrated vector storage and row-level security. The platform serves as an LLM orchestration gateway, offering a unified endpoint to route requests across various AI providers through an OpenAI-compatible interface. It enables AI-driven application generation and connects AI agents to backend resources using a standardize
InsForge acts as an LLM orchestration gateway providing a unified endpoint for multiple AI providers alongside its broader backend-as-a-service features, though it lacks some specialized proxy features like detailed cost tracking.
Helicone is an AI gateway and observability platform designed to intercept, manage, and monitor interactions with large language models. By acting as a reverse-proxy, it provides a centralized layer for routing requests across multiple AI providers, allowing developers to maintain consistent application logic while gaining deep visibility into model performance, usage, and costs. The platform distinguishes itself through a robust suite of traffic management and prompt engineering tools. It enables policy-driven control, including automatic failover between providers, rate limiting, and edge-b
Helicone is an AI gateway and observability platform that acts as a reverse proxy for multiple providers, though its primary emphasis is on monitoring and analytics rather than self-hosted request routing.
LocalAI is a self-hosted inference server that enables the execution of machine learning models directly on local hardware. By providing a unified interface for text, image, and audio processing, it allows users to maintain full control over data privacy and infrastructure costs while eliminating dependencies on external network services. The platform functions as an API gateway that mimics standard cloud-based artificial intelligence interfaces, allowing existing applications to integrate local models as drop-in replacements. It utilizes a container-based architecture to package runtimes and
LocalAI is a self-hosted inference server that acts as an API proxy and gateway for local models, though its primary focus is on running models locally rather than aggregating multiple external cloud providers under a unified API.
This project is an HTTP request forwarder designed to act as a middle layer for routing traffic to OpenAI services. It functions as a proxy that manages network connectivity, allowing users to bypass restrictions while centralizing API access and request handling. The service distinguishes itself by supporting real-time stream forwarding, which delivers incremental data chunks to clients as they are generated. It also incorporates middleware for content moderation, enabling the inspection and filtering of data streams to enforce safety guidelines before content reaches the end user. The arch
This is a self-hostable OpenAI API proxy deployed via Docker that supports streaming responses, though it focuses primarily on proxying and forwarding rather than full multi-provider aggregation with advanced cost tracking.
| Repository | Stars | Language | License | Last push |
|---|---|---|---|---|
| diegosouzapw/omniroute | 6.4K | TypeScript | MIT | |
| coaidev/coai | 9.2K | TypeScript | Apache-2.0 | |
| looplj/axonhub | 4.4K | Go | NOASSERTION | |
| portkey-ai/gateway | 12.1K | TypeScript | MIT | |
| winfunc/deepreasoning | 5.4K | Rust | MIT | |
| berriai/litellm | 50.6K | Python | NOASSERTION | |
| quantumnous/new-api | 39.7K | Go | AGPL-3.0 | |
| xtekky/gpt4free | 66.3K | Python | GPL-3.0 | |
| nguyenphutrong/quotio | 3.6K | Swift | mit | |
| snailyp/gemini-balance | 5.8K | Python | other |