For a self hosted alternative to ChatGPT, the first results are quivrhq/quivr, chatgptnextweb/nextchat (NextChat is a self-hostable conversational web application supporting multiple AI providers and persistent chat histories, though it lacks built-in RAG and web search integration out of the box) and danny-avila/librechat (LibreChat is a comprehensive open-source AI conversational interface that supports self-hosting, multiple LLM providers, persistent chat memory, RAG, web search, RBAC, and rich markdown and code rendering). onyx-dot-app/onyx and open-webui/open-webui round out the shortlist. Compare the match explanations and check the project documentation against your requirements.
Self-hosted conversational AI models and chatbot interfaces that provide private alternatives to proprietary language services.
Quivr is a retrieval-augmented generation platform designed to transform raw documents into searchable knowledge bases. It functions as a centralized environment where users can ingest files, index them into vector databases, and interact with language models to receive contextually relevant, data-backed responses. The platform distinguishes itself through an agentic workflow orchestrator that sequences retrieval tasks, tool execution, and model interactions to resolve complex, multi-step queries. This engine is entirely configuration-driven, allowing users to define document ingestion, chunk
Quivr is a self-hostable retrieval-augmented generation platform featuring document ingestion, vector databases, and conversational chat interfaces for language models, though it is primarily optimized as a knowledge base and RAG assistant rather than a general-purpose multi-provider chat UI.
NextChat is a self-hosted web application that provides a unified interface for interacting with multiple large language models. It functions as a conversational platform where users can manage and switch between diverse AI providers through configurable API backends, maintaining full control over their data and infrastructure. The platform features a persistent session layer designed to handle long-running dialogues by managing message history and context. It distinguishes itself through a structured prompt engineering environment that allows for the development and application of templates
NextChat is a self-hostable conversational web application supporting multiple AI providers and persistent chat histories, though it lacks built-in RAG and web search integration out of the box.
LibreChat is an artificial intelligence orchestration platform that provides a unified interface for interacting with multiple language models. It functions as a centralized workspace where users can switch between different intelligence engines, manage complex conversational workflows, and maintain persistent memory across sessions through a vector-database-backed storage system. The platform distinguishes itself through an extensible agent framework that supports autonomous task execution and the integration of external tools. It features a secure, containerized environment for executing co
LibreChat is a comprehensive open-source AI conversational interface that supports self-hosting, multiple LLM providers, persistent chat memory, RAG, web search, RBAC, and rich markdown and code rendering.
Onyx is an enterprise-grade AI platform designed for knowledge management, search, and autonomous agent orchestration. It functions as a centralized system that aggregates unstructured organizational data, enabling secure, context-aware retrieval and interaction across internal documents and communication history. By integrating retrieval-augmented generation with multi-model orchestration, the platform provides a unified interface for teams to query internal knowledge bases and execute complex, multi-step business processes. The platform distinguishes itself through a focus on private infras
Onyx is an enterprise-grade AI platform and chat interface supporting self-hosting, document retrieval, and multi-model orchestration, matching your need for a customizable assistant platform despite focusing heavily on knowledge management.
Open WebUI is a self-hosted, web-based platform designed for interacting with local and remote artificial intelligence models. It functions as a unified interface and orchestration suite, enabling users to build, deploy, and manage specialized AI agents equipped with custom instructions, external tool access, and private knowledge bases. The platform distinguishes itself through a modular architecture that supports complex AI workflows. It features a plugin-based framework for custom logic and pipeline-based request processing, allowing developers to filter or transform data streams before th
Open WebUI is a self-hostable conversational AI platform providing a feature-rich web interface for multiple language models, complete with RAG support, chat history, web search, and markdown rendering.
Danswer is an LLM application framework and RAG engine that provides a self-hosted interface for connecting large language models to private data. It serves as an enterprise AI chat interface and agent orchestrator, enabling the creation of specialized assistants with custom instructions and knowledge bases. The platform differentiates itself through an observability dashboard for tracking query history and token consumption, as well as a white-labeled interface for customized branding. It includes a multi-step research workflow for producing long-form reports and a sandboxed environment for
Danswer is a self-hostable conversational AI platform focused on RAG and enterprise document search, offering role-based access control and multi-user chat capabilities, though it is more heavily tailored as a search and retrieval engine than a general-purpose multi-model chat UI.
This project is a self-hosted large language model chat interface and AI model aggregator. It provides a unified web environment for interacting with multiple AI providers and local models, acting as a provider-agnostic API gateway to standardize requests across different endpoints. The platform functions as an agentic AI framework and generative UI workspace, enabling the construction of specialized assistants with custom instructions and subagents. It features a sandboxed code interpreter for secure execution of multiple programming languages and a generative UI system that renders interact
This repository provides a self-hosted chat interface and AI aggregator that supports multiple model providers, chat history, document uploads, web search, and code rendering, perfectly matching your need for a customizable conversational AI platform.
This project is a self-hosted web interface and desktop application designed for interacting with language models. It provides a private platform for managing conversational sessions, allowing users to connect to external AI services while maintaining control over their interaction history and configuration settings. The application distinguishes itself by offering a unified interface that supports multimodal inputs and outputs, including voice interaction processing and generative image creation. It secures sensitive credentials by routing requests through a backend proxy and ensures data pr
This project is a self-hostable chat interface that connects to external language models, providing conversation history management and a desktop application wrapper to serve your conversational AI needs.
Chatbot-ui is a self-hosted AI dashboard and LLM chat interface that serves as a centralized hub for interacting with multiple artificial intelligence providers. It functions as a multi-provider AI client and model orchestrator, allowing users to send prompts and receive responses from various large language models through a unified conversational interface. The project enables multi-model AI chat within a single workspace, allowing for the comparison of outputs and capabilities across different backends. It provides a private frontend for AI workspace management, where users can organize cha
This repository is a self-hostable LLM chat interface that provides a unified dashboard for interacting with multiple AI providers, meeting all core requirements for a custom conversational AI platform.
Khoj is a self-hosted artificial intelligence platform designed for personal knowledge management and semantic information retrieval. It functions as a private assistant that indexes your local documents, notes, and external workspaces, allowing you to interact with your data through natural language queries and conversational chat. By maintaining a local-first architecture, the system ensures that your information remains under your control while providing context-aware responses grounded in your personal knowledge base. The platform distinguishes itself through a modular, cross-platform int
Khoj is a self-hostable conversational AI and semantic search platform that connects to various LLM providers and supports document upload and RAG, though it focuses more heavily on personal knowledge management than a general-purpose chatbot interface.
SillyTavern is a comprehensive interface and orchestration platform designed for immersive AI roleplay and interactive chat experiences. It functions as a unified gateway that connects users to a wide array of local and cloud-based large language models, providing a centralized environment to manage complex character personas, narrative context, and model-driven interactions. The platform distinguishes itself through its advanced prompt engineering and automation capabilities. It utilizes a sophisticated macro-based templating engine and vector-database retrieval to dynamically inject lore, c
SillyTavern is a feature-rich, self-hostable conversational UI supporting multiple LLM backends and vector-based retrieval, making it a strong fit for custom-deployed chat and roleplay interfaces despite lacking built-in enterprise web search and RBAC.
DeepReasoning is a self-hosted AI gateway and chat interface that provides an LLM inference API. It functions as a bridge that merges reasoning traces from DeepSeek R1 with the generative capabilities of Claude models to facilitate complex problem solving. The system is delivered as a dockerized application, allowing for deployment on private infrastructure. This architecture enables private LLM inference and secure local management of API keys and authentication tokens on user-controlled hardware. The project covers multi-model orchestration by combining chain-of-thought reasoning and gener
DeepReasoning is a self-hosted AI gateway and chat interface that supports private LLM inference and model orchestration, though it focuses more on reasoning-generative model bridges and APIs rather than offering a full suite of user-facing assistant features like RAG or web search.
Serge is a self-hosted web chat interface for running large language models locally using the llama.cpp inference engine. It loads GGUF-format model files directly on your own machine, removing the need for internet connectivity or external API keys, and streams responses to the browser in real time via WebSocket connections. The project is packaged for containerized deployment using Docker and Docker Compose, with a Traefik reverse proxy that handles HTTP and WebSocket routing along with automatic TLS certificate management. Ready-made Kubernetes manifests are also provided, enabling deploym
Serge is a self-hosted web chat interface designed for running local models via llama.cpp, making it a strong fit for self-hosted LLM deployment, though it lacks some advanced features like multi-provider support or web search out of the box.
Chatbox is a cross-platform desktop application that provides a unified interface for interacting with a wide range of artificial intelligence models. It functions as a model-agnostic client, allowing users to connect to various third-party AI providers or execute open-source models directly on their own hardware. By centralizing these diverse services into a single workspace, the application enables users to manage multiple chat sessions, adjust model parameters, and switch between different AI backends with ease. The project distinguishes itself through a local-first architecture that prior
Chatbox is a model-agnostic desktop client that connects to multiple LLM providers and supports local knowledge bases, though it is primarily a client application rather than a fully self-hostable server platform.
wukong-robot is an open-source, Chinese-language voice assistant platform that integrates ChatGPT for multi-turn conversational AI. It is built around a plugin-based smart speaker framework, combining offline wake word detection with local speech synthesis to enable hands-free, voice-controlled interactions without requiring a constant internet connection. The platform distinguishes itself through its modular architecture, supporting custom wake word training via the command line and a plugin system that routes user intents using regular expressions for extensible functionality. It offers mul
wzpan/wukong-robot is an open-source voice assistant platform that integrates conversational AI models for smart speaker interactions, though it focuses primarily on voice-controlled hardware and speech processing rather than a standard web-based LLM chat interface.
Jan is a desktop application that functions as a local artificial intelligence model runtime and an open-standard API server. It enables the execution of large language models directly on local hardware, ensuring that data remains private and accessible offline while providing a unified interface for managing model weights and inference runtimes. The platform distinguishes itself by offering a modular inference backend that allows users to swap execution engines based on hardware compatibility and performance needs. It acts as a cross-platform orchestrator, providing the ability to switch bet
Jan is a local AI desktop application that lets you run and chat with language models privately on your hardware, missing online features like web search and document RAG out of the box but serving as a strong self-hosted chat interface.
ChatGPT-Next-Web is a cross-platform web interface and frontend for interacting with large language models. It functions as a self-hosted client that allows users to connect to various AI model providers through a unified chat interface compatible with web browsers and desktop operating systems. The project includes a prompt template manager for creating and organizing reusable masks to standardize interactions. It supports self-hosting on private clouds to maintain data security and provides a centralized administrative panel for managing API resources and member access permissions. The app
ChatGPT-Next-Web is a self-hostable cross-platform chat interface supporting multiple LLM providers, conversation history, and prompt management, though it lacks built-in RAG document uploads and direct web search integration.
Create chatbots with ease
Dialoqbase is a self-hostable conversational AI platform that lets you build chatbots with document upload and retrieval capabilities, making it a strong fit for custom deployment despite lacking some advanced features like web search.
Conversations with your files! Manage and run your AI presets!
This TypeScript repository provides a file-based conversational AI interface for managing AI presets and interacting with language models, fitting the category well despite missing some advanced enterprise features like robust role-based access control.
Rag-stack is an enterprise knowledge retrieval system designed to deploy private generative artificial intelligence environments. It functions as a retrieval-augmented generation stack, orchestrating the connection between internal document repositories and open-source language models to enable natural language querying of private organizational data. The platform distinguishes itself by providing a complete infrastructure for private large language model hosting and vector database management. By utilizing infrastructure-as-code provisioning and containerized microservices, it allows organiz
Finic-ai/rag-stack is a self-hostable conversational AI platform designed for private corporate deployment with knowledge base integration, though it lacks explicit mention of web search or role-based access control.
| Repository | Stars | Language | License | Last push |
|---|---|---|---|---|
| quivrhq/quivr | 39.2K | Python | NOASSERTION | |
| chatgptnextweb/nextchat | 88.3K | TypeScript | MIT | |
| danny-avila/librechat | 39.3K | TypeScript | MIT | |
| onyx-dot-app/onyx | 17.5K | Python | other | |
| open-webui/open-webui | 142.7K | Python | NOASSERTION | |
| danswer-ai/danswer | 30.6K | Python | NOASSERTION | |
| danny-avila/chatgpt-clone | 39.3K | TypeScript | MIT | |
| niek/chatgpt-web | 2K | Svelte | GPL-3.0 | |
| mckaywrigley/chatbot-ui | 33.3K | TypeScript | MIT | |
| khoj-ai/khoj | 35.2K | Python | AGPL-3.0 |