For 轻量级自托管 LLM 聊天界面, the strongest matches are huggingface/chat-ui (This is Hugging Face’s official self-hostable chat UI for), ramon-victor/freegpt-webui (freegpt-webui is a self-hosted web interface for LLMs that) and gaizhenbiao/chuanhuchatgpt (Chuanhuchatgpt is a web-based UI and API gateway for). chatgptnextweb/nextchat and nomic-ai/gpt4all round out the shortlist. Each is ranked by relevance to your query, popularity and recent activity.
我们为您精选了匹配 “nanochat llm” 的开源 GitHub 仓库。结果按与您查询的相关性进行排名 — 您可以使用下方筛选器缩小范围,或通过 AI 进行优化。
This project is a web-based user interface for interacting with large language models, featuring streaming responses and persistent conversation history. It functions as an orchestration gateway that directs user prompts to specific language models and acts as a Model Context Protocol client to execute external tools and incorporate live data into conversations. The application includes a routing layer that analyzes input signals and tool requirements to dynamically direct messages to the most appropriate specialized model. It also provides customization settings for brand identity, allowing
This is Hugging Face’s official self-hostable chat UI for LLMs, supporting multiple backends, streaming responses, conversation history, and markdown rendering — a comprehensive and deployable solution that exactly matches what you’re looking for.
freegpt-webui is a self-hosted web interface for interacting with large language models. It provides a chat-based frontend designed for communicating with GPT 3.5 and GPT 4 models. The application enables API keyless chat, allowing users to access conversational AI for text generation and information retrieval without providing or managing personal authentication keys. The system handles model integration through a reverse-proxy gateway and supports asynchronous stream processing for real-time text generation. User preferences and conversation history are persisted via client-side session st
freegpt-webui is a self-hosted web interface for LLMs that provides streaming responses and conversation history via client-side storage, fitting the search for a minimal chat UI, though its multi-model support is limited to OpenAI models.
This project is a web-based user interface and multi-model API gateway for interacting with various large language model providers and local inference services. It functions as a retrieval-augmented generation chatbot for private document questioning, a manager for model fine-tuning, and an autonomous agent framework. The system distinguishes itself by integrating an autonomous assistant mode that uses web search and external tools to solve complex, multi-step tasks without manual prompting. It also features an API gateway capable of rotating multiple authentication keys to balance usage and
Chuanhuchatgpt is a web-based UI and API gateway for LLMs that supports multiple backends, streaming, conversation history, and markdown rendering, and is self-hostable; it is a fitting chat interface, though its feature set is broad rather than minimal.
NextChat is a self-hosted web application that provides a unified interface for interacting with multiple large language models. It functions as a conversational platform where users can manage and switch between diverse AI providers through configurable API backends, maintaining full control over their data and infrastructure. The platform features a persistent session layer designed to handle long-running dialogues by managing message history and context. It distinguishes itself through a structured prompt engineering environment that allows for the development and application of templates
NextChat delivers exactly what you're looking for: a self-hostable, minimal web UI that connects to multiple LLM providers like Ollama, Claude GPT-4o Gemini Groq streaming conversation history is built in markdown rendering is supported the interface is designed to be lightweight and fast makes it a clean minimal chat experience fitting your self-hosted requirement perfectly
GPT4All is a cross-platform runtime environment designed to execute large language models directly on local consumer hardware. By leveraging an optimized C++ inference backend, it enables private, offline AI interactions without requiring an internet connection or external cloud services. The project provides a comprehensive ecosystem for managing the entire model lifecycle, including discovery, downloading, and configuration of local weights. What distinguishes the platform is its integrated retrieval-augmented generation engine, which allows users to index local documents into semantic vect
GPT4All is a self-hostable runtime environment that runs LLMs locally on consumer hardware, providing a chat interface with conversation history, streaming, and local document indexing, which squarely matches the need for a lightweight, private AI chat tool.
This platform serves as a comprehensive environment for managing private language models, document knowledge bases, and automated agent workflows within secure local infrastructure. It functions as a document-aware workspace that enables users to ingest diverse file formats into searchable repositories, ensuring that all data processing and model inference remain within private, local environments to maintain data sovereignty. The system distinguishes itself through a modular agentic engine that allows for the definition of custom skills and external tool execution. By utilizing a multi-model
AnythingLLM is a comprehensive self-hosted LLM chat platform that supports multiple backend models (via Ollama, LocalAI, etc.), streaming, conversation history, and document-aware chat—all within a local, private environment, making it a fitting but feature-rich choice for your lightweight chat interface search.
Open WebUI is a self-hosted, web-based platform designed for interacting with local and remote artificial intelligence models. It functions as a unified interface and orchestration suite, enabling users to build, deploy, and manage specialized AI agents equipped with custom instructions, external tool access, and private knowledge bases. The platform distinguishes itself through a modular architecture that supports complex AI workflows. It features a plugin-based framework for custom logic and pipeline-based request processing, allowing developers to filter or transform data streams before th
Open WebUI is a self-hosted, web-based interface for interacting with local and remote LLMs, fully aligning with the need for a self-hosted LLM chat interface with multi-model support, streaming, and conversation history.
Lobe Chat is a self-hosted AI platform that provides a web-based interface for interacting with multiple large language models. It functions as an AI agent orchestrator, allowing for the design, scheduling, and management of autonomous agent teams to perform operational tasks. The platform features an extensible plugin framework and SDK to integrate external tools and custom function calls into workflows. It utilizes a provider-agnostic model layer to unify various AI APIs and includes a context-aware memory system to store structured user information for personalized interactions. The syste
Lobe Chat is a self-hosted web platform that provides a direct interface for chatting with multiple large language models, covering the core needs of multi-model support, streaming, and conversation history while remaining fully self-hostable.
This project is a comprehensive platform for hosting and interacting with large language models directly on local hardware. It provides a web-based graphical interface that allows users to manage model loading, configure generation parameters, and execute text or chat interactions entirely offline. By running models locally, the software ensures complete data privacy and eliminates reliance on external cloud services for generative tasks. Beyond basic inference, the platform functions as a versatile workbench for generative AI development. It includes an integrated pipeline for fine-tuning mo
text-generation-webui is a comprehensive web interface for locally hosting and interacting with large language models, supporting multiple backends, streaming responses, conversation history, and markdown rendering — exactly the self-hosted, minimal-UI chat interface this search targets.
Jan is a desktop application that functions as a local artificial intelligence model runtime and an open-standard API server. It enables the execution of large language models directly on local hardware, ensuring that data remains private and accessible offline while providing a unified interface for managing model weights and inference runtimes. The platform distinguishes itself by offering a modular inference backend that allows users to swap execution engines based on hardware compatibility and performance needs. It acts as a cross-platform orchestrator, providing the ability to switch bet
Jan is a desktop application that runs LLMs locally with a modular inference backend and a user interface, directly matching the need for a self-hostable chat interface—it supports multiple models, streaming, and conversation management while keeping data private and offline.
Jan is a local language model desktop application and AI assistant orchestrator. It provides a unified interface for interacting with both resident models and remote cloud AI providers. The project functions as a host for the Model Context Protocol, connecting AI models to external tools and data sources. It also operates as an OpenAI compatible API server, exposing local models through a standardized server endpoint for other applications to query. The system supports the creation of specialized AI personas with custom instructions and allows for the management of hybrid model environments,
Jan is a local desktop app that unifies interactions with local and cloud LLMs, supporting multiple models, streaming, and conversation management — exactly the self-hosted chat interface for LLMs the visitor is looking for.
Khoj is a self-hosted artificial intelligence platform designed for personal knowledge management and semantic information retrieval. It functions as a private assistant that indexes your local documents, notes, and external workspaces, allowing you to interact with your data through natural language queries and conversational chat. By maintaining a local-first architecture, the system ensures that your information remains under your control while providing context-aware responses grounded in your personal knowledge base. The platform distinguishes itself through a modular, cross-platform int
Khoj is a self-hosted AI platform that provides a chat interface for interacting with LLMs alongside RAG over your personal documents — it fits the core request for a self-hosted LLM chat tool, though its focus on knowledge management makes it less of a minimalist standalone chat app than the visitor might expect.
The free AI already on your Mac. CLI tool, OpenAI-compatible server, and interactive chat — all on-device via Apple Intelligence. No API keys, no cloud, no downloads.
Apfel provides an interactive chat and an OpenAI-compatible server that runs locally on macOS via Apple Intelligence, making it a self-hosted chat interface for LLMs, though its platform and model scope are narrower than a general-purpose solution.
| 仓库 | Star 数 | 语言 | 许可证 | 最后推送 |
|---|---|---|---|---|
| huggingface/chat-ui | 10.8K | TypeScript | Apache-2.0 | |
| ramon-victor/freegpt-webui | 5.6K | Python | GPL-3.0 | |
| gaizhenbiao/chuanhuchatgpt | 15.3K | Python | GPL-3.0 | |
| chatgptnextweb/nextchat | 88.3K | TypeScript | MIT | |
| nomic-ai/gpt4all | 77.4K | C++ | MIT | |
| mintplex-labs/anything-llm | 61.7K | JavaScript | MIT | |
| open-webui/open-webui | 142.7K | Python | NOASSERTION | |
| lobehub/lobe-chat | 78.8K | TypeScript | NOASSERTION | |
| oobabooga/text-generation-webui | 47.3K | Python | AGPL-3.0 | |
| janhq/jan | 43K | TypeScript | NOASSERTION |