For an open source desktop application for running large language models, the strongest matches are nomic-ai/gpt4all (GPT4All is a cross-platform desktop application and inference runtime), josstorer/rwkv-runner (This desktop application provides chat interfaces, model management, and) and arthur-ficial/apfel (Apfel is a macOS-native local LLM application offering an). janhq/jan and yidadaa/chatgpt-next-web round out the shortlist. Each is ranked by relevance to your query, popularity and recent activity.
We curate open-source GitHub repositories matching “open source alternatives to lm studio”. Results are ranked by relevance to your query — pick filters below to narrow, or refine with AI.
GPT4All is a cross-platform runtime environment designed to execute large language models directly on local consumer hardware. By leveraging an optimized C++ inference backend, it enables private, offline AI interactions without requiring an internet connection or external cloud services. The project provides a comprehensive ecosystem for managing the entire model lifecycle, including discovery, downloading, and configuration of local weights. What distinguishes the platform is its integrated retrieval-augmented generation engine, which allows users to index local documents into semantic vect
GPT4All is a cross-platform desktop application and inference runtime that allows you to run, download, and chat with large language models locally on your hardware.
This desktop application provides chat interfaces, model management, and an OpenAI-compatible server specifically tailored for running RWKV models locally, making it a strong fit despite its focus on a specific model architecture rather than general-purpose backends.
The free AI already on your Mac. CLI tool, OpenAI-compatible server, and interactive chat — all on-device via Apple Intelligence. No API keys, no cloud, no downloads.
Apfel is a macOS-native local LLM application offering an interactive chat interface and an OpenAI-compatible server powered by Apple Intelligence, though it is limited to the Apple ecosystem rather than being cross-platform.
Jan is a desktop application that functions as a local artificial intelligence model runtime and an open-standard API server. It enables the execution of large language models directly on local hardware, ensuring that data remains private and accessible offline while providing a unified interface for managing model weights and inference runtimes. The platform distinguishes itself by offering a modular inference backend that allows users to swap execution engines based on hardware compatibility and performance needs. It acts as a cross-platform orchestrator, providing the ability to switch bet
Jan is a cross-platform desktop application that runs large language models locally with a chat interface, model management, and an OpenAI-compatible API server.
ChatGPT-Next-Web is a web-based chat interface for interacting with large language models via API or self-hosted model runners. It functions as a prompt management tool and a cross-platform application available for web, mobile, and desktop environments. The project distinguishes itself through a plugin integration gateway that extends model capabilities with external tools like network search and calculators. It includes a self-hosted administrative dashboard for controlling model lists, member permissions, and access passwords on private infrastructure. The application covers prompt engine
This project is a cross-platform web and desktop interface for interacting with large language models, though it is primarily designed to connect to APIs or external model runners rather than executing models locally itself.
Chatbox is a cross-platform desktop application that provides a unified interface for interacting with a wide range of artificial intelligence models. It functions as a model-agnostic client, allowing users to connect to various third-party AI providers or execute open-source models directly on their own hardware. By centralizing these diverse services into a single workspace, the application enables users to manage multiple chat sessions, adjust model parameters, and switch between different AI backends with ease. The project distinguishes itself through a local-first architecture that prior
Chatbox is a cross-platform desktop application for interacting with local and remote language models, featuring a chat interface, prompt management, and local model connectivity.
Cherry Studio is a cross-platform desktop application that serves as a centralized workspace for managing and interacting with multiple artificial intelligence models. It functions as a local-first orchestrator, prioritizing user privacy by storing all conversation history and knowledge bases directly on your device. By providing a unified interface for both cloud-based and local AI services, the platform simplifies API key management and allows for consistent model interaction across different operating systems. The application distinguishes itself through a robust retrieval-augmented genera
Cherry Studio is a cross-platform desktop application that provides a unified chat interface, local model management, and prompt handling for interacting with both local and cloud-based LLMs.
Open WebUI is a self-hosted, web-based platform designed for interacting with local and remote artificial intelligence models. It functions as a unified interface and orchestration suite, enabling users to build, deploy, and manage specialized AI agents equipped with custom instructions, external tool access, and private knowledge bases. The platform distinguishes itself through a modular architecture that supports complex AI workflows. It features a plugin-based framework for custom logic and pipeline-based request processing, allowing developers to filter or transform data streams before th
Open WebUI is a feature-rich, self-hosted web interface and desktop-friendly platform for interacting with local LLMs, offering model management, RAG support, and an OpenAI-compatible API server.
KoboldAI-Client is a web-based interface and toolkit for interacting with large language models. It functions as a local AI text generator for storytelling and conversational AI, providing a front end for models hosted either on local hardware or within cloud-provisioned environments. The system includes a persona manager that uses external modules and soft-prompting to guide AI responses toward specific characters and writing styles. It also provides an API wrapper that exposes a standardized, OpenAI-compatible REST API, allowing external applications to communicate with the hosted models.
KoboldAI-Client is a web-based local LLM frontend and text generation interface that supports persona management and an OpenAI-compatible API, though it acts as a browser-based tool rather than a standalone cross-platform desktop application.
This project is a comprehensive platform for hosting and interacting with large language models directly on local hardware. It provides a web-based graphical interface that allows users to manage model loading, configure generation parameters, and execute text or chat interactions entirely offline. By running models locally, the software ensures complete data privacy and eliminates reliance on external cloud services for generative tasks. Beyond basic inference, the platform functions as a versatile workbench for generative AI development. It includes an integrated pipeline for fine-tuning mo
This project is a web-based local LLM interface that handles model downloading, chat interactions, and API serving, though it runs as a browser UI rather than a native desktop app.
SillyTavern is a comprehensive interface and orchestration platform designed for immersive AI roleplay and interactive chat experiences. It functions as a unified gateway that connects users to a wide array of local and cloud-based large language models, providing a centralized environment to manage complex character personas, narrative context, and model-driven interactions. The platform distinguishes itself through its advanced prompt engineering and automation capabilities. It utilizes a sophisticated macro-based templating engine and vector-database retrieval to dynamically inject lore, c
SillyTavern is a feature-rich desktop-oriented web interface and chat application designed to connect with local and cloud large language models, offering robust prompt templates and character management, though it relies on external backends rather than executing models natively.
NextChat is a self-hosted web application that provides a unified interface for interacting with multiple large language models. It functions as a conversational platform where users can manage and switch between diverse AI providers through configurable API backends, maintaining full control over their data and infrastructure. The platform features a persistent session layer designed to handle long-running dialogues by managing message history and context. It distinguishes itself through a structured prompt engineering environment that allows for the development and application of templates
NextChat is a cross-platform desktop and web chat client supporting local model execution via Ollama and prompt templates, though it acts primarily as a frontend interface rather than an embedded local runner or OpenAI-compatible server.
This project is a self-hosted large language model chat interface and AI model aggregator. It provides a unified web environment for interacting with multiple AI providers and local models, acting as a provider-agnostic API gateway to standardize requests across different endpoints. The platform functions as an agentic AI framework and generative UI workspace, enabling the construction of specialized assistants with custom instructions and subagents. It features a sandboxed code interpreter for secure execution of multiple programming languages and a generative UI system that renders interact
This repository provides a web-based chat interface and aggregator for interacting with both local models and external providers, fitting the functional scope though designed as a web application rather than a native desktop app.
Lobe Chat is a self-hosted AI platform that provides a web-based interface for interacting with multiple large language models. It functions as an AI agent orchestrator, allowing for the design, scheduling, and management of autonomous agent teams to perform operational tasks. The platform features an extensible plugin framework and SDK to integrate external tools and custom function calls into workflows. It utilizes a provider-agnostic model layer to unify various AI APIs and includes a context-aware memory system to store structured user information for personalized interactions. The syste
Lobe Chat provides a feature-rich web-based chat interface for interacting with large language models, though it is primarily a web platform rather than a native desktop app.
| Repository | Stars | Language | License | Last push |
|---|---|---|---|---|
| nomic-ai/gpt4all | 77.4K | C++ | MIT | |
| josstorer/rwkv-runner | 6.2K | TypeScript | mit | |
| arthur-ficial/apfel | 5.9K | Swift | MIT | |
| janhq/jan | 43K | TypeScript | NOASSERTION | |
| yidadaa/chatgpt-next-web | 88.3K | TypeScript | MIT | |
| chatboxai/chatbox | 40.5K | TypeScript | GPL-3.0 | |
| cherryhq/cherry-studio | 47.4K | TypeScript | AGPL-3.0 | |
| open-webui/open-webui | 142.7K | Python | NOASSERTION | |
| koboldai/koboldai-client | 3.9K | Python | AGPL-3.0 | |
| oobabooga/text-generation-webui | 47.3K | Python | AGPL-3.0 |