11 रिपॉजिटरी
Unified abstractions for interacting with chat-based LLMs.
Explore 11 awesome GitHub repositories matching artificial intelligence & ml · Chat Model Interfaces. Refine with filters or upvote what's useful.
LangChain is an orchestration framework designed for building, managing, and deploying applications powered by large language models. It provides a unified integration layer that normalizes disparate model provider APIs into a consistent set of primitives, enabling developers to build complex, multi-step AI workflows that manage state, memory, and tool execution. The project distinguishes itself through a durable execution runtime that maintains persistent state across long-running processes by checkpointing progress to external storage. It models agent workflows as directed graphs, allowing
Exposes unified interfaces for initializing and interacting with various chat-based language models.
Ktransformers is a comprehensive framework designed for the operation, fine-tuning, and serving of large language models. It functions as a heterogeneous inference engine and quantized execution runtime, enabling the deployment of massive models by distributing computational workloads across both CPU and GPU resources. This architecture allows users to bypass local memory constraints, making it possible to run and train models that exceed the capacity of a single device. The project distinguishes itself through specialized support for sparse architectures, particularly mixture-of-experts mode
Offers an interactive command-line interface for direct chat-based testing and validation of loaded models.
Mistral Inference is a library for running Mistral large language models on a GPU, generating text from prompts with token streaming. It loads pretrained model weights from local disk or a remote registry into GPU memory, then produces output tokens one by one for real-time display in interactive applications. The library supports multimodal prompts that accept image URLs alongside text, enabling visual description and reasoning. It includes content safety guardrails that scan generated text against predefined policies to block or flag policy violations. For structured interactions, it provid
Provides a command-line session that accepts user prompts and streams model responses.
MiniCPM is a collection of small language models designed for local, on-device deployment in resource-constrained environments. The project focuses on running dense Transformer models on consumer hardware, including GPUs, CPUs, and Apple Silicon, without requiring custom code forks. The project distinguishes itself through heavy optimization for edge hardware, utilizing quantized weight compression in GGUF and MLX formats to reduce memory overhead. It implements advanced inference techniques such as speculative sampling and radix-tree prefix caching to accelerate generation speed and throughp
Implements an interactive chat interface allowing back-and-forth conversations with response interruption via CLI.
CopilotForXcode is an AI-powered coding assistant integrated directly into Xcode as a source editor extension. It functions as an agent that can automate multi-step project tasks, such as editing files, running terminal commands, and searching across the entire codebase, all while understanding the full context of the current Xcode project. The assistant provides a context-aware chat interface that answers coding questions based on open files, symbols, and recent edits. It also offers diff-based code review, analyzing changes to provide feedback on code quality and potential issues before mer
Answers coding questions and provides suggestions through a chat interface integrated into the development environment.
CodeCompanion is a Neovim plugin that brings large language model capabilities directly into the editor, enabling turn-based conversations with AI models in a dedicated chat buffer. It provides a comprehensive interface for interacting with LLMs, supporting multiple providers through a flexible adapter system that can route requests to various hosted or local language model services. The plugin distinguishes itself through its extensive context-sharing capabilities, allowing users to send buffer contents, visual selections, git diffs, LSP diagnostics, terminal output, quickfix lists, and view
Opens a chat buffer inside the editor to converse with language models and receive coding assistance.
freegpt-webui लार्ज लैंग्वेज मॉडल के साथ बातचीत करने के लिए एक सेल्फ-होस्टेड वेब इंटरफेस है। यह GPT 3.5 और GPT 4 मॉडल के साथ संचार करने के लिए डिज़ाइन किया गया एक चैट-आधारित फ्रंटएंड प्रदान करता है। एप्लिकेशन API कीलेस चैट को सक्षम बनाता है, जिससे उपयोगकर्ता व्यक्तिगत प्रमाणीकरण कुंजियों को प्रदान या प्रबंधित किए बिना टेक्स्ट जनरेशन और सूचना पुनर्प्राप्ति के लिए संवादात्मक AI तक पहुंच सकते हैं। सिस्टम रिवर्स-प्रॉक्सी गेटवे के माध्यम से मॉडल एकीकरण को संभालता है और रियल-टाइम टेक्स्ट जनरेशन के लिए एसिंक्रोनस स्ट्रीम प्रोसेसिंग का समर्थन करता है।
Provides a dedicated chat-based web interface for interacting with large language models.
AIOS is an LLM agent operating system and orchestration kernel designed to manage memory, resource scheduling, and tool execution for multiple autonomous AI agents. It serves as a comprehensive framework for developing and deploying agents, featuring a dedicated resource manager that coordinates model backends, GPU memory, and isolated kernel instances. The system distinguishes itself through a semantic memory engine that uses vector search and autonomous clustering for long-term knowledge management, and a semantic file system that allows users to control computer files and system operations
Executes chat completions with integrated support for JSON output, tool calls, and file operations.
AICommand एक LLM-संचालित एडिटर कंट्रोलर और ऑटोमेशन टूल है जिसे Unity Editor के लिए डिज़ाइन किया गया है। यह प्राकृतिक भाषा प्रॉम्प्ट्स को निष्पादन योग्य कमांड्स में बदलने के लिए लार्ज लैंग्वेज मॉडल्स को एकीकृत करता है, जिससे उपयोगकर्ताओं को टेक्स्ट-आधारित निर्देशों के माध्यम से डेवलपमेंट वातावरण सेटिंग्स और ऑब्जेक्ट्स में हेरफेर करने की अनुमति मिलती है। सिस्टम एक टेक्स्ट-टू-एक्शन पाइपलाइन का उपयोग करता है जो LLM-आधारित कमांड मैपिंग और रिफ्लेक्शन-आधारित API डिस्कवरी का उपयोग करके प्राकृतिक भाषा इनपुट को विशिष्ट एडिटर फंक्शन्स में हल करता है। यह प्रक्रिया रनटाइम पर स्ट्रिंग आइडेंटिफायर्स को कोड हैंडल्स में हल करके Unity एडिटर फंक्शन्स के निष्पादन को सक्षम बनाती है। टूल उपयोगकर्ता के प्रॉम्प्ट्स को कैप्चर करने के लिए डेवलपमेंट वातावरण के भीतर एक एकीकृत चैट इंटरफेस प्रदान करता है। यह दोहराव वाले सेटअप और कॉन्फ़िगरेशन कार्यों को स्वचालित करने और मैन्युअल मेनू नेविगेशन को प्राकृतिक भाषा नियंत्रण से बदलकर गेम इंजन वर्कफ़्लो को सुव्यवस्थित करने पर केंद्रित है।
Provides a custom chat window embedded within the Unity editor to capture and send user prompts.
ChatGPT.nvim एक OpenAI LLM Neovim प्लगइन है जो कोड स्निपेट्स को रिफैक्टर करने, पूरा करने और ठीक करने के लिए AI-संचालित कोड सहायक के रूप में कार्य करता है। यह एक कस्टम प्रॉम्प्ट ऑटोमेशन टूल और मार्कडाउन रेंडरर के रूप में कार्य करता है, जो उपयोगकर्ताओं को सीधे एडिटर के भीतर भाषा मॉडल के साथ बातचीत करने की अनुमति देता है। यह प्रोजेक्ट पुन: प्रयोज्य AI व्यक्तित्वों और सिस्टम प्रॉम्प्ट और टेम्प्लेट का उपयोग करके स्वचालित टेक्स्ट प्रोसेसिंग कार्यों को परिभाषित करने के लिए एक फ्रेमवर्क के माध्यम से खुद को अलग करता है। इसमें एक विशेष रेंडरिंग सिस्टम शामिल है जो प्रतिक्रियाओं को फोल्ड करने योग्य कोड ब्लॉक के साथ स्टाइल किए गए मार्कडाउन में बदल देता है। यह प्लगइन इंटरैक्टिव चैट इंटरफ़ेस, प्रोजेक्ट फ़ाइलों के लिए संदर्भ-जागरूक प्रॉम्प्टिंग, और व्याकरण सुधार और बग फिक्सिंग के लिए स्वचालित कोड क्रियाओं को कवर करता है। यह स्थिर भंडारण के बजाय बाहरी कमांड के माध्यम से API कुंजियाँ प्राप्त करके सुरक्षित क्रेडेंशियल प्रबंधन भी लागू करता है।
Provides a chat interface embedded directly within the Neovim editor for interactive AI-assisted development.
Claudian is a framework that combines AI coding agents, knowledge base integration, and a multi-provider orchestrator for managed interactions with large language models. It functions as a browser extension that connects users to AI services through a sidebar and inline editing interface, providing a system for integrating agents into local directories to perform file operations, bash commands, and workspace searches. The project distinguishes itself with a multi-provider orchestrator that allows switching between different AI backends while maintaining separate conversation states and config
Integrates a chat interface into a sidebar and inline-edit flow for real-time interaction with LLMs.