12 مستودعات
Ready-to-use interfaces and tools for interacting with AI models and performing specific tasks.
Explore 12 awesome GitHub repositories matching part of an awesome list · End-User Applications. Refine with filters or upvote what's useful.
Open WebUI is a self-hosted, web-based platform designed for interacting with local and remote artificial intelligence models. It functions as a unified interface and orchestration suite, enabling users to build, deploy, and manage specialized AI agents equipped with custom instructions, external tool access, and private knowledge bases. The platform distinguishes itself through a modular architecture that supports complex AI workflows. It features a plugin-based framework for custom logic and pipeline-based request processing, allowing developers to filter or transform data streams before th
Web interface for interacting with various LLMs.
ComfyUI is a modular generative AI workflow orchestrator and node-based GUI for designing and executing complex diffusion model pipelines. It functions as both a visual interface for building generative logic graphs and a programmable backend API that exposes diffusion model operations for external integration. The system distinguishes itself through a graph-based execution model that supports differential workflow execution, re-running only modified nodes to reduce computation. It features dynamic model offloading to manage memory between system RAM and GPU VRAM and utilizes metadata-embedde
Visual node-based interface for Stable Diffusion.
Lobe Chat is a self-hosted AI platform that provides a web-based interface for interacting with multiple large language models. It functions as an AI agent orchestrator, allowing for the design, scheduling, and management of autonomous agent teams to perform operational tasks. The platform features an extensible plugin framework and SDK to integrate external tools and custom function calls into workflows. It utilizes a provider-agnostic model layer to unify various AI APIs and includes a context-aware memory system to store structured user information for personalized interactions. The syste
Modern AI conversation interface.
Upscayl is a cross-platform desktop application designed to increase the resolution and visual quality of digital images using artificial intelligence. By executing all processing tasks locally on the user's machine, the software ensures that sensitive media files remain private and never leave the host system for cloud-based services. The application distinguishes itself through a hardware-agnostic architecture that offloads intensive rendering workloads directly to the local graphics unit. It utilizes a hardware abstraction layer to translate enhancement commands into instructions compatibl
AI-powered image upscaling tool.
LibreChat is an artificial intelligence orchestration platform that provides a unified interface for interacting with multiple language models. It functions as a centralized workspace where users can switch between different intelligence engines, manage complex conversational workflows, and maintain persistent memory across sessions through a vector-database-backed storage system. The platform distinguishes itself through an extensible agent framework that supports autonomous task execution and the integration of external tools. It features a secure, containerized environment for executing co
Open-source ChatGPT alternative.
Quivr is a retrieval-augmented generation platform designed to transform raw documents into searchable knowledge bases. It functions as a centralized environment where users can ingest files, index them into vector databases, and interact with language models to receive contextually relevant, data-backed responses. The platform distinguishes itself through an agentic workflow orchestrator that sequences retrieval tasks, tool execution, and model interactions to resolve complex, multi-step queries. This engine is entirely configuration-driven, allowing users to define document ingestion, chunk
Personal second brain and AI assistant.
Facefusion is a modular framework designed for automated image and video manipulation, specializing in tasks such as face swapping, enhancement, and restoration. It functions as a computer vision processing pipeline that chains independent machine learning modules to perform complex transformations, including facial animation, age modification, and lip synchronization. The system is built to handle both real-time interactive feeds and large-scale batch processing tasks. The platform distinguishes itself through a highly extensible architecture that supports custom processing modules and inter
AI face swapping and enhancement tool.
Screenpipe is a local screen and audio recorder that captures and indexes digital activity to create a searchable archive of computer usage. It functions as an AI context engine, providing a local database of visual and auditory history to ground large language models. The system serves as a Model Context Protocol server, delivering screen history and meeting transcriptions to external AI assistants. It utilizes an OCR screen search tool to extract text from visual data and a speech-to-text transcription tool for identifying speakers in system and microphone audio. The software includes capa
Local AI that records, searches, and automates tasks based on your screen and audio.
This project is an AI research tool designed for autonomous web information gathering and automated topic research. It utilizes agent orchestration to combine search engines and web scraping, enabling the system to discover detailed information and build a comprehensive understanding of complex subjects without manual step-by-step guidance. The tool employs an iterative research execution model that recursively generates targeted search queries and refines directions based on previous results. It includes a feedback loop that compares current findings against initial objectives to identify kn
AI-powered research assistant for iterative, deep research on any topic.
DocsGPT is a retrieval-augmented generation platform and private knowledge base used to build AI agents that perform grounded search and analysis. It functions as a multi-model AI orchestrator and enterprise agent builder, allowing for the integration of various local and cloud language models to customize reasoning and text generation. The project provides a visual environment for developing automated assistants using conditional logic and third-party API connectivity. It enables the creation of private AI agents capable of performing enterprise search and detailed document analysis using pr
Documentation-based question answering system.
DeepTutor is a framework for personalized AI tutoring and educational content generation. It functions as an agentic workflow system that executes reasoning loops to complete multi-step tasks, transforming raw sources into structured learning materials such as interactive books, quizzes, and concept graphs. The platform distinguishes itself through an extensible skill architecture that allows the installation and auditing of third-party capability packages from community registries. It utilizes persona-driven tool policies to deploy persistent AI companions with unique behavioral profiles and
AI-powered personalized learning assistant with document Q&A, exercise generation, and deep research capabilities.
Jaaz هي مجموعة أدوات تصميم بالذكاء الاصطناعي ذاتية الاستضافة ومساحة عمل متعددة الوسائط تُستخدم لإنشاء وتحرير الصور ومقاطع الفيديو. تعمل كمساحة عمل للتصميم حيث يمكن للمستخدمين إنتاج محتوى مرئي وأصول من خلال مزيج من نماذج الذكاء الاصطناعي المحلية والقائمة على السحابة. يتميز المشروع بمنسق نماذج هجين يوجه الطلبات بين مشغلات النماذج المحلية وواجهات برمجة التطبيقات البعيدة لتحقيق التوازن بين خصوصية البيانات وأداء المعالجة. ويستخدم أداة تعاونية ذات لوحة غير محدودة لتنظيم لوحات القصة والأصول، ويتضمن مُحسّن مطالبات الصور لترجمة الأفكار الأولية إلى مطالبات توليد مفصلة. تغطي المنصة التصميم متعدد الوسائط وإنتاج الوسائط، بما في ذلك ترجمة الرسم إلى صورة وتحرير الصور القائم على الدردشة للحفاظ على اتساق الشخصية. وتوفر أدوات تنظيم مكانية لتخطيط السرد المرئي وتدعم عمليات النشر الخاصة ومتعددة المستأجرين على البنية التحتية للمؤسسة لضمان ملكية البيانات.
Open-source multimodal creative assistant and privacy-focused alternative to Canva/Manus for local image/video generation.