awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

58 个仓库

Awesome GitHub RepositoriesAI Application Frameworks

Frameworks for developing AI-native applications.

Distinguishing note: Focuses on application development frameworks rather than simple tools.

Explore 58 awesome GitHub repositories matching artificial intelligence & ml · AI Application Frameworks. Refine with filters or upvote what's useful.

Awesome AI Application Frameworks GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • ottermind/chat2dbOtterMind 的头像

    OtterMind/Chat2DB

    25,784在 GitHub 上查看↗

    Chat2DB is an AI-powered SQL client and multi-database GUI manager designed for managing various relational and NoSQL database systems. It serves as a visual database management tool and a natural language to SQL interface, allowing users to convert plain text descriptions into executable and optimized queries. The platform distinguishes itself through automated business intelligence capabilities, which include the generation of real-time data visualization dashboards and AI-driven data analysis from spreadsheets. To ensure data privacy, it supports secure local AI deployment, enabling large

    Runs large language models on local hardware to process sensitive database metadata without external uploads.

    Javaaibichatgpt
    在 GitHub 上查看↗25,784
  • vercel-labs/aivercel-labs 的头像

    vercel-labs/ai

    24,918在 GitHub 上查看↗

    This project is a TypeScript SDK and application framework for integrating large language models into software. It provides a unified interface and multi-provider model wrapper to interact with various AI model providers through a single, consistent API. The toolkit includes a generative UI framework and an AI agent orchestrator. These tools enable the creation of autonomous agents capable of executing functions and the development of AI-driven user interfaces with specialized state management for streaming chatbot components. The framework covers broad capability areas including stream-base

    Provides a comprehensive framework for building AI-native applications with a focus on intelligent user interfaces and agents.

    TypeScript
    在 GitHub 上查看↗24,918
  • datawhalechina/prompt-engineering-for-developersdatawhalechina 的头像

    datawhalechina/prompt-engineering-for-developers

    24,267在 GitHub 上查看↗

    This project is a technical curriculum and development guide focused on large language model prompt engineering, fine-tuning, and the creation of retrieval augmented generation applications. It serves as a comprehensive resource for developers to master crafting precise instructions and textual patterns to improve the quality and predictability of model outputs. The material covers the end-to-end workflow of adapting open-source models to specific datasets and integrating language models with vector databases to generate responses based on private information. It also provides a systematic ap

    Guides the creation of AI-native applications by combining orchestration frameworks and user interfaces.

    Jupyter Notebook
    在 GitHub 上查看↗24,267
  • jina-ai/jinajina-ai 的头像

    jina-ai/jina

    21,858在 GitHub 上查看↗

    Jina is a cloud-native framework for building and deploying multimodal AI applications that process text, images, and audio across distributed microservices. It functions as an inference orchestrator and a distributed model gateway, providing a containerized stack to organize AI executors into operational pipelines. The system manages large language model workloads through token-streamed response delivery and dynamic batching to increase hardware throughput. It utilizes a protocol-agnostic communication layer to route data across different machine learning frameworks. The framework covers hi

    Provides a cloud-native framework for building and deploying AI applications that integrate text, images, and audio across distributed microservices.

    Python
    在 GitHub 上查看↗21,858
  • ymcui/chinese-llama-alpacaymcui 的头像

    ymcui/Chinese-LLaMA-Alpaca

    18,944在 GitHub 上查看↗

    This project is a comprehensive toolkit for adapting large language models to the Chinese language, providing a specialized framework for fine-tuning, inference, and local deployment. It serves as a coordinated suite for language-specific adaptation, including tools for expanding tokenizers and implementing retrieval-augmented generation. The project distinguishes itself through a complete pipeline for model adaptation, featuring multilingual tokenizer expansion and a fine-tuning framework that supports instruction-based supervised training and adapter merging. It also includes a dedicated de

    Integrates models into frameworks to create end-to-end tools for question answering, summarization, and chatbots.

    Pythonalpacaalpaca-2large-language-models
    在 GitHub 上查看↗18,944
  • apple/ml-stable-diffusionapple 的头像

    apple/ml-stable-diffusion

    17,901在 GitHub 上查看↗

    This project is a framework for running Stable Diffusion image generation models on Apple Silicon using Core ML hardware acceleration. It provides a local generative AI pipeline for producing images from text prompts using Swift and Python without relying on external cloud APIs. The system includes a model converter to transform deep learning checkpoints into Core ML formats and a model optimizer to quantize weights and activations. It features a ControlNet integration layer to guide image generation using external signals such as edge and depth maps. Capabilities cover text-to-image generat

    Provides dedicated libraries for integrating generative image pipelines into native macOS and iOS applications.

    Python
    在 GitHub 上查看↗17,901
  • deepseek-ai/janusdeepseek-ai 的头像

    deepseek-ai/Janus

    17,746在 GitHub 上查看↗

    Janus is a multimodal large language model and unified framework that integrates visual understanding and image generation within a single neural network. It functions as both a visual understanding model for analyzing images and a text-to-image generator. The system uses a unified transformer backbone and a multimodal latent space to bridge the gap between text and visual data. This architecture employs decoupled visual encoding and cross-modal tokenization to separate the paths for discriminative understanding and generative tasks, representing images as grids of discrete codes. The projec

    Provides a unified framework capable of both interpreting and synthesizing visual content.

    Pythonany-to-anyfoundation-modelsllm
    在 GitHub 上查看↗17,746
  • nvidia/nemoNVIDIA 的头像

    NVIDIA/NeMo

    17,394在 GitHub 上查看↗

    NeMo is a multimodal AI framework and toolkit designed for the development, training, and scaling of large language models, generative AI systems, and speech-based models. It functions as an automatic speech recognition toolkit, a text-to-speech engine, and a framework for building models that process and generate combinations of text, image, and audio data. The project serves as a conversational AI orchestrator capable of managing real-time, interruptible voice interactions. It provides specialized workflows for speech translation, converting spoken audio from one language into text or speec

    Provides a framework to build and manage models that process and generate combinations of text, image, and audio data.

    Python
    在 GitHub 上查看↗17,394
  • microsoft/ai-edumicrosoft 的头像

    microsoft/ai-edu

    14,065在 GitHub 上查看↗

    ai-edu is a comprehensive AI education curriculum and machine learning courseware collection. It provides theoretical tutorials, deep learning lab exercises, and project blueprints designed to teach artificial intelligence fundamentals through a combination of study and practical implementation. The project focuses on a learning-by-doing approach, guiding users from Python programming and neural network basics to advanced topics. It includes specialized instructional content on distributed AI training, MLOps educational guides for model quantization and pruning, and detailed frameworks for im

    Provides operational instructions and practical cases for developing vision, language, and speech applications.

    HTML
    在 GitHub 上查看↗14,065
  • fujiwarachoki/moneyprinterFujiwaraChoki 的头像

    FujiwaraChoki/MoneyPrinter

    13,571在 GitHub 上查看↗

    MoneyPrinter is an automated short-form video creation pipeline that generates complete YouTube Shorts from a given topic. It combines local LLM-powered script generation with programmatic video assembly, all managed through a database-backed job queue for reliable, restart-tolerant processing. The system uses an Ollama-powered local language model to write video scripts and metadata entirely on-device, keeping data private and offline. It then produces the final video clip using MoviePy for compositing clips, text, and audio, creating a complete YouTube Shorts video without manual editing. V

    Writes video scripts and metadata by querying a local Ollama language model, keeping all data processing on-device.

    Pythonautomationchatgptmoviepy
    在 GitHub 上查看↗13,571
  • cocktailpeanut/dalaicocktailpeanut 的头像

    cocktailpeanut/dalai

    12,920在 GitHub 上查看↗

    The simplest way to run LLaMA on your local machine

    Downloads specific model variants by name from a CDN for local use.

    CSSaillamallm
    在 GitHub 上查看↗12,920
  • chainlit/chainlitChainlit 的头像

    Chainlit/chainlit

    12,213在 GitHub 上查看↗

    Chainlit is a Python framework designed for building and deploying interactive, stateful conversational AI interfaces. It provides a backend-driven platform that connects language models and agent frameworks to a web-based chat frontend, managing the complexities of session state, message history, and real-time communication. The framework distinguishes itself by offering a component-based UI builder that allows developers to inject interactive widgets, rich media, and data visualizations directly into the chat stream. It supports the visualization of complex agent workflows, enabling users t

    Implements secure user authentication and session management for conversational AI applications.

    Pythonchatgptlangchainllm
    在 GitHub 上查看↗12,213
  • yaofanguk/video-subtitle-removerYaoFANGUK 的头像

    YaoFANGUK/video-subtitle-remover

    11,493在 GitHub 上查看↗

    This project is a local AI inpainting tool designed to erase hard-coded subtitles and watermarks from videos and images. It functions as a content-aware media restorer that uses deep learning to reconstruct missing pixels and preserve the original resolution of the source files. The software is distinguished by its local execution model, running inference on host hardware to process media without relying on external cloud APIs. It employs content-aware model selection, allowing the use of different generative algorithms based on media types, such as animation or live action, to optimize visua

    Runs generative filling models on host hardware for local media inpainting without cloud APIs.

    Pythonaideepleanringsub-remove
    在 GitHub 上查看↗11,493
  • salesforce/lavissalesforce 的头像

    salesforce/LAVIS

    11,236在 GitHub 上查看↗

    LAVIS is a multimodal large language model framework and vision-language model library. It provides tools for training and evaluating models that integrate visual, textual, and audio data, serving as a cross-modal feature extractor and a zero-shot visual reasoning engine. The framework distinguishes itself by using frozen-backbone integration, where pretrained encoders remain non-trainable while lightweight adapter layers are updated. It employs cross-modal feature alignment to map different representations into a shared embedding space and utilizes a modular model wrapper to swap vision and

    Provides a comprehensive framework for training and evaluating large language models that integrate visual, textual, and audio data.

    Jupyter Notebook
    在 GitHub 上查看↗11,236
  • mistralai/mistral-srcmistralai 的头像

    mistralai/mistral-src

    10,821在 GitHub 上查看↗

    该项目是一个大语言模型推理库和框架,旨在运行用于文本生成、问题解决和编码辅助的模型。它包括一个用于处理图像和文本组合输入的多模态框架,以及一个基于模型推理执行外部工具的工具调用实现。 该系统具有分布式 GPU 推理引擎,可将大型模型工作负载分散到多个图形处理器上,以提高处理速度并满足内存需求。它还通过预打包的镜像和依赖项提供容器化模型部署,以便在隔离环境中运行推理引擎。 该库涵盖了一系列功能,包括多模态输入分析、函数调用集成,以及用于预测缺失代码段的“中间填充”(fill-in-the-middle)编码。它还支持通过命令行界面进行交互式模型聊天,以维持对话会话。

    Ships a framework for processing combined image and text inputs to describe visual content and answer questions.

    Jupyter Notebook
    在 GitHub 上查看↗10,821
  • wasmedge/wasmedgeWasmEdge 的头像

    WasmEdge/WasmEdge

    10,665在 GitHub 上查看↗

    WasmEdge is an extensible WebAssembly runtime that executes WebAssembly bytecode in a secure sandbox for cloud, edge, and embedded applications. It functions as a multi-language compiler, compiling applications written in Rust, JavaScript, Go, and Python into WebAssembly bytecode for sandboxed execution, and as a server-side JavaScript runtime that runs JavaScript programs with ES6 modules, NPM packages, and Node.js-compatible APIs. The runtime also serves as an AI inference runtime, executing AI models from JavaScript using WASI-NN plug-ins for inference tasks on personal devices and edge har

    Executes AI models on smart devices by running them inside a WebAssembly sandbox with GPU access.

    C++artificial-intelligencecloudcloud-native
    在 GitHub 上查看↗10,665
  • langchain-ai/local-deep-researcherL

    langchain-ai/local-deep-researcher

    9,223在 GitHub 上查看↗

    Local Deep Researcher is a fully local web research assistant that uses any LLM hosted by Ollama or LMStudio. Give it a topic and it will generate a web search query, gather web search results, summarize the results of web search, reflect on the summary to examine knowledge gaps, generate a new…

    Provides a research agent that runs entirely on local hardware using Ollama-hosted LLMs.

    Python
    在 GitHub 上查看↗9,223
  • spring-projects/spring-aispring-projects 的头像

    spring-projects/spring-ai

    9,001在 GitHub 上查看↗

    Spring AI is an application framework for Java that provides a portable, fluent API for integrating AI models, tools, and vector stores into applications. It wraps multiple AI providers behind a common interface, allowing developers to switch between chat, embedding, image, and speech models without changing application code. The framework includes a chainable chat client API similar to WebClient or RestClient, supports both synchronous and streaming interactions, and offers structured output conversion that transforms unstructured AI responses into strongly-typed Java objects. The framework

    Ships a portable, fluent Java framework for integrating AI models, tools, and vector stores into applications.

    Javaartificial-intelligencejavaspring-ai
    在 GitHub 上查看↗9,001
  • reorproject/reorreorproject 的头像

    reorproject/reor

    8,560在 GitHub 上查看↗

    Reor is a local AI knowledge management application that stores, links, and searches personal notes using large language models and vector embeddings entirely on the user's device. It functions as a private AI note assistant, keeping all data and processing local for full privacy without relying on external cloud services. The application integrates with Ollama to manage the lifecycle of local LLMs and embedding models, handling downloads, updates, and execution. Notes are imported from markdown files, preserving existing file structure, and are automatically linked through vector-similarity

    Downloads, updates, and executes LLMs and embedding models through the Ollama runtime for local AI processing.

    JavaScriptailancedbllama
    在 GitHub 上查看↗8,560
  • optimalscale/lmflowOptimalScale 的头像

    OptimalScale/LMFlow

    8,488在 GitHub 上查看↗

    LMFlow is a comprehensive suite for large language model fine-tuning, context extension, multimodal processing, and inference execution. It provides a toolkit for updating model parameters through full tuning or memory-efficient adapter algorithms, alongside an inference engine for executing tuned models via command-line or web-based interfaces. The framework includes a dedicated alignment suite for supervised tuning and reward model training to refine model behavior. It features a context window extender to increase maximum input lengths and a multimodal framework for building chatbots that

    Provides a framework for building chatbots that process combined image and text inputs.

    Pythonchatgptdeep-learninginstruction-following
    在 GitHub 上查看↗8,488
上一个123下一个
  1. Home
  2. Artificial Intelligence & ML
  3. AI Application Frameworks

探索子标签

  • Local On-Device AI8 个子标签Development of AI applications that run search and inference locally on the user's hardware. **Distinct from AI Application Frameworks:** Distinct from general AI frameworks: focuses specifically on the local, on-device execution environment.
  • Multimodal FrameworksFrameworks specifically designed to process and integrate multiple data modalities like text, image, and audio. **Distinct from AI Application Frameworks:** Specializes AI application frameworks for multimodal data processing rather than general AI application development.