awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

16 个仓库

Awesome GitHub RepositoriesContent Summarization

Utilities for generating and truncating content excerpts for use in listings and previews.

Distinct from Custom Page Frameworks: Distinct from Custom Page Frameworks: focuses on the specific logic for generating content synopses rather than general page rendering.

Explore 16 awesome GitHub repositories matching web development · Content Summarization. Refine with filters or upvote what's useful.

Awesome Content Summarization GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • panniantong/agent-reachPanniantong 的头像

    Panniantong/Agent-Reach

    31,610在 GitHub 上查看↗

    Agent-Reach is an AI agent web gateway and search tool that provides language models with the ability to search and read content from the open web, social media, and community forums without using official APIs. It functions as a routing layer that connects large language models to various internet backends while managing content parsing and connection health. The system enables API-free information retrieval by using open-source backends to extract text and metadata from platforms such as Twitter, Reddit, and YouTube. It converts unstructured website content, RSS feeds, and video transcripts

    Extracts transcripts and metadata from video platforms to allow AI agents to summarize video material.

    Pythonagent-infrastructureai-agentai-search
    在 GitHub 上查看↗31,610
  • getzola/zolagetzola 的头像

    getzola/zola

    17,196在 GitHub 上查看↗

    Zola is a static site generator that compiles Markdown and templates into a standalone website. It is distributed as a single binary, removing the need for external runtimes or package managers to build the final site. The project includes a built-in Sass compiler to transform styles into compressed CSS and a dedicated Markdown rendering engine that supports task lists and footnotes. It also features a client-side search indexer, enabling full-text site search without a backend server, and a multilingual content manager for organizing translated content. Additional capabilities cover asset o

    Splits long lists of content into multiple sequential paginated pages to improve readability.

    Rustblog-enginecmscontent-management-system
    在 GitHub 上查看↗17,196
  • getgrav/gravgetgrav 的头像

    getgrav/grav

    15,395在 GitHub 上查看↗

    Grav is a flat-file content management system that eliminates the need for a traditional database by storing site content and configuration in human-readable Markdown and YAML files. Built as a modular PHP web framework, it uses a hierarchical page routing system where the physical directory structure directly determines the site's URL paths. The platform is distinguished by its event-driven plugin architecture and a command-line interface that prioritizes system administration, deployment, and maintenance tasks. It utilizes a blueprint-driven system to generate administrative forms from stru

    Provides configurable rules for generating content synopses and excerpts for site listings.

    PHPcmscontentcontent-management
    在 GitHub 上查看↗15,395
  • smol-ai/developersmol-ai 的头像

    smol-ai/developer

    12,188在 GitHub 上查看↗

    This project is an AI software engineering tool and framework for building autonomous coding agents. It provides a system for automating program synthesis and bug fixing by integrating large language models with codebase analysis and iterative refinement loops. The framework features an agentic development server that exposes task execution interfaces to remote agents through a structured protocol. This allows for the remote execution of development tasks and the embedding of autonomous program synthesis capabilities into external software projects. The toolset covers AI-driven project scaff

    Extracts text and titles from browser tabs to generate structured content summaries.

    Python
    在 GitHub 上查看↗12,188
  • miromindai/mirothinkerMiroMindAI 的头像

    MiroMindAI/MiroThinker

    8,315在 GitHub 上查看↗

    MiroThinker 是一个自主研究系统,使用大语言模型通过迭代推理进行深度研究和预测。它充当一个 Web 搜索 AI 框架,能够检索实时互联网数据并抓取 Web 内容,为复杂查询提供可验证的来源。 该系统包括一个多模态内容处理器,可将图像、音频和视频转换为文本描述,供基于文本的模型进行分析。为了确保计算准确性,它利用沙盒代码执行器来运行 Python 代码和数据分析。性能通过 AI 基准测试工具进行管理,该工具使用自动化评判对标准化数据集上的代理响应的准确性和质量进行评估。 该项目提供了代理工作流功能,包括迭代推理循环、研究报告生成和研究文档导入。它还结合了内存管理策略来优化上下文窗口,并记录交互历史以供模型训练。

    Uses language models to generate concise summaries of large volumes of retrieved web content.

    Python
    在 GitHub 上查看↗8,315
  • prefecthq/marvinPrefectHQ 的头像

    PrefectHQ/marvin

    6,170在 GitHub 上查看↗

    an ambient intelligence library

    Produces concise summaries of any provided text or content using a language model.

    Python
    在 GitHub 上查看↗6,170
  • jimmylv/bibigpt-v1JimmyLv 的头像

    JimmyLv/BibiGPT-v1

    6,116在 GitHub 上查看↗

    BibiGPT-v1 is an AI-powered media summarizer that generates concise summaries and enables interactive Q&A for audio and video content from multiple platforms. It uses large language models to process transcripts from sources like YouTube, Bilibili, and local files, delivering real-time streaming responses for an interactive chat experience. The project distinguishes itself by combining multi-platform content aggregation with a conversational learning assistant capability, allowing users to query audio and video content through AI-driven dialogue. It also includes export functionality for savi

    Generates concise summaries from audio and video content across multiple platforms for quick comprehension.

    TypeScriptbilibilichatgptgpt
    在 GitHub 上查看↗6,116
  • jimmylv/bibigptJimmyLv 的头像

    JimmyLv/BibiGPT

    6,111在 GitHub 上查看↗

    BibiGPT v1 · one-Click AI Summary for Audio/Video & Chat with Learning Content: Bilibili | YouTube | Tweet丨TikTok丨Dropbox丨Google Drive丨Local files | Websites丨Podcasts | Meetings | Lectures, etc. 音视频内容 AI 一键总结 & 对话:哔哩哔哩丨YouTube丨推特丨小红书丨抖音丨快手丨百度网盘丨阿里云盘丨网页丨播客丨会议丨本地文件等 (原 BiliGPT 省流神器 & AI课代表)

    Generates concise summaries of audio and video content from platforms like YouTube and Bilibili using AI.

    TypeScript
    在 GitHub 上查看↗6,111
  • sylinko/everywhereSylinko 的头像

    Sylinko/Everywhere

    6,093在 GitHub 上查看↗

    Everywhere is a desktop AI assistant that understands whatever is on your screen and can act across applications without requiring screenshots or manual context switching. It reads structured UI data through accessibility and automation APIs to perceive the active application and visible content, then provides context-aware help, summaries, translations, and answers to natural language questions about what you are viewing. The tool distinguishes itself by combining on-screen content analysis with a multi-LLM agent platform that routes requests to providers like OpenAI, Anthropic, and local mo

    Extracts key points from a webpage the user is viewing and delivers a concise summary on demand.

    C#aiai-agentsai-assistant
    在 GitHub 上查看↗6,093
  • richasy/bili.copilotRichasy 的头像

    Richasy/Bili.Copilot

    5,220在 GitHub 上查看↗

    Bili.Copilot is a native Windows desktop client for Bilibili that integrates large language models to provide AI-enhanced media browsing and video summarization. It functions as a media browser and video summarizer, enabling users to generate concise overviews of videos and articles by processing subtitles and text. The application allows users to interact with AI models to query specific information and evaluate content within videos. These capabilities are delivered through a native Windows interface built with the Windows App SDK and WinUI. The software covers media management and content

    Generates concise summaries from video subtitles and articles using artificial intelligence.

    GLSLbilibiliwindows-app-sdkwinui3
    在 GitHub 上查看↗5,220
  • worldbrain/memexWorldBrain 的头像

    WorldBrain/Memex

    4,691在 GitHub 上查看↗

    Memex 是一个浏览器扩展知识库和个人信息管理器,旨在将 Web 内容索引、注释并组织到可搜索的个人档案中。它作为 Web 注释工具,允许用户直接向网页和 PDF 添加高亮和笔记。 该系统具有 AI 驱动的文档摘要器,可根据索引材料使用引用的参考资料生成简洁的答案和摘要。它包括一个加密内容同步器,使用端到端加密在多个设备之间镜像档案和注释。 该平台提供跨书签网站和文档的全文搜索和过滤内容检索功能。它涵盖了通过浏览器标签组织和自动化标记进行的内容管理,以及通过本地优先存储和云备份进行的数据持久化。 该项目提供标准化 API,将保存的知识连接到外部工具和 AI 代理,以实现可互操作的信息共享。

    Generates concise summaries and answers questions based on indexed web materials using cited references.

    TypeScriptannotateannotationannotations
    在 GitHub 上查看↗4,691
  • madawei2699/mygptreadermadawei2699 的头像

    madawei2699/myGPTReader

    4,418在 GitHub 上查看↗

    myGPTReader 是一套大型语言模型应用套件,包含聊天界面、文档分析工具和新闻聚合器。该系统专注于从数字文件和网页内容中提取信息,以实现对话式分析和自动化内容摘要。 该项目具有提示词模板管理器,用于构建对话流程并提高响应准确性。它还包含一个多语言语音聊天客户端,集成了语音转文字(STT)和文字转语音(TTS)功能,用于实时交互式辅导和语言练习。 该平台涵盖了检索增强对话、每日新闻自动推送,以及网站和视频内容摘要等更广泛的功能。

    Extracts and condenses information from both websites and videos into conversational summaries.

    Python
    在 GitHub 上查看↗4,418
  • wendy7756/ai-video-transcriberwendy7756 的头像

    wendy7756/AI-Video-Transcriber

    2,799在 GitHub 上查看↗

    AI-Video-Transcriber is an automated media processing platform that converts audio and video files into structured, searchable text documents. It utilizes speech-to-text recognition and external language models to perform transcription, summarization, and translation of media content. The system distinguishes itself through a modular pipeline that orchestrates media extraction, processing, and storage. It features automated media monitoring that tracks channels to compile periodic content digests, alongside a vector-based knowledge retrieval engine that allows users to query their stored tran

    Provides AI-driven summarization of long-form media content to extract key takeaways.

    Pythonaitooltiktoktranscribe
    在 GitHub 上查看↗2,799
  • any4ai/anycrawlany4ai 的头像

    any4ai/AnyCrawl

    2,742在 GitHub 上查看↗

    AnyCrawl is an AI-powered data extractor, automated web crawler, and headless browser orchestrator. It serves as a web content extraction API and a gateway that connects crawling and scraping tools to language models using a standardized API protocol. The project specializes in converting unstructured website content into structured JSON or markdown optimized for AI assistants. It utilizes language models and JSON schemas to pull specific information into validated formats and provides capabilities for AI page summarization and LLM-optimized content extraction. The system manages comprehensi

    Uses artificial intelligence to generate a concise abstract of a webpage's main information.

    TypeScriptai-scrapingaitoolscrawl
    在 GitHub 上查看↗2,742
  • spacecowboy/feederspacecowboy 的头像

    spacecowboy/Feeder

    2,641在 GitHub 上查看↗

    Feeder is an RSS and Atom feed reader that aggregates content into a single interface. It functions as a full-text content extractor that removes website clutter to isolate the main body of articles, and a self-hosted feed synchronizer that maintains subscription lists and read statuses across devices via a private backend server. The application integrates AI services and external API keys to translate and generate concise summaries of long-form articles. It also features a text-to-speech reader that uses system engines with automatic language detection to convert written content into spoken

    Uses AI services to generate concise summaries of long-form feed articles.

    Kotlinandroidatomjetpack-compose
    在 GitHub 上查看↗2,641
  • byjlw/video-analyzerbyjlw 的头像

    byjlw/video-analyzer

    1,464在 GitHub 上查看↗

    Video analyzer is a toolkit that processes video files through computer vision and automatic speech recognition to produce structured JSON data and natural language summaries. The system extracts visual frames, samples key moments based on pixel differences, and transcribes soundtrack audio into written text to generate comprehensive descriptions across chronological timelines. The software coordinates sequential processing stages that combine frame-by-frame visual analysis with audio transcripts using local or cloud AI models. It supports adaptive and uniform frame sampling, hardware-accele

    Synthesizes chronological frame analyses and audio transcripts into complete natural language video descriptions.

    Pythonasrllmsvideo
    在 GitHub 上查看↗1,464
  1. Home
  2. Web Development
  3. Custom Page Frameworks
  4. Content Summarization

探索子标签

  • AI-Powered Web Summarization2 个子标签Uses large language models to generate concise abstracts of a webpage's primary information. **Distinct from Content Summarization:** Specifies the use of AI/LLMs for synthesis rather than simple truncation or excerpting.
  • Audio and Video SummarizersAI tools that generate concise summaries from audio and video content. **Distinct from Content Summarization:** Distinct from Content Summarization: specifically targets audio/video media rather than text or web pages.
  • PaginationLogic for splitting long lists of content into multiple sequential pages. **Distinct from Content Summarization:** Distinct from Content Summarization: focuses on the structural division of content into pages rather than the generation of excerpts.
  • VideoGenerating textual summaries specifically from extracted video transcripts. **Distinct from Content Summarization:** Focuses on video-derived text summaries rather than generic page excerpt generation.