awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

62 个仓库

Awesome GitHub RepositoriesStreaming and Network Frameworks

Systems designed for real-time data transmission, network-based audio protocols, and engine-level streaming logic.

Explore 62 awesome GitHub repositories matching graphics & multimedia · Streaming and Network Frameworks. Refine with filters or upvote what's useful.

Awesome Streaming and Network Frameworks GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • sindresorhus/awesomesindresorhus 的头像

    sindresorhus/awesome

    476,211在 GitHub 上查看↗

    这是一个由社区维护的目录,作为软件工具、框架和教育资源的综合索引。它充当开源知识库,将不同的工程领域和技术资源组织成结构化的分类体系,以帮助开发者发现高质量内容。 该目录通过去中心化的同行评审模型脱颖而出,由独立贡献者策划、验证和更新条目,以确保准确性和相关性。所有信息均以版本控制的纯文本 Markdown 格式存储,确保了整个集合的平台独立性、透明度和可审计性。 该项目涵盖了广泛的能力领域,包括技术资源发现、职业发展和软件开发知识管理。它提供结构化的学习路径、基础设施和安全工具、数据管理实用程序,以及从医疗保健到数字人文等领域的专业资源。 该仓库作为公共版本控制集合进行维护,支持程序化访问和社区驱动的数据更新。

    Supports the transmission of low-latency audio signals across IP networks.

    awesomeawesome-listlists
    在 GitHub 上查看↗476,211
  • obsproject/obs-studioobsproject 的头像

    obsproject/obs-studio

    73,384在 GitHub 上查看↗

    This project is a professional live video production suite designed for capturing, encoding, and broadcasting high-quality media. At its core, it features a real-time media processing engine that utilizes hardware acceleration to composite multiple audio and video sources with minimal latency. The application provides a centralized studio interface for managing complex scene transitions, layering visual sources through a hierarchical scene-graph engine, and streaming content to multiple platforms simultaneously. The software is built on a cross-platform abstraction layer that ensures consiste

    Processes multiple audio and video inputs into a single, synchronized output stream in real time.

    Ccc-plus-plusdirectshow
    在 GitHub 上查看↗73,384
  • ffmpeg/ffmpegFFmpeg 的头像

    FFmpeg/FFmpeg

    61,176在 GitHub 上查看↗

    FFmpeg is a cross-platform multimedia framework designed for the recording, conversion, and streaming of audio and video content. It functions as a comprehensive toolkit that provides both a command-line utility for direct media manipulation and a collection of low-level libraries for integration into custom applications. At its core, the project utilizes a packet-based stream engine and a format-agnostic abstraction layer to handle diverse media standards, containers, and network protocols. The framework distinguishes itself through a modular, graph-based filter execution model that allows f

    Records audio and video directly from hardware devices, screen displays, or network streams for immediate processing or storage.

    Caudiocffmpeg
    在 GitHub 上查看↗61,176
  • microsoft/vibevoicemicrosoft 的头像

    microsoft/VibeVoice

    49,394在 GitHub 上查看↗

    VibeVoice is a generative artificial intelligence platform designed for text-to-speech synthesis. It functions as a neural audio generation framework that converts written text into natural-sounding spoken audio, specifically engineered to maintain consistent vocal characteristics and narrative prosody across extended passages of content. The system distinguishes itself through its ability to generate long-form conversational speech while preserving speaker identity and linguistic content. By utilizing latent space disentanglement, the model separates speaker traits from the input text, allow

    Enables streaming audio inference for real-time delivery of synthesized speech in interactive applications.

    Python
    在 GitHub 上查看↗49,394
  • predidit/kazumiPredidit 的头像

    Predidit/Kazumi

    26,597在 GitHub 上查看↗

    Kazumi is a cross-platform media player and streaming platform that centralizes video content from diverse third-party web sources. It functions as an automated scraping tool, utilizing configurable path patterns and selectors to extract and aggregate media streams into a unified interface. The platform distinguishes itself through its focus on synchronized group viewing and real-time state management. Users can participate in shared virtual rooms where playback progress and controls are aligned across multiple devices. Additionally, the application includes integrated image processing capabi

    The application applies real-time image processing to video streams to improve visual clarity and detail during playback for a better viewing experience.

    Dartandroidcross-platformdanmaku
    在 GitHub 上查看↗26,597
  • arendst/tasmotaarendst 的头像

    arendst/Tasmota

    24,502在 GitHub 上查看↗

    Tasmota is a universal firmware platform for ESP8266 and ESP32 microcontrollers, designed to provide local control and management of smart home hardware. It functions as an event-driven automation controller that replaces proprietary factory firmware, allowing users to manage relays, sensors, and lighting systems without relying on external cloud services. The system is built on a modular driver architecture that enables dynamic hardware configuration and peripheral support through a web-based management interface. The platform distinguishes itself through a template-driven hardware mapping s

    Transmits real-time audio between devices over UDP to create intercom systems.

    Carduinoautomationesp32
    在 GitHub 上查看↗24,502
  • spotdl/spotify-downloaderspotDL 的头像

    spotDL/spotify-downloader

    23,996在 GitHub 上查看↗

    Spotify-downloader is a command-line utility designed to archive music from Spotify by matching track URLs to external video sources. It functions as a high-fidelity downloader that retrieves audio content and saves it as local files, ensuring optimal sound quality by selecting the highest available bitrate from the source media. The tool distinguishes itself through its ability to maintain local music collections by mirroring remote playlist states. It performs local-remote synchronization to determine which tracks require downloading or removal, while utilizing a modular architecture to dec

    "Manages multiple simultaneous download and conversion tasks using asynchronous workers to maximize throughput and reduce total processing time."

    Pythondownload-musichacktoberfestmp3
    在 GitHub 上查看↗23,996
  • danielgatis/rembgdanielgatis 的头像

    danielgatis/rembg

    21,911在 GitHub 上查看↗

    Rembg is a machine learning-based toolkit designed for automated image background removal and subject segmentation. It functions as a versatile engine that identifies and extracts subjects from images, supporting diverse input methods including individual files, directory-based batch processing, and live binary data streams. The project distinguishes itself through its flexible integration options, offering a command-line interface for local automation, a library for programmatic access, and an HTTP service for remote requests. It utilizes deep learning architectures to classify pixels and ge

    Supports real-time processing of binary pixel data streams for dynamic background removal.

    Pythonbackground-removalimage-processingpython
    在 GitHub 上查看↗21,911
  • w-okada/voice-changerw-okada 的头像

    w-okada/voice-changer

    19,729在 GitHub 上查看↗

    This software is a real-time voice changer that utilizes machine learning inference to transform live microphone input into target vocal characteristics. It functions as an artificial intelligence audio processing tool designed to modify vocal identity during active communication or live broadcasts. The application distinguishes itself by executing neural network models directly within the browser environment. It leverages web-based compute acceleration and dedicated audio threading to maintain low-latency performance, allowing users to switch between different voice profiles while processing

    Injects processed audio directly into communication platforms by replacing standard microphone input streams.

    Python
    在 GitHub 上查看↗19,729
  • bluenviron/mediamtxB

    bluenviron/mediamtx

    19,226在 GitHub 上查看↗

    MediaMTX is a multi-protocol media server designed for routing, proxying, and recording real-time audio and video streams. It functions as a programmable media router and a gateway between streaming standards such as RTMP, RTSP, and WebRTC, enabling the conversion of live media between different protocols. The server distinguishes itself through on-the-fly format transmuxing and protocol-agnostic routing, which decouples input and output protocols via an internal media bus. It features a programmable automation system that executes external shell commands via event hooks triggered by client c

    Saves incoming live streams to disk using specific container formats for later playback.

    Go
    在 GitHub 上查看↗19,226
  • putyy/res-downloaderputyy 的头像

    putyy/res-downloader

    18,158在 GitHub 上查看↗

    Res-downloader is a network proxy utility designed to intercept, analyze, and extract multimedia assets from web traffic. It functions as a gateway that captures video, audio, and image files directly from data streams for local storage and offline access. The tool employs man-in-the-middle interception to decrypt and inspect network packets, allowing it to identify media resources through pattern matching and content type filtering. It integrates proxy-based routing to manage outgoing requests, enabling the retrieval of content that may be subject to regional restrictions or network-level ac

    Processes large media streams asynchronously to ensure efficient file reconstruction without blocking the main execution flow.

    Godouyinkuaishoures-downloader
    在 GitHub 上查看↗18,158
  • xifangczy/cat-catchxifangczy 的头像

    xifangczy/cat-catch

    18,106在 GitHub 上查看↗

    Cat-catch is a browser-based media utility designed to detect, capture, and manage web-based video and audio resources. It functions as a comprehensive sniffing and download management system, enabling users to identify hidden or protected media assets directly from active web pages. The tool specializes in reconstructing fragmented streaming protocols, such as DASH and M3U8, into complete files while providing options for real-time stream recording and playback control. The project distinguishes itself through its deep integration with local system environments and external automation tools.

    Provides frameworks for capturing, encoding, and processing audio and video data streams in real time.

    JavaScriptchromechrome-extensionfirefox
    在 GitHub 上查看↗18,106
  • ffmpegwasm/ffmpeg.wasmffmpegwasm 的头像

    ffmpegwasm/ffmpeg.wasm

    17,184在 GitHub 上查看↗

    ffmpeg.wasm is a browser-based multimedia processing engine that brings the capabilities of the FFmpeg library directly to the client environment. By utilizing WebAssembly, it enables audio and video transcoding, format conversion, and stream recording to occur entirely within the browser without requiring server-side infrastructure. The library distinguishes itself by executing resource-intensive media tasks in background threads, ensuring that the main user interface remains responsive during complex operations. It manages data through an isolated, in-memory virtual file system, allowing fo

    Captures audio and video input from browser sources and encodes the data into files for local storage.

    Caudioexperimental-featuresffmpeg
    在 GitHub 上查看↗17,184
  • pion/webrtcpion 的头像

    pion/webrtc

    16,571在 GitHub 上查看↗

    This project is a cross-platform implementation of the WebRTC standard, providing a comprehensive library for building real-time audio, video, and data communication applications. It functions as a peer-to-peer networking framework and media processing engine, enabling direct, low-latency connections between devices without relying on central servers. By strictly adhering to official protocol specifications, the library ensures interoperability with browsers and other native communication software across mobile, desktop, and server environments. The engine distinguishes itself through a modul

    Captures and transmits local audio and video content across peer connections using portable interfaces for consistent media handling.

    Goaudiogogolang
    在 GitHub 上查看↗16,571
  • imagemagick/imagemagickImageMagick 的头像

    ImageMagick/ImageMagick

    15,742在 GitHub 上查看↗

    ImageMagick is a comprehensive software suite for the creation, editing, composition, and conversion of digital images. It functions as both a command-line utility for batch processing and automation, and as a programming library that allows developers to integrate advanced image manipulation capabilities into external applications. The project is distinguished by its modular architecture, which supports hundreds of image formats through a pluggable coder system and external delegate libraries. It is designed for high-performance environments, utilizing memory-mapped pixel caching, stream-ori

    Improves local detail and edge definition using contrast-limited adaptive histogram equalization.

    Ccommand-line-image-tooldigital-image-editingimage-conversion
    在 GitHub 上查看↗15,742
  • vercel/vercelvercel 的头像

    vercel/vercel

    15,738在 GitHub 上查看↗

    Vercel is a cloud platform for building, deploying, and scaling web applications. It provides a unified infrastructure that automates the build process by detecting project frameworks and distributing static and dynamic content through a global content delivery network. The platform executes application logic using serverless functions that scale automatically based on real-time traffic demand. The platform distinguishes itself through a centralized AI gateway that proxies requests to multiple model providers, enabling standardized authentication, observability, and cost tracking. It supports

    Enables real-time processing of visual data streams for interactive applications.

    TypeScriptclicloudcommand
    在 GitHub 上查看↗15,738
  • webrtc/sampleswebrtc 的头像

    webrtc/samples

    14,624在 GitHub 上查看↗

    This repository provides a collection of reference implementations and practical demonstrations for using WebRTC to establish real-time audio, video, and data communication. It contains code samples for negotiating peer-to-peer connections, managing media streams, and utilizing low-latency data channels. The project demonstrates the capture of audio and video from hardware devices, as well as the redirection of canvas element content into media streams. It includes examples of transferring arbitrary text and binary data between peers and managing the negotiation of direct connections. The sa

    Demonstrates routing raw media frames through processing layers before transmission or playback.

    JavaScript
    在 GitHub 上查看↗14,624
  • alibaba/mnnalibaba 的头像

    alibaba/MNN

    14,242在 GitHub 上查看↗

    MNN is a high-performance inference engine and framework designed for on-device machine learning. It provides a comprehensive environment for executing, optimizing, and deploying neural network models directly on mobile and resource-constrained edge devices. The framework distinguishes itself through a robust model optimization toolkit that supports quantization, compression, and structural graph manipulation to minimize memory footprint and maximize execution speed. It features a modular architecture that abstracts hardware-specific backends, allowing models to run efficiently across diverse

    Executes lightweight image processing and codec functions to prepare visual data for high-performance inference pipelines.

    C++armconvolutiondeep-learning
    在 GitHub 上查看↗14,242
  • arut/nginx-rtmp-modulearut 的头像

    arut/nginx-rtmp-module

    13,982在 GitHub 上查看↗

    The NGINX RTMP module is a server-side extension that functions as a live video streaming engine. It enables the ingestion, processing, and distribution of real-time audio and video feeds, supporting both RTMP and HLS protocols to facilitate media delivery to multiple clients. The module distinguishes itself by integrating directly into the host server event loop, allowing for high-concurrency network input and output without blocking the main thread. It provides a toolkit for managing media streams through event-driven callbacks, which can trigger external process invocations for custom tran

    Processes incoming media packets through event-driven callbacks triggered by stream lifecycle events.

    C
    在 GitHub 上查看↗13,982
  • mltframework/shotcutmltframework 的头像

    mltframework/shotcut

    13,460在 GitHub 上查看↗

    Shotcut is a professional-grade, cross-platform non-linear video editor built on the MLT multimedia framework. It provides a comprehensive suite for post-production, supporting multi-track timeline editing, high-fidelity color processing, and complex visual effects. The application is designed to handle diverse audio and video formats natively, ensuring high-resolution and HDR workflows are managed within a unified environment. The software distinguishes itself through a modular architecture that emphasizes performance and precision. It utilizes a GPU-accelerated rendering pipeline and proxy-

    Captures video and audio from hardware devices, network streams, and screen inputs for direct integration into editing workflows.

    C++cross-platformgplv3mlt
    在 GitHub 上查看↗13,460
上一个123…4下一个
  1. Home
  2. Graphics & Multimedia
  3. Media Processing and Analysis
  4. Media Manipulation
  5. Media Processing
  6. Streaming and Network Frameworks

探索子标签

  • Audio Over IP1 个子标签Resources and frameworks for transmitting audio data over internet protocol networks.
  • Media Stream Processing7 个子标签Frameworks for capturing, encoding, and processing audio and video data streams in real time.
  • Media Streaming Engines1 个子标签Engines that stream audio and video content over networks while performing real-time format conversion.
  • NDI Protocol StreamingReal-time transmission of high-quality audio and video using the Network Device Interface protocol. **Distinct from Streaming and Network Frameworks:** Specifically implements the NDI protocol for professional broadcast streaming rather than general streaming frameworks.