awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

121 个仓库

Awesome GitHub RepositoriesVideo Analysis and Processing

Specialized tools for frame-level manipulation, metadata retrieval, and hardware-accelerated video pipeline management.

Explore 121 awesome GitHub repositories matching graphics & multimedia · Video Analysis and Processing. Refine with filters or upvote what's useful.

Awesome Video Analysis and Processing GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • obsproject/obs-studioobsproject 的头像

    obsproject/obs-studio

    73,384在 GitHub 上查看↗

    This project is a professional live video production suite designed for capturing, encoding, and broadcasting high-quality media. At its core, it features a real-time media processing engine that utilizes hardware acceleration to composite multiple audio and video sources with minimal latency. The application provides a centralized studio interface for managing complex scene transitions, layering visual sources through a hierarchical scene-graph engine, and streaming content to multiple platforms simultaneously. The software is built on a cross-platform abstraction layer that ensures consiste

    Leverages dedicated graphics hardware to perform real-time encoding, compositing, and filtering of video streams with minimal latency.

    Ccc-plus-plusdirectshow
    在 GitHub 上查看↗73,384
  • ffmpeg/ffmpegFFmpeg 的头像

    FFmpeg/FFmpeg

    61,176在 GitHub 上查看↗

    FFmpeg is a cross-platform multimedia framework designed for the recording, conversion, and streaming of audio and video content. It functions as a comprehensive toolkit that provides both a command-line utility for direct media manipulation and a collection of low-level libraries for integration into custom applications. At its core, the project utilizes a packet-based stream engine and a format-agnostic abstraction layer to handle diverse media standards, containers, and network protocols. The framework distinguishes itself through a modular, graph-based filter execution model that allows f

    Delegates intensive encoding and decoding tasks to dedicated graphics or video hardware to improve performance.

    Caudiocffmpeg
    在 GitHub 上查看↗61,176
  • deepfakes/faceswapdeepfakes 的头像

    deepfakes/faceswap

    55,289在 GitHub 上查看↗

    Faceswap is a comprehensive framework for automated media manipulation and neural face synthesis. It provides a modular pipeline that manages the entire lifecycle of facial feature extraction, deep learning model training, and image conversion. By coordinating complex computer vision workflows, the system enables users to map facial identities between source and destination datasets while maintaining structural alignment and lighting consistency across video frames. The project distinguishes itself through a highly extensible plugin-based architecture that handles hardware-accelerated process

    Automates the extraction of video frames, metadata retrieval, and the reconstruction of video files.

    Pythondeep-face-swapdeep-learningdeep-neural-networks
    在 GitHub 上查看↗55,289
  • lizardbyte/sunshineLizardByte 的头像

    LizardByte/Sunshine

    38,332在 GitHub 上查看↗

    Sunshine is a self-hosted remote desktop and game streaming server designed to broadcast desktop environments and applications over a network. It functions as a host application that captures system display output and encodes it into low-latency video streams for transmission to remote client devices. The system distinguishes itself through hardware-accelerated media encoding, which utilizes graphics processor pipelines to compress high-resolution video in real time. To ensure interactive control, it performs virtual input emulation by translating remote controller and keyboard signals into n

    Encodes and transmits high-resolution video streams using graphics processor acceleration to maintain performance and responsiveness during remote sessions.

    C++cppdockerflathub-pkg
    在 GitHub 上查看↗38,332
  • panniantong/agent-reachPanniantong 的头像

    Panniantong/Agent-Reach

    31,610在 GitHub 上查看↗

    Agent-Reach is an AI agent web gateway and search tool that provides language models with the ability to search and read content from the open web, social media, and community forums without using official APIs. It functions as a routing layer that connects large language models to various internet backends while managing content parsing and connection health. The system enables API-free information retrieval by using open-source backends to extract text and metadata from platforms such as Twitter, Reddit, and YouTube. It converts unstructured website content, RSS feeds, and video transcripts

    Retrieves subtitles and metadata from video platforms to enable content summarization.

    Pythonagent-infrastructureai-agentai-search
    在 GitHub 上查看↗31,610
  • iawia002/annieiawia002 的头像

    iawia002/annie

    31,414在 GitHub 上查看↗

    Annie is a command-line video downloader and web video extraction library written in Go. It functions as a concurrent media downloader designed to fetch video files and playlists from websites via URLs. The tool distinguishes itself through a proxy-aware network layer that supports SOCKS5 and HTTP proxies to bypass regional content restrictions. It also incorporates session cookie integration and referrer spoofing to facilitate the download of authenticated or age-gated content. The project provides capabilities for bulk media acquisition, including batch downloading from text files and extr

    Extracts and saves video files from websites to a local device.

    Go
    在 GitHub 上查看↗31,414
  • iawia002/luxiawia002 的头像

    iawia002/lux

    31,412在 GitHub 上查看↗

    Lux is a command line video downloader written in Go designed for extracting and saving video and audio from various websites. It functions as a concurrent media downloader that increases transfer speeds by splitting files into fragments and downloading them using multiple threads. The tool serves as a playlist download manager capable of retrieving entire video collections or specific ranges of items. It also operates as a proxy-enabled media client, supporting HTTP and SOCKS5 proxies and session cookies to access region-locked, private, or age-gated content. Additional capabilities include

    Retrieves videos, images, and audio from supported websites using a command line interface.

    Gobilibilicrawlerdownload
    在 GitHub 上查看↗31,412
  • facefusion/facefusionfacefusion 的头像

    facefusion/facefusion

    28,806在 GitHub 上查看↗

    Facefusion is a modular framework designed for automated image and video manipulation, specializing in tasks such as face swapping, enhancement, and restoration. It functions as a computer vision processing pipeline that chains independent machine learning modules to perform complex transformations, including facial animation, age modification, and lip synchronization. The system is built to handle both real-time interactive feeds and large-scale batch processing tasks. The platform distinguishes itself through a highly extensible architecture that supports custom processing modules and inter

    Enables complex media transformation tasks through headless command-line operations.

    Pythonaideep-fakedeepfake
    在 GitHub 上查看↗28,806
  • maaassistantarknights/maaassistantarknightsMaaAssistantArknights 的头像

    MaaAssistantArknights/MaaAssistantArknights

    21,583在 GitHub 上查看↗

    MaaAssistantArknights is a cross-platform automation engine designed for mobile games, utilizing computer vision and input simulation to perform routine tasks. It functions as an Android emulator controller, managing game lifecycles, resource farming, and infrastructure optimization through structured, scripted workflows. The project distinguishes itself through a modular configuration system that allows users to define complex automation logic via external instruction files. This framework supports dynamic task modification, configuration inheritance, and schema validation, ensuring that cus

    Provides a command-line interface for programmatic game task execution and status callbacks.

    C++arknightscomputer-visionmaa
    在 GitHub 上查看↗21,583
  • xbmc/xbmcxbmc 的头像

    xbmc/xbmc

    20,868在 GitHub 上查看↗

    This project is a cross-platform media center, player, and digital media library manager. It serves as a centralized home theater hub for organizing, managing, and playing digital audio and video files across multiple operating systems. The application features a skinable media interface designed for remote control and ten-foot interface optimization. This is supported by a skinning engine that separates visual layout from application logic, allowing for custom user interface designs. The system provides automated media library organization by scanning folders to generate structured database

    Automates the processing and organization of media libraries, including the generation of structured databases with cover art.

    C++androidc-plus-plusentertainment-hub
    在 GitHub 上查看↗20,868
  • iv-org/invidiousiv-org 的头像

    iv-org/invidious

    20,471在 GitHub 上查看↗

    Invidious is a privacy-focused, self-hosted alternative frontend for mainstream video platforms. It operates as a decentralized network of independent instances that provide a lightweight, ad-free interface for consuming media. By acting as a proxy between the user and the content provider, the platform prevents tracking and data collection while maintaining a familiar browsing experience. The project distinguishes itself through its robust suite of network-level traffic management and anonymization tools. It employs techniques such as IP rotation, reverse proxy stream routing, and integratio

    Fetches detailed video metadata including duration, resolution, and caption availability.

    Crystalagplv3hacktoberfestinvidious
    在 GitHub 上查看↗20,471
  • bradlarson/gpuimageBradLarson 的头像

    BradLarson/GPUImage

    20,299在 GitHub 上查看↗

    GPUImage is a GPU-accelerated image processing framework for iOS designed to apply real-time filters and effects to images and video. It functions as a processing engine and fragment shader library that manages textures and shaders for efficient visual data manipulation. The framework utilizes a chainable filter architecture and a texture-based data pipeline to pass image data between processing stages without expensive memory transfers. It enables the creation of bespoke visual effects through the authoring of custom fragment shaders and provides mechanisms to synchronize texture data with e

    Loads movie files from disk to apply a sequence of filters to frames and encode the results into a new video file.

    Objective-C
    在 GitHub 上查看↗20,299
  • k4yt3x/video2xk4yt3x 的头像

    k4yt3x/video2x

    18,754在 GitHub 上查看↗

    Video2x is a modular processing framework designed for AI-enhanced video upscaling and frame rate conversion. It functions as a comprehensive toolset for increasing the resolution and visual clarity of media files while generating intermediate frames to improve motion smoothness. The system is built to handle intensive media transformation tasks by leveraging hardware acceleration and custom encoding pipelines. The project distinguishes itself through a plugin-based architecture that allows for the integration of custom machine learning models and specialized algorithms. It utilizes a modular

    Utilizes dedicated graphics hardware to accelerate intensive video transformation and rendering pipelines.

    C++anime4kframe-interpolationmachine-learning
    在 GitHub 上查看↗18,754
  • chidiwilliams/buzzchidiwilliams 的头像

    chidiwilliams/buzz

    17,903在 GitHub 上查看↗

    Buzz is a desktop application that provides a local speech-to-text engine for transcribing and translating audio and video files. By leveraging local machine inference, the software ensures data privacy and offline performance, removing the need for cloud connectivity during media processing. The application distinguishes itself through a modular plugin architecture that allows for the integration of custom functionality, such as content summarization and automated text formatting, without modifying the core codebase. It also features a speaker diarization pipeline that identifies and labels

    Automates transcription and translation tasks by monitoring directories for new media assets.

    Pythonwhisper
    在 GitHub 上查看↗17,903
  • capsoftware/capCapSoftware 的头像

    CapSoftware/Cap

    17,026在 GitHub 上查看↗

    Cap is a self-hosted screen recording and video collaboration platform designed for teams to replace synchronous meetings with asynchronous video updates. It provides a comprehensive suite for capturing high-resolution desktop activity, including system audio, microphone input, and camera overlays, which are then processed through an integrated post-production workflow. The platform distinguishes itself by offering full data sovereignty through containerized deployment and object storage abstractions, allowing users to host their media assets on private infrastructure or S3-compatible buckets

    Automates the organization, distribution, and management of video media libraries.

    TypeScriptappcapcoss
    在 GitHub 上查看↗17,026
  • aaronfeng753/waifu2x-extension-guiAaronFeng753 的头像

    AaronFeng753/Waifu2x-Extension-GUI

    16,146在 GitHub 上查看↗

    Waifu2x-Extension-GUI is a desktop application designed for high-fidelity media restoration and enhancement. It functions as a graphical interface that orchestrates specialized deep learning engines to upscale, denoise, and interpolate images and videos, improving visual clarity and motion smoothness. The software distinguishes itself through its ability to manage complex, automated media processing pipelines. Users can chain multiple tasks—such as format conversion, scene detection, and frame rate interpolation—into sequential workflows that execute without manual intervention. It provides g

    Restores high-fidelity visual quality in grainy or low-resolution media using specialized deep learning engines.

    C++animeanime4kesrgan
    在 GitHub 上查看↗16,146
  • zulko/moviepyZulko 的头像

    Zulko/moviepy

    14,699在 GitHub 上查看↗

    MoviePy is a Python video editing library and automated video processor designed for programmatically cutting, concatenating, and manipulating video and audio files. It serves as a non-linear video editor and an interface for FFmpeg to handle the reading, writing, and conversion of diverse media formats and codecs. The library enables automated video composition through the layering of multiple video and audio streams using transparency and coordinate-based positioning. It supports dynamic content generation by inserting text overlays and performing custom video frame processing where raw fra

    Ships a scriptable framework for batch processing video files and applying pixel-level effects.

    Pythonanimationgifhacktoberfest
    在 GitHub 上查看↗14,699
  • chocobozzz/peertubeChocobozzz 的头像

    Chocobozzz/PeerTube

    14,520在 GitHub 上查看↗

    PeerTube is a decentralized, open-source video hosting platform that enables users to operate independent, interoperable servers. By utilizing the ActivityPub protocol, it connects these servers into a global, federated network where users can follow channels, discover content, and interact across different instances. The platform is designed to function as a self-hosted video content management system, providing a community-driven alternative to centralized media services. What distinguishes PeerTube is its hybrid approach to content delivery and infrastructure management. It integrates peer

    Supports uploading, modifying, and merging video streams with automated editing capabilities.

    TypeScriptactivitypubangulardecentralized
    在 GitHub 上查看↗14,520
  • clsid2/mpc-hcclsid2 的头像

    clsid2/mpc-hc

    14,378在 GitHub 上查看↗

    This project is an open-source multimedia player for Windows designed for high-performance audio and video playback. It functions as a DirectShow-based media renderer that utilizes hardware-accelerated graphics APIs to perform color space conversion and high-quality scaling directly on the display adapter. The application distinguishes itself through granular control over playback dynamics and visual output. Users can manipulate video orientation through rotation, flipping, and zooming, while also leveraging support for high dynamic range rendering. The player supports automated playback sequ

    Utilizes dedicated graphics hardware to accelerate real-time video rendering and color space conversion.

    C++
    在 GitHub 上查看↗14,378
  • lucksiege/pictureselectorLuckSiege 的头像

    LuckSiege/PictureSelector

    13,592在 GitHub 上查看↗

    PictureSelector is an Android media selection library and toolkit for browsing and picking images, videos, and audio files from a device album. It provides a comprehensive framework for capturing new photos and videos via system hardware, extracting media metadata, and managing the resulting files. The library features a modular architecture that allows for custom media engine implementations to replace default image loading, file compression, and video playback logic. It offers extensive UI customization, enabling the replacement of default layout resources and theme configurations to modify

    Reduces the file size of chosen images or videos using a custom compression engine.

    Javaandroidandroid-image-selectorcamera
    在 GitHub 上查看↗13,592
上一个123456…7下一个
  1. Home
  2. Graphics & Multimedia
  3. Media Processing and Analysis
  4. Media Manipulation
  5. Media Processing
  6. Video Analysis and Processing

探索子标签

  • Hardware-Accelerated Video Pipelines3 个子标签Video processing pipelines that utilize dedicated graphics hardware to accelerate real-time encoding and compositing tasks.
  • Media Automation4 个子标签Tools for automating the processing, organization, and management of media libraries. **Distinct from Video Analysis and Processing:** Distinct from Video Analysis: focuses on automated library maintenance and episode processing rather than frame-level analysis.
  • Semantic Video Understanding ToolsTools that use vision-language AI models to understand video content, extracting scene descriptions, key events, and summaries. **Distinct from Video Analysis and Processing:** Distinct from Video Analysis and Processing: focuses on semantic understanding and natural language descriptions of video content, rather than low-level frame manipulation or metadata extraction.
  • Video File Processors6 个子标签Tools for extracting frames, generating video from frames, and retrieving metadata from video files.
  • Video Metadata Extraction6 个子标签Utilities for retrieving technical information from video files, including frame counts, duration, and keyframe data.
  • Video Metadata RetrievalFetching detailed information about a video, including available formats, quality labels, and thumbnails. **Distinct from Video Analysis and Processing:** Focuses on retrieving public-facing metadata from platforms rather than frame-level pixel analysis or hardware pipeline management.
  • Video Metadata Scraping1 个子标签Extracting unique identifiers from filenames to aggregate detailed content information from external web databases. **Distinct from Video Metadata Extraction:** Distinct from technical video extraction; it focuses on scraping descriptive web data based on filename identifiers.
  • Video Project File ManipulatorsTools for programmatically modifying the internal structure and properties of video project files. **Distinct from Video File Processors:** Distinct from Video File Processors: modifies the project metadata/timeline rather than processing raw video frames.