awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

47 个仓库

Awesome GitHub RepositoriesCodec and Encoding Utilities

Tools focused on low-level stream compression, container manipulation, and hardware-accelerated encoding configurations.

Explore 47 awesome GitHub repositories matching graphics & multimedia · Codec and Encoding Utilities. Refine with filters or upvote what's useful.

Awesome Codec and Encoding Utilities GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • ffmpeg/ffmpegFFmpeg 的头像

    FFmpeg/FFmpeg

    61,176在 GitHub 上查看↗

    FFmpeg is a cross-platform multimedia framework designed for the recording, conversion, and streaming of audio and video content. It functions as a comprehensive toolkit that provides both a command-line utility for direct media manipulation and a collection of low-level libraries for integration into custom applications. At its core, the project utilizes a packet-based stream engine and a format-agnostic abstraction layer to handle diverse media standards, containers, and network protocols. The framework distinguishes itself through a modular, graph-based filter execution model that allows f

    Extends format and codec support by linking against third-party libraries for specialized encoding, decoding, or processing tasks.

    Caudiocffmpeg
    在 GitHub 上查看↗61,176
  • deepfakes/faceswapdeepfakes 的头像

    deepfakes/faceswap

    55,289在 GitHub 上查看↗

    Faceswap is a comprehensive framework for automated media manipulation and neural face synthesis. It provides a modular pipeline that manages the entire lifecycle of facial feature extraction, deep learning model training, and image conversion. By coordinating complex computer vision workflows, the system enables users to map facial identities between source and destination datasets while maintaining structural alignment and lighting consistency across video frames. The project distinguishes itself through a highly extensible plugin-based architecture that handles hardware-accelerated process

    Combines processed video streams and audio tracks into final files using configurable codec settings.

    Pythondeep-face-swapdeep-learningdeep-neural-networks
    在 GitHub 上查看↗55,289
  • lizardbyte/sunshineLizardByte 的头像

    LizardByte/Sunshine

    38,332在 GitHub 上查看↗

    Sunshine is a self-hosted remote desktop and game streaming server designed to broadcast desktop environments and applications over a network. It functions as a host application that captures system display output and encodes it into low-latency video streams for transmission to remote client devices. The system distinguishes itself through hardware-accelerated media encoding, which utilizes graphics processor pipelines to compress high-resolution video in real time. To ensure interactive control, it performs virtual input emulation by translating remote controller and keyboard signals into n

    Utilizes dedicated graphics processor pipelines to compress high-resolution video frames into efficient network-ready bitstreams in real time.

    C++cppdockerflathub-pkg
    在 GitHub 上查看↗38,332
  • qarmin/czkawkaqarmin 的头像

    qarmin/czkawka

    31,526在 GitHub 上查看↗

    Czkawka is a cross-platform utility designed for storage optimization and filesystem maintenance. It functions as a comprehensive file analysis engine that identifies redundant data, including duplicate files, empty directories, broken symbolic links, and temporary files. By utilizing hash-based content verification, the tool ensures accurate identification of duplicates regardless of file names or metadata. The project distinguishes itself by offering both a native graphical user interface and a command-line interface, allowing for both interactive management and automated, headless system m

    Re-encodes video files using efficient codecs and crops static bars to reduce total disk usage.

    Fluentcleanerduplicatesmultiplatform
    在 GitHub 上查看↗31,526
  • ossrs/srsossrs 的头像

    ossrs/srs

    28,971在 GitHub 上查看↗

    SRS is a real-time media server designed to ingest, route, and distribute live audio and video streams across various transport protocols. It functions as a multi-protocol stream relay, including a dedicated RTMP media gateway and a WebRTC signaling server to coordinate peer-to-peer media exchanges. The system features a multi-protocol relay engine that transforms incoming media packets between different transport formats without re-encoding. This allows it to serve as a video delivery proxy that routes live media from a single source to multiple concurrent viewers using diverse delivery prot

    Processes various audio and video encoding formats through a unified internal pipeline to maintain device compatibility.

    C++audiocc-plus-plus
    在 GitHub 上查看↗28,971
  • handbrake/handbrakeHandBrake 的头像

    HandBrake/HandBrake

    22,419在 GitHub 上查看↗

    HandBrake is an open-source media converter and video transcoding application designed to process digital video and audio files. It functions as a desktop utility that converts media from nearly any format into widely supported codecs, facilitating video format conversion and the optimization of files for specific playback requirements. The software serves as a tool for digital media archiving, allowing users to compress and preserve high-quality video into manageable formats. It also functions as a DVD and Blu-ray ripper, enabling the extraction and conversion of content from physical optica

    Uses core libraries to decode, filter, and re-encode digital video and audio streams.

    Cgplv2multi-platformvideo-transcoding
    在 GitHub 上查看↗22,419
  • iv-org/invidiousiv-org 的头像

    iv-org/invidious

    20,471在 GitHub 上查看↗

    Invidious is a privacy-focused, self-hosted alternative frontend for mainstream video platforms. It operates as a decentralized network of independent instances that provide a lightweight, ad-free interface for consuming media. By acting as a proxy between the user and the content provider, the platform prevents tracking and data collection while maintaining a familiar browsing experience. The project distinguishes itself through its robust suite of network-level traffic management and anonymization tools. It employs techniques such as IP rotation, reverse proxy stream routing, and integratio

    Combines separate audio and video tracks on the server side to enable high-quality playback without client-side tracking.

    Crystalagplv3hacktoberfestinvidious
    在 GitHub 上查看↗20,471
  • 11ty/eleventy11ty 的头像

    11ty/eleventy

    19,670在 GitHub 上查看↗

    Eleventy is a JavaScript-based static site generator designed to transform templates, data files, and markdown into optimized HTML. It functions as a versatile template rendering engine and content management framework, allowing developers to aggregate data from diverse sources—including local files, databases, and external APIs—to populate structured web content. The project is distinguished by its template-engine-agnostic pipeline, which decouples the build process from specific rendering languages. This allows users to integrate multiple template formats, such as Liquid, Nunjucks, Handleba

    Automates image processing, lazy loading, and responsive asset serving to improve page load performance.

    JavaScriptblog-enginedocumentation-tooleleventy
    在 GitHub 上查看↗19,670
  • k4yt3x/video2xk4yt3x 的头像

    k4yt3x/video2x

    18,754在 GitHub 上查看↗

    Video2x is a modular processing framework designed for AI-enhanced video upscaling and frame rate conversion. It functions as a comprehensive toolset for increasing the resolution and visual clarity of media files while generating intermediate frames to improve motion smoothness. The system is built to handle intensive media transformation tasks by leveraging hardware acceleration and custom encoding pipelines. The project distinguishes itself through a plugin-based architecture that allows for the integration of custom machine learning models and specialized algorithms. It utilizes a modular

    Integrates specialized hardware encoders to optimize output quality and compression efficiency.

    C++anime4kframe-interpolationmachine-learning
    在 GitHub 上查看↗18,754
  • alyssaxuu/screenityalyssaxuu 的头像

    alyssaxuu/screenity

    18,321在 GitHub 上查看↗

    Screenity is a browser-based screen recorder designed to capture screen activity and audio directly within a web browser. It functions as a privacy-focused capture tool that handles data locally and includes a web-based video editor for basic media refinement. The project distinguishes itself through real-time screen annotation tools, allowing users to draw shapes and arrows or zoom into specific areas during a recording. It also provides specialized privacy controls to blur sensitive information and apply backgrounds to camera feeds. The tool covers a broad range of media capabilities, incl

    Converts captured streams into downloadable video formats using the browser's internal encoding capabilities.

    JavaScriptannotationannotation-toolaudio
    在 GitHub 上查看↗18,321
  • jianchang512/pyvideotransjianchang512 的头像

    jianchang512/pyvideotrans

    17,991在 GitHub 上查看↗

    Pyvideotrans is an automated video localization platform designed to transcribe, translate, and dub media content for international distribution. It functions as an end-to-end workflow that combines speech recognition, text translation, and synthetic voice generation to process video files into localized versions. The system distinguishes itself by offering a choice between local model inference for privacy and integration with third-party cloud services via user-provided credentials. This architecture allows users to maintain control over their billing and data security while utilizing modul

    Maintains visual integrity by bypassing re-encoding processes during video export.

    Pythonspeech-to-texttext-to-speechvideo-transition
    在 GitHub 上查看↗17,991
  • pytorch/visionpytorch 的头像

    pytorch/vision

    17,743在 GitHub 上查看↗

    This project is a comprehensive computer vision library for the PyTorch ecosystem, providing a standardized collection of neural network architectures, datasets, and high-performance transformation utilities. It serves as a foundational framework for building, training, and deploying deep learning models, offering a centralized model registry that allows developers to instantiate architectures with pre-trained weights for tasks such as image classification, object detection, and semantic segmentation. The library distinguishes itself through its modular approach to data and compute management

    Allows selection of underlying image and video decoding libraries to optimize performance based on system requirements.

    Pythoncomputer-visionmachine-learning
    在 GitHub 上查看↗17,743
  • aaronfeng753/waifu2x-extension-guiAaronFeng753 的头像

    AaronFeng753/Waifu2x-Extension-GUI

    16,146在 GitHub 上查看↗

    Waifu2x-Extension-GUI is a desktop application designed for high-fidelity media restoration and enhancement. It functions as a graphical interface that orchestrates specialized deep learning engines to upscale, denoise, and interpolate images and videos, improving visual clarity and motion smoothness. The software distinguishes itself through its ability to manage complex, automated media processing pipelines. Users can chain multiple tasks—such as format conversion, scene detection, and frame rate interpolation—into sequential workflows that execute without manual intervention. It provides g

    Utilizes specialized hardware to accelerate resource-intensive media processing and enhancement tasks.

    C++animeanime4kesrgan
    在 GitHub 上查看↗16,146
  • c4illin/convertxC4illin 的头像

    C4illin/ConvertX

    15,905在 GitHub 上查看↗

    ConvertX is a web-based file conversion management platform designed to transform documents, images, and video files between various formats. It utilizes system-level binary orchestration to execute conversion tasks, leveraging background worker threads to handle concurrent, high-volume bulk processing without blocking the user interface. The platform distinguishes itself through a comprehensive security and access control framework, which includes multi-user account management, session-based token authentication, and role-based permissions. Users can secure their output files with passwords

    Allows granular configuration of media encoding parameters for specific processing requirements.

    TypeScriptbunconversionconvert
    在 GitHub 上查看↗15,905
  • clsid2/mpc-hcclsid2 的头像

    clsid2/mpc-hc

    14,378在 GitHub 上查看↗

    This project is an open-source multimedia player for Windows designed for high-performance audio and video playback. It functions as a DirectShow-based media renderer that utilizes hardware-accelerated graphics APIs to perform color space conversion and high-quality scaling directly on the display adapter. The application distinguishes itself through granular control over playback dynamics and visual output. Users can manipulate video orientation through rotation, flipping, and zooming, while also leveraging support for high dynamic range rendering. The player supports automated playback sequ

    Integrates modular codec libraries to support a wide range of audio and video formats.

    C++
    在 GitHub 上查看↗14,378
  • mltframework/shotcutmltframework 的头像

    mltframework/shotcut

    13,460在 GitHub 上查看↗

    Shotcut is a professional-grade, cross-platform non-linear video editor built on the MLT multimedia framework. It provides a comprehensive suite for post-production, supporting multi-track timeline editing, high-fidelity color processing, and complex visual effects. The application is designed to handle diverse audio and video formats natively, ensuring high-resolution and HDR workflows are managed within a unified environment. The software distinguishes itself through a modular architecture that emphasizes performance and precision. It utilizes a GPU-accelerated rendering pipeline and proxy-

    Converts media files between various formats and codecs while utilizing graphics hardware to improve processing speed.

    C++cross-platformgplv3mlt
    在 GitHub 上查看↗13,460
  • owncast/owncastowncast 的头像

    owncast/owncast

    10,950在 GitHub 上查看↗

    Owncast is a self-hosted live streaming server that provides full control over broadcast infrastructure and audience data. It functions as an RTMP video streaming server, accepting incoming video feeds and distributing them to viewers through HLS-based segmented streaming. The platform includes a built-in, stateful web-based chat interface that enables real-time viewer engagement during broadcasts. The project distinguishes itself through deep integration with the decentralized Fediverse, allowing servers to automatically broadcast stream status updates and notify followers across distributed

    Utilizes dedicated hardware to perform video encoding tasks, reducing processor load and improving performance during live broadcasts.

    Goactivitypubbroadcastingchat
    在 GitHub 上查看↗10,950
  • imgproxy/imgproxyimgproxy 的头像

    imgproxy/imgproxy

    10,876在 GitHub 上查看↗

    This project is a high-performance image transformation server and media optimization proxy designed to process, resize, and convert assets on the fly. It functions as a secure pipeline that fetches remote source files and applies transformations—such as cropping, watermarking, and visual filtering—directly through parameters defined in the request URL. The service distinguishes itself through a focus on secure, resource-aware delivery. It protects infrastructure by validating incoming requests with cryptographic signatures to prevent unauthorized access and enforces strict limits on file dim

    Acts as a proxy to reduce image sizes and select optimal formats based on browser support.

    Goavifcrop-imagedocker
    在 GitHub 上查看↗10,876
  • pikvm/pikvmpikvm 的头像

    pikvm/pikvm

    10,109在 GitHub 上查看↗

    PiKVM is an IP-KVM remote access solution for out-of-band server management. It enables remote control of a host computer's hardware and operating system through a web browser or VNC client, allowing access to the BIOS or boot sequence when the main operating system is unresponsive. The system provides remote hardware control by managing physical power switches and GPIO triggers. It emulates human interface devices, such as keyboards and mice, and supports the emulation of mass storage to mount virtual disk images for bare metal operating system installation and recovery. Additional capabili

    Uses on-chip hardware acceleration to encode HDMI capture streams for low-latency remote viewing.

    atxhardwarehdmi
    在 GitHub 上查看↗10,109
  • aandrew-me/ytdownloaderaandrew-me 的头像

    aandrew-me/ytDownloader

    9,775在 GitHub 上查看↗

    ytDownloader is a video downloader and media extraction tool that uses the yt-dlp engine to retrieve video and audio files from various social media and video sharing platforms. It functions as a utility for capturing full media files, specific segments or ranges of tracks, and entire video playlists. The project includes a hardware-accelerated video compressor to reduce file sizes while maintaining visual quality. It also features a subtitle downloader capable of retrieving both text captions and embedded subtitle tracks for accessibility and translation. The system handles broad media task

    Offloads video compression to GPUs and dedicated hardware encoders to increase processing speed.

    JavaScriptappimagecompressordownloader
    在 GitHub 上查看↗9,775
上一个123下一个
  1. Home
  2. Graphics & Multimedia
  3. Media Processing and Analysis
  4. Media Manipulation
  5. Media Processing
  6. Codec and Encoding Utilities

探索子标签

  • Agnostic PipelinesUnified processing paths that handle multiple media codecs without requiring format-specific logic at each stage. **Distinct from Codec and Encoding Utilities:** Distinct from generic codec utilities by focusing on the agnostic architectural pipeline rather than individual encoding/decoding tools.
  • Hardware Accelerated Media Encoders1 个子标签Utilities that utilize specialized computer hardware to accelerate resource-intensive media compression and encoding tasks.
  • Media Codec Libraries2 个子标签Software libraries providing low-level, high-performance capabilities for encoding and decoding multimedia data.
  • Media Encoding Configurations2 个子标签Tools and settings for configuring codecs, filters, and muxers to meet specific audio and video processing requirements.
  • Video Muxing2 个子标签Utilities for combining processed video streams and audio into final media files using specified codecs.