14 रिपॉजिटरी
Automated workflows for performing the same operations across multiple video files.
Distinct from Batch Processing: Existing batch candidates are for audio, cloud trials, or data inputs, not video file sets.
Explore 14 awesome GitHub repositories matching graphics & multimedia · Batch Video Processing. Refine with filters or upvote what's useful.
Perfect Green Screen Keys
Processes multiple video files in one run, applying background removal to each clip automatically.
This project is an optical character recognition tool designed to extract hardcoded subtitles from video frames and convert them into synchronized subtitle files. It functions as a text processor that transforms embedded visual text into a written format to improve video accessibility and translation. The system uses graphics processing units to increase the speed and accuracy of text recognition. It includes a subtitle cleaning tool that applies custom mapping configurations to filter out watermarks, channel logos, and duplicate lines from the extracted text. The tool supports batch process
Enables subtitles to be extracted from multiple video files simultaneously when resolution and text regions are identical.
Gyroflow is a gyroscope video stabilization software and IMU telemetry processor designed to remove camera shake from video files. It functions as a hardware-accelerated video renderer and lens calibration tool, utilizing embedded or external gyroscope and accelerometer data to perform pixel-level stabilization. The system is distinguished by its ability to integrate with professional non-linear video editing software via plugins, allowing stabilization to be applied directly to timelines without transcoding original footage. It supports diverse telemetry ingestion from camera brands, flight
Applies uniform smoothness, horizon locking, and zoom configurations across multiple video files.
Backgroundremover is an AI-powered tool that removes backgrounds from both images and videos, accessible through a command-line interface and a Python API. At its core, it uses a pre-trained deep learning model to classify each pixel as foreground or background, producing a binary mask for removal. The tool distinguishes itself through multiple integration methods and output capabilities. It can process images and videos via Unix pipeline data streams, operate as an HTTP API server, or be called programmatically within Python scripts. Users can choose among different AI models to balance proc
Processes images and videos from the terminal with batch operations, Unix pipes, and model selection.
MoneyPrinterPlus is an automated video production system designed for the mass creation of short-form AI content. It functions as an end-to-end pipeline that uses large language models to generate scripts, synthesize voiceovers, and produce visual assets to assemble complete videos. The project is distinguished by its ability to batch-process high volumes of unique content through automated mixing and randomized asset pairing. It includes a social media auto-publisher that uses browser simulation to automate the upload and distribution of generated videos to platforms such as TikTok and Xiaoh
Produces large quantities of non-duplicate short videos through automated mixing and batch editing.
Manages multiple encoding jobs in a queue for efficient batch processing of video files.
यह ऑटोमैटिक स्पीच रिकग्निशन के लिए एक Windows एप्लिकेशन है जो वीडियो फाइलों से बोले गए ऑडियो को टाइमस्टैम्प वाली SRT सबटाइटल फाइलों में ट्रांसक्राइब करता है। यह एक सबटाइटल जनरेटर और अनुवाद टूल के रूप में कार्य करता है जो मीडिया स्पीच को सिंक्रोनाइज़्ड टेक्स्ट में बदलता है। सॉफ्टवेयर एक बैच मीडिया ट्रांसक्राइबर के रूप में कार्य करता है, जो बल्क में सबटाइटल उत्पन्न करने के लिए कई ऑडियो और वीडियो फाइलों के एक साथ प्रसंस्करण की अनुमति देता है। इसमें द्विभाषी या स्थानीयकृत फाइलों के निर्माण के लिए विभिन्न भाषाओं के बीच ट्रांसक्रिप्शन को बदलने के लिए एक अनुवाद वर्कफ़्लो शामिल है। सिस्टम टेक्स्ट रिफाइनमेंट क्षमताएं भी प्रदान करता है, जो फिलर शब्दों और अवांछित पैटर्न को हटाकर ट्रांसक्रिप्ट को साफ करने के लिए रेगुलर एक्सप्रेशन और कस्टम फिल्टर का उपयोग करता है।
Automates the creation of subtitles across multiple video files using batch workflows.
node-ytdl-core Node.js के लिए एक JavaScript लाइब्रेरी है जिसे YouTube से मेटाडेटा निकालने और वीडियो और ऑडियो कंटेंट को स्ट्रीम करने के लिए डिज़ाइन किया गया है। यह एक मीडिया डाउनलोडर और स्ट्रीम फेचर के रूप में कार्य करता है, जो यूज़र्स को रिमोट सोर्सेज से वीडियो विवरण और मीडिया डेटा प्राप्त करने की अनुमति देता है। यह लाइब्रेरी वीडियो एक्सट्रैक्शन के लिए विशेष क्षमताएं प्रदान करती है, जिसमें अद्वितीय आइडेंटिफायर्स के लिए मीडिया URLs को पार्स करने और उपलब्ध फॉर्मेट्स का विश्लेषण करने की क्षमता शामिल है। यह गुणवत्ता और रिज़ॉल्यूशन मानदंडों के आधार पर विशिष्ट वीडियो और ऑडियो स्ट्रीम्स के चयन और फिल्टरिंग की अनुमति देती है। यह प्रोजेक्ट रेट लिमिट से बचने और प्रमाणीकरण कुकी मैनेजमेंट के माध्यम से नेटवर्क ट्रैफिक को मैनेज करता है। यह डेटा पाइपिंग के लिए रीडेबल स्ट्रीम्स का उपयोग करता है और मीडिया फाइलों के विशिष्ट सेगमेंट प्राप्त करने के लिए बाइट-रेंज रिक्वेस्टिंग का समर्थन करता है।
Parses URLs and strings to retrieve one or more unique video identifiers for batch retrieval.
Automatic Optical Disc Ripping Server is a headless system that detects inserted CDs, DVDs, and Blu-rays to automatically extract media, transcode video, and eject discs. It functions as a multi-drive media digitizer using a concurrent processing pipeline to rip and transcode media from several optical drives simultaneously without queuing. The system includes an asynchronous video transcoding pipeline that batches conversion tasks to run during scheduled off-peak hours. It also serves as a media server automation tool, fetching metadata from online APIs to name folders and trigger library re
Implements a batch processing system that converts ripped video files into target formats during scheduled off-peak hours.
SmartSub AI-संचालित वीडियो ट्रांसक्रिप्शन और सबटाइटल जनरेशन के लिए एक क्रॉस-प्लेटफ़ॉर्म डेस्कटॉप एप्लिकेशन है। यह स्थानीय AI मॉडल का उपयोग करके ऑडियो और वीडियो फ़ाइलों को टेक्स्ट सबटाइटल में परिवर्तित करता है और प्रसंस्करण गति बढ़ाने के लिए हार्डवेयर त्वरण को शामिल करता है। इस उपकरण में एक सबटाइटल अनुवादक है जो OpenAI और DeepSeek जैसे बड़े भाषा मॉडल का लाभ उठाकर सबटाइटल को विभिन्न भाषाओं के बीच परिवर्तित करता है। इसमें प्रूफरीडिंग और ट्रांसक्राइब्ड टेक्स्ट को पॉलिश करने के लिए एक विज़ुअल एडिटर शामिल है, जिसे फ़्रेम-सटीक सिंक्रनाइज़ेशन के लिए वीडियो पूर्वावलोकन के साथ जोड़ा गया है। यह सॉफ़्टवेयर कई मीडिया फ़ाइलों के बैच प्रसंस्करण का समर्थन करता है और सबटाइटल को स्विच करने योग्य सॉफ्ट ट्रैक के रूप में एम्बेड करने या उन्हें स्थायी रूप से वीडियो फ़्रेम में बर्न करने के लिए उपयोगिताएं प्रदान करता है। इसमें ट्रांसक्रिप्शन मॉडल फ़ाइलों को प्रबंधित करने और बाहरी AI सेवा मापदंडों को कॉन्फ़िगर करने के लिए सिस्टम भी शामिल हैं।
Automates the transcription and translation process across multiple video files simultaneously.
Squirrel-RIFE is a GPU-accelerated video processing tool that uses a neural network to generate intermediate frames between existing video frames, enabling smooth slow-motion effects and frame rate conversion. It is built around the RIFE (Real-Time Intermediate Flow Estimation) model, which analyzes motion between consecutive frames to predict and insert new frames, and leverages NVIDIA CUDA for parallel processing to achieve high-speed inference. The tool distinguishes itself by combining neural frame interpolation with practical video preprocessing features, including pixel-level duplicate
Processes multiple video frames in parallel batches on NVIDIA GPUs to maximize throughput and reduce per-frame overhead.
Short video factory is a local AI content generator and automated video editing tool. It provides a production pipeline that uses large language models to transform text prompts into marketing scripts and rendered short-form videos. The system is designed for local-first execution, running all processing and asset management on the host machine to maintain data privacy. It distinguishes itself through a batch-processing workflow that can sequentially execute copywriting and rendering for multiple items using predefined presets. The software covers a broad range of media capabilities, includi
Automatically produces a sequence of videos by batching the copywriting and rendering processes.
ComfyUI-SeedVR2_VideoUpscaler is an AI video upscaling tool that uses diffusion models to increase the resolution of videos and images while maintaining visual consistency across frames. The project implements distributed video rendering by splitting datasets into chunks for parallel processing across multiple GPUs. It utilizes model compilation and specialized attention backends to reduce inference latency and increase throughput. Additional capabilities include video color correction using wavelet and LAB matching methods to preserve color fidelity. Hardware memory is managed through block
Offers automated workflows for performing resolution enhancement across multiple video files via CLI.
This project is a deep learning framework built for detecting and tracking human body keypoints in images and video streams. It functions as both a real-time motion tracking system and a machine learning environment for training and evaluating pose estimation models. The system utilizes a two-branch convolutional neural network to predict body part locations and their directional connections simultaneously. It employs multi-stage feature refinement to improve keypoint localization accuracy and uses greedy parsing and bipartite matching algorithms to associate detected parts into individual sk
Executes parallel tensor-based batch processing to maintain high frame rates during real-time video analysis.