14 个仓库
Tools for automating repetitive image processing tasks.
Distinguishing note: Focuses on batch automation workflows.
Explore 14 awesome GitHub repositories matching development tools & productivity · Batch Image Processors. Refine with filters or upvote what's useful.
Aseprite is a specialized graphics editor and animation suite designed for the creation of pixel-based artwork. It provides a comprehensive environment for managing multi-layered animation sequences, offering tools for frame-by-frame design, onion skinning, and real-time motion previews. The application is built to handle both indexed color palettes and full-color RGB editing, allowing users to maintain precise control over pixel data and transparency. What distinguishes Aseprite is its focus on programmable workflows and game asset production. It features a scriptable command architecture th
Executes command-line scripts to batch convert, resize, or export large volumes of files.
Rembg is a machine learning-based toolkit designed for automated image background removal and subject segmentation. It functions as a versatile engine that identifies and extracts subjects from images, supporting diverse input methods including individual files, directory-based batch processing, and live binary data streams. The project distinguishes itself through its flexible integration options, offering a command-line interface for local automation, a library for programmatic access, and an HTTP service for remote requests. It utilizes deep learning architectures to classify pixels and ge
Automates background removal across directories and watch folders for high-volume image processing.
This project is an extension for Stable Diffusion that provides an image-to-image control framework. It serves as a multi-control constraint manager and structural data preprocessor, allowing users to guide the layout and composition of generated images through spatial maps and structural constraints. The system enables multi-constraint image generation by combining several different control inputs to enforce multiple stylistic or spatial rules within a single generation pass. It provides tools for visual image referencing and precise geometric or anatomical templating to ensure generated ima
Automates the execution of sequential control tasks across directories of image files for batch generation.
ImageMagick is a comprehensive software suite for the creation, editing, composition, and conversion of digital images. It functions as both a command-line utility for batch processing and automation, and as a programming library that allows developers to integrate advanced image manipulation capabilities into external applications. The project is distinguished by its modular architecture, which supports hundreds of image formats through a pluggable coder system and external delegate libraries. It is designed for high-performance environments, utilizing memory-mapped pixel caching, stream-ori
Provides a scriptable engine for automating complex, high-volume image workflows at scale.
This is a blind image watermarking and steganography tool designed to embed and extract hidden data from images without requiring the original source file. It functions as a framework for concealing text or bit arrays within images using mathematical transforms to ensure the marks remain invisible to the viewer. The system is designed for robust watermark extraction, allowing hidden information to be recovered even after images have undergone rotations, cropping, resizing, noise injection, or brightness changes. It utilizes a blind extraction mechanism that retrieves data using a shared passw
Applies invisible marks to large volumes of images using parallel processing for efficiency.
ImageToolbox is an open-source Android application designed for comprehensive image manipulation and batch processing. It provides a toolkit for performing advanced visual edits, including background removal, geometric transformations, and the application of complex filter chains to prepare image assets. The application distinguishes itself through a modular, pipeline-based architecture that allows for the integration of new processing algorithms as isolated plugins. It leverages native hardware acceleration to handle intensive pixel manipulation tasks and supports asynchronous execution to m
Streamlines repetitive tasks by applying complex editing operations and filter chains to multiple files simultaneously.
SD.Next is an all-in-one web interface and multi-backend inference engine for generating, editing, and processing images and videos using diffusion models. It functions as a comprehensive tool for diffusion model management and an automated image processing pipeline for bulk operations. The project is distinguished by its hardware-backend abstraction layer, which provides automatic detection and acceleration for NVIDIA CUDA, AMD ROCm, Intel OpenVINO, and DirectML. It features a headless generative API and a programmatic command interface, allowing users to trigger tasks via REST API or CLI wi
Provides batch image processing to apply generation or editing operations to multiple files simultaneously.
Caesium is an image compression tool that reduces file sizes for JPG, PNG, WebP, and TIFF images while preserving visual quality and metadata. It operates as a cross-platform desktop application with a graphical interface, a command-line tool for scripting and automation, and a web-based interface for browser uploads, all supporting batch processing of multiple images at once. The tool distinguishes itself by offering multiple interaction modes — desktop, terminal, and web — each capable of handling the same core compression tasks. It preserves folder structure when saving compressed images,
Compresses multiple images at once, maintaining folder structure and supporting automated workflows.
chaiNNer is a GPU-accelerated AI image upscaling application that uses a visual node-based interface for constructing image processing pipelines. At its core, it provides a node-based visual programming environment where users connect processing nodes in a directed acyclic graph, with a graph execution scheduler that traverses the pipeline in topological order. The application includes an iterator-based batch processing system that automatically applies the same pipeline to multiple files, and a model format conversion pipeline that transforms neural network models between PyTorch, ONNX, and N
Processes multiple files through a visual pipeline using iterator nodes for uniform operations.
Compresses multiple PNG files in a single command with recursive directory traversal and shell script integration.
Thumbnailator 是一个 Java 图像缩略图库,专为生成保持纵横比的高质量缩放图像而设计。它作为一个工具包,用于在 Java 应用程序中旋转、裁剪和调整图像的不透明度。 该库的特色在于能够作为感知 Exif 的图像处理器,根据嵌入的方向元数据自动旋转缩略图。它还提供了用于数字水印的专门实用程序,允许以可调节的透明度叠加辅助图像和品牌标记。 核心功能涵盖了广泛的图像处理任务,包括焦点裁剪、尺寸调整和添加图像边框。该项目还使用可配置的命名方案和压缩设置处理已处理图像到文件或流的导出。
Enables automated programmatic workflows for applying rotations, borders, and opacity changes to collections of images.
本软件是一个水印去除系统,使用机器学习和图像修复 (inpainting) 技术从图像中删除不需要的文本或徽标。它通过预训练模型重建缺失的像素以匹配原始背景,从而确保视觉一致性。 该项目包括一个掩码实用程序,用于通过二进制掩码、边界框或画笔笔触隔离特定区域以进行内容替换。它还具有一个批处理处理器,通过预定义的文件列表将这些清理任务应用于大量图像。 该系统通过将尺寸和纵横比归一化为张量来处理图像准备,从而使图像与其对应的掩码对齐,以进行神经网络处理。
Provides a system to automate the removal of designated areas across multiple images.
Text-Grab is a desktop utility that captures text from screen regions, images, PDFs, and native user interface elements using on-device optical character recognition (OCR) and Windows UI Automation. It processes text entirely locally without sending data to external services, and extracts text directly from UI controls with perfect accuracy by reading the accessibility tree. The application also includes a persistent snippet dictionary for instant retrieval of frequently used text via a configurable system-wide hotkey. The tool supports building reusable extraction workflows by saving capture
Applies saved capture regions and pattern rules to automatically extract text from entire folders of images or PDFs.
Imageflow 是一个高性能图像操作库和合成引擎,可作为 C 兼容库、命令行图像处理器和动态图像处理服务器使用。它通过编程接口、JSON 作业文件或即时 URL 查询字符串,提供了解码、编码和对图像应用复杂视觉变换的方法。 该系统通过基于图的处理流水线脱颖而出,允许单次多格式编码,从单次解码中生成多种图像尺寸和格式,从而减少开销。它还具有资源受限的解码引擎,强制执行严格的内存和尺寸限制,以防止资源耗尽和拒绝服务攻击。 该项目涵盖了广泛的操作能力,包括尺寸调整、裁剪、旋转和颜色过滤。它支持高级合成任务,如水印、空白画布生成和几何形状渲染,以及使用直方图分析的自动色彩校正和白平衡调整。 核心逻辑通过外部函数接口绑定暴露,以实现跨语言集成。
Ships a command-line tool for automating batch image processing using JSON job files and operation graphs.