1 रिपॉजिटरी
Use of the Olive toolkit to compress and compile models for improved inference speed.
Distinct from Model Performance Optimization: Focuses specifically on the Olive optimization toolkit, whereas Model Performance Optimization is a general category.
Explore 1 awesome GitHub repository matching artificial intelligence & ml · Olive Model Compression. Refine with filters or upvote what's useful.
SD.Next is an all-in-one web interface and multi-backend inference engine for generating, editing, and processing images and videos using diffusion models. It functions as a comprehensive tool for diffusion model management and an automated image processing pipeline for bulk operations. The project is distinguished by its hardware-backend abstraction layer, which provides automatic detection and acceleration for NVIDIA CUDA, AMD ROCm, Intel OpenVINO, and DirectML. It features a headless generative API and a programmatic command interface, allowing users to trigger tasks via REST API or CLI wi
Utilizes the Olive toolkit to compress and compile models into optimized formats for faster inference.