17 repositorios
Web-based platforms for managing and executing local diffusion models.
Distinct from Diffusion Models: Distinct from Diffusion Models: focuses on the web-based interface and workflow system rather than the model architecture itself.
Explore 17 awesome GitHub repositories matching artificial intelligence & ml · Stable Diffusion Web Interfaces. Refine with filters or upvote what's useful.
InvokeAI is a self-hosted, professional-grade platform designed for managing generative models and performing complex image synthesis. It provides a local application environment that allows users to execute diffusion models directly on their own hardware, ensuring data privacy and complete ownership of all generated assets. The platform distinguishes itself through a node-based workflow system that enables the construction of reproducible and automated image generation pipelines. By chaining modular functional units into directed acyclic graphs, users can automate intricate production tasks
Provides a professional-grade web interface for generating and editing images using local diffusion models.
IOPaint is an AI image editor and Stable Diffusion inpainting tool providing a web interface for removing objects and replacing image content. It utilizes latent diffusion image processing to synthesize high-resolution replacements for erased sections of an image. The project features a specialized AI background remover for isolating subjects and an AI image upscaler that employs super-resolution models for general photos and anime artwork. The software covers a broad range of capabilities including image segmentation for object isolation, face restoration for improving facial details, and t
Ships a web interface for removing objects and replacing image content using Stable Diffusion.
Lama Cleaner is an AI-powered image editing application focused on inpainting, object removal, and generative filling. It provides a suite of tools for erasing unwanted elements from photos and filling the resulting gaps using generative artificial intelligence. The project includes specialized capabilities for image outpainting to extend borders, background removal through object segmentation, and face restoration to fix visual defects. It also features an image upscaler to increase resolution and clarity via super-resolution AI, as well as a Stable Diffusion-based editor for replacing speci
Offers a web interface for performing generative filling and border extension via Stable Diffusion.
This project provides a cloud-based notebook configuration for deploying a Stable Diffusion web interface. It functions as a specialized environment for image generation, incorporating a model trainer for fine-tuning weights and creating training datasets. The system emphasizes infrastructure persistence by saving software installations and model files to cloud storage, avoiding repetitive setups between sessions. It uses a tunnel-based interface to expose the web dashboard to a public URL for remote interaction. The project covers end-to-end AI workflows, including dataset preparation and t
Deploys a web-based platform for managing and executing Stable Diffusion models on Google Colab.
Stable Diffusion WebUI Forge is a web-based interface and inference engine designed for the generation of AI media. It functions as a platform for executing diffusion-based models, providing a centralized environment to manage image preprocessors, custom generation logic, and hardware-accelerated sampling. The project distinguishes itself through a neural network patching framework that allows for the modification of model layers and the application of spatial conditioning during inference. By injecting custom logic and adapters directly into the network, users can influence output behaviors
Functions as a high-performance web interface and inference engine for executing and optimizing diffusion-based image generation models.
kohya_ss is a graphical user interface and workbench for fine-tuning diffusion models, specifically designed for Stable Diffusion. It provides a suite of tools for training generative AI models, including specialized interfaces for creating Low-Rank Adaptation weights and training ControlNet spatial control networks. The project distinguishes itself through integrated VRAM usage optimization and hardware acceleration, featuring specific support for Intel GPUs via XPU-accelerated libraries. It implements parameter-efficient training methods and memory-saving techniques like gradient checkpoint
Provides a graphical web interface for fine-tuning Stable Diffusion models using custom hyperparameters.
Sygil-webui is a web interface for Stable Diffusion latent diffusion models, providing a creative suite for text-to-image and text-to-video synthesis. It functions as an image generation tool and a latent diffusion image editor, allowing users to create visuals and video sequences from textual descriptions. The project includes a dedicated model training interface for creating custom textual inversion embeddings, which introduces specific new concepts or styles into the diffusion models. It also features specialized tools for generative image editing, including mask-based inpainting, image-to
Provides a comprehensive web-based platform for managing and executing local Stable Diffusion models.
Stable Diffusion Web UI is a browser-based interface for generating, editing, and upscaling images and videos using latent diffusion models. It functions as a text-to-image generator, an AI image editor, and a tool for increasing image resolution and clarity. The system includes capabilities for custom model training, specifically allowing the creation of textual inversion embeddings to teach a model new concepts and visual styles from user photos. It also provides tools for AI video production, generating short clips from text prompts. The software covers image-to-image transformation, imag
Provides a browser-based interface for managing and executing local latent diffusion models for image and video generation.
This project is a cloud-based AI deployment system and latent diffusion model trainer. It provides a framework for launching image generation interfaces and training pipelines on remote GPU infrastructure, specifically serving as a text-to-image model fine-tuner. The system features a specialized training interface for fine-tuning Stable Diffusion models on custom image datasets. It allows for the creation of personalized visual outputs by training models on specific subjects or artistic styles using a small set of reference images. The software covers generative AI deployment, custom style
Provides a web-based platform for managing and fine-tuning diffusion models on custom image datasets.
StabilityMatrix is a centralized installer and orchestrator for Stable Diffusion web interfaces and their dependencies. It functions as a generative AI workspace and portable runtime, providing a unified interface to install and update AI image generation packages within isolated environments to prevent global system conflicts. The project distinguishes itself through a shared model manager that imports, organizes, and shares checkpoints across different installations. It utilizes a central model repository and filesystem mapping to allow multiple packages to access the same large binary asse
Functions as a centralized installer and orchestrator for various Stable Diffusion web interfaces and their dependencies.
This project is a containerized deployment for running Stable Diffusion web interfaces. It provides a portable runtime for generative AI that manages dependencies and hardware acceleration to enable text-to-image generation and image-to-image transformations via a browser-based interface. The system uses hardware-specific image tags to support both GPU-accelerated synthesis and CPU-only execution. It ensures environment isolation across different operating systems while utilizing bind-mount data persistence to keep heavy model weights and generated outputs on the host machine. The deployment
Deploys Stable Diffusion web interfaces within Docker containers for consistent cross-platform execution.
qrbtf is an AI QR code generator and image synthesis system that blends machine-readable data with artistic imagery. It uses a latent diffusion model and spatial control networks to produce functional QR codes that incorporate visual art generated from descriptive text prompts. The system provides a dedicated interface and programmatic API for tuning visual output, allowing for the adjustment of control strength, padding ratios, and error correction levels. It supports deterministic sampling via random seeds and the use of negative prompts to refine the final aesthetic of the generated assets
Ships a web-based interface for managing prompts, seeds, and restoration rates to generate stylized QR codes via latent diffusion.
StableSwarmUI es una interfaz web y orquestador de backend para la generación de imágenes con Stable Diffusion. Funciona como un generador de imágenes GPU distribuido y un pipeline de imágenes de IA modular, proporcionando un controlador centralizado para gestionar las solicitudes de generación de imágenes. El sistema se distingue por la capacidad de dividir las tareas de generación entre múltiples procesadores gráficos para aumentar el rendimiento por lotes. Utiliza una interfaz agnóstica al backend para conectarse a servidores locales, servidores remotos y APIs en la nube, e incluye un diseñador de flujos de trabajo visual basado en grafos para definir operaciones complejas de procesamiento de imágenes. La plataforma incluye un sistema de extensión de plugins dinámico para añadir funciones personalizadas y utilidades automatizadas para el aprovisionamiento de dependencias a nivel de sistema. Combina herramientas de generación modulares e interfaces de edición rápida con la capacidad de enrutar cargas de trabajo a través de hardware distribuido.
Provides a web-based interface for configuring and executing image generation workflows using Stable Diffusion models.
SwarmUI is a web-based interface and orchestrator for Stable Diffusion, designed to generate images and video. It functions as a modular workflow manager and an API gateway that allows for the configuration and execution of generative AI pipelines. The system is characterized by its ability to distribute generation workloads across multiple graphics cards to increase processing speed and total throughput. It employs a decoupled client-server architecture and a backend-agnostic interface, allowing the user interface to remain separate from the model execution environment. The platform support
Provides a comprehensive web-based platform for managing and executing Stable Diffusion models for image and video generation.
Este proyecto es una extensión de Stable Diffusion WebUI para la interfaz Forge que implementa un modelo de difusión de capa latente. Funciona como un generador de imágenes transparentes por IA diseñado para producir imágenes con canales alfa, permitiendo la separación automática de elementos en primer plano de sus fondos. El software permite la generación de capas de primer plano y fondo compatibles que pueden ser compuestas para edición de postproducción. Admite la creación de primeros planos transparentes basados en un fondo proporcionado, la generación de fondos que se ajusten a un primer plano transparente existente y la producción simultánea de primer plano, fondo e imágenes combinadas.
Ships as an extension for the Forge web interface to add layer-based generation and blending.
Riffusion-hobby is a generative AI tool that creates music by producing spectrogram images via Stable Diffusion and converting them into playable audio. It functions as a spectrogram audio synthesizer, utilizing deep learning to transform image-based frequency representations of sound into audio files. The project operates as an AI music inference server, providing a web-based API endpoint to generate audio from text prompts and seed images. It also includes a command line interface for executing music generation tasks and configuring diffusion models for automated audio creation, as well as
Deployes a web server to manage and execute diffusion models for music and image generation.
VoltaML-fast-stable-diffusion is a generative system designed for high-performance image synthesis from text prompts. It provides a comprehensive environment for executing inference tasks, managing pre-trained machine learning models, and integrating visual asset creation into external applications and workflows. The project distinguishes itself through multi-modal interaction capabilities, including a browser-based web interface for direct generation and a messaging platform integration that allows users to trigger and monitor tasks via chat commands. It supports automated creative workflows
Provides a browser-based platform for generating AI images using high-performance inference engines.