17 repository-uri
Web-based platforms for managing and executing local diffusion models.
Distinct from Diffusion Models: Distinct from Diffusion Models: focuses on the web-based interface and workflow system rather than the model architecture itself.
Explore 17 awesome GitHub repositories matching artificial intelligence & ml · Stable Diffusion Web Interfaces. Refine with filters or upvote what's useful.
InvokeAI is a self-hosted, professional-grade platform designed for managing generative models and performing complex image synthesis. It provides a local application environment that allows users to execute diffusion models directly on their own hardware, ensuring data privacy and complete ownership of all generated assets. The platform distinguishes itself through a node-based workflow system that enables the construction of reproducible and automated image generation pipelines. By chaining modular functional units into directed acyclic graphs, users can automate intricate production tasks
Provides a professional-grade web interface for generating and editing images using local diffusion models.
IOPaint is an AI image editor and Stable Diffusion inpainting tool providing a web interface for removing objects and replacing image content. It utilizes latent diffusion image processing to synthesize high-resolution replacements for erased sections of an image. The project features a specialized AI background remover for isolating subjects and an AI image upscaler that employs super-resolution models for general photos and anime artwork. The software covers a broad range of capabilities including image segmentation for object isolation, face restoration for improving facial details, and t
Ships a web interface for removing objects and replacing image content using Stable Diffusion.
Lama Cleaner is an AI-powered image editing application focused on inpainting, object removal, and generative filling. It provides a suite of tools for erasing unwanted elements from photos and filling the resulting gaps using generative artificial intelligence. The project includes specialized capabilities for image outpainting to extend borders, background removal through object segmentation, and face restoration to fix visual defects. It also features an image upscaler to increase resolution and clarity via super-resolution AI, as well as a Stable Diffusion-based editor for replacing speci
Offers a web interface for performing generative filling and border extension via Stable Diffusion.
This project provides a cloud-based notebook configuration for deploying a Stable Diffusion web interface. It functions as a specialized environment for image generation, incorporating a model trainer for fine-tuning weights and creating training datasets. The system emphasizes infrastructure persistence by saving software installations and model files to cloud storage, avoiding repetitive setups between sessions. It uses a tunnel-based interface to expose the web dashboard to a public URL for remote interaction. The project covers end-to-end AI workflows, including dataset preparation and t
Deploys a web-based platform for managing and executing Stable Diffusion models on Google Colab.
Stable Diffusion WebUI Forge is a web-based interface and inference engine designed for the generation of AI media. It functions as a platform for executing diffusion-based models, providing a centralized environment to manage image preprocessors, custom generation logic, and hardware-accelerated sampling. The project distinguishes itself through a neural network patching framework that allows for the modification of model layers and the application of spatial conditioning during inference. By injecting custom logic and adapters directly into the network, users can influence output behaviors
Functions as a high-performance web interface and inference engine for executing and optimizing diffusion-based image generation models.
kohya_ss is a graphical user interface and workbench for fine-tuning diffusion models, specifically designed for Stable Diffusion. It provides a suite of tools for training generative AI models, including specialized interfaces for creating Low-Rank Adaptation weights and training ControlNet spatial control networks. The project distinguishes itself through integrated VRAM usage optimization and hardware acceleration, featuring specific support for Intel GPUs via XPU-accelerated libraries. It implements parameter-efficient training methods and memory-saving techniques like gradient checkpoint
Provides a graphical web interface for fine-tuning Stable Diffusion models using custom hyperparameters.
Sygil-webui is a web interface for Stable Diffusion latent diffusion models, providing a creative suite for text-to-image and text-to-video synthesis. It functions as an image generation tool and a latent diffusion image editor, allowing users to create visuals and video sequences from textual descriptions. The project includes a dedicated model training interface for creating custom textual inversion embeddings, which introduces specific new concepts or styles into the diffusion models. It also features specialized tools for generative image editing, including mask-based inpainting, image-to
Provides a comprehensive web-based platform for managing and executing local Stable Diffusion models.
Stable Diffusion Web UI is a browser-based interface for generating, editing, and upscaling images and videos using latent diffusion models. It functions as a text-to-image generator, an AI image editor, and a tool for increasing image resolution and clarity. The system includes capabilities for custom model training, specifically allowing the creation of textual inversion embeddings to teach a model new concepts and visual styles from user photos. It also provides tools for AI video production, generating short clips from text prompts. The software covers image-to-image transformation, imag
Provides a browser-based interface for managing and executing local latent diffusion models for image and video generation.
This project is a cloud-based AI deployment system and latent diffusion model trainer. It provides a framework for launching image generation interfaces and training pipelines on remote GPU infrastructure, specifically serving as a text-to-image model fine-tuner. The system features a specialized training interface for fine-tuning Stable Diffusion models on custom image datasets. It allows for the creation of personalized visual outputs by training models on specific subjects or artistic styles using a small set of reference images. The software covers generative AI deployment, custom style
Provides a web-based platform for managing and fine-tuning diffusion models on custom image datasets.
StabilityMatrix is a centralized installer and orchestrator for Stable Diffusion web interfaces and their dependencies. It functions as a generative AI workspace and portable runtime, providing a unified interface to install and update AI image generation packages within isolated environments to prevent global system conflicts. The project distinguishes itself through a shared model manager that imports, organizes, and shares checkpoints across different installations. It utilizes a central model repository and filesystem mapping to allow multiple packages to access the same large binary asse
Functions as a centralized installer and orchestrator for various Stable Diffusion web interfaces and their dependencies.
This project is a containerized deployment for running Stable Diffusion web interfaces. It provides a portable runtime for generative AI that manages dependencies and hardware acceleration to enable text-to-image generation and image-to-image transformations via a browser-based interface. The system uses hardware-specific image tags to support both GPU-accelerated synthesis and CPU-only execution. It ensures environment isolation across different operating systems while utilizing bind-mount data persistence to keep heavy model weights and generated outputs on the host machine. The deployment
Deploys Stable Diffusion web interfaces within Docker containers for consistent cross-platform execution.
qrbtf is an AI QR code generator and image synthesis system that blends machine-readable data with artistic imagery. It uses a latent diffusion model and spatial control networks to produce functional QR codes that incorporate visual art generated from descriptive text prompts. The system provides a dedicated interface and programmatic API for tuning visual output, allowing for the adjustment of control strength, padding ratios, and error correction levels. It supports deterministic sampling via random seeds and the use of negative prompts to refine the final aesthetic of the generated assets
Ships a web-based interface for managing prompts, seeds, and restoration rates to generate stylized QR codes via latent diffusion.
StableSwarmUI este o interfață web și un orchestrator backend pentru generarea de imagini Stable Diffusion. Acesta funcționează ca un generator de imagini GPU distribuit și un pipeline modular de imagini AI, oferind un controler centralizat pentru a gestiona cererile de generare de imagini. Sistemul se distinge prin abilitatea de a împărți sarcinile de generare între mai multe procesoare grafice pentru a crește throughput-ul batch-urilor. Utilizează o interfață agnostică față de backend pentru a se conecta la servere locale, servere la distanță și API-uri cloud, și include un designer de flux de lucru vizual bazat pe grafuri pentru definirea operațiunilor complexe de procesare a imaginilor. Platforma include un sistem dinamic de extensii plugin pentru adăugarea de funcționalități personalizate și utilitare automatizate pentru provizionarea dependențelor la nivel de sistem. Combină instrumente de generare modulare și interfețe de editare rapidă cu capacitatea de a ruta sarcinile de lucru pe hardware distribuit.
Provides a web-based interface for configuring and executing image generation workflows using Stable Diffusion models.
SwarmUI este o interfață web și un orchestrator pentru Stable Diffusion, conceput pentru a genera imagini și video. Funcționează ca un manager de fluxuri de lucru modular și un gateway API care permite configurarea și execuția pipeline-urilor AI generative. Sistemul se caracterizează prin capacitatea de a distribui sarcinile de generare pe mai multe plăci grafice pentru a crește viteza de procesare și throughput-ul total. Utilizează o arhitectură decuplată client-server și o interfață agnostică față de backend, permițând interfeței utilizator să rămână separată de mediul de execuție al modelului. Platforma suportă extensibilitatea printr-o arhitectură bazată pe plugin-uri pentru adăugarea de noi componente UI și handlere de logică. Oferă control programatic prin endpoint-uri HTTP și WebSocket pentru ca aplicațiile externe să declanșeze generări și să sincronizeze starea în timp real. Sunt incluse instrumente de administrare pentru a acorda și gestiona accesul utilizatorilor remote la mediul generativ prin rețea.
Provides a comprehensive web-based platform for managing and executing Stable Diffusion models for image and video generation.
Acest proiect este o extensie Stable Diffusion WebUI pentru interfața Forge, care implementează un model de difuzie cu straturi latente. Acesta servește drept generator de imagini transparente AI, conceput pentru a produce imagini cu canale alfa, permițând separarea automată a elementelor din prim-plan de fundalul lor. Software-ul permite generarea de straturi compatibile de prim-plan și fundal care pot fi compuse pentru editare post-producție. Suportă crearea de prim-planuri transparente pe baza unui fundal furnizat, generarea de fundaluri care să se potrivească unui prim-plan transparent existent și producerea simultană a prim-planului, fundalului și a imaginilor combinate.
Ships as an extension for the Forge web interface to add layer-based generation and blending.
Riffusion-hobby este un instrument AI generativ care creează muzică prin producerea de imagini spectrogramă via Stable Diffusion și convertirea lor în audio redabil. Funcționează ca un sintetizator audio de spectrogramă, utilizând deep learning pentru a transforma reprezentările de frecvență bazate pe imagini ale sunetului în fișiere audio. Proiectul operează ca un server de inferență muzicală AI, oferind un endpoint API bazat pe web pentru a genera audio din prompt-uri text și imagini seed. Include, de asemenea, o interfață de linie de comandă pentru executarea sarcinilor de generare muzicală și configurarea modelelor de difuzie pentru crearea automată de audio, precum și un generator audio în timp real pentru manipularea reprezentărilor sonore. Sistemul acoperă o gamă largă de capabilități, inclusiv implementarea modelelor în cloud, găzduirea inferenței la distanță și procesarea semnalului digital pentru conversia imagine-audio. Oferă, de asemenea, un playground interactiv bazat pe web pentru experimentarea cu parametrii modelului și explorarea setărilor de generare muzicală.
Deployes a web server to manage and execute diffusion models for music and image generation.
VoltaML-fast-stable-diffusion is a generative system designed for high-performance image synthesis from text prompts. It provides a comprehensive environment for executing inference tasks, managing pre-trained machine learning models, and integrating visual asset creation into external applications and workflows. The project distinguishes itself through multi-modal interaction capabilities, including a browser-based web interface for direct generation and a messaging platform integration that allows users to trigger and monitor tasks via chat commands. It supports automated creative workflows
Provides a browser-based platform for generating AI images using high-performance inference engines.