awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

6 Repos

Awesome GitHub RepositoriesImage-Conditioned Generation

Creating new images using another image as a structural or stylistic reference.

Distinct from Image Generation: Distinct from general generation by requiring an image as a primary prompt/reference.

Explore 6 awesome GitHub repositories matching artificial intelligence & ml · Image-Conditioned Generation. Refine with filters or upvote what's useful.

Awesome Image-Conditioned Generation GitHub Repositories

Finde die besten Repos mit KI.Wir suchen mit KI nach den am besten passenden Repositories.
  • vladmandic/sdnextAvatar von vladmandic

    vladmandic/sdnext

    7,139Auf GitHub ansehen↗

    SD.Next is an all-in-one web interface and multi-backend inference engine for generating, editing, and processing images and videos using diffusion models. It functions as a comprehensive tool for diffusion model management and an automated image processing pipeline for bulk operations. The project is distinguished by its hardware-backend abstraction layer, which provides automatic detection and acceleration for NVIDIA CUDA, AMD ROCm, Intel OpenVINO, and DirectML. It features a headless generative API and a programmatic command interface, allowing users to trigger tasks via REST API or CLI wi

    Generates or edits images using other images as starting points or visual references.

    Pythonai-artcaptiondiffusers
    Auf GitHub ansehen↗7,139
  • tencent-ailab/ip-adapterAvatar von tencent-ailab

    tencent-ailab/IP-Adapter

    6,604Auf GitHub ansehen↗

    IP-Adapter is a framework for conditioning pretrained text-to-image diffusion models to use image prompts as visual guides. It serves as a text-to-image model extension that transforms a text-based diffusion model to accept and process image inputs as primary generation sources. The system implements identity preservation to maintain consistent facial features across multiple outputs using a reference photo. It also enables style transfer workflows to produce image variations that preserve the artistic characteristics of a source image. Capabilities cover multi-modal prompting, including the

    Creates new visual content using an existing image as the primary structural or stylistic reference.

    Jupyter Notebook
    Auf GitHub ansehen↗6,604
  • levihsu/ootdiffusionAvatar von levihsu

    levihsu/OOTDiffusion

    6,556Auf GitHub ansehen↗

    OOTDiffusion is an AI virtual try-on system designed for controllable image synthesis. It generates images of people wearing specific clothing items by superimposing garments onto human figures for both half-body and full-body compositions. The project facilitates digital fashion prototyping and virtual clothing fitting by creating garment-to-person overlays. It aims to maintain the original identity of the wearer and the specific details of the clothing during the synthesis process. The system utilizes a latent diffusion model and conditioning-based image generation to control the output. I

    Implements generation guided by garment images and human poses as structural and stylistic references.

    Python
    Auf GitHub ansehen↗6,556
  • cubiq/comfyui_ipadapter_plusAvatar von cubiq

    cubiq/ComfyUI_IPAdapter_plus

    6,031Auf GitHub ansehen↗

    ComfyUIIPAdapterplus ist eine knotenbasierte Erweiterung für ComfyUI, die IPAdapter-Modelle implementiert, um die Bildgenerierung unter Verwendung von Referenzbildern zu steuern. Sie fungiert als Bild-Prompting-Tool und Stable-Diffusion-Bildadapter, der es ermöglicht, Referenzdateien als visuelle Prompts zur Steuerung von Stil, Komposition und Subjektidentität zu verwenden. Das Projekt bietet spezialisierte Funktionen zur Wahrung der Gesichtsidentität und hochauflösender Merkmale über generierte Porträts hinweg. Es ermöglicht die Übertragung visueller Eigenschaften und künstlerischer Stile von Referenzbildern sowie die Extraktion räumlicher Layouts, um die Anordnung von Objekten in neuen Generationen zu steuern. Die Erweiterung deckt breite Funktionsbereiche ab, einschließlich KI-Bildkonditionierung, konsistenter Charaktergenerierung und Bildkompositionskontrolle.

    Enables the generation of new images using reference files as structural or stylistic baselines.

    Python
    Auf GitHub ansehen↗6,031
  • fanghua-yu/supirAvatar von Fanghua-Yu

    Fanghua-Yu/SUPIR

    5,587Auf GitHub ansehen↗

    SUPIR ist ein KI-Bild-Upscaler und Restaurierungssystem, das darauf ausgelegt ist, Artefakte zu entfernen und die Qualität realer Fotos wiederherzustellen. Es fungiert als diffusionsbasiertes Bildverbesserungs- und Restaurierungstool, das großskalige Modellskalierung verwendet, um hochauflösende Ergebnisse mit fotorealistischen Details zu erzielen. Das System gleicht visuelle Ästhetik mit Eingabetreue ab und ermöglicht einen Kompromiss zwischen strikter Einhaltung des Originalbildes und der allgemeinen visuellen Attraktivität der Ausgabe. Es nutzt großskalige Modell-Inferenz, um die Bildklarheit zu verbessern und realistische Details während des Upscaling-Prozesses beizubehalten.

    Uses the original low-resolution image as a structural reference to guide the generation of high-resolution output.

    Python
    Auf GitHub ansehen↗5,587
  • nunchaku-ai/nunchakuAvatar von nunchaku-ai

    nunchaku-ai/nunchaku

    3,883Auf GitHub ansehen↗

    Nunchaku is a 4-bit model quantization library and diffusion model inference engine designed to run large-scale neural networks on consumer GPUs. It functions as a GPU-accelerated optimizer that reduces VRAM usage and increases inference speed through weight compression and memory management. The project utilizes low-rank weight decomposition and SVD weight quantization to compress models to four-bit precision while maintaining visual fidelity. It employs kernel-level operator fusion to minimize data movement and hardware-aware precision mapping to adjust numerical precision based on the unde

    Supports image-to-image generation by using existing images as structural or stylistic references alongside text prompts.

    Pythoncomfyuidiffusion-modelsflux
    Auf GitHub ansehen↗3,883
  1. Home
  2. Artificial Intelligence & ML
  3. Image Generation
  4. Image-Conditioned Generation