awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
ali-vilab avatar

ali-vilab/AnyDoor

0
View on GitHub↗
4,229 stars·371 forks·Python·MIT·44 viewsali-vilab.github.io/AnyDoor-Page↗

AnyDoor

AnyDoor is a zero-shot image customization framework designed to transfer specific objects from reference images into new scenes without requiring additional model training. It functions as a diffusion-based object insertion tool that enables the placement of objects into target environments while preserving their original identity, lighting, and posture.

The system supports both single and multi-object insertion, allowing several distinct objects from different references to be composed into a single target image. It utilizes a segmentation mechanism for mask refinement to clean and sharpen object boundaries, ensuring precise blending between the inserted objects and the background.

The project provides capabilities for object-level image editing and regional generation guided by spatial masks. It also includes utilities for custom model training on specific datasets using configurable hyperparameters to improve object transfer results.

Features

  • Object Customizations - Provides a framework for transferring specific objects from reference images into new scenes without requiring additional training.
  • Attention Layer Injectors - Implements mechanisms for injecting visual control signals into diffusion model attention layers to preserve object identity.
  • Reference-Guided Generation - Uses reference image features to guide the latent diffusion process for identity-preserving generation.
  • Zero-Shot Identity Synthesis - Implements a system for zero-shot identity synthesis to transfer objects without per-object training.
  • Mask-Guided Image Editors - Provides a tool for synthesizing specific objects into target regions guided by spatial masks.
  • Diffusion-Based Object Insertions - Ships a tool that blends target objects into new environments while preserving their original identity, lighting, and posture.
  • Generative Object Insertions - Allows the placement of multiple specific objects from reference images into a new scene with adaptive lighting.
  • Generative Object Compositions - Enables the placement of multiple distinct objects from various reference images into a single target image.
  • Multi-Reference Blending - Combines visual features from multiple reference images to compose several distinct objects into a single scene.
  • Mask Refinement Loops - Employs iterative processes to refine and sharpen segmentation boundaries for precise object blending.
  • Mask Refinements - Implements mask refinement techniques to clean object boundaries for higher quality image customization.
  • Image Editing - Provides capabilities for modifying specific image regions by inserting or replacing objects.

Star history

Star history chart for ali-vilab/anydoorStar history chart for ali-vilab/anydoor

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Projects sharing features with AnyDoor

These projects share indexed features with AnyDoor. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • ant-research/magicquillant-research avatar

    ant-research/MagicQuill

    3,682View on GitHub↗

    MagicQuill is a suite of interactive tools for image segmentation, diffusion-based editing, layered composition, and prompt-guided visual synthesis. It functions as a diffusion model image editor and a layered visual composition tool, enabling the addition, removal, and recoloring of image elements through a combination of sketches and text prompts. The system features a prompt-guided image generator that predicts editing instructions by analyzing user drawings to automatically populate text prompts. It allows for visual style control by swapping generative model weights to shift outputs betw

    Pythonaigcgradioimage-editing
    View on GitHub↗3,682
  • qwenlm/qwen-imageQwenLM avatar

    QwenLM/Qwen-Image

    7,379View on GitHub↗

    Qwen-Image is a text-to-image model and large language model image generation framework. It functions as an AI image editing suite and a personalized image trainer, capable of producing high-fidelity visuals and accurate typography from natural language descriptions. The system is distinguished by its precision text rendering engine, which integrates multi-script calligraphy and layout-coherent alphabetic text into images. It provides specialized capabilities for subject identity preservation and consistent subject generation across different poses and viewpoints, alongside a training pipelin

    Python
    View on GitHub↗7,379
  • sanster/iopaintSanster avatar

    Sanster/IOPaint

    23,244View on GitHub↗

    IOPaint is an AI image editor and Stable Diffusion inpainting tool providing a web interface for removing objects and replacing image content. It utilizes latent diffusion image processing to synthesize high-resolution replacements for erased sections of an image. The project features a specialized AI background remover for isolating subjects and an AI image upscaler that employs super-resolution models for general photos and anime artwork. The software covers a broad range of capabilities including image segmentation for object isolation, face restoration for improving facial details, and t

    Pythoninpaintinglamalatent-diffusion
    View on GitHub↗23,244
  • sygil-dev/sygil-webuiSygil-Dev avatar

    Sygil-Dev/sygil-webui

    7,879View on GitHub↗

    Sygil-webui is a web interface for Stable Diffusion latent diffusion models, providing a creative suite for text-to-image and text-to-video synthesis. It functions as an image generation tool and a latent diffusion image editor, allowing users to create visuals and video sequences from textual descriptions. The project includes a dedicated model training interface for creating custom textual inversion embeddings, which introduces specific new concepts or styles into the diffusion models. It also features specialized tools for generative image editing, including mask-based inpainting, image-to

    Python
    View on GitHub↗7,879
Compare all 30 related projects→

Frequently asked questions

What does ali-vilab/anydoor do?

AnyDoor is a zero-shot image customization framework designed to transfer specific objects from reference images into new scenes without requiring additional model training. It functions as a diffusion-based object insertion tool that enables the placement of objects into target environments while preserving their original identity, lighting, and posture.

What are the main features of ali-vilab/anydoor?

The main features of ali-vilab/anydoor are: Object Customizations, Attention Layer Injectors, Reference-Guided Generation, Zero-Shot Identity Synthesis, Mask-Guided Image Editors, Diffusion-Based Object Insertions, Generative Object Insertions, Generative Object Compositions.

Which projects share features with ali-vilab/anydoor?

Projects with overlapping indexed features include: ant-research/magicquill — MagicQuill is a suite of interactive tools for image segmentation, diffusion-based editing, layered composition, and… qwenlm/qwen-image — Qwen-Image is a text-to-image model and large language model image generation framework. It functions as an AI image… sanster/iopaint — IOPaint is an AI image editor and Stable Diffusion inpainting tool providing a web interface for removing objects and… sygil-dev/sygil-webui — Sygil-webui is a web interface for Stable Diffusion latent diffusion models, providing a creative suite for… vectorspacelab/omnigen2 — OmniGen2 is a unified image generation model and multimodal large language model designed to handle text-to-image… geekyutao/inpaint-anything — Inpaint-Anything is a diffusion-based image editor and inpainting tool designed to remove or replace objects in…