awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetServeur MCPÀ proposNotre méthodologiePresse
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

3 dépôts

Awesome GitHub RepositoriesVisual Guidance Inputs

Using sparse image or sketch data to guide the generation of visual elements in AI models.

Distinct from Sparse Visual Mapping: No candidates cover the use of sparse RGB/sketches as generative guidance specifically for video synthesis.

Explore 3 awesome GitHub repositories matching artificial intelligence & ml · Visual Guidance Inputs. Refine with filters or upvote what's useful.

Awesome Visual Guidance Inputs GitHub Repositories

Trouvez les meilleurs dépôts grâce à l'IA.Nous recherchons les dépôts les plus pertinents grâce à l'IA.
  • guoyww/animatediffAvatar de guoyww

    guoyww/AnimateDiff

    12,144Voir sur GitHub↗

    AnimateDiff is a latent diffusion video generator and text-to-video diffusion framework. It converts existing text-to-image diffusion models into animation generators by applying specialized motion modules, allowing for the creation of video sequences without modifying the original base model. The project provides an image-to-video animation framework that uses sparse RGB images, sketches, or structural keyframe constraints to guide generation. It further distinguishes itself with a motion adapter system that injects cinematic camera movements, such as zooming, panning, and tilting, into anim

    Guides video generation using sparse RGB images or sketch inputs to define specific visual elements.

    Python
    Voir sur GitHub↗12,144
  • nvlabs/spadeAvatar de NVlabs

    NVlabs/SPADE

    7,718Voir sur GitHub↗

    SPADE is a semantic image synthesis framework and generative adversarial network designed to transform semantic label maps into photorealistic images. It uses a spatially-adaptive normalization model to modulate activations based on semantic maps, ensuring that spatial layouts and details are preserved throughout the synthesis process. The project enables the generation of diverse image variations from a single semantic layout by integrating variational autoencoders and latent vector style control. These mechanisms allow for the adjustment of visual appearances and textures while keeping the

    Uses semantic label maps as visual guidance to direct the placement and structure of generated objects.

    Python
    Voir sur GitHub↗7,718
  • yolain/comfyui-easy-useAvatar de yolain

    yolain/ComfyUI-Easy-Use

    2,567Voir sur GitHub↗

    ComfyUI-Easy-Use is a custom node suite and workflow optimizer designed to simplify Stable Diffusion generation pipelines. It provides a set of integrated tools to reduce visual clutter and streamline the process of creating images from text and existing image references. The project distinguishes itself through a pipeline manager that consolidates models, conditioning, and latents into unified data pipes, eliminating complex wiring in the node graph. It also introduces a logical operator set that enables conditional if-else branching and for-loop structures directly within the visual program

    Integrates visual guidance inputs such as pose maps or edge maps via ControlNet to steer image generation.

    Python
    Voir sur GitHub↗2,567
  1. Home
  2. Artificial Intelligence & ML
  3. Visual Guidance Inputs