awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectServer MCPDespreCum realizăm clasamentulPresă
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
hustvl avatar

hustvl/DiffusionVL

0
View on GitHub↗
149 stele·9 fork-uri·Python·Apache-2.0·5 vizualizări

DiffusionVL

DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models

Features

  • Multimodal Diffusion Models - Translating autoregressive models into vision-language diffusion models.

Istoric stele

Graficul istoricului de stele pentru hustvl/diffusionvlGraficul istoricului de stele pentru hustvl/diffusionvl

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Alternative open-source pentru DiffusionVL

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu DiffusionVL.
  • ml-gsai/lladaAvatar ML-GSAI

    ML-GSAI/LLaDA

    3,580Vezi pe GitHub↗

    LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining masked tokens through a diffusion process rather than predicting the next token in a sequence. The project functions as a vision-language diffusion model, converting visual inputs into text responses. It also serves as a preference optimization framework that uses log-likelihood estimation and evidence lower bounds to tune model responses. The system supports multi-round conversational AI and text sequence evaluation. It integrates vision-language embedding for cross-modal con

    Python
    Vezi pe GitHub↗3,580
  • vectorspacelab/omnigenAvatar VectorSpaceLab

    VectorSpaceLab/OmniGen

    4,326Vezi pe GitHub↗

    OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks through a single system. It functions as a multimodal diffusion framework that treats diverse vision operations as unified image synthesis problems using shared model weights, removing the need for external adapter modules. The system supports subject-driven image generation to preserve the identity of objects from reference photos and allows for multi-reference image synthesis. It also operates as an instruction-based image editor, modifying visual content through natural languag

    Jupyter Notebookdiffusionimageimage-edit
    Vezi pe GitHub↗4,326
  • alpha-vllm/lumina-dimooAvatar Alpha-VLLM

    Alpha-VLLM/Lumina-DiMOO

    1,001Vezi pe GitHub↗

    Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding

    Python
    Vezi pe GitHub↗1,001
  • fudoki-hku/fudokiAvatar fudoki-hku

    fudoki-hku/FUDOKI

    76Vezi pe GitHub↗

    This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities.

    Python
    Vezi pe GitHub↗76
Vezi toate cele 15 alternative pentru DiffusionVL→

Întrebări frecvente

Ce face hustvl/diffusionvl?

DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models

Care sunt principalele funcționalități ale hustvl/diffusionvl?

Principalele funcționalități ale hustvl/diffusionvl sunt: Multimodal Diffusion Models.

Care sunt câteva alternative open-source pentru hustvl/diffusionvl?

Alternativele open-source pentru hustvl/diffusionvl includ: ml-gsai/llada — LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining… vectorspacelab/omnigen — OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks… alpha-vllm/lumina-dimoo — Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding. gen-verse/mmada — Multimodal Large Diffusion Language Models (NeurIPS 2025). jacklishufan/lavida — [[Paper]](paper/paper.pdf) [[Arxiv]](https://arxiv.org/abs/2505.16839) [[Checkpoints]](https://huggingface.co/collectio… fudoki-hku/fudoki — This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via…