awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to adobe-research/lavida-o

Open-source alternatives to LaVida O

15 open-source projects similar to adobe-research/lavida-o, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best LaVida O alternative.

  • ml-gsai/lladaML-GSAI 的头像

    ML-GSAI/LLaDA

    3,580在 GitHub 上查看↗

    LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining masked tokens through a diffusion process rather than predicting the next token in a sequence. The project functions as a vision-language diffusion model, converting visual inputs into text responses. It also serves as a preference optimization framework that uses log-likelihood estimation and evidence lower bounds to tune model responses. The system supports multi-round conversational AI and text sequence evaluation. It integrates vision-language embedding for cross-modal con

    Python
    在 GitHub 上查看↗3,580
  • vectorspacelab/omnigenVectorSpaceLab 的头像

    VectorSpaceLab/OmniGen

    4,326在 GitHub 上查看↗

    OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks through a single system. It functions as a multimodal diffusion framework that treats diverse vision operations as unified image synthesis problems using shared model weights, removing the need for external adapter modules. The system supports subject-driven image generation to preserve the identity of objects from reference photos and allows for multi-reference image synthesis. It also operates as an instruction-based image editor, modifying visual content through natural languag

    Jupyter Notebookdiffusionimageimage-edit
    在 GitHub 上查看↗4,326
  • fudoki-hku/fudokifudoki-hku 的头像

    fudoki-hku/FUDOKI

    76在 GitHub 上查看↗

    This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities.

    Python
    在 GitHub 上查看↗76
  • gen-verse/mmadaGen-Verse 的头像

    Gen-Verse/MMaDA

    1,656在 GitHub 上查看↗

    Multimodal Large Diffusion Language Models (NeurIPS 2025)

    Python
    在 GitHub 上查看↗1,656

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Find more with AI search
  • hustvl/diffusionvlhustvl 的头像

    hustvl/DiffusionVL

    149在 GitHub 上查看↗

    DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models

    Python
    在 GitHub 上查看↗149
  • jacklishufan/lavidajacklishufan 的头像

    jacklishufan/LaViDa

    221在 GitHub 上查看↗

    [Paper](paper/paper.pdf) [Arxiv](https://arxiv.org/abs/2505.16839) [Checkpoints](https://huggingface.co/collections/jacklishufan/lavida-10-682ecf5a5fa8c5df85c61ded) [Data](https://huggingface.co/datasets/jacklishufan/lavida-train) [Website](https://homepage.jackli.org/projects/lavida/)

    Python
    在 GitHub 上查看↗221
  • jiyt17/rediffjiyt17 的头像

    jiyt17/ReDiff

    45在 GitHub 上查看↗

    We introduce ReDiff, a refining-enhanced vision-language diffusion model.

    Python
    在 GitHub 上查看↗45
  • m-e-agi-lab/mudditM-E-AGI-Lab 的头像

    M-E-AGI-Lab/Muddit

    117在 GitHub 上查看↗

    Muddit is the 2nd generation Meissonic. It is built upon discrete diffusion for unified and efficient multimodal generation.

    Python
    在 GitHub 上查看↗117
  • ml-gsai/llada-vML-GSAI 的头像

    ML-GSAI/LLaDA-V

    345在 GitHub 上查看↗

    2026.03.23 We are excited to introduce LLaDA-o, the latest model in the LLaDA series. As an effective and length-adaptive omni diffusion model for unified multimodal understanding and generation, LLaDA-o extends the LLaDA line to broader multimodal settings, supporting visual understanding,…

    Python
    在 GitHub 上查看↗345
  • openhelix-team/unified-diffusion-vlaOpenHelix-Team 的头像

    OpenHelix-Team/Unified-Diffusion-VLA

    182在 GitHub 上查看↗

    Jiayi Chen¹\,Wenxuan Song¹†\, Pengxiang Ding²˒³, Ziyang Zhou¹, Han Zhao²˒³, Feilong Tang⁴,Donglin Wang², Haoang Li¹‡

    Python
    在 GitHub 上查看↗182
  • tyfeld/mmada-paralleltyfeld 的头像

    tyfeld/MMaDA-Parallel

    300在 GitHub 上查看↗

    ICLR 2026 MMaDA-Parallel: Parallel Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation

    Python
    在 GitHub 上查看↗300
  • yu-rp/dimpleyu-rp 的头像

    yu-rp/Dimple

    117在 GitHub 上查看↗

    Dimple, the first Discrete Diffusion Multimodal Large Language Model

    Python
    在 GitHub 上查看↗117
  • alexanderswerdlow/unidiscalexanderswerdlow 的头像

    alexanderswerdlow/unidisc

    141在 GitHub 上查看↗

    Unified Multimodal Discrete Diffusion

    Python
    在 GitHub 上查看↗141
  • zihohe/vidladaziHoHe 的头像

    ziHoHe/VidLaDA

    10在 GitHub 上查看↗
    Python
    在 GitHub 上查看↗10
  • alpha-vllm/lumina-dimooAlpha-VLLM 的头像

    Alpha-VLLM/Lumina-DiMOO

    1,001在 GitHub 上查看↗

    Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding

    Python
    在 GitHub 上查看↗1,001