awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
tyfeld avatar

tyfeld/MMaDA-Parallel

0
View on GitHub↗
300 星标·10 分支·Python·Apache-2.0·3 次浏览arxiv.org/abs/2511.09611↗

MMaDA Parallel

[ICLR 2026] MMaDA-Parallel: Parallel Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation

Features

  • Multimodal Diffusion Models - Multimodal diffusion model for thinking-aware editing and generation.

Star 历史

tyfeld/mmada-parallel 的 Star 历史图表tyfeld/mmada-parallel 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

MMaDA Parallel 的开源替代方案

相似的开源项目,按与 MMaDA Parallel 的功能重合度排序。
  • ml-gsai/lladaML-GSAI 的头像

    ML-GSAI/LLaDA

    3,580在 GitHub 上查看↗

    LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining masked tokens through a diffusion process rather than predicting the next token in a sequence. The project functions as a vision-language diffusion model, converting visual inputs into text responses. It also serves as a preference optimization framework that uses log-likelihood estimation and evidence lower bounds to tune model responses. The system supports multi-round conversational AI and text sequence evaluation. It integrates vision-language embedding for cross-modal con

    Python
    在 GitHub 上查看↗3,580
  • vectorspacelab/omnigenVectorSpaceLab 的头像

    VectorSpaceLab/OmniGen

    4,326在 GitHub 上查看↗

    OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks through a single system. It functions as a multimodal diffusion framework that treats diverse vision operations as unified image synthesis problems using shared model weights, removing the need for external adapter modules. The system supports subject-driven image generation to preserve the identity of objects from reference photos and allows for multi-reference image synthesis. It also operates as an instruction-based image editor, modifying visual content through natural languag

    Jupyter Notebookdiffusionimageimage-edit
    在 GitHub 上查看↗4,326
  • alpha-vllm/lumina-dimooAlpha-VLLM 的头像

    Alpha-VLLM/Lumina-DiMOO

    1,001在 GitHub 上查看↗

    Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding

    Python
    在 GitHub 上查看↗1,001
  • fudoki-hku/fudokifudoki-hku 的头像

    fudoki-hku/FUDOKI

    76在 GitHub 上查看↗

    This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities.

    Python
    在 GitHub 上查看↗76
查看 MMaDA Parallel 的所有 15 个替代方案→

常见问题解答

tyfeld/mmada-parallel 是做什么的?

[ICLR 2026] MMaDA-Parallel: Parallel Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation

tyfeld/mmada-parallel 的主要功能有哪些?

tyfeld/mmada-parallel 的主要功能包括:Multimodal Diffusion Models。

tyfeld/mmada-parallel 有哪些开源替代品?

tyfeld/mmada-parallel 的开源替代品包括: ml-gsai/llada — LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining… vectorspacelab/omnigen — OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks… alpha-vllm/lumina-dimoo — Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding. gen-verse/mmada — Multimodal Large Diffusion Language Models (NeurIPS 2025). hustvl/diffusionvl — DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models. fudoki-hku/fudoki — This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via…