awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Back to alexanderswerdlow/unidisc

Projects sharing features with Unidisc

15 open-source projects similar to alexanderswerdlow/unidisc, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.

  • ml-gsai/lladaML-GSAI avatar

    ML-GSAI/LLaDA

    3,580View on GitHub↗

    LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining masked tokens through a diffusion process rather than predicting the next token in a sequence. The project functions as a vision-language diffusion model, converting visual inputs into text responses. It also serves as a preference optimization framework that uses log-likelihood estimation and evidence lower bounds to tune model responses. The system supports multi-round conversational AI and text sequence evaluation. It integrates vision-language embedding for cross-modal con

    Python
    View on GitHub↗3,580
  • vectorspacelab/omnigenVectorSpaceLab avatar

    VectorSpaceLab/OmniGen

    4,326View on GitHub↗

    OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks through a single system. It functions as a multimodal diffusion framework that treats diverse vision operations as unified image synthesis problems using shared model weights, removing the need for external adapter modules. The system supports subject-driven image generation to preserve the identity of objects from reference photos and allows for multi-reference image synthesis. It also operates as an instruction-based image editor, modifying visual content through natural languag

    Jupyter Notebookdiffusionimageimage-edit
    View on GitHub↗4,326
  • fudoki-hku/fudokifudoki-hku avatar

    fudoki-hku/FUDOKI

    76View on GitHub↗

    This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities.

    Python
    View on GitHub↗76
  • gen-verse/mmadaGen-Verse avatar

    Gen-Verse/MMaDA

    1,656View on GitHub↗

    Multimodal Large Diffusion Language Models (NeurIPS 2025)

    Python
    View on GitHub↗1,656

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • hustvl/diffusionvlhustvl avatar

    hustvl/DiffusionVL

    149View on GitHub↗

    DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models

    Python
    View on GitHub↗149
  • jacklishufan/lavidajacklishufan avatar

    jacklishufan/LaViDa

    221View on GitHub↗

    [Paper](paper/paper.pdf) [Arxiv](https://arxiv.org/abs/2505.16839) [Checkpoints](https://huggingface.co/collections/jacklishufan/lavida-10-682ecf5a5fa8c5df85c61ded) [Data](https://huggingface.co/datasets/jacklishufan/lavida-train) [Website](https://homepage.jackli.org/projects/lavida/)

    Python
    View on GitHub↗221
  • jiyt17/rediffjiyt17 avatar

    jiyt17/ReDiff

    45View on GitHub↗

    We introduce ReDiff, a refining-enhanced vision-language diffusion model.

    Python
    View on GitHub↗45
  • m-e-agi-lab/mudditM-E-AGI-Lab avatar

    M-E-AGI-Lab/Muddit

    117View on GitHub↗

    Muddit is the 2nd generation Meissonic. It is built upon discrete diffusion for unified and efficient multimodal generation.

    Python
    View on GitHub↗117
  • ml-gsai/llada-vML-GSAI avatar

    ML-GSAI/LLaDA-V

    345View on GitHub↗

    2026.03.23 We are excited to introduce LLaDA-o, the latest model in the LLaDA series. As an effective and length-adaptive omni diffusion model for unified multimodal understanding and generation, LLaDA-o extends the LLaDA line to broader multimodal settings, supporting visual understanding,…

    Python
    View on GitHub↗345
  • openhelix-team/unified-diffusion-vlaOpenHelix-Team avatar

    OpenHelix-Team/Unified-Diffusion-VLA

    182View on GitHub↗

    Jiayi Chen¹\,Wenxuan Song¹†\, Pengxiang Ding²˒³, Ziyang Zhou¹, Han Zhao²˒³, Feilong Tang⁴,Donglin Wang², Haoang Li¹‡

    Python
    View on GitHub↗182
  • tyfeld/mmada-paralleltyfeld avatar

    tyfeld/MMaDA-Parallel

    300View on GitHub↗

    ICLR 2026 MMaDA-Parallel: Parallel Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation

    Python
    View on GitHub↗300
  • yu-rp/dimpleyu-rp avatar

    yu-rp/Dimple

    117View on GitHub↗

    Dimple, the first Discrete Diffusion Multimodal Large Language Model

    Python
    View on GitHub↗117
  • adobe-research/lavida-oadobe-research avatar

    adobe-research/LaVida-O

    21View on GitHub↗

    [Paper](https://arxiv.org/abs/2509.19244) [Project Site](https://homepage.jackli.org/projects/lavida_o/index.html) [Huggingface](https://huggingface.co/jacklishufan/LaViDa-O-v1.0/tree/main)

    Python
    View on GitHub↗21
  • zihohe/vidladaziHoHe avatar

    ziHoHe/VidLaDA

    10View on GitHub↗
    Python
    View on GitHub↗10
  • alpha-vllm/lumina-dimooAlpha-VLLM avatar

    Alpha-VLLM/Lumina-DiMOO

    1,001View on GitHub↗

    Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding

    Python
    View on GitHub↗1,001