awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
hustvl avatar

hustvl/DiffusionVL

0
View on GitHub↗
149 نجوم·9 تفرعات·Python·Apache-2.0·6 مشاهدات

DiffusionVL

DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models

Features

  • Multimodal Diffusion Models - Translating autoregressive models into vision-language diffusion models.

سجل النجوم

مخطط تاريخ النجوم لـ hustvl/diffusionvlمخطط تاريخ النجوم لـ hustvl/diffusionvl

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

بدائل مفتوحة المصدر لـ DiffusionVL

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع DiffusionVL.
  • ml-gsai/lladaالصورة الرمزية لـ ML-GSAI

    ML-GSAI/LLaDA

    3,580عرض على GitHub↗

    LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining masked tokens through a diffusion process rather than predicting the next token in a sequence. The project functions as a vision-language diffusion model, converting visual inputs into text responses. It also serves as a preference optimization framework that uses log-likelihood estimation and evidence lower bounds to tune model responses. The system supports multi-round conversational AI and text sequence evaluation. It integrates vision-language embedding for cross-modal con

    Python
    عرض على GitHub↗3,580
  • vectorspacelab/omnigenالصورة الرمزية لـ VectorSpaceLab

    VectorSpaceLab/OmniGen

    4,326عرض على GitHub↗

    OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks through a single system. It functions as a multimodal diffusion framework that treats diverse vision operations as unified image synthesis problems using shared model weights, removing the need for external adapter modules. The system supports subject-driven image generation to preserve the identity of objects from reference photos and allows for multi-reference image synthesis. It also operates as an instruction-based image editor, modifying visual content through natural languag

    Jupyter Notebookdiffusionimageimage-edit
    عرض على GitHub↗4,326
  • alpha-vllm/lumina-dimooالصورة الرمزية لـ Alpha-VLLM

    Alpha-VLLM/Lumina-DiMOO

    1,001عرض على GitHub↗

    Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding

    Python
    عرض على GitHub↗1,001
  • fudoki-hku/fudokiالصورة الرمزية لـ fudoki-hku

    fudoki-hku/FUDOKI

    76عرض على GitHub↗

    This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities.

    Python
    عرض على GitHub↗76
عرض جميع البدائل الـ 15 لـ DiffusionVL→

الأسئلة الشائعة

ما هي وظيفة hustvl/diffusionvl؟

DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models

ما هي الميزات الرئيسية لـ hustvl/diffusionvl؟

الميزات الرئيسية لـ hustvl/diffusionvl هي: Multimodal Diffusion Models.

ما هي البدائل مفتوحة المصدر لـ hustvl/diffusionvl؟

تشمل البدائل مفتوحة المصدر لـ hustvl/diffusionvl: ml-gsai/llada — LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining… vectorspacelab/omnigen — OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks… alpha-vllm/lumina-dimoo — Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding. gen-verse/mmada — Multimodal Large Diffusion Language Models (NeurIPS 2025). jacklishufan/lavida — [[Paper]](paper/paper.pdf) [[Arxiv]](https://arxiv.org/abs/2505.16839) [[Checkpoints]](https://huggingface.co/collectio… fudoki-hku/fudoki — This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via…