awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
adobe-research avatar

adobe-research/LaVida-O

0
View on GitHub↗
21 स्टार्स·2 फोर्क्स·Python·4 व्यूज़

LaVida O

[Paper] [Project Site] [Huggingface]

Features

  • Multimodal Diffusion Models - Elastic large masked diffusion for multimodal understanding and generation.

स्टार हिस्ट्री

adobe-research/lavida-o के लिए स्टार हिस्ट्री चार्टadobe-research/lavida-o के लिए स्टार हिस्ट्री चार्ट

AI सर्च

और अधिक बेहतरीन रिपॉजिटरी खोजें

अपनी ज़रूरत को सरल भाषा में बताएं — AI हजारों क्यूरेटेड ओपन-सोर्स प्रोजेक्ट्स को प्रासंगिकता के आधार पर रैंक करता है।

Start searching with AI

LaVida O के ओपन-सोर्स विकल्प

समान ओपन-सोर्स प्रोजेक्ट्स, जो LaVida O के साथ साझा की गई सुविधाओं के आधार पर रैंक किए गए हैं।
  • ml-gsai/lladaML-GSAI का अवतार

    ML-GSAI/LLaDA

    3,580GitHub पर देखें↗

    LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining masked tokens through a diffusion process rather than predicting the next token in a sequence. The project functions as a vision-language diffusion model, converting visual inputs into text responses. It also serves as a preference optimization framework that uses log-likelihood estimation and evidence lower bounds to tune model responses. The system supports multi-round conversational AI and text sequence evaluation. It integrates vision-language embedding for cross-modal con

    Python
    GitHub पर देखें↗3,580
  • vectorspacelab/omnigenVectorSpaceLab का अवतार

    VectorSpaceLab/OmniGen

    4,326GitHub पर देखें↗

    OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks through a single system. It functions as a multimodal diffusion framework that treats diverse vision operations as unified image synthesis problems using shared model weights, removing the need for external adapter modules. The system supports subject-driven image generation to preserve the identity of objects from reference photos and allows for multi-reference image synthesis. It also operates as an instruction-based image editor, modifying visual content through natural languag

    Jupyter Notebookdiffusionimageimage-edit
    GitHub पर देखें↗4,326
  • fudoki-hku/fudokifudoki-hku का अवतार

    fudoki-hku/FUDOKI

    76GitHub पर देखें↗

    This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities.

    Python
    GitHub पर देखें↗76
  • gen-verse/mmadaGen-Verse का अवतार

    Gen-Verse/MMaDA

    1,656GitHub पर देखें↗

    Multimodal Large Diffusion Language Models (NeurIPS 2025)

    Python
    GitHub पर देखें↗1,656
LaVida O के सभी 15 विकल्प देखें→

अक्सर पूछे जाने वाले प्रश्न

adobe-research/lavida-o क्या करता है?

[[Paper]](https://arxiv.org/abs/2509.19244) [[Project Site]](https://homepage.jackli.org/projects/lavida_o/index.html) [[Huggingface]](https://huggingface.co/jacklishufan/LaViDa-O-v1.0/tree/main)

adobe-research/lavida-o की मुख्य विशेषताएं क्या हैं?

adobe-research/lavida-o की मुख्य विशेषताएं हैं: Multimodal Diffusion Models।

adobe-research/lavida-o के कुछ ओपन-सोर्स विकल्प क्या हैं?

adobe-research/lavida-o के ओपन-सोर्स विकल्पों में शामिल हैं: ml-gsai/llada — LLaDA is a masked diffusion language model and conditional text generator. It generates text by iteratively refining… vectorspacelab/omnigen — OmniGen is a unified image generation model and diffusion framework that processes text, images, and vision tasks… fudoki-hku/fudoki — This repository is the official implementation of FUDOKI: Discrete Flow-based Unified Understanding and Generation via… hustvl/diffusionvl — DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models. jacklishufan/lavida — [[Paper]](paper/paper.pdf) [[Arxiv]](https://arxiv.org/abs/2505.16839) [[Checkpoints]](https://huggingface.co/collectio… gen-verse/mmada — Multimodal Large Diffusion Language Models (NeurIPS 2025).