awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
SkyworkAI avatar

SkyworkAI/Vitron

0
View on GitHub↗
577 Stars·34 Forks·Python·6 Aufrufevitron-llm.github.io↗

Vitron

NeurIPS 2024 Paper

Features

  • Unified Models - Unified vision-language model for generation and editing.
  • Unified Multimodal Models - Multimodal model for vision and language tasks.
  • Unified Understanding and Generation - Hybrid instruction method for discrete text and continuous visual signals.

Star-Verlauf

Star-Verlauf für skyworkai/vitronStar-Verlauf für skyworkai/vitron

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Open-Source-Alternativen zu Vitron

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit Vitron.
  • deepseek-ai/janusAvatar von deepseek-ai

    deepseek-ai/Janus

    17,746Auf GitHub ansehen↗

    Janus is a multimodal large language model and unified framework that integrates visual understanding and image generation within a single neural network. It functions as both a visual understanding model for analyzing images and a text-to-image generator. The system uses a unified transformer backbone and a multimodal latent space to bridge the gap between text and visual data. This architecture employs decoupled visual encoding and cross-modal tokenization to separate the paths for discriminative understanding and generative tasks, representing images as grids of discrete codes. The projec

    Pythonany-to-anyfoundation-modelsllm
    Auf GitHub ansehen↗17,746
  • alpha-vllm/lumina-dimooAvatar von Alpha-VLLM

    Alpha-VLLM/Lumina-DiMOO

    1,001Auf GitHub ansehen↗

    Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding

    Python
    Auf GitHub ansehen↗1,001
  • bytedance/lanceAvatar von bytedance

    bytedance/Lance

    1,250Auf GitHub ansehen↗

    A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.

    Python
    Auf GitHub ansehen↗1,250
  • byteflow-ai/tokenflowAvatar von ByteFlow-AI

    ByteFlow-AI/TokenFlow

    465Auf GitHub ansehen↗

    CVPR 2025 🔥 Official impl. of "TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation".

    Python
    Auf GitHub ansehen↗465
Alle 15 Alternativen zu Vitron anzeigen→

Häufig gestellte Fragen

Was macht skyworkai/vitron?

NeurIPS 2024 Paper

Was sind die Hauptfunktionen von skyworkai/vitron?

Die Hauptfunktionen von skyworkai/vitron sind: Unified Models, Unified Multimodal Models, Unified Understanding and Generation.

Welche Open-Source-Alternativen gibt es zu skyworkai/vitron?

Open-Source-Alternativen zu skyworkai/vitron sind unter anderem: deepseek-ai/janus — Janus is a multimodal large language model and unified framework that integrates visual understanding and image… alpha-vllm/lumina-dimoo — Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding. bytedance/lance — A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing. byteflow-ai/tokenflow — [CVPR 2025] 🔥 Official impl. of "TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation". facebookresearch/tuna-2 — Official implementation of Tuna-2: Pixel Embeddings Beat Vision Encoders for Unified Understanding and Generation. lehduong/onediffusion.