awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
ByteFlow-AI avatar

ByteFlow-AI/TokenFlow

0
View on GitHub↗
465 स्टार्स·10 फोर्क्स·Python·Apache-2.0·11 व्यूज़arxiv.org/abs/2412.03069↗

TokenFlow

[CVPR 2025] 🔥 Official impl. of "TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation".

Features

  • Computer Vision Research - Unified image tokenizer for multimodal understanding and generation.
  • Generative AI - Unified tokenizer for multimodal understanding and generation.
  • Unified Models - Unified framework for multimodal token processing.
  • Unified Multimodal Models - Framework for multimodal token-based generation.

स्टार हिस्ट्री

byteflow-ai/tokenflow के लिए स्टार हिस्ट्री चार्टbyteflow-ai/tokenflow के लिए स्टार हिस्ट्री चार्ट

AI सर्च

और अधिक बेहतरीन रिपॉजिटरी खोजें

अपनी ज़रूरत को सरल भाषा में बताएं — AI हजारों क्यूरेटेड ओपन-सोर्स प्रोजेक्ट्स को प्रासंगिकता के आधार पर रैंक करता है।

Start searching with AI

TokenFlow के ओपन-सोर्स विकल्प

समान ओपन-सोर्स प्रोजेक्ट्स, जो TokenFlow के साथ साझा की गई सुविधाओं के आधार पर रैंक किए गए हैं।
  • bytedance/lancebytedance का अवतार

    bytedance/Lance

    1,250GitHub पर देखें↗

    A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.

    Python
    GitHub पर देखें↗1,250
  • bytedance/x-dynabytedance का अवतार

    bytedance/X-Dyna

    269GitHub पर देखें↗

    CVPR 2025 Highlight X-Dyna: Expressive Dynamic Human Image Animation

    Python
    GitHub पर देखें↗269
  • alpha-vllm/lumina-dimooAlpha-VLLM का अवतार

    Alpha-VLLM/Lumina-DiMOO

    1,001GitHub पर देखें↗

    Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding

    Python
    GitHub पर देखें↗1,001
  • deepseek-ai/janusdeepseek-ai का अवतार

    deepseek-ai/Janus

    17,746GitHub पर देखें↗

    Janus is a multimodal large language model and unified framework that integrates visual understanding and image generation within a single neural network. It functions as both a visual understanding model for analyzing images and a text-to-image generator. The system uses a unified transformer backbone and a multimodal latent space to bridge the gap between text and visual data. This architecture employs decoupled visual encoding and cross-modal tokenization to separate the paths for discriminative understanding and generative tasks, representing images as grids of discrete codes. The projec

    Pythonany-to-anyfoundation-modelsllm
    GitHub पर देखें↗17,746
TokenFlow के सभी 30 विकल्प देखें→

अक्सर पूछे जाने वाले प्रश्न

byteflow-ai/tokenflow क्या करता है?

[CVPR 2025] 🔥 Official impl. of "TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation".

byteflow-ai/tokenflow की मुख्य विशेषताएं क्या हैं?

byteflow-ai/tokenflow की मुख्य विशेषताएं हैं: Computer Vision Research, Generative AI, Unified Models, Unified Multimodal Models।

byteflow-ai/tokenflow के कुछ ओपन-सोर्स विकल्प क्या हैं?

byteflow-ai/tokenflow के ओपन-सोर्स विकल्पों में शामिल हैं: bytedance/x-dyna — [CVPR 2025 Highlight] X-Dyna: Expressive Dynamic Human Image Animation. facebookresearch/tuna-2 — Official implementation of Tuna-2: Pixel Embeddings Beat Vision Encoders for Unified Understanding and Generation. alpha-vllm/lumina-dimoo — Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding. bytedance/lance — A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing. deepseek-ai/janus — Janus is a multimodal large language model and unified framework that integrates visual understanding and image… hustvl/lightningdit — [CVPR 2025 Oral] Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models.