awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Back to tulerfeng/video-r1

Projects sharing features with Video R1

30 open-source projects similar to tulerfeng/video-r1, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.

  • osilly/vision-r1Osilly avatar

    Osilly/Vision-R1

    1,475View on GitHub↗

    The official repo for "Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models".

    Python
    View on GitHub↗1,475
  • opengvlab/videochat-r1OpenGVLab avatar

    OpenGVLab/VideoChat-R1

    267View on GitHub↗

    x 2025/09/26:🔥🔥🔥 We release our VideoChat-R1.5 model at Huggingface, paper, and eval code. - x 2025/09/22: 🎉🎉🎉 Our VideoChat-R1.5 is accepted by NIPS2025. - x 2025/04/22:🔥🔥🔥 We release our VideoChat-R1-caption at Huggingface. - x 2025/04/14:🔥🔥🔥 We release our VideoChat-R1 and…

    Python
    View on GitHub↗267
  • om-ai-lab/vlm-r1om-ai-lab avatar

    om-ai-lab/VLM-R1

    5,991View on GitHub↗

    VLM-R1 is a reasoning vision-language model and embodied AI framework designed to map visual inputs and language instructions into physical navigation waypoints and robotic actions. It functions as a multimodal policy optimizer and an open vocabulary detector capable of locating objects based on arbitrary natural language descriptions. The system distinguishes itself through the use of chain-of-thought reasoning and reinforcement learning to solve complex visual and spatial tasks. It utilizes a video semantic memory system, which employs a visual cache to maintain a history of live video for

    Python
    View on GitHub↗5,991
  • liuziyu77/visual-rftLiuziyu77 avatar

    Liuziyu77/Visual-RFT

    2,250View on GitHub↗

    Visual-RFT: Visual Reinforcement Fine-Tuning Ziyu Liu · Zeyi Sun · Yuhang Zang · Xiaoyi Dong · Yuhang Cao · Haodong Duan · Dahua Lin · Jiaqi Wang Accepted By ICCV 2025! 📖 Paper | 🤗 Datasets | 🤗 Daily Paper 🌈We introduce Visual Reinforcement Fine-tuning (Visual-RFT) , the first comprehensive…

    Jupyter Notebook
    View on GitHub↗2,250

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • qwenlm/qwen2.5QwenLM avatar

    QwenLM/Qwen2.5

    27,307View on GitHub↗

    Qwen2.5 is a suite of large language model foundation models designed for natural language generation, code production, and complex mathematical reasoning. The project encompasses a multilingual language model capable of processing dozens of languages and a specialized code generation model for technical problem solving and debugging. The framework is distinguished by its long context capabilities, enabling the analysis of massive inputs ranging from 256K up to 1 million tokens. It further functions as an agentic framework, utilizing standardized templates and parsers to execute autonomous wo

    Python
    View on GitHub↗27,307
  • deepseek-ai/janusdeepseek-ai avatar

    deepseek-ai/Janus

    17,746View on GitHub↗

    Janus is a multimodal large language model and unified framework that integrates visual understanding and image generation within a single neural network. It functions as both a visual understanding model for analyzing images and a text-to-image generator. The system uses a unified transformer backbone and a multimodal latent space to bridge the gap between text and visual data. This architecture employs decoupled visual encoding and cross-modal tokenization to separate the paths for discriminative understanding and generative tasks, representing images as grids of discrete codes. The projec

    Pythonany-to-anyfoundation-modelsllm
    View on GitHub↗17,746
  • appletea233/temporal-r1appletea233 avatar

    appletea233/Temporal-R1

    62View on GitHub↗

    Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency

    Python
    View on GitHub↗62
  • aliyun/qwen-dianjinA

    aliyun/qwen-dianjin

    0View on GitHub↗
    View on GitHub↗0
  • agentica-project/deepscalerA

    agentica-project/deepscaler

    0View on GitHub↗
    View on GitHub↗0
  • bklieger-groq/g1B

    bklieger-groq/g1

    0View on GitHub↗
    View on GitHub↗0
  • brendanhogan/deepseekrl-extendedbrendanhogan avatar

    brendanhogan/DeepSeekRL-Extended

    252View on GitHub↗

    Exploring Applications of GRPO

    Python
    View on GitHub↗252
  • bytedance-seed/seed-thinking-v1.5B

    ByteDance-Seed/Seed-Thinking-v1.5

    0View on GitHub↗
    View on GitHub↗0
  • csfufu/revisual-r1C

    CSfufu/Revisual-R1

    0View on GitHub↗
    View on GitHub↗0
  • datawhalechina/unlock-deepseekdatawhalechina avatar

    datawhalechina/unlock-deepseek

    733View on GitHub↗

    DeepSeek 系列工作解读、扩展和复现。

    Python
    View on GitHub↗733
  • baichuan-inc/baichuan-m1-14bbaichuan-inc avatar

    baichuan-inc/Baichuan-M1-14B

    219View on GitHub↗

    Baichuan-M1-14B

    View on GitHub↗219
  • aidc-ai/marco-o1AIDC-AI avatar

    AIDC-AI/Marco-o1

    1,540View on GitHub↗

    An Open Large Reasoning Model for Real-World Solutions

    Python
    View on GitHub↗1,540
  • deepseek-ai/deepseek-v4D

    deepseek-ai/DeepSeek-V4

    0View on GitHub↗
    View on GitHub↗0
  • baibizhe/efficient-r1-vllmB

    baibizhe/Efficient-R1-VLLM

    0View on GitHub↗
    View on GitHub↗0
  • dhcode-cpp/x-r1D

    dhcode-cpp/X-R1

    0View on GitHub↗
    View on GitHub↗0
  • diankun-wu/spatial-mllmdiankun-wu avatar

    diankun-wu/Spatial-MLLM

    470View on GitHub↗

    Yi-Hsin Hung 1 , Yueqi Duan 1 , Equal Contribution. 1 Tsinghua University NeurIPS 2025 (Spotlight)

    Python
    View on GitHub↗470
  • dvlab-research/seg-zerodvlab-research avatar

    dvlab-research/Seg-Zero

    632View on GitHub↗

    Project Page For "Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement"

    Python
    View on GitHub↗632
  • egolife-ai/ego-r1egolife-ai avatar

    egolife-ai/Ego-R1

    158View on GitHub↗

    TPAMI 2026 Ego-R1: Agentic Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning

    Python
    View on GitHub↗158
  • eric-ai-lab/griteric-ai-lab avatar

    eric-ai-lab/GRIT

    190View on GitHub↗

    Grounded Reasoning wiht Texts and Images (GRIT) is a novel method for training Multimodal Large Language Models (MLLMs) to perform grounded reasoning by generating reasoning chains that interleave natural language and explicit bounding box coordinates. This approach can use as few as 20 training…

    Python
    View on GitHub↗190
  • evelinehong/3d-clr-officialevelinehong avatar

    evelinehong/3D-CLR-Official

    85View on GitHub↗

    Checkpoints take up a lot of space. Please email yninghong@gmail.com if you need them.

    Python
    View on GitHub↗85
  • evolvinglmms-lab/open-r1-multimodalEvolvingLMMs-Lab avatar

    EvolvingLMMs-Lab/open-r1-multimodal

    1,484View on GitHub↗
    Python
    View on GitHub↗1,484
  • facebookresearch/swe-rlfacebookresearch avatar

    facebookresearch/swe-rl

    704View on GitHub↗

    🧐 About | 🚀 Quick Start | 🐣 Agentless Mini | 📝 Citation | 🙏 Acknowledgements

    Python
    View on GitHub↗704
  • fancy-mllm/r1-onevisionFancy-MLLM avatar

    Fancy-MLLM/R1-Onevision

    581View on GitHub↗

    R1-onevision, a visual language model capable of deep CoT reasoning.

    Python
    View on GitHub↗581
  • fanqingm/r1-multimodal-journeyF

    FanqingM/R1-Multimodal-Journey

    0View on GitHub↗
    View on GitHub↗0
  • flagai-open/openseekFlagAI-Open avatar

    FlagAI-Open/OpenSeek

    262View on GitHub↗

    OpenSeek aims to unite the global open source community to drive collaborative innovation in algorithms, data and systems to develop next-generation models.

    Python
    View on GitHub↗262
  • deepseek-ai/deepseek-r1deepseek-ai avatar

    deepseek-ai/DeepSeek-R1

    91,996View on GitHub↗

    DeepSeek-R1 is an open-weights large language model focused on advanced reasoning. It uses chain-of-thought processing and internal monologues to solve complex mathematical and logical problems by breaking tasks into sequential, verifiable thought processes. The model is developed using reinforcement learning to optimize reasoning patterns and verify logical steps. It employs a distillation process to transfer these high-performance logic capabilities from a large teacher model into smaller, computationally efficient versions. The training framework incorporates group relative policy optimiz

    View on GitHub↗91,996