awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
tulerfeng avatar

tulerfeng/Video-R1

0
View on GitHub↗
878 stars·46 forks·Python·13 views

Video R1

[📖 Paper] [🤗 Video-R1-7B-model] [🤗 Video-R1-train-data] [🤖 Video-R1-7B-model] [🤖 Video-R1-train-data]

Features

  • Multimodal Understanding - Reinforcing video reasoning capabilities in multimodal models.
  • Reasoning Models - Video-centric reasoning model implementation.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Star history

Star history chart for tulerfeng/video-r1Star history chart for tulerfeng/video-r1

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

Projects sharing features with Video R1

These projects share indexed features with Video R1. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • opengvlab/videochat-r1OpenGVLab avatar

    OpenGVLab/VideoChat-R1

    267View on GitHub↗

    x 2025/09/26:🔥🔥🔥 We release our VideoChat-R1.5 model at Huggingface, paper, and eval code. - x 2025/09/22: 🎉🎉🎉 Our VideoChat-R1.5 is accepted by NIPS2025. - x 2025/04/22:🔥🔥🔥 We release our VideoChat-R1-caption at Huggingface. - x 2025/04/14:🔥🔥🔥 We release our VideoChat-R1 and…

    Python
    View on GitHub↗267
  • om-ai-lab/vlm-r1om-ai-lab avatar

    om-ai-lab/VLM-R1

    5,991View on GitHub↗

    VLM-R1 is a reasoning vision-language model and embodied AI framework designed to map visual inputs and language instructions into physical navigation waypoints and robotic actions. It functions as a multimodal policy optimizer and an open vocabulary detector capable of locating objects based on arbitrary natural language descriptions. The system distinguishes itself through the use of chain-of-thought reasoning and reinforcement learning to solve complex visual and spatial tasks. It utilizes a video semantic memory system, which employs a visual cache to maintain a history of live video for

    Python
    View on GitHub↗5,991
  • liuziyu77/visual-rftLiuziyu77 avatar

    Liuziyu77/Visual-RFT

    2,250View on GitHub↗

    Visual-RFT: Visual Reinforcement Fine-Tuning Ziyu Liu · Zeyi Sun · Yuhang Zang · Xiaoyi Dong · Yuhang Cao · Haodong Duan · Dahua Lin · Jiaqi Wang Accepted By ICCV 2025! 📖 Paper | 🤗 Datasets | 🤗 Daily Paper 🌈We introduce Visual Reinforcement Fine-tuning (Visual-RFT) , the first comprehensive…

    Jupyter Notebook
    View on GitHub↗2,250
  • osilly/vision-r1Osilly avatar

    Osilly/Vision-R1

    1,475View on GitHub↗

    The official repo for "Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models".

    Python
    View on GitHub↗1,475
Compare all 30 related projects→

Frequently asked questions

What does tulerfeng/video-r1 do?

[📖 Paper] [🤗 Video-R1-7B-model] [🤗 Video-R1-train-data] [🤖 Video-R1-7B-model] [🤖 Video-R1-train-data]

What are the main features of tulerfeng/video-r1?

The main features of tulerfeng/video-r1 are: Multimodal Understanding, Reasoning Models.

Which projects share features with tulerfeng/video-r1?

Projects with overlapping indexed features include: opengvlab/videochat-r1 — [x] 2025/09/26:🔥🔥🔥 We release our VideoChat-R1.5 model at Huggingface, paper, and eval code. - [x] 2025/09/22:… osilly/vision-r1 — The official repo for "Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models". om-ai-lab/vlm-r1 — VLM-R1 is a reasoning vision-language model and embodied AI framework designed to map visual inputs and language… liuziyu77/visual-rft — Visual-RFT: Visual Reinforcement Fine-Tuning Ziyu Liu · Zeyi Sun · Yuhang Zang · Xiaoyi Dong · Yuhang Cao · Haodong… deepseek-ai/janus — Janus is a multimodal large language model and unified framework that integrates visual understanding and image… qwenlm/qwen2.5 — Qwen2.5 is a suite of large language model foundation models designed for natural language generation, code…