awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
Β© 2026 Bringes Technology SRLΒ·VAT RO45896025Β·hello@awesome-repositories.com
OpenGVLab avatar

OpenGVLab/VideoChat-R1

0
View on GitHub↗
267 starsΒ·10 forksΒ·PythonΒ·13 views

VideoChat R1

[x] 2025/09/26:πŸ”₯πŸ”₯πŸ”₯ We release our VideoChat-R1.5 model at Huggingface, paper, and eval code. - [x] 2025/09/22: πŸŽ‰πŸŽ‰πŸŽ‰ Our VideoChat-R1.5 is accepted by NIPS2025. - [x] 2025/04/22:πŸ”₯πŸ”₯πŸ”₯ We release our VideoChat-R1-caption at Huggingface. - [x] 2025/04/14:πŸ”₯πŸ”₯πŸ”₯ We release our VideoChat-R1 and…

Features

  • Multimodal Understanding - Enhancing spatio-temporal perception via reinforcement fine-tuning.
  • Reasoning Models - Video-based reasoning model.

Star history

Star history chart for opengvlab/videochat-r1Star history chart for opengvlab/videochat-r1

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English β€” the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does opengvlab/videochat-r1 do?

[x] 2025/09/26:πŸ”₯πŸ”₯πŸ”₯ We release our VideoChat-R1.5 model at Huggingface, paper, and eval code. - [x] 2025/09/22: πŸŽ‰πŸŽ‰πŸŽ‰ Our VideoChat-R1.5 is accepted by NIPS2025. - [x] 2025/04/22:πŸ”₯πŸ”₯πŸ”₯ We release our VideoChat-R1-caption at Huggingface. - [x] 2025/04/14:πŸ”₯πŸ”₯πŸ”₯ We release our VideoChat-R1 and…

What are the main features of opengvlab/videochat-r1?

The main features of opengvlab/videochat-r1 are: Multimodal Understanding, Reasoning Models.

Which projects share features with opengvlab/videochat-r1?

Projects with overlapping indexed features include: tulerfeng/video-r1 β€” [πŸ“– Paper] [πŸ€— Video-R1-7B-model] [πŸ€— Video-R1-train-data] [πŸ€– Video-R1-7B-model] [πŸ€– Video-R1-train-data]. osilly/vision-r1 β€” The official repo for "Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models". om-ai-lab/vlm-r1 β€” VLM-R1 is a reasoning vision-language model and embodied AI framework designed to map visual inputs and language… liuziyu77/visual-rft β€” Visual-RFT: Visual Reinforcement Fine-Tuning Ziyu Liu Β· Zeyi Sun Β· Yuhang Zang Β· Xiaoyi Dong Β· Yuhang Cao Β· Haodong… qwenlm/qwen2.5 β€” Qwen2.5 is a suite of large language model foundation models designed for natural language generation, code… deepseek-ai/janus β€” Janus is a multimodal large language model and unified framework that integrates visual understanding and image…

Projects sharing features with VideoChat R1

These projects share indexed features with VideoChat R1. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • osilly/vision-r1Osilly avatar

    Osilly/Vision-R1

    1,475View on GitHub↗

    The official repo for "Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models".

    Python
    View on GitHub↗1,475
  • om-ai-lab/vlm-r1om-ai-lab avatar

    om-ai-lab/VLM-R1

    5,991View on GitHub↗

    VLM-R1 is a reasoning vision-language model and embodied AI framework designed to map visual inputs and language instructions into physical navigation waypoints and robotic actions. It functions as a multimodal policy optimizer and an open vocabulary detector capable of locating objects based on arbitrary natural language descriptions. The system distinguishes itself through the use of chain-of-thought reasoning and reinforcement learning to solve complex visual and spatial tasks. It utilizes a video semantic memory system, which employs a visual cache to maintain a history of live video for

    Python
    View on GitHub↗5,991
  • liuziyu77/visual-rftLiuziyu77 avatar

    Liuziyu77/Visual-RFT

    2,250View on GitHub↗

    Visual-RFT: Visual Reinforcement Fine-Tuning Ziyu Liu Β· Zeyi Sun Β· Yuhang Zang Β· Xiaoyi Dong Β· Yuhang Cao Β· Haodong Duan Β· Dahua Lin Β· Jiaqi Wang Accepted By ICCV 2025! πŸ“– Paper | πŸ€— Datasets | πŸ€— Daily Paper 🌈We introduce Visual Reinforcement Fine-tuning (Visual-RFT) , the first comprehensive…

    Jupyter Notebook
    View on GitHub↗2,250
  • tulerfeng/video-r1tulerfeng avatar

    tulerfeng/Video-R1

    878View on GitHub↗

    πŸ“– Paper πŸ€— Video-R1-7B-model πŸ€— Video-R1-train-data πŸ€– Video-R1-7B-model πŸ€– Video-R1-train-data

    Python
    View on GitHub↗878
Compare all 30 related projects→