awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Back to wangqinsi1/vision-zero

Projects sharing features with Vision Zero

30 open-source projects similar to wangqinsi1/vision-zero, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.

  • chengsong-huang/r-zeroChengsong-Huang avatar

    Chengsong-Huang/R-Zero

    822View on GitHub↗

    Check out our paper or webpage for the details

    Python
    View on GitHub↗822
  • wantbook-book/serlwantbook-book avatar

    wantbook-book/SeRL

    23View on GitHub↗

    SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Data

    Python
    View on GitHub↗23
  • deepseek-ai/janusdeepseek-ai avatar

    deepseek-ai/Janus

    17,746View on GitHub↗

    Janus is a multimodal large language model and unified framework that integrates visual understanding and image generation within a single neural network. It functions as both a visual understanding model for analyzing images and a text-to-image generator. The system uses a unified transformer backbone and a multimodal latent space to bridge the gap between text and visual data. This architecture employs decoupled visual encoding and cross-modal tokenization to separate the paths for discriminative understanding and generative tasks, representing images as grids of discrete codes. The projec

    Pythonany-to-anyfoundation-modelsllm
    View on GitHub↗17,746
  • egolife-ai/ego-r1egolife-ai avatar

    egolife-ai/Ego-R1

    158View on GitHub↗

    TPAMI 2026 Ego-R1: Agentic Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning

    Python
    View on GitHub↗158
  • iamhankai/forest-of-thoughtiamhankai avatar

    iamhankai/Forest-of-Thought

    55View on GitHub↗

    Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning

    Python
    View on GitHub↗55

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • eric-ai-lab/griteric-ai-lab avatar

    eric-ai-lab/GRIT

    190View on GitHub↗

    Grounded Reasoning wiht Texts and Images (GRIT) is a novel method for training Multimodal Large Language Models (MLLMs) to perform grounded reasoning by generating reasoning chains that interleave natural language and explicit bounding box coordinates. This approach can use as few as 20 training…

    Python
    View on GitHub↗190
  • diankun-wu/spatial-mllmdiankun-wu avatar

    diankun-wu/Spatial-MLLM

    470View on GitHub↗

    Yi-Hsin Hung 1 , Yueqi Duan 1 , Equal Contribution. 1 Tsinghua University NeurIPS 2025 (Spotlight)

    Python
    View on GitHub↗470
  • gpoesia/minimogpoesia avatar

    gpoesia/minimo

    36View on GitHub↗

    This is the implementation of the following paper:

    Rust
    View on GitHub↗36
  • hitsz-tmg/veripoHITsz-TMG avatar

    HITsz-TMG/VerIPO

    10View on GitHub↗

    📄 Paper Link 🤗 VerIPO-7B-v1.0

    Python
    View on GitHub↗10
  • hkust-nlp/mstarH

    hkust-nlp/mstar

    0View on GitHub↗

    :star: Project Page

    View on GitHub↗0
  • microsoft/codetM

    microsoft/CodeT

    0View on GitHub↗

    This repository contains projects that aims to equip large-scale pretrained language models with better programming and reasoning skills. These projects are presented by Microsoft Research Asia and Microsoft Azure AI.

    View on GitHub↗0
  • insightllm/rl-without-gtinsightLLM avatar

    insightLLM/rl-without-gt

    8View on GitHub↗

    //: # (![Hugging Face Collection(https://img.shields.io/badge/Models-fcd022?style=for-the-badge&logo=huggingface&logoColor=000)]())

    Jupyter Notebook
    View on GitHub↗8
  • jun297/v1jun297 avatar

    jun297/v1

    20View on GitHub↗

    Jiwan Chung   Junhyeok Kim   Siyeol Kim   Jaeyoung Lee   Minsoo Kim   Youngjae Yu

    Python
    View on GitHub↗20
  • kelaxon/ssr-zeroKelaxon avatar

    Kelaxon/SSR-Zero

    9View on GitHub↗

    💫SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation

    View on GitHub↗9
  • leaplabthu/absolute-zero-reasonerLeapLabTHU avatar

    LeapLabTHU/Absolute-Zero-Reasoner

    1,869View on GitHub↗

    ⚙️ Algorithm Flow • 📊 Results ✨ Getting Started • 🏋️ Training • 🔧 Usage • 📃 Evaluation 🎈 Citation • 🌻 Acknowledgement • 📧 Contact • 📈 Star History

    Python
    View on GitHub↗1,869
  • kwai-keye/keyeKwai-Keye avatar

    Kwai-Keye/Keye

    797View on GitHub↗
    Python
    View on GitHub↗797
  • lili-chen/self-questioning-lmlili-chen avatar

    lili-chen/self-questioning-lm

    57View on GitHub↗

    Self-Questioning Language Models

    Python
    View on GitHub↗57
  • liuziyu77/visual-rftLiuziyu77 avatar

    Liuziyu77/Visual-RFT

    2,250View on GitHub↗

    Visual-RFT: Visual Reinforcement Fine-Tuning Ziyu Liu · Zeyi Sun · Yuhang Zang · Xiaoyi Dong · Yuhang Cao · Haodong Duan · Dahua Lin · Jiaqi Wang Accepted By ICCV 2025! 📖 Paper | 🤗 Datasets | 🤗 Daily Paper 🌈We introduce Visual Reinforcement Fine-tuning (Visual-RFT) , the first comprehensive…

    Jupyter Notebook
    View on GitHub↗2,250
  • longmalongma/tw-grpolongmalongma avatar

    longmalongma/TW-GRPO

    36View on GitHub↗

    🤗 Model &nbsp&nbsp | &nbsp&nbsp 📑 Paper &nbsp&nbsp

    Python
    View on GitHub↗36
  • lucidrains/self-rewarding-lm-pytorchlucidrains avatar

    lucidrains/self-rewarding-lm-pytorch

    1,410View on GitHub↗

    Implementation of the training framework proposed in Self-Rewarding Language Model , from MetaAI

    Python
    View on GitHub↗1,410
  • maifoundations/visionary-r1maifoundations avatar

    maifoundations/Visionary-R1

    44View on GitHub↗

    Visionary-R1: Mitigating Shortcuts in Visual Reasoning with Reinforcement Learning A new RL method for visual reasoning, which significantly outperforms vanilla GRPO, and bypasses the need for explicit chain-of-thought supervision during training.

    Python
    View on GitHub↗44
  • masworks/mas-gptMASWorks avatar

    MASWorks/MAS-GPT

    80View on GitHub↗

    📑 ArXiv Paper | 🤗 Model Weights

    Python
    View on GitHub↗80
  • ezelikman/starezelikman avatar

    ezelikman/STaR

    227View on GitHub↗

    1. STaR 2. Mesh Transformer JAX 1. Updates 3. Pretrained Models 1. GPT-J-6B 1. Links 2. Acknowledgments 3. License 4. Model Details 5. Zero-Shot Evaluations 4. Architecture and Usage 1. Fine-tuning 2. JAX Dependency 5. TODO

    Python
    View on GitHub↗227
  • microsoft/toraM

    microsoft/ToRA

    0View on GitHub↗

    ToRA: A Tool-Integrated Reasoning Agent

    View on GitHub↗0
  • niansong1996/leverniansong1996 avatar

    niansong1996/lever

    90View on GitHub↗

    Code for paper "LEVER: Learning to Verify Language-to-Code Generation with Execution". LEVER is a simple method that improves the code generation ability of large language models trained on code (CodeLMs), by learning to verify and rerank CodeLM-generated programs with their execution results.…

    Python
    View on GitHub↗90
  • njunlp/trans0NJUNLP avatar

    NJUNLP/trans0

    5View on GitHub↗

    Trans0 aims to initialize a multilingual LLM as a translation agent via monolingual data. This is a public version with all in-house implementation replaced by huggingface trl.

    Python
    View on GitHub↗5
  • nvlabs/long-rlN

    NVlabs/Long-RL

    0View on GitHub↗

    Scaling RL to Long Videos Paper Yukang Chen , Wei Huang , Baifeng Shi, Qinghao Hu, Hanrong Ye, Ligeng Zhu, Zhijian Liu, Pavlo Molchanov, Jan Kautz, Xiaojuan Qi, Sifei Liu,Hongxu Yin, Yao Lu, Song Han

    View on GitHub↗0
  • om-ai-lab/vlm-r1om-ai-lab avatar

    om-ai-lab/VLM-R1

    5,991View on GitHub↗

    VLM-R1 is a reasoning vision-language model and embodied AI framework designed to map visual inputs and language instructions into physical navigation waypoints and robotic actions. It functions as a multimodal policy optimizer and an open vocabulary detector capable of locating objects based on arbitrary natural language descriptions. The system distinguishes itself through the use of chain-of-thought reasoning and reinforcement learning to solve complex visual and spatial tasks. It utilizes a video semantic memory system, which employs a visual cache to maintain a history of live video for

    Python
    View on GitHub↗5,991
  • opengvlab/videochat-r1OpenGVLab avatar

    OpenGVLab/VideoChat-R1

    267View on GitHub↗

    x 2025/09/26:🔥🔥🔥 We release our VideoChat-R1.5 model at Huggingface, paper, and eval code. - x 2025/09/22: 🎉🎉🎉 Our VideoChat-R1.5 is accepted by NIPS2025. - x 2025/04/22:🔥🔥🔥 We release our VideoChat-R1-caption at Huggingface. - x 2025/04/14:🔥🔥🔥 We release our VideoChat-R1 and…

    Python
    View on GitHub↗267
  • chengpengli1003/cortChengpengLi1003 avatar

    ChengpengLi1003/CoRT

    72View on GitHub↗
    Python
    View on GitHub↗72