awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
yihedeng9 avatar

yihedeng9/OpenVLThinker

0
View on GitHub↗
152 stars·7 forks·Python·Apache-2.0·8 views

OpenVLThinker

Yihe Deng , Nanyun Peng , Kai-Wei Chang

Features

  • Critic-Free Algorithms - Iterative SFT-RL cycles for complex vision-language reasoning.
  • Multimodal Understanding - Complex vision-language reasoning via iterative SFT-RL cycles.
  • Reasoning Datasets - Vision-language reasoning dataset for iterative training.

Star history

Star history chart for yihedeng9/openvlthinkerStar history chart for yihedeng9/openvlthinker

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to OpenVLThinker

Similar open-source projects, ranked by how many features they share with OpenVLThinker.
  • deepseek-ai/janusdeepseek-ai avatar

    deepseek-ai/Janus

    17,746View on GitHub↗

    Janus is a multimodal large language model and unified framework that integrates visual understanding and image generation within a single neural network. It functions as both a visual understanding model for analyzing images and a text-to-image generator. The system uses a unified transformer backbone and a multimodal latent space to bridge the gap between text and visual data. This architecture employs decoupled visual encoding and cross-modal tokenization to separate the paths for discriminative understanding and generative tasks, representing images as grids of discrete codes. The projec

    Pythonany-to-anyfoundation-modelsllm
    View on GitHub↗17,746
  • alibaba-nlp/zerosearchAlibaba-NLP avatar

    Alibaba-NLP/ZeroSearch

    1,296View on GitHub↗

    ZeroSearch: Incentivize the Search Capability of LLMs without Searching

    Python
    View on GitHub↗1,296
  • bytedtsinghua-sia/dapoBytedTsinghua-SIA avatar

    BytedTsinghua-SIA/DAPO

    1,831View on GitHub↗

    DAPO: an Open-source RL System from ByteDance Seed and Tsinghua AIR

    Python
    View on GitHub↗1,831
  • agentica-project/rllmagentica-project avatar

    agentica-project/rllm

    400View on GitHub↗

    🚀 Reinforcement Learning for Language Agents🌟

    Jupyter Notebook
    View on GitHub↗400
See all 30 alternatives to OpenVLThinker→

Frequently asked questions

What does yihedeng9/openvlthinker do?

Yihe Deng , Nanyun Peng , Kai-Wei Chang

What are the main features of yihedeng9/openvlthinker?

The main features of yihedeng9/openvlthinker are: Critic-Free Algorithms, Multimodal Understanding, Reasoning Datasets.

What are some open-source alternatives to yihedeng9/openvlthinker?

Open-source alternatives to yihedeng9/openvlthinker include: deepseek-ai/janus — Janus is a multimodal large language model and unified framework that integrates visual understanding and image… alibaba-nlp/zerosearch — ZeroSearch: Incentivize the Search Capability of LLMs without Searching. bytedtsinghua-sia/dapo — DAPO: an Open-source RL System from ByteDance Seed and Tsinghua AIR. camel-ai/loong — Community | Paper | Cookbook | Datasets | Loong Blog | Contributing | CAMEL-AI. deepseek-ai/deepseek-math — DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models. agentica-project/rllm — 🚀 Reinforcement Learning for Language Agents🌟.