awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoServidor MCPAcerca deCómo clasificamosPrensa
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
sail-sg avatar

sail-sg/understand-r1-zero

0
View on GitHub↗
1,214 estrellas·56 forks·Python·mit·4 vistasarxiv.org/pdf/2503.20783↗

Understand R1 Zero

Features

  • Critic-Free Algorithms - Critical analysis and implementation of reasoning-focused training.
  • Reasoning Models - Analysis of zero-shot reasoning models.
  • Reinforcement Learning Frameworks - Codebase for reproducing and understanding reasoning training.

Historial de estrellas

Gráfico del historial de estrellas de sail-sg/understand-r1-zeroGráfico del historial de estrellas de sail-sg/understand-r1-zero

Búsqueda con IA

Explora más repositorios increíbles

Describe lo que necesitas en lenguaje sencillo: la IA clasifica miles de proyectos open-source curados por relevancia.

Start searching with AI

Alternativas open-source a Understand R1 Zero

Proyectos open-source similares, clasificados según cuántas características comparten con Understand R1 Zero.
  • modalminds/mm-eurekaAvatar de ModalMinds

    ModalMinds/MM-EUREKA

    771Ver en GitHub↗

    MM-EUREKA: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

    Python
    Ver en GitHub↗771
  • hiyouga/easyr1Avatar de hiyouga

    hiyouga/EasyR1

    5,034Ver en GitHub↗

    EasyR1 is a distributed model training system and reinforcement learning framework for large language and vision-language models. It functions as a multimodal trainer and an implementation of a Proximal Policy Optimization pipeline designed to refine the reasoning and perception capabilities of models that process both text and images. The system specializes in distributing reinforcement learning workloads across multiple compute nodes to manage high memory requirements. It optimizes hardware utilization through padding-free training and fine-tuning to fit large models onto available graphics

    Python
    Ver en GitHub↗5,034
  • agentica-project/rllmAvatar de agentica-project

    agentica-project/rllm

    400Ver en GitHub↗

    🚀 Reinforcement Learning for Language Agents🌟

    Jupyter Notebook
    Ver en GitHub↗400
  • deep-agent/r1-vD

    Deep-Agent/R1-V

    0Ver en GitHub↗
    Ver en GitHub↗0
Ver las 30 alternativas a Understand R1 Zero→

Preguntas frecuentes

¿Cuáles son las características principales de sail-sg/understand-r1-zero?

Las características principales de sail-sg/understand-r1-zero son: Critic-Free Algorithms, Reasoning Models, Reinforcement Learning Frameworks.

¿Qué alternativas de código abierto existen para sail-sg/understand-r1-zero?

Las alternativas de código abierto para sail-sg/understand-r1-zero incluyen: modalminds/mm-eureka — MM-EUREKA: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning. jiayi-pan/tinyzero — TinyZero is a reinforcement learning framework and implementation designed to train language models to develop… agentica-project/rllm — 🚀 Reinforcement Learning for Language Agents🌟. deep-agent/r1-v. hiyouga/easyr1 — EasyR1 is a distributed model training system and reinforcement learning framework for large language and… inclusionai/areal — AReaL is a system for agent orchestration, distributed model training, and parameter-efficient tuning. It provides a…