awesome-repositories.comCatégoriesBlog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetServeur MCPÀ proposNotre méthodologiePresse
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

6 dépôts

Awesome GitHub RepositoriesCritic-Based Algorithms

Reinforcement learning approaches that utilize value functions or process verifiers.

Explore 6 awesome GitHub repositories matching part of an awesome list · Critic-Based Algorithms. Refine with filters or upvote what's useful.

Awesome Critic-Based Algorithms GitHub Repositories

Trouvez les meilleurs dépôts grâce à l'IA.Nous recherchons les dépôts les plus pertinents grâce à l'IA.
  • open-reasoner-zero/open-reasoner-zeroAvatar de Open-Reasoner-Zero

    Open-Reasoner-Zero/Open-Reasoner-Zero

    2,095Voir sur GitHub↗

    An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

    Scaling reinforcement learning on base models using open approaches.

    Python
    Voir sur GitHub↗2,095
  • prime-rl/primeAvatar de PRIME-RL

    PRIME-RL/PRIME

    1,863Voir sur GitHub↗

    ✨ Getting Started • 📖 Introduction 🔧 Usage • 📃 Evaluation • 🎈 Citation • 🌻 Acknowledgement • 📈 Star History

    Process reinforcement learning using implicit reward signals.

    Python
    Voir sur GitHub↗1,863
  • lifan-yuan/implicitprmAvatar de lifan-yuan

    lifan-yuan/ImplicitPRM

    171Voir sur GitHub↗

    Free Process Rewards without Process Labels

    Generating process rewards without requiring explicit process labels.

    Python
    Voir sur GitHub↗171
  • rookie-joe/autopsvAvatar de rookie-joe

    rookie-joe/AutoPSV

    50Voir sur GitHub↗

    This repository contains the official implementation of AutoPSV: Automated Process-Supervised Verifier, accepted at NeurIPS 2024 (poster).

    Automated verification for process-supervised reinforcement learning.

    Python
    Voir sur GitHub↗50
  • hitsz-tmg/veripoAvatar de HITsz-TMG

    HITsz-TMG/VerIPO

    10Voir sur GitHub↗

    📄 Paper Link 🤗 VerIPO-7B-v1.0

    Iterative policy optimization for long-reasoning video models.

    Python
    Voir sur GitHub↗10
  • corl-team/vl-dacAvatar de corl-team

    corl-team/VL-DAC

    12Voir sur GitHub↗

    Enhancing Vision-Language Model Training with Reinforcement Learning in Synthetic Worlds for Real-World Success

    Reinforcement learning for vision-language models in synthetic environments.

    Python
    Voir sur GitHub↗12
  1. Home
  2. Part of an Awesome List
  3. AI & Machine Learning
  4. Critic-Based Algorithms