awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Chenluye99 avatar

Chenluye99/PROF

0
View on GitHub↗
11 stars·0 forks·7 views

PROF

Introduction

Features

  • Dense Reward Optimization - Harmonizing process and outcome rewards in training.

Star history

Star history chart for chenluye99/profStar history chart for chenluye99/prof

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to PROF

Similar open-source projects, ranked by how many features they share with PROF.
  • amap-ml/tree-grpoAMAP-ML avatar

    AMAP-ML/Tree-GRPO

    378View on GitHub↗

    Tree Search for LLM Agent Reinforcement Learning

    Python
    View on GitHub↗378
  • cjreinforce/pureCJReinforce avatar

    CJReinforce/PURE

    169View on GitHub↗

    2025/10/23 🔥🔥Our paper is accepted by NeurIPS 2025.🔥🔥 - 2025/04/22 Released our Paper on arXiv. See here - 2025/03/24 We re-implement our algorithm based on verl. ✨✨ Key features: (1) add ~50 additional metrics to comprehensively monitor the training process and stability, (2) add a…

    Python
    View on GitHub↗169
  • cmu-aire/mrtCMU-AIRe avatar

    CMU-AIRe/MRT

    119View on GitHub↗

    This repository contains the code for our paper titled "Optimizing Test-Time Compute via Meta Reinforcement Finetuning." In this work, we introduce a novel approach to optimizing test-time compute through meta reinforcement learning, aiming to balance the efficiency and discovery capabilities of…

    Python
    View on GitHub↗119
  • aiframeresearch/spoAIFrameResearch avatar

    AIFrameResearch/SPO

    53View on GitHub↗

    🚀 Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models 🌟

    Python
    View on GitHub↗53
See all 19 alternatives to PROF→

Frequently asked questions

What does chenluye99/prof do?

Introduction

What are the main features of chenluye99/prof?

The main features of chenluye99/prof are: Dense Reward Optimization.

What are some open-source alternatives to chenluye99/prof?

Open-source alternatives to chenluye99/prof include: amap-ml/tree-grpo — Tree Search for LLM Agent Reinforcement Learning. cjreinforce/pure — [2025/10/23] 🔥🔥Our paper is accepted by NeurIPS 2025.🔥🔥 - [2025/04/22] Released our Paper on arXiv. See here -… cmu-aire/mrt — This repository contains the code for our paper titled "Optimizing Test-Time Compute via Meta Reinforcement… dongguanting/tool-star — 🔧✨Tool-Star: Empowering Multi-Tool Collaborative Web Agent via Reinforcement Learning. gen-verse/reasonflux — Princeton University \& PKU \& UIUC \& University of Chicago \& ByteDance Seed. aiframeresearch/spo — 🚀 Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models 🌟.