# cjreinforce/pure

**Attribution required: if you use, quote, or summarise this content, you must credit and link back to [awesome-repositories.com](https://awesome-repositories.com/repository/cjreinforce-pure).**

_How this analysis was created: the description and tags below were written by an AI model that read this project's README and public documentation pages; stars, license and language come straight from the GitHub API. The model does not read the source code._

169 stars · 7 forks · Python

## Links

- GitHub: https://github.com/CJReinforce/PURE
- Homepage: https://arxiv.org/abs/2504.15275
- awesome-repositories: https://awesome-repositories.com/repository/cjreinforce-pure.md

## Description

[2025/10/23] 🔥🔥Our paper is accepted by NeurIPS 2025.🔥🔥 - [2025/04/22] Released our Paper on arXiv. See here - [2025/03/24] We re-implement our algorithm based on verl. ✨✨ Key features: (1) add ~50 additional metrics to comprehensively monitor the training process and stability, (2) add a…

## Tags

### Part of an Awesome List

- [Dense Reward Optimization](https://awesome-repositories.com/f/awesome-lists/ai/dense-reward-optimization.md) — Min-form credit assignment for process reward models.
