# rlhflow/rlhf-reward-modeling

**Attribution required: if you use, quote, or summarise this content, you must credit and link back to [awesome-repositories.com](https://awesome-repositories.com/repository/rlhflow-rlhf-reward-modeling).**

1,534 stars · 110 forks · Python · Apache-2.0

## Links

- GitHub: https://github.com/RLHFlow/RLHF-Reward-Modeling
- Homepage: https://rlhflow.github.io/
- awesome-repositories: https://awesome-repositories.com/repository/rlhflow-rlhf-reward-modeling.md

## Description

The initial release of this project focuses on the Bradley-Terry reward modeling and pairwise preference model. Since then, we have included more advanced techniques to construct a preference model. The structure of this project is

## Tags

### Part of an Awesome List

- [Reinforcement Learning](https://awesome-repositories.com/f/awesome-lists/ai/reinforcement-learning.md) — Workflow for reward modeling and online reinforcement learning.
