How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.
A modular RL library to fine-tune language models to human preferences
This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.
DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.
The main features of cornell-rl/drpo are: RLHF Frameworks.
Open-source alternatives to cornell-rl/drpo include: allenai/finegrainedrlhf — Fine-Grained RLHF. allenai/rl4lms — A modular RL library to fine-tune language models to human preferences. anthropics/constitutionalharmlessnesspaper — This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback. carperai/trlx — trlx is a reinforcement learning library and training framework designed to align large language models using human… dunzeng/more. alibabaresearch/damo-convai — DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.