How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
A modular RL library to fine-tune language models to human preferences
This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.
DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.
The main features of cornell-rl/drpo are: RLHF Frameworks.
Projects with overlapping indexed features include: allenai/finegrainedrlhf — Fine-Grained RLHF. allenai/rl4lms — A modular RL library to fine-tune language models to human preferences. anthropics/constitutionalharmlessnesspaper — This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback. carperai/trlx — trlx is a reinforcement learning library and training framework designed to align large language models using human… dunzeng/more. alibabaresearch/damo-convai — DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.