This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.
Paper linkLINKTOPAPER
DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.
Die Hauptfunktionen von alibabaresearch/damo-convai sind: Reinforcement Learning, RLHF Frameworks, Text-to-SQL Models, Text to SQL Datasets, Tool and API Benchmarks.
Open-Source-Alternativen zu alibabaresearch/damo-convai sind unter anderem: ganjinzero/rrhf — Arxiv. openrlhf/openrlhf — OpenRLHF is a training framework and alignment library designed for reinforcement learning from human feedback across… anthropics/constitutionalharmlessnesspaper — This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback. chia-hsuan-lee/kaggledbqa. openai/following-instructions-human-feedback — [Paper link][LINKTOPAPER]. rucaibox/rlmec — This repo provides the source code & data of our paper: Improving Large Language Models via Fine-grained Reinforcement…