awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoServidor MCPAcerca deCómo clasificamosPrensa
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
anthropics avatar

anthropics/ConstitutionalHarmlessnessPaperArchived

0
View on GitHub↗
263 estrellas·33 forks·5 vistas

ConstitutionalHarmlessnessPaper

This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.

Features

  • Reinforcement Learning - Aligning models with harmlessness via AI feedback.
  • RLHF Frameworks - Framework for harmlessness via AI feedback.

Historial de estrellas

Gráfico del historial de estrellas de anthropics/constitutionalharmlessnesspaperGráfico del historial de estrellas de anthropics/constitutionalharmlessnesspaper

Búsqueda con IA

Explora más repositorios increíbles

Describe lo que necesitas en lenguaje sencillo: la IA clasifica miles de proyectos open-source curados por relevancia.

Start searching with AI

Preguntas frecuentes

¿Qué hace anthropics/constitutionalharmlessnesspaper?

This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.

¿Cuáles son las características principales de anthropics/constitutionalharmlessnesspaper?

Las características principales de anthropics/constitutionalharmlessnesspaper son: Reinforcement Learning, RLHF Frameworks.

¿Qué alternativas de código abierto existen para anthropics/constitutionalharmlessnesspaper?

Las alternativas de código abierto para anthropics/constitutionalharmlessnesspaper incluyen: openai/following-instructions-human-feedback — [Paper link][LINKTOPAPER]. rucaibox/rlmec — This repo provides the source code & data of our paper: Improving Large Language Models via Fine-grained Reinforcement… alibabaresearch/damo-convai — DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI. ganjinzero/rrhf — Arxiv. openrlhf/openrlhf — OpenRLHF is a training framework and alignment library designed for reinforcement learning from human feedback across… uclaml/spin — The official implementation of Self-Play Fine-Tuning (SPIN).

Alternativas open-source a ConstitutionalHarmlessnessPaper

Proyectos open-source similares, clasificados según cuántas características comparten con ConstitutionalHarmlessnessPaper.
  • ganjinzero/rrhfAvatar de GanjinZero

    GanjinZero/RRHF

    806Ver en GitHub↗

    Arxiv

    Python
    Ver en GitHub↗806
  • openai/following-instructions-human-feedbackAvatar de openai

    openai/following-instructions-human-feedback

    1,258Ver en GitHub↗

    Paper linkLINKTOPAPER

    Ver en GitHub↗1,258
  • alibabaresearch/damo-convaiAvatar de AlibabaResearch

    AlibabaResearch/DAMO-ConvAI

    1,561Ver en GitHub↗

    DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.

    Pythonconversational-aideep-learningdialog
    Ver en GitHub↗1,561
  • openrlhf/openrlhfAvatar de OpenRLHF

    OpenRLHF/OpenRLHF

    9,675Ver en GitHub↗

    OpenRLHF is a training framework and alignment library designed for reinforcement learning from human feedback across distributed GPU clusters. It provides tools for aligning large language models and multimodal vision-language models using algorithms such as PPO, GRPO, and DPO. The framework distinguishes itself through a distributed inference engine that overlaps sample rollout with training to increase throughput. It supports scaling to models exceeding 70 billion parameters via parameter sharding and handles long-context sequences through ring-attention sequence parallelism. The project

    Pythonlarge-language-modelsopenai-o1proximal-policy-optimization
    Ver en GitHub↗9,675
Ver las 30 alternativas a ConstitutionalHarmlessnessPaper→