awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
anthropics avatar

anthropics/ConstitutionalHarmlessnessPaperArchived

0
View on GitHub↗
263 Stars·33 Forks·5 Aufrufe

ConstitutionalHarmlessnessPaper

This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.

Features

  • Reinforcement Learning - Aligning models with harmlessness via AI feedback.
  • RLHF Frameworks - Framework for harmlessness via AI feedback.

Star-Verlauf

Star-Verlauf für anthropics/constitutionalharmlessnesspaperStar-Verlauf für anthropics/constitutionalharmlessnesspaper

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Häufig gestellte Fragen

Was macht anthropics/constitutionalharmlessnesspaper?

This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.

Was sind die Hauptfunktionen von anthropics/constitutionalharmlessnesspaper?

Die Hauptfunktionen von anthropics/constitutionalharmlessnesspaper sind: Reinforcement Learning, RLHF Frameworks.

Welche Open-Source-Alternativen gibt es zu anthropics/constitutionalharmlessnesspaper?

Open-Source-Alternativen zu anthropics/constitutionalharmlessnesspaper sind unter anderem: openai/following-instructions-human-feedback — [Paper link][LINKTOPAPER]. rucaibox/rlmec — This repo provides the source code & data of our paper: Improving Large Language Models via Fine-grained Reinforcement… alibabaresearch/damo-convai — DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI. ganjinzero/rrhf — Arxiv. openrlhf/openrlhf — OpenRLHF is a training framework and alignment library designed for reinforcement learning from human feedback across… uclaml/spin — The official implementation of Self-Play Fine-Tuning (SPIN).

Open-Source-Alternativen zu ConstitutionalHarmlessnessPaper

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit ConstitutionalHarmlessnessPaper.
  • ganjinzero/rrhfAvatar von GanjinZero

    GanjinZero/RRHF

    806Auf GitHub ansehen↗

    Arxiv

    Python
    Auf GitHub ansehen↗806
  • openai/following-instructions-human-feedbackAvatar von openai

    openai/following-instructions-human-feedback

    1,258Auf GitHub ansehen↗

    Paper linkLINKTOPAPER

    Auf GitHub ansehen↗1,258
  • alibabaresearch/damo-convaiAvatar von AlibabaResearch

    AlibabaResearch/DAMO-ConvAI

    1,561Auf GitHub ansehen↗

    DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.

    Pythonconversational-aideep-learningdialog
    Auf GitHub ansehen↗1,561
  • openrlhf/openrlhfAvatar von OpenRLHF

    OpenRLHF/OpenRLHF

    9,675Auf GitHub ansehen↗

    OpenRLHF is a training framework and alignment library designed for reinforcement learning from human feedback across distributed GPU clusters. It provides tools for aligning large language models and multimodal vision-language models using algorithms such as PPO, GRPO, and DPO. The framework distinguishes itself through a distributed inference engine that overlaps sample rollout with training to increase throughput. It supports scaling to models exceeding 70 billion parameters via parameter sharding and handles long-context sequences through ring-attention sequence parallelism. The project

    Pythonlarge-language-modelsopenai-o1proximal-policy-optimization
    Auf GitHub ansehen↗9,675
Alle 30 Alternativen zu ConstitutionalHarmlessnessPaper anzeigen→