awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetServeur MCPÀ proposNotre méthodologiePresse
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
R

RUCAIBox/RLMEC

0
View on GitHub↗
0 stars·0 forks·5 vues

RLMEC

This repo provides the source code & data of our paper: Improving Large Language Models via Fine-grained Reinforcement Learning with Minimum Editing Constraint (arXiv 2024)

Features

  • Feedback Alignment - Improves models via fine-grained reinforcement learning with editing constraints.
  • Reinforcement Learning - Improves models via fine-grained reinforcement learning with editing constraints.
  • RLHF Frameworks - Framework for fine-grained token-level reinforcement learning.

Historique des stars

Graphique de l'historique des stars pour rucaibox/rlmecGraphique de l'historique des stars pour rucaibox/rlmec

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Start searching with AI

Questions fréquentes

Que fait rucaibox/rlmec ?

This repo provides the source code & data of our paper: Improving Large Language Models via Fine-grained Reinforcement Learning with Minimum Editing Constraint (arXiv 2024)

Quelles sont les fonctionnalités principales de rucaibox/rlmec ?

Les fonctionnalités principales de rucaibox/rlmec sont : Feedback Alignment, Reinforcement Learning, RLHF Frameworks.

Quelles sont les alternatives open-source à rucaibox/rlmec ?

Les alternatives open-source à rucaibox/rlmec incluent : facebookresearch/motif — This repository contains PyTorch code for Motif, training AI agents on NetHack with reward functions derived from an… openai/following-instructions-human-feedback — [Paper link][LINKTOPAPER]. alibabaresearch/damo-convai — DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI. anthropics/constitutionalharmlessnesspaper — This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback. ganjinzero/rrhf — Arxiv. openrlhf/openrlhf — OpenRLHF is a training framework and alignment library designed for reinforcement learning from human feedback across…

Alternatives open source à RLMEC

Projets open source similaires, classés selon le nombre de fonctionnalités partagées avec RLMEC.
  • anthropics/constitutionalharmlessnesspaperAvatar de anthropics

    anthropics/ConstitutionalHarmlessnessPaper

    263Voir sur GitHub↗

    This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.

    Voir sur GitHub↗263
  • facebookresearch/motifAvatar de facebookresearch

    facebookresearch/motif

    136Voir sur GitHub↗

    This repository contains PyTorch code for Motif, training AI agents on NetHack with reward functions derived from an LLM's preferences.

    Python
    Voir sur GitHub↗136
  • alibabaresearch/damo-convaiAvatar de AlibabaResearch

    AlibabaResearch/DAMO-ConvAI

    1,561Voir sur GitHub↗

    DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.

    Pythonconversational-aideep-learningdialog
    Voir sur GitHub↗1,561
  • ganjinzero/rrhfAvatar de GanjinZero

    GanjinZero/RRHF

    806Voir sur GitHub↗

    Arxiv

    Python
    Voir sur GitHub↗806
Voir les 30 alternatives à RLMEC→