awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectServer MCPDespreCum realizăm clasamentulPresă
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to breakend/deepreinforcementlearningthatmatters

Open-source alternatives to DeepReinforcementLearningThatMatters

30 open-source projects similar to breakend/deepreinforcementlearningthatmatters, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best DeepReinforcementLearningThatMatters alternative.

  • ai4co/rl4coAvatar ai4co

    ai4co/rl4co

    806Vezi pe GitHub↗
    Pythonattentionattention-modelbenchmark
    Vezi pe GitHub↗806
  • ai4finance-foundation/finrlAvatar AI4Finance-Foundation

    AI4Finance-Foundation/FinRL

    13,964Vezi pe GitHub↗

    FinRL is a reinforcement learning framework designed for the development, training, and backtesting of automated trading strategies. It functions as a quantitative finance toolkit that integrates deep learning algorithms with financial market simulations to address complex portfolio management and asset allocation tasks. The platform provides an end-to-end pipeline for transforming raw market data into actionable trading models. The project distinguishes itself through a layered, modular architecture that separates data processing, environment simulation, and agent training. This design allow

    Jupyter Notebookalgorithmic-tradingdeep-reinforcement-learningdrl-algorithms
    Vezi pe GitHub↗13,964
  • aikorea/awesome-rlAvatar aikorea

    aikorea/awesome-rl

    9,812Vezi pe GitHub↗

    Reinforcement learning resources curated

    Vezi pe GitHub↗9,812
  • airlab-polimi/mushroomA

    AIRLab-POLIMI/mushroom

    0Vezi pe GitHub↗
    Vezi pe GitHub↗0
  • alessiodm/drl-zhAvatar alessiodm

    alessiodm/drl-zh

    2,291Vezi pe GitHub↗

    Welcome to drlzh.ai: a hands-on deep reinforcement learning course where you build the algorithms, not just read about them.

    Jupyter Notebook
    Vezi pe GitHub↗2,291

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Find more with AI search
  • alexis-jacq/pytorch-dppoA

    alexis-jacq/Pytorch-DPPO

    0Vezi pe GitHub↗
    Vezi pe GitHub↗0
  • alibaba/chatlearnAvatar alibaba

    alibaba/ChatLearn

    452Vezi pe GitHub↗

    A flexible and efficient training framework for large-scale alignment tasks

    Python
    Vezi pe GitHub↗452
  • alibaba/rollAvatar alibaba

    alibaba/ROLL

    2,844Vezi pe GitHub↗

    ROLL is a distributed reinforcement learning framework and model alignment toolkit designed for large language models. It serves as a scalable training pipeline and GPU cluster manager, providing the infrastructure to align model behavior using reinforcement learning algorithms and preference optimization techniques. The project distinguishes itself through an agentic rollout orchestrator that generates and collects multi-turn interaction trajectories between AI agents and simulated environments. It supports specialized alignment methods including Direct Preference Optimization, reinforcement

    Pythonagenticrlhfrlvr
    Vezi pe GitHub↗2,844
  • alibabaresearch/damo-convaiAvatar AlibabaResearch

    AlibabaResearch/DAMO-ConvAI

    1,561Vezi pe GitHub↗

    DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.

    Pythonconversational-aideep-learningdialog
    Vezi pe GitHub↗1,561
  • alihassanijr/discernAvatar alihassanijr

    alihassanijr/DISCERN

    0Vezi pe GitHub↗

    This repository contains the implementation of DISCERN in Python. You can download the manuscript from my website or arXiv.

    Python
    Vezi pe GitHub↗0
  • amenra/ranxAvatar AmenRa

    AmenRa/ranx

    681Vezi pe GitHub↗

    ⚡️A Blazing-Fast Python Library for Ranking Evaluation, Comparison, and Fusion 🐍

    Python
    Vezi pe GitHub↗681
  • anthropics/constitutionalharmlessnesspaperAvatar anthropics

    anthropics/ConstitutionalHarmlessnessPaper

    263Vezi pe GitHub↗

    This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.

    Vezi pe GitHub↗263
  • applieddatasciencepartners/deepreinforcementlearningAvatar AppliedDataSciencePartners

    AppliedDataSciencePartners/DeepReinforcementLearning

    2,032Vezi pe GitHub↗

    A replica of the AlphaZero methodology for deep reinforcement learning in Python

    Jupyter Notebook
    Vezi pe GitHub↗2,032
  • ars-ashuha/quantile-regression-dqn-pytorchAvatar ars-ashuha

    ars-ashuha/quantile-regression-dqn-pytorch

    97Vezi pe GitHub↗

    A short and easy implementation of Quantile Regression DQN | Distributional Reinforcement Learning

    Jupyter Notebook
    Vezi pe GitHub↗97
  • astooke/rlpytAvatar astooke

    astooke/rlpyt

    2,280Vezi pe GitHub↗

    Reinforcement Learning in PyTorch

    Python
    Vezi pe GitHub↗2,280
  • atgambardella/pytorch-esA

    atgambardella/pytorch-es

    0Vezi pe GitHub↗
    Vezi pe GitHub↗0
  • aunum/goldAvatar aunum

    aunum/gold

    351Vezi pe GitHub↗

    Reinforcement Learning in Go

    Go
    Vezi pe GitHub↗351
  • awjuliani/deeprl-agentsAvatar awjuliani

    awjuliani/DeepRL-Agents

    2,277Vezi pe GitHub↗

    A set of Deep Reinforcement Learning Agents implemented in Tensorflow.

    Jupyter Notebookreinforcement-learningtensorflow
    Vezi pe GitHub↗2,277
  • breakend/reproducibilityincontinuouspolicygradientmethodsB

    Breakend/ReproducibilityInContinuousPolicyGradientMethods

    0Vezi pe GitHub↗
    Vezi pe GitHub↗0
  • carpedm20/deep-rl-tensorflowAvatar carpedm20

    carpedm20/deep-rl-tensorflow

    1,581Vezi pe GitHub↗

    TensorFlow implementation of Deep Reinforcement Learning papers

    Pythondeep-reinforcement-learningdqntensorflow
    Vezi pe GitHub↗1,581
  • catalyst-team/catalyst-rlAvatar catalyst-team

    catalyst-team/catalyst-rl

    48Vezi pe GitHub↗

    Accelerated RL

    Python
    Vezi pe GitHub↗48
  • chenmientan/rl2Avatar ChenmienTan

    ChenmienTan/RL2

    1,293Vezi pe GitHub↗
    Python
    Vezi pe GitHub↗1,293
  • coax-dev/coaxAvatar coax-dev

    coax-dev/coax

    185Vezi pe GitHub↗

    |tests| |pypi| |docs| |License|

    Python
    Vezi pe GitHub↗185
  • contextualai/halosAvatar ContextualAI

    ContextualAI/HALOs

    906Vezi pe GitHub↗

    A library with extensible implementations of DPO, KTO, PPO, ORPO, and other human-aware loss functions (HALOs).

    Pythonalignmentdpohalos
    Vezi pe GitHub↗906
  • coreylynch/async-rlAvatar coreylynch

    coreylynch/async-rl

    1,006Vezi pe GitHub↗

    Tensorflow Keras OpenAI Gym implementation of 1-step Q Learning from "Asynchronous Methods for Deep Reinforcement Learning"

    Python
    Vezi pe GitHub↗1,006
  • cpnota/autonomous-learning-libraryC

    cpnota/autonomous-learning-library

    0Vezi pe GitHub↗
    Vezi pe GitHub↗0
  • daveonwave/gym4realAvatar Daveonwave

    Daveonwave/gym4ReaL

    48Vezi pe GitHub↗

    Gymnasium-based benchmarking suite for testing RL algorithms on real-world scenarios

    Python
    Vezi pe GitHub↗48
  • deepmind/labAvatar deepmind

    deepmind/lab

    7,365Vezi pe GitHub↗

    Lab is a customizable 3D platform and research testbed designed for training and testing autonomous agents using reinforcement learning. It serves as a spatial AI training simulator where agents can be evaluated through navigation and puzzle-solving tasks. The environment allows for the definition of complex layouts and task behaviors through external scripting, enabling the generation of specific challenges for AI research. It supports both automated training via standard API bindings and manual agent control to validate simulation dynamics. The system utilizes a grid-based spatial represen

    C
    Vezi pe GitHub↗7,365
  • deepmind/mctxAvatar deepmind

    deepmind/mctx

    2,638Vezi pe GitHub↗

    Mctx is a library with a JAX-native implementation of Monte Carlo tree search (MCTS) algorithms such as AlphaZero, MuZero, and Gumbel MuZero. For computation speed up, the implementation fully supports JIT-compilation. Search algorithms in Mctx are defined for and operate on batches of inputs,…

    Python
    Vezi pe GitHub↗2,638
  • 2toinf/uniactAvatar 2toinf

    2toinf/UniAct

    241Vezi pe GitHub↗

    Project Page Paper

    Python
    Vezi pe GitHub↗241