awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetÀ proposNotre méthodologiePresseServeur MCP
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to deepmind/mctx

Open-source alternatives to Mctx

30 open-source projects similar to deepmind/mctx, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best Mctx alternative.

  • ai4co/rl4coAvatar de ai4co

    ai4co/rl4co

    806Voir sur GitHub↗
    Pythonattentionattention-modelbenchmark
    Voir sur GitHub↗806
  • ai4finance-foundation/finrlAvatar de AI4Finance-Foundation

    AI4Finance-Foundation/FinRL

    13,964Voir sur GitHub↗

    FinRL is a reinforcement learning framework designed for the development, training, and backtesting of automated trading strategies. It functions as a quantitative finance toolkit that integrates deep learning algorithms with financial market simulations to address complex portfolio management and asset allocation tasks. The platform provides an end-to-end pipeline for transforming raw market data into actionable trading models. The project distinguishes itself through a layered, modular architecture that separates data processing, environment simulation, and agent training. This design allow

    Jupyter Notebookalgorithmic-tradingdeep-reinforcement-learningdrl-algorithms
    Voir sur GitHub↗13,964
  • aikorea/awesome-rlAvatar de aikorea

    aikorea/awesome-rl

    9,812Voir sur GitHub↗

    Reinforcement learning resources curated

    Voir sur GitHub↗9,812
  • airlab-polimi/mushroomA

    AIRLab-POLIMI/mushroom

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • alessiodm/drl-zhAvatar de alessiodm

    alessiodm/drl-zh

    2,291Voir sur GitHub↗

    Welcome to drlzh.ai: a hands-on deep reinforcement learning course where you build the algorithms, not just read about them.

    Jupyter Notebook
    Voir sur GitHub↗2,291

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Find more with AI search
  • alexis-jacq/pytorch-dppoA

    alexis-jacq/Pytorch-DPPO

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • alibaba/chatlearnAvatar de alibaba

    alibaba/ChatLearn

    452Voir sur GitHub↗

    A flexible and efficient training framework for large-scale alignment tasks

    Python
    Voir sur GitHub↗452
  • alibaba/rollAvatar de alibaba

    alibaba/ROLL

    2,844Voir sur GitHub↗

    ROLL is a distributed reinforcement learning framework and model alignment toolkit designed for large language models. It serves as a scalable training pipeline and GPU cluster manager, providing the infrastructure to align model behavior using reinforcement learning algorithms and preference optimization techniques. The project distinguishes itself through an agentic rollout orchestrator that generates and collects multi-turn interaction trajectories between AI agents and simulated environments. It supports specialized alignment methods including Direct Preference Optimization, reinforcement

    Pythonagenticrlhfrlvr
    Voir sur GitHub↗2,844
  • alibabaresearch/damo-convaiAvatar de AlibabaResearch

    AlibabaResearch/DAMO-ConvAI

    1,561Voir sur GitHub↗

    DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.

    Pythonconversational-aideep-learningdialog
    Voir sur GitHub↗1,561
  • alihassanijr/discernAvatar de alihassanijr

    alihassanijr/DISCERN

    0Voir sur GitHub↗

    This repository contains the implementation of DISCERN in Python. You can download the manuscript from my website or arXiv.

    Python
    Voir sur GitHub↗0
  • amenra/ranxAvatar de AmenRa

    AmenRa/ranx

    681Voir sur GitHub↗

    ⚡️A Blazing-Fast Python Library for Ranking Evaluation, Comparison, and Fusion 🐍

    Python
    Voir sur GitHub↗681
  • anthropics/constitutionalharmlessnesspaperAvatar de anthropics

    anthropics/ConstitutionalHarmlessnessPaper

    263Voir sur GitHub↗

    This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.

    Voir sur GitHub↗263
  • applieddatasciencepartners/deepreinforcementlearningAvatar de AppliedDataSciencePartners

    AppliedDataSciencePartners/DeepReinforcementLearning

    2,032Voir sur GitHub↗

    A replica of the AlphaZero methodology for deep reinforcement learning in Python

    Jupyter Notebook
    Voir sur GitHub↗2,032
  • ars-ashuha/quantile-regression-dqn-pytorchAvatar de ars-ashuha

    ars-ashuha/quantile-regression-dqn-pytorch

    97Voir sur GitHub↗

    A short and easy implementation of Quantile Regression DQN | Distributional Reinforcement Learning

    Jupyter Notebook
    Voir sur GitHub↗97
  • astooke/rlpytAvatar de astooke

    astooke/rlpyt

    2,280Voir sur GitHub↗

    Reinforcement Learning in PyTorch

    Python
    Voir sur GitHub↗2,280
  • atgambardella/pytorch-esA

    atgambardella/pytorch-es

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • aunum/goldAvatar de aunum

    aunum/gold

    351Voir sur GitHub↗

    Reinforcement Learning in Go

    Go
    Voir sur GitHub↗351
  • awjuliani/deeprl-agentsAvatar de awjuliani

    awjuliani/DeepRL-Agents

    2,277Voir sur GitHub↗

    A set of Deep Reinforcement Learning Agents implemented in Tensorflow.

    Jupyter Notebookreinforcement-learningtensorflow
    Voir sur GitHub↗2,277
  • breakend/deepreinforcementlearningthatmattersB

    Breakend/DeepReinforcementLearningThatMatters

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • breakend/reproducibilityincontinuouspolicygradientmethodsB

    Breakend/ReproducibilityInContinuousPolicyGradientMethods

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • carpedm20/deep-rl-tensorflowAvatar de carpedm20

    carpedm20/deep-rl-tensorflow

    1,581Voir sur GitHub↗

    TensorFlow implementation of Deep Reinforcement Learning papers

    Pythondeep-reinforcement-learningdqntensorflow
    Voir sur GitHub↗1,581
  • catalyst-team/catalyst-rlAvatar de catalyst-team

    catalyst-team/catalyst-rl

    48Voir sur GitHub↗

    Accelerated RL

    Python
    Voir sur GitHub↗48
  • chenmientan/rl2Avatar de ChenmienTan

    ChenmienTan/RL2

    1,293Voir sur GitHub↗
    Python
    Voir sur GitHub↗1,293
  • coax-dev/coaxAvatar de coax-dev

    coax-dev/coax

    185Voir sur GitHub↗

    |tests| |pypi| |docs| |License|

    Python
    Voir sur GitHub↗185
  • contextualai/halosAvatar de ContextualAI

    ContextualAI/HALOs

    906Voir sur GitHub↗

    A library with extensible implementations of DPO, KTO, PPO, ORPO, and other human-aware loss functions (HALOs).

    Pythonalignmentdpohalos
    Voir sur GitHub↗906
  • coreylynch/async-rlAvatar de coreylynch

    coreylynch/async-rl

    1,006Voir sur GitHub↗

    Tensorflow Keras OpenAI Gym implementation of 1-step Q Learning from "Asynchronous Methods for Deep Reinforcement Learning"

    Python
    Voir sur GitHub↗1,006
  • cpnota/autonomous-learning-libraryC

    cpnota/autonomous-learning-library

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • daveonwave/gym4realAvatar de Daveonwave

    Daveonwave/gym4ReaL

    48Voir sur GitHub↗

    Gymnasium-based benchmarking suite for testing RL algorithms on real-world scenarios

    Python
    Voir sur GitHub↗48
  • deepmind/labAvatar de deepmind

    deepmind/lab

    7,365Voir sur GitHub↗

    Lab is a customizable 3D platform and research testbed designed for training and testing autonomous agents using reinforcement learning. It serves as a spatial AI training simulator where agents can be evaluated through navigation and puzzle-solving tasks. The environment allows for the definition of complex layouts and task behaviors through external scripting, enabling the generation of specific challenges for AI research. It supports both automated training via standard API bindings and manual agent control to validate simulation dynamics. The system utilizes a grid-based spatial represen

    C
    Voir sur GitHub↗7,365
  • 2toinf/uniactAvatar de 2toinf

    2toinf/UniAct

    241Voir sur GitHub↗

    Project Page Paper

    Python
    Voir sur GitHub↗241