awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to instadeepai/jumanji

Open-source alternatives to Jumanji

30 open-source projects similar to instadeepai/jumanji, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best Jumanji alternative.

  • shangtongzhang/deeprlالصورة الرمزية لـ ShangtongZhang

    ShangtongZhang/DeepRL

    3,428عرض على GitHub↗

    Modularized Implementation of Deep RL Algorithms in PyTorch

    Python
    عرض على GitHub↗3,428
  • openai/baselinesالصورة الرمزية لـ openai

    openai/baselines

    16,733عرض على GitHub↗

    Baselines is a comprehensive suite of frameworks for reinforcement learning algorithm implementation, imitation learning, and training orchestration. It provides a library of standardized learning algorithms used to benchmark and replicate research results, alongside a deep learning policy framework for constructing neural network architectures such as multi-layer perceptrons, convolutional networks, and long short-term memory networks. The project includes a specialized imitation learning toolkit that enables agents to mimic expert behavior through behavior cloning and generative adversarial

    Python
    عرض على GitHub↗16,733
  • google/dopamineالصورة الرمزية لـ google

    google/dopamine

    10,879عرض على GitHub↗

    Dopamine is a reinforcement learning research framework designed for prototyping and testing algorithms across diverse simulated environments. It provides an agent development toolkit that utilizes a flat class hierarchy to facilitate the creation and extension of learning agents. The framework includes a standardization layer via environment wrappers that connect agents to various physics simulations and gaming environments. It also features a high-performance experience replay buffer for storing and sampling transition data to improve training stability, alongside a dedicated hyperparameter

    Jupyter Notebook
    عرض على GitHub↗10,879
  • novasky-ai/skyrlالصورة الرمزية لـ NovaSky-AI

    NovaSky-AI/SkyRL

    1,611عرض على GitHub↗
    Python
    عرض على GitHub↗1,611

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Find more with AI search
  • chainer/chainerrlالصورة الرمزية لـ chainer

    chainer/chainerrl

    1,200عرض على GitHub↗

    ChainerRL is a deep reinforcement learning library built on top of Chainer.

    Python
    عرض على GitHub↗1,200
  • openrlhf/openrlhfالصورة الرمزية لـ OpenRLHF

    OpenRLHF/OpenRLHF

    9,675عرض على GitHub↗

    OpenRLHF is a training framework and alignment library designed for reinforcement learning from human feedback across distributed GPU clusters. It provides tools for aligning large language models and multimodal vision-language models using algorithms such as PPO, GRPO, and DPO. The framework distinguishes itself through a distributed inference engine that overlaps sample rollout with training to increase throughput. It supports scaling to models exceeding 70 billion parameters via parameter sharding and handles long-context sequences through ring-attention sequence parallelism. The project

    Pythonlarge-language-modelsopenai-o1proximal-policy-optimization
    عرض على GitHub↗9,675
  • rlinf/rlinfالصورة الرمزية لـ RLinf

    RLinf/RLinf

    2,502عرض على GitHub↗

    RLinf is a distributed reinforcement learning orchestrator and embodied AI training framework. It provides the infrastructure to train vision-language-action models and robotic policies using a combination of reinforcement learning and supervised fine-tuning. The system is designed for scaling workloads across GPU clusters, managing the placement of actors, rollout workers, and environment components. It features a specialized robotics data collection pipeline for gathering teleoperated demonstrations and simulation trajectories into standardized replay buffers, alongside a hardware interface

    Pythonagentic-aiembodied-aireinforcement-learning
    عرض على GitHub↗2,502
  • kaixhin/atariالصورة الرمزية لـ Kaixhin

    Kaixhin/Atari

    264عرض على GitHub↗

    Persistent advantage learning dueling double DQN for the Arcade Learning Environment

    Lua
    عرض على GitHub↗264
  • alibaba/rollالصورة الرمزية لـ alibaba

    alibaba/ROLL

    2,844عرض على GitHub↗

    ROLL is a distributed reinforcement learning framework and model alignment toolkit designed for large language models. It serves as a scalable training pipeline and GPU cluster manager, providing the infrastructure to align model behavior using reinforcement learning algorithms and preference optimization techniques. The project distinguishes itself through an agentic rollout orchestrator that generates and collects multi-turn interaction trajectories between AI agents and simulated environments. It supports specialized alignment methods including Direct Preference Optimization, reinforcement

    Pythonagenticrlhfrlvr
    عرض على GitHub↗2,844
  • rlcode/reinforcement-learningالصورة الرمزية لـ rlcode

    rlcode/reinforcement-learning

    3,642عرض على GitHub↗

    Minimal and Clean Reinforcement Learning Examples

    Python
    عرض على GitHub↗3,642
  • volcengine/verlالصورة الرمزية لـ volcengine

    volcengine/verl

    22,015عرض على GitHub↗

    verl is a distributed training system designed for large language model alignment and reinforcement learning. It provides a framework for executing post-training pipelines, including supervised fine-tuning and reinforcement learning from human feedback, to refine model behavior and agentic capabilities. The system utilizes a hybrid training and inference engine that optimizes memory and communication when switching between model generation and gradient updates. It supports multi-modal reinforcement learning for models processing both image and text data, and implements algorithms such as PPO

    Python
    عرض على GitHub↗22,015
  • yandexdataschool/agentnetالصورة الرمزية لـ yandexdataschool

    yandexdataschool/AgentNet

    299عرض على GitHub↗

    Deep Reinforcement Learning library for humans

    Python
    عرض على GitHub↗299
  • thudm/slimeالصورة الرمزية لـ THUDM

    THUDM/slime

    4,259عرض على GitHub↗

    SLIME is a distributed reinforcement learning framework for large language model post-training that bridges Megatron training with SGLang inference servers. It orchestrates scalable RL loops across GPU clusters, decoupling training and inference into independent processes that communicate over HTTP and NCCL for independent scaling and fault tolerance. The system supports multi-agent reinforcement learning workflows with parallel agent instances, customizable rollout strategies, and personalized agent serving that improves models from prior conversations without disrupting API serving. The fra

    Python
    عرض على GitHub↗4,259
  • resibots/blackdropsالصورة الرمزية لـ resibots

    resibots/blackdrops

    66عرض على GitHub↗

    Code for the Black-DROPS algorithm: "Black-Box Data-efficient Policy Search for Robotics", IROS 2017/ICRA 2018

    C++
    عرض على GitHub↗66
  • huggingface/trlالصورة الرمزية لـ huggingface

    huggingface/trl

    18,653عرض على GitHub↗

    This library provides a comprehensive framework for fine-tuning, aligning, and distilling transformer-based language models. It serves as a toolkit for adapting models to specialized domains through supervised learning, while offering advanced methodologies to improve output quality and reasoning capabilities. The project distinguishes itself through specialized alignment and optimization techniques, including direct preference optimization and reinforcement learning, which allow models to be tuned against human preferences without complex reward modeling. It further supports training efficie

    Python
    عرض على GitHub↗18,653
  • nivwusquorum/tensorflow-deepqالصورة الرمزية لـ nivwusquorum

    nivwusquorum/tensorflow-deepq

    1,166عرض على GitHub↗

    A deep Q learning demonstration using Google Tensorflow

    Jupyter Notebook
    عرض على GitHub↗1,166
  • inclusionai/arealالصورة الرمزية لـ inclusionAI

    inclusionAI/AReaL

    3,559عرض على GitHub↗

    AReaL is a system for agent orchestration, distributed model training, and parameter-efficient tuning. It provides a framework for developing multi-turn reasoning agents and training large models using reinforcement learning from human feedback. The project implements a toolkit for improving the visual reasoning and geometry problem solving capabilities of vision-language models. It utilizes a memory-efficient tuning system to optimize mathematical and reasoning models across different inference backends. The infrastructure supports large-scale training through tensor, pipeline, and expert p

    Pythonagentllmllm-agent
    عرض على GitHub↗3,559
  • aunum/goldالصورة الرمزية لـ aunum

    aunum/gold

    351عرض على GitHub↗

    Reinforcement Learning in Go

    Go
    عرض على GitHub↗351
  • facebookresearch/habitat-labالصورة الرمزية لـ facebookresearch

    facebookresearch/habitat-lab

    2,848عرض على GitHub↗

    Habitat-Lab is an open-source platform for training and evaluating embodied AI agents in photorealistic 3D indoor environments. It functions as a high-performance 3D indoor environment simulator that supports physics-based interaction, enabling research into navigation and manipulation tasks. The platform provides a modular task-environment abstraction that separates task logic from environment simulation, using configuration-driven pipeline assembly to compose simulation and training pipelines. It includes a hierarchical sensor-actuator architecture for mixing and matching perception and act

    Pythonaicomputer-visiondeep-learning
    عرض على GitHub↗2,848
  • applieddatasciencepartners/deepreinforcementlearningالصورة الرمزية لـ AppliedDataSciencePartners

    AppliedDataSciencePartners/DeepReinforcementLearning

    2,032عرض على GitHub↗

    A replica of the AlphaZero methodology for deep reinforcement learning in Python

    Jupyter Notebook
    عرض على GitHub↗2,032
  • anthropics/constitutionalharmlessnesspaperالصورة الرمزية لـ anthropics

    anthropics/ConstitutionalHarmlessnessPaper

    263عرض على GitHub↗

    This repository provides supplementary material for our paper Constitutional AI: Harmlessness from AI Feedback.

    عرض على GitHub↗263
  • alessiodm/drl-zhالصورة الرمزية لـ alessiodm

    alessiodm/drl-zh

    2,291عرض على GitHub↗

    Welcome to drlzh.ai: a hands-on deep reinforcement learning course where you build the algorithms, not just read about them.

    Jupyter Notebook
    عرض على GitHub↗2,291
  • chenmientan/rl2الصورة الرمزية لـ ChenmienTan

    ChenmienTan/RL2

    1,293عرض على GitHub↗
    Python
    عرض على GitHub↗1,293
  • amenra/ranxالصورة الرمزية لـ AmenRa

    AmenRa/ranx

    681عرض على GitHub↗

    ⚡️A Blazing-Fast Python Library for Ranking Evaluation, Comparison, and Fusion 🐍

    Python
    عرض على GitHub↗681
  • catalyst-team/catalyst-rlالصورة الرمزية لـ catalyst-team

    catalyst-team/catalyst-rl

    48عرض على GitHub↗

    Accelerated RL

    Python
    عرض على GitHub↗48
  • coax-dev/coaxالصورة الرمزية لـ coax-dev

    coax-dev/coax

    185عرض على GitHub↗

    |tests| |pypi| |docs| |License|

    Python
    عرض على GitHub↗185
  • contextualai/halosالصورة الرمزية لـ ContextualAI

    ContextualAI/HALOs

    906عرض على GitHub↗

    A library with extensible implementations of DPO, KTO, PPO, ORPO, and other human-aware loss functions (HALOs).

    Pythonalignmentdpohalos
    عرض على GitHub↗906
  • coreylynch/async-rlالصورة الرمزية لـ coreylynch

    coreylynch/async-rl

    1,006عرض على GitHub↗

    Tensorflow Keras OpenAI Gym implementation of 1-step Q Learning from "Asynchronous Methods for Deep Reinforcement Learning"

    Python
    عرض على GitHub↗1,006
  • cpnota/autonomous-learning-libraryC

    cpnota/autonomous-learning-library

    0عرض على GitHub↗
    عرض على GitHub↗0
  • alihassanijr/discernالصورة الرمزية لـ alihassanijr

    alihassanijr/DISCERN

    0عرض على GitHub↗

    This repository contains the implementation of DISCERN in Python. You can download the manuscript from my website or arXiv.

    Python
    عرض على GitHub↗0