26 open-source projects similar to agent-rl/recall, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.
This repository contains the official implementation for the paper "AutoLogi: Automated Generation of Logic Puzzles for Evaluating Reasoning Abilities of Large Language Models".
Aligning Text and Embodied Environments for Interactive Learning Mohit Shridhar, Xingdi (Eric) Yuan, Marc-Alexandre Côté, Yonatan Bisk, Adam Trischler, Matthew Hausknecht ICLR 2021
We introduce Enigmata, the first comprehensive suite tailored for improving LLMs with puzzle reasoning skills, which integrates seamlessly with reinforcement learning using verifiable rule-based rewards.
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
MLGym A New Framework and Benchmark for Advancing AI Research Agents
Optimus-3: Dual-Router Aligned Mixture-of-Experts Agent with Dual-Granularity Reasoning-Aware Policy Optimization Zaijing Li 1 2 , Yuquan Xie 1 , Rui Shao 1✉ , Gongwei Chen 1 , Weili Guan 1 , Dongmei Jiang 2 , Liqiang Nie 1✉ 1 Harbin Institute of…
Repository with environment and training scripts for paper Cross-Environment Cooperation Enables Zero-shot Multi-agent Coordination (Jha et. al, 2025)
⚙️ Algorithm Flow • 📊 Results ✨ Getting Started • 🏋️ Training • 🔧 Usage • 📃 Evaluation 🎈 Citation • 🌻 Acknowledgement • 📧 Contact • 📈 Star History
A Collection of Competitive Text-Based Games for Language Model Evaluation and Reinforcement Learning
LMGame Bench and Gaming Agent 📜 Paper | 🏆 Leaderboard | 📺 Gallery | 🌐 Website
NeurIPS 2025 The official repo of SynLogic: Synthesizing Verifiable Reasoning Data at Scale for Learning Logical Reasoning and Beyond
MLE-Dojo: Interactive Environments for Empowering LLM Agents in Machine Learning Engineering
NeurIPS 2025 Spotlight Reasoning Environments for Reinforcement Learning with Verifiable Rewards
We propose ZeroGUI, a fully automated online reinforcement learning framework that enables GUI agents to train and adapt in interactive environments at zero human cost.
R2E-Gym: Procedural Environment Generation and Hybrid Verifiers for Scaling Open-Weights SWE Agents
This repository contains the code for PuzzleJAX, a GPU-accelerated implementation of PuzzleScript (https://www.puzzlescript.net)
Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning
🌍 AppWorld: A Controllable World of Apps and People for Benchmarking Function Calling and Interactive Coding Agent, ACL'24 Best Resource Paper.
Official repository for paper "Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning". This is the first work, to the best of our knowledge, that adapts game code to synthesize multimodal game data for training VLMs. When we apply Game-RL, which is simple GRPO on…
ICLR'26 MedAgentGYM: Training LLM Agents for Code-Based Medical Reasoning at Scale