awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
long-horizon-execution avatar

long-horizon-execution/measuring-execution

0
View on GitHub↗
56 stars·9 forks·Python·10 views

Measuring Execution

This project contains the code accompanying the ICLR 2026 paper The Illusion of Diminishing Returns: Measuring Long Horizon Execution in LLMs (OpenReview, arXiv).

Features

  • Agentic Reasoning Applications - Framework for measuring long-horizon execution in language models.

Star history

Star history chart for long-horizon-execution/measuring-executionStar history chart for long-horizon-execution/measuring-execution

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Projects sharing features with Measuring Execution

These projects share indexed features with Measuring Execution. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • dongguanting/arpodongguanting avatar

    dongguanting/ARPO

    1,061View on GitHub↗

    ✨ Agentic Reinforced Policy Optimization

    Python
    View on GitHub↗1,061
  • facebookresearch/swe-rlfacebookresearch avatar

    facebookresearch/swe-rl

    704View on GitHub↗

    🧐 About | 🚀 Quick Start | 🐣 Agentless Mini | 📝 Citation | 🙏 Acknowledgements

    Python
    View on GitHub↗704
  • gair-nlp/torlGAIR-NLP avatar

    GAIR-NLP/ToRL

    351View on GitHub↗

    #

    Python
    View on GitHub↗351
  • chengpengli1003/cortChengpengLi1003 avatar

    ChengpengLi1003/CoRT

    72View on GitHub↗
    Python
    View on GitHub↗72
Compare all 11 related projects→

Frequently asked questions

What does long-horizon-execution/measuring-execution do?

This project contains the code accompanying the ICLR 2026 paper The Illusion of Diminishing Returns: Measuring Long Horizon Execution in LLMs (OpenReview, arXiv).

What are the main features of long-horizon-execution/measuring-execution?

The main features of long-horizon-execution/measuring-execution are: Agentic Reasoning Applications.

Which projects share features with long-horizon-execution/measuring-execution?

Projects with overlapping indexed features include: chengpengli1003/cort. dongguanting/arpo — ✨ Agentic Reinforced Policy Optimization. facebookresearch/swe-rl — 🧐  About | 🚀  Quick Start | 🐣  Agentless Mini | 📝  Citation | 🙏  Acknowledgements. gair-nlp/torl — #. internlm/internbootcamp — 1. 架构概述 - 1.1 InternBootcampv2核心改进 - 1.2 Multi-round Toolcall的实现原理 - 1.2.1 与Bootcampv1的差异 - 1.2.2 复杂Bootcamp的代码逻辑 - 2.… ltzheng/simpletir — SimpleTIR: End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning Zhenghai Xue · Longtao Zheng ·…