awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
BytedTsinghua-SIA avatar

BytedTsinghua-SIA/DAPO

0
View on GitHub↗
1,831 stars·84 forks·Python·8 views

DAPO

DAPO: an Open-source RL System from ByteDance Seed and Tsinghua AIR

Features

  • Critic-Free Algorithms - Large-scale open-source reinforcement learning system for LLMs.
  • Policy Optimization - Scalable reinforcement learning system for language models.

Star history

Star history chart for bytedtsinghua-sia/dapoStar history chart for bytedtsinghua-sia/dapo

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does bytedtsinghua-sia/dapo do?

DAPO: an Open-source RL System from ByteDance Seed and Tsinghua AIR

What are the main features of bytedtsinghua-sia/dapo?

The main features of bytedtsinghua-sia/dapo are: Critic-Free Algorithms, Policy Optimization.

Which projects share features with bytedtsinghua-sia/dapo?

Projects with overlapping indexed features include: trademaster-ntu/trademaster — TradeMaster is a reinforcement learning trading framework and algorithmic trading simulator designed for designing and… deepseek-ai/deepseek-math — DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models. liaomengqi/e3-rl4llms — [25/08/20] : Aceept as EMNLP 2025 Main Conference paper. liziniu/remax — ReMax is a reinforcement learning method, tailored for reward maximization in RLHF. mcgill-nlp/vineppo — Paper - Abstract - Updates - Quick Start - Installation - Download the datasets - Create Experiment Script - Single… chenxinan-fdu/polaris — 🌠 A PO st-training recipe for scaling R L on A dvanced R eason I ng model S 🚀.

Projects sharing features with DAPO

These projects share indexed features with DAPO. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • trademaster-ntu/trademasterTradeMaster-NTU avatar

    TradeMaster-NTU/TradeMaster

    2,484View on GitHub↗

    TradeMaster is a reinforcement learning trading framework and algorithmic trading simulator designed for designing and testing quantitative trading strategies. The system provides a platform for developing reinforcement learning agents, managing quantitative portfolios, and optimizing trade execution using financial market data. The project features specialized components for multi-modality data preprocessing, a high-fidelity market environment simulation for strategy backtesting, and a quantitative portfolio manager for capital reallocation across multiple assets. It includes a trade executi

    Jupyter Notebookfinancefintechinvestment-strategies
    View on GitHub↗2,484
  • deepseek-ai/deepseek-mathdeepseek-ai avatar

    deepseek-ai/DeepSeek-Math

    3,346View on GitHub↗

    DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

    Python
    View on GitHub↗3,346
  • liaomengqi/e3-rl4llmsLiaoMengqi avatar

    LiaoMengqi/E3-RL4LLMs

    17View on GitHub↗

    25/08/20 : Aceept as EMNLP 2025 Main Conference paper

    Python
    View on GitHub↗17
  • chenxinan-fdu/polarisChenxinAn-fdu avatar

    ChenxinAn-fdu/POLARIS

    688View on GitHub↗

    🌠 A PO st-training recipe for scaling R L on A dvanced R eason I ng model S 🚀

    Python
    View on GitHub↗688
Compare all 25 related projects→