awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
yaof20 avatar

yaof20/Flash-RL

0
View on GitHub↗
305 stars·23 forks·Python·MIT·6 viewsfengyao.notion.site/flash-rl↗

Flash RL

Fast RL training with Quantized Rollouts ( Blog )

Features

  • Critic-Free Algorithms - Accelerated reinforcement learning training using quantized rollouts.

Star history

Star history chart for yaof20/flash-rlStar history chart for yaof20/flash-rl

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Flash RL

Similar open-source projects, ranked by how many features they share with Flash RL.
  • deepseek-ai/deepseek-mathdeepseek-ai avatar

    deepseek-ai/DeepSeek-Math

    3,346View on GitHub↗

    DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

    Python
    View on GitHub↗3,346
  • liziniu/remaxliziniu avatar

    liziniu/ReMax

    202View on GitHub↗

    ReMax is a reinforcement learning method, tailored for reward maximization in RLHF.

    Python
    View on GitHub↗202
  • mcgill-nlp/vineppoMcGill-NLP avatar

    McGill-NLP/VinePPO

    192View on GitHub↗

    Paper - Abstract - Updates - Quick Start - Installation - Download the datasets - Create Experiment Script - Single GPU Training (Only for Rho models) - Running the experiments - Code Structure - Initial SFT Checkpoints - Acknowledgement - Citation

    Python
    View on GitHub↗192
  • bytedtsinghua-sia/dapoBytedTsinghua-SIA avatar

    BytedTsinghua-SIA/DAPO

    1,831View on GitHub↗

    DAPO: an Open-source RL System from ByteDance Seed and Tsinghua AIR

    Python
    View on GitHub↗1,831
See all 11 alternatives to Flash RL→

Frequently asked questions

What does yaof20/flash-rl do?

Fast RL training with Quantized Rollouts ( Blog )

What are the main features of yaof20/flash-rl?

The main features of yaof20/flash-rl are: Critic-Free Algorithms.

What are some open-source alternatives to yaof20/flash-rl?

Open-source alternatives to yaof20/flash-rl include: bytedtsinghua-sia/dapo — DAPO: an Open-source RL System from ByteDance Seed and Tsinghua AIR. deepseek-ai/deepseek-math — DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models. liziniu/remax — ReMax is a reinforcement learning method, tailored for reward maximization in RLHF. mcgill-nlp/vineppo — Paper - Abstract - Updates - Quick Start - Installation - Download the datasets - Create Experiment Script - Single… minimax-ai/minimax-m1 — MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model. modalminds/mm-eureka — MM-EUREKA: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.