awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
wantbook-book avatar

wantbook-book/SeRL

0
View on GitHub↗
23 stars·4 forks·Python·Apache-2.0·11 views

SeRL

SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Data

Features

  • Single Agent Optimization - Self-play reinforcement learning for data-constrained environments.
  • Unsupervised Reward Methods - Self-play reinforcement learning optimized for limited data scenarios.

Star history

Star history chart for wantbook-book/serlStar history chart for wantbook-book/serl

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does wantbook-book/serl do?

SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Data

What are the main features of wantbook-book/serl?

The main features of wantbook-book/serl are: Single Agent Optimization, Unsupervised Reward Methods.

Which projects share features with wantbook-book/serl?

Projects with overlapping indexed features include: chengsong-huang/r-zero — Check out our paper or webpage for the details. wangqinsi1/vision-zero — A domain-agnostic framework enabling VLM self-improvement through competitive visual games. hkust-nlp/mstar — :star: Project Page    . gpoesia/minimo — This is the implementation of the following paper:. ezelikman/star — 1. STaR 2. Mesh Transformer JAX 1. Updates 3. Pretrained Models 1. GPT-J-6B 1. Links 2. Acknowledgments 3. License 4.… chengpengli1003/cort.

Projects sharing features with SeRL

These projects share indexed features with SeRL. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • wangqinsi1/vision-zerowangqinsi1 avatar

    wangqinsi1/Vision-Zero

    136View on GitHub↗

    A domain-agnostic framework enabling VLM self-improvement through competitive visual games

    Python
    View on GitHub↗136
  • chengsong-huang/r-zeroChengsong-Huang avatar

    Chengsong-Huang/R-Zero

    822View on GitHub↗

    Check out our paper or webpage for the details

    Python
    View on GitHub↗822
  • ezelikman/starezelikman avatar

    ezelikman/STaR

    227View on GitHub↗

    1. STaR 2. Mesh Transformer JAX 1. Updates 3. Pretrained Models 1. GPT-J-6B 1. Links 2. Acknowledgments 3. License 4. Model Details 5. Zero-Shot Evaluations 4. Architecture and Usage 1. Fine-tuning 2. JAX Dependency 5. TODO

    Python
    View on GitHub↗227
  • chengpengli1003/cortChengpengLi1003 avatar

    ChengpengLi1003/CoRT

    72View on GitHub↗
    Python
    View on GitHub↗72
Compare all 30 related projects
→