awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
AIFrameResearch avatar

AIFrameResearch/SPO

0
View on GitHub↗
53 stars·7 forks·Python·MIT·3 viewsarxiv.org/abs/2505.23564↗

SPO

🚀 Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models 🌟

Features

  • Dense Reward Optimization - Segment-level credit assignment for reinforcement learning.
  • Off-Policy Optimization - Online off-policy reinforcement learning for sequence models.

Star history

Star history chart for aiframeresearch/spoStar history chart for aiframeresearch/spo

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to SPO

Similar open-source projects, ranked by how many features they share with SPO.
  • rlinf/rlinfRLinf avatar

    RLinf/RLinf

    2,502View on GitHub↗

    RLinf is a distributed reinforcement learning orchestrator and embodied AI training framework. It provides the infrastructure to train vision-language-action models and robotic policies using a combination of reinforcement learning and supervised fine-tuning. The system is designed for scaling workloads across GPU clusters, managing the placement of actors, rollout workers, and environment components. It features a specialized robotics data collection pipeline for gathering teleoperated demonstrations and simulation trajectories into standardized replay buffers, alongside a hardware interface

    Pythonagentic-aiembodied-aireinforcement-learning
    View on GitHub↗2,502
  • morvanzhou/pytorch-tutorialMorvanZhou avatar

    MorvanZhou/PyTorch-Tutorial

    8,458View on GitHub↗

    This project is a collection of PyTorch learning resources and educational guides designed to teach the construction and training of neural networks. It serves as a comprehensive deep learning tutorial covering various model architectures and practical implementation strategies. The resources provide specific guidance on implementing computer vision tasks, such as image classification and synthetic imagery generation, as well as reinforcement learning agents using value networks and experience replay. It also covers sequential data modeling through recurrent networks and generative modeling u

    Jupyter Notebookautoencoderbatchbatch-normalization
    View on GitHub↗8,458
  • chanliang/bridgeChanLiang avatar

    ChanLiang/BRIDGE

    6View on GitHub↗

    The code for BRIDGE (Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning).

    View on GitHub↗6
  • chenluye99/profChenluye99 avatar

    Chenluye99/PROF

    11View on GitHub↗

    Introduction

    View on GitHub↗11
See all 30 alternatives to SPO→

Frequently asked questions

What does aiframeresearch/spo do?

🚀 Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models 🌟

What are the main features of aiframeresearch/spo?

The main features of aiframeresearch/spo are: Dense Reward Optimization, Off-Policy Optimization.

What are some open-source alternatives to aiframeresearch/spo?

Open-source alternatives to aiframeresearch/spo include: rlinf/rlinf — RLinf is a distributed reinforcement learning orchestrator and embodied AI training framework. It provides the… morvanzhou/pytorch-tutorial — This project is a collection of PyTorch learning resources and educational guides designed to teach the construction… chanliang/bridge — The code for BRIDGE (Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning). cjreinforce/pure — [2025/10/23] 🔥🔥Our paper is accepted by NeurIPS 2025.🔥🔥 - [2025/04/22] Released our Paper on arXiv. See here -… cmu-aire/mrt — This repository contains the code for our paper titled "Optimizing Test-Time Compute via Meta Reinforcement… chenluye99/prof — Introduction.