awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
xfactlab avatar

xfactlab/orpo

0
View on GitHub↗
481 stars·44 forks·Python·Apache-2.0·0 views

Orpo

[X] Sample script for ORPOTrainer in 🤗 TRL is added to trl/testorpotrainerdemo.py - [X] New model, 🤗 kaist-ai/mistral-orpo-capybara-7k , is added to 🤗 ORPO Collection - [X] Now you can try ORPO in 🤗 TRL , Axolotl and LLaMA-Factory 🔥 - [X] We are making general guideline for training LLMs…

Features

  • Reinforcement Learning - Monolithic preference optimization without a separate reference model.

Star history

Star history chart for xfactlab/orpoStar history chart for xfactlab/orpo

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Orpo

Similar open-source projects, ranked by how many features they share with Orpo.
  • ai4co/rl4coai4co avatar

    ai4co/rl4co

    806View on GitHub↗
    Pythonattentionattention-modelbenchmark
    View on GitHub↗806
  • ai4finance-foundation/finrlAI4Finance-Foundation avatar

    AI4Finance-Foundation/FinRL

    13,964View on GitHub↗

    FinRL is a reinforcement learning framework designed for the development, training, and backtesting of automated trading strategies. It functions as a quantitative finance toolkit that integrates deep learning algorithms with financial market simulations to address complex portfolio management and asset allocation tasks. The platform provides an end-to-end pipeline for transforming raw market data into actionable trading models. The project distinguishes itself through a layered, modular architecture that separates data processing, environment simulation, and agent training. This design allow

    Jupyter Notebookalgorithmic-tradingdeep-reinforcement-learningdrl-algorithms
    View on GitHub↗13,964
  • aikorea/awesome-rlaikorea avatar

    aikorea/awesome-rl

    9,812View on GitHub↗

    Reinforcement learning resources curated

    View on GitHub↗9,812
  • 2toinf/uniact2toinf avatar

    2toinf/UniAct

    241View on GitHub↗

    Project Page Paper

    Python
    View on GitHub↗241
See all 30 alternatives to Orpo→

Frequently asked questions

What does xfactlab/orpo do?

[X] Sample script for ORPOTrainer in 🤗 TRL is added to trl/testorpotrainerdemo.py - [X] New model, 🤗 kaist-ai/mistral-orpo-capybara-7k , is added to 🤗 ORPO Collection - [X] Now you can try ORPO in 🤗 TRL , Axolotl and LLaMA-Factory 🔥 - [X] We are making general guideline for training LLMs…

What are the main features of xfactlab/orpo?

The main features of xfactlab/orpo are: Reinforcement Learning.

What are some open-source alternatives to xfactlab/orpo?

Open-source alternatives to xfactlab/orpo include: ai4co/rl4co. ai4finance-foundation/finrl — FinRL is a reinforcement learning framework designed for the development, training, and backtesting of automated… aikorea/awesome-rl — Reinforcement learning resources curated. airlab-polimi/mushroom. alessiodm/drl-zh — Welcome to drlzh.ai: a hands-on deep reinforcement learning course where you build the algorithms, not just read about… 2toinf/uniact — [Project Page] [Paper].