awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Back to liumy2010/uft

Projects sharing features with Liumy2010 UFT

12 open-source projects similar to liumy2010/uft, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.

  • morvanzhou/pytorch-tutorialMorvanZhou avatar

    MorvanZhou/PyTorch-Tutorial

    8,458View on GitHub↗

    This project is a collection of PyTorch learning resources and educational guides designed to teach the construction and training of neural networks. It serves as a comprehensive deep learning tutorial covering various model architectures and practical implementation strategies. The resources provide specific guidance on implementing computer vision tasks, such as image classification and synthetic imagery generation, as well as reinforcement learning agents using value networks and experience replay. It also covers sequential data modeling through recurrent networks and generative modeling u

    Jupyter Notebookautoencoderbatchbatch-normalization
    View on GitHub↗8,458
  • rlinf/rlinfRLinf avatar

    RLinf/RLinf

    2,502View on GitHub↗

    RLinf is a distributed reinforcement learning orchestrator and embodied AI training framework. It provides the infrastructure to train vision-language-action models and robotic policies using a combination of reinforcement learning and supervised fine-tuning. The system is designed for scaling workloads across GPU clusters, managing the placement of actors, rollout workers, and environment components. It features a specialized robotics data collection pipeline for gathering teleoperated demonstrations and simulation trajectories into standardized replay buffers, alongside a hardware interface

    Pythonagentic-aiembodied-aireinforcement-learning
    View on GitHub↗2,502
  • chanliang/bridgeChanLiang avatar

    ChanLiang/BRIDGE

    6View on GitHub↗

    The code for BRIDGE (Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning).

    View on GitHub↗6

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • elliottyan/luffyElliottYan avatar

    ElliottYan/LUFFY

    455View on GitHub↗

    LUFFY: Learning to Reason Under Off‑Policy Guidance A general framework for off-policy learning in large reasoning models.

    Python
    View on GitHub↗455
  • millioniron/openrlhf-millioniron-millioniron avatar

    millioniron/OpenRLHF-Millioniron-

    2View on GitHub↗

    Open-source / Comprehensive / Lightweight / Easy-to-use

    View on GitHub↗2
  • mozerwang/ampoMozerWang avatar

    MozerWang/AMPO

    51View on GitHub↗

    Minzheng Wang 1,2 , Yongbin Li 3 , Haobo Wang 4 , Xinghua Zhang 3🌟 , Nan Xu 1 , Bingli Wu 3 , Fei Huang 3 , Haiyang Yu 3 , Wenji Mao 1,2🌟

    Python
    View on GitHub↗51
  • theroadqaq/reliftTheRoadQaQ avatar

    TheRoadQaQ/ReLIFT

    84View on GitHub↗

    Learning What Reinforcement Learning Can't

    Python
    View on GitHub↗84
  • tsinghuac3i/intuitive-fine-tuningTsinghuaC3I avatar

    TsinghuaC3I/Intuitive-Fine-Tuning

    30View on GitHub↗

    This repository contains the code for the paper "Intuitive Fine-Tuning: Towards Simplifying Alignment into a Single Process".

    Python
    View on GitHub↗30
  • yaof20/verlyaof20 avatar

    yaof20/verl

    21View on GitHub↗

    👋 Hi, everyone! verl is a RL training library initiated by ByteDance Seed team and maintained by the verl community.

    Python
    View on GitHub↗21
  • aiframeresearch/spoAIFrameResearch avatar

    AIFrameResearch/SPO

    53View on GitHub↗

    🚀 Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models 🌟

    Python
    View on GitHub↗53
  • yongliang-wu/dftyongliang-wu avatar

    yongliang-wu/DFT

    581View on GitHub↗

    Yizhou Zhou*   Zhou Ziheng   Yingzhe Peng   Xinyu Ye   Xinting Hu   Wenbo Zhu   Lu Qi   Ming-Hsuan Yang   Xu Yang  

    Python
    View on GitHub↗581
  • anitaleungxx/remix-reincarnated-mix-policy-proximal-policy-gradientAnitaLeungxx avatar

    AnitaLeungxx/ReMix-Reincarnated-Mix-policy-Proximal-Policy-Gradient

    12View on GitHub↗

    🧽 Squeeze the Soaked Sponge 🌊 Efficient Off-policy Reinforcement Finetuning for Large Language Model

    Python
    View on GitHub↗12