awesome-repositories.com
博客
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目关于排名机制媒体报道MCP 服务器
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to dhcode-cpp/x-r1

Open-source alternatives to X R1

30 open-source projects similar to dhcode-cpp/x-r1, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best X R1 alternative.

  • qwenlm/qwen2.5QwenLM 的头像

    QwenLM/Qwen2.5

    27,307在 GitHub 上查看↗

    Qwen2.5 is a suite of large language model foundation models designed for natural language generation, code production, and complex mathematical reasoning. The project encompasses a multilingual language model capable of processing dozens of languages and a specialized code generation model for technical problem solving and debugging. The framework is distinguished by its long context capabilities, enabling the analysis of massive inputs ranging from 256K up to 1 million tokens. It further functions as an agentic framework, utilizing standardized templates and parsers to execute autonomous wo

    Python
    在 GitHub 上查看↗27,307
  • agentica-project/deepscalerA

    agentica-project/deepscaler

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • deep-agent/r1-vD

    Deep-Agent/R1-V

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • aidc-ai/marco-o1AIDC-AI 的头像

    AIDC-AI/Marco-o1

    1,540在 GitHub 上查看↗

    An Open Large Reasoning Model for Real-World Solutions

    Python
    在 GitHub 上查看↗1,540
  • aliyun/qwen-dianjinA

    aliyun/qwen-dianjin

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • appletea233/temporal-r1appletea233 的头像

    appletea233/Temporal-R1

    62在 GitHub 上查看↗

    Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency

    Python
    在 GitHub 上查看↗62

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Find more with AI search
  • atfortes/awesome-llm-reasoningatfortes 的头像

    atfortes/Awesome-LLM-Reasoning

    3,640在 GitHub 上查看↗

    From Chain-of-Thought prompting to OpenAI o1 and DeepSeek-R1 🍓

    awesomechain-of-thoughtchatgpt
    在 GitHub 上查看↗3,640
  • baibizhe/efficient-r1-vllmB

    baibizhe/Efficient-R1-VLLM

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • baichuan-inc/baichuan-m1-14bbaichuan-inc 的头像

    baichuan-inc/Baichuan-M1-14B

    219在 GitHub 上查看↗

    Baichuan-M1-14B

    在 GitHub 上查看↗219
  • bklieger-groq/g1B

    bklieger-groq/g1

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • brendanhogan/deepseekrl-extendedbrendanhogan 的头像

    brendanhogan/DeepSeekRL-Extended

    252在 GitHub 上查看↗

    Exploring Applications of GRPO

    Python
    在 GitHub 上查看↗252
  • bytedance-seed/seed-thinking-v1.5B

    ByteDance-Seed/Seed-Thinking-v1.5

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • csfufu/revisual-r1C

    CSfufu/Revisual-R1

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • datawhalechina/unlock-deepseekdatawhalechina 的头像

    datawhalechina/unlock-deepseek

    733在 GitHub 上查看↗

    DeepSeek 系列工作解读、扩展和复现。

    Python
    在 GitHub 上查看↗733
  • agentica-project/rllmagentica-project 的头像

    agentica-project/rllm

    400在 GitHub 上查看↗

    🚀 Reinforcement Learning for Language Agents🌟

    Jupyter Notebook
    在 GitHub 上查看↗400
  • deepseek-ai/deepseek-r1deepseek-ai 的头像

    deepseek-ai/DeepSeek-R1

    91,996在 GitHub 上查看↗

    DeepSeek-R1 is an open-weights large language model focused on advanced reasoning. It uses chain-of-thought processing and internal monologues to solve complex mathematical and logical problems by breaking tasks into sequential, verifiable thought processes. The model is developed using reinforcement learning to optimize reasoning patterns and verify logical steps. It employs a distillation process to transfer these high-performance logic capabilities from a large teacher model into smaller, computationally efficient versions. The training framework incorporates group relative policy optimiz

    在 GitHub 上查看↗91,996
  • deepseek-ai/deepseek-v4D

    deepseek-ai/DeepSeek-V4

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • dvlab-research/seg-zerodvlab-research 的头像

    dvlab-research/Seg-Zero

    632在 GitHub 上查看↗

    Project Page For "Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement"

    Python
    在 GitHub 上查看↗632
  • evelinehong/3d-clr-officialevelinehong 的头像

    evelinehong/3D-CLR-Official

    85在 GitHub 上查看↗

    Checkpoints take up a lot of space. Please email yninghong@gmail.com if you need them.

    Python
    在 GitHub 上查看↗85
  • evolvinglmms-lab/open-r1-multimodalEvolvingLMMs-Lab 的头像

    EvolvingLMMs-Lab/open-r1-multimodal

    1,484在 GitHub 上查看↗
    Python
    在 GitHub 上查看↗1,484
  • facebookresearch/swe-rlfacebookresearch 的头像

    facebookresearch/swe-rl

    704在 GitHub 上查看↗

    🧐 About | 🚀 Quick Start | 🐣 Agentless Mini | 📝 Citation | 🙏 Acknowledgements

    Python
    在 GitHub 上查看↗704
  • fancy-mllm/r1-onevisionFancy-MLLM 的头像

    Fancy-MLLM/R1-Onevision

    581在 GitHub 上查看↗

    R1-onevision, a visual language model capable of deep CoT reasoning.

    Python
    在 GitHub 上查看↗581
  • fanqingm/r1-multimodal-journeyF

    FanqingM/R1-Multimodal-Journey

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0
  • flagai-open/openseekFlagAI-Open 的头像

    FlagAI-Open/OpenSeek

    262在 GitHub 上查看↗

    OpenSeek aims to unite the global open source community to drive collaborative innovation in algorithms, data and systems to develop next-generation models.

    Python
    在 GitHub 上查看↗262
  • freedomintelligence/huatuogpt-o1FreedomIntelligence 的头像

    FreedomIntelligence/HuatuoGPT-o1

    1,331在 GitHub 上查看↗

    Medical o1, Towards medical complex reasoning with LLMs

    Python
    在 GitHub 上查看↗1,331
  • gair-nlp/limoGAIR-NLP 的头像

    GAIR-NLP/LIMO

    1,077在 GitHub 上查看↗

    📄 Paper | 🌐 Dataset (v2) | 📘 Model (v2)

    Python
    在 GitHub 上查看↗1,077
  • gair-nlp/o1-journeyGAIR-NLP 的头像

    GAIR-NLP/O1-Journey

    2,000在 GitHub 上查看↗

    O1 Replication Journey

    在 GitHub 上查看↗2,000
  • hijkzzz/awesome-llm-strawberryhijkzzz 的头像

    hijkzzz/Awesome-LLM-Strawberry

    6,896在 GitHub 上查看↗

    A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.

    chain-of-thoughtcodingllm
    在 GitHub 上查看↗6,896
  • hiyouga/easyr1hiyouga 的头像

    hiyouga/EasyR1

    5,034在 GitHub 上查看↗

    EasyR1 is a distributed model training system and reinforcement learning framework for large language and vision-language models. It functions as a multimodal trainer and an implementation of a Proximal Policy Optimization pipeline designed to refine the reasoning and perception capabilities of models that process both text and images. The system specializes in distributing reinforcement learning workloads across multiple compute nodes to manage high memory requirements. It optimizes hardware utilization through padding-free training and fine-tuning to fit large models onto available graphics

    Python
    在 GitHub 上查看↗5,034
  • adam-bjtu/openrftA

    ADaM-BJTU/OpenRFT

    0在 GitHub 上查看↗
    在 GitHub 上查看↗0