awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعحولكيفية ترتيب النتائجالصحافةخادم MCP
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
brendanhogan avatar

brendanhogan/DeepSeekRL-Extended

0
View on GitHub↗
252 نجوم·34 تفرعات·Python·MIT·3 مشاهدات

DeepSeekRL Extended

Exploring Applications of GRPO

Features

  • Reasoning Models - Extended reinforcement learning for reasoning models.

سجل النجوم

مخطط تاريخ النجوم لـ brendanhogan/deepseekrl-extendedمخطط تاريخ النجوم لـ brendanhogan/deepseekrl-extended

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

بدائل مفتوحة المصدر لـ DeepSeekRL Extended

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع DeepSeekRL Extended.
  • qwenlm/qwen2.5الصورة الرمزية لـ QwenLM

    QwenLM/Qwen2.5

    27,307عرض على GitHub↗

    Qwen2.5 is a suite of large language model foundation models designed for natural language generation, code production, and complex mathematical reasoning. The project encompasses a multilingual language model capable of processing dozens of languages and a specialized code generation model for technical problem solving and debugging. The framework is distinguished by its long context capabilities, enabling the analysis of massive inputs ranging from 256K up to 1 million tokens. It further functions as an agentic framework, utilizing standardized templates and parsers to execute autonomous wo

    Python
    عرض على GitHub↗27,307
  • agentica-project/deepscalerA

    agentica-project/deepscaler

    0عرض على GitHub↗
    عرض على GitHub↗0
  • agentica-project/rllmالصورة الرمزية لـ agentica-project

    agentica-project/rllm

    400عرض على GitHub↗

    🚀 Reinforcement Learning for Language Agents🌟

    Jupyter Notebook
    عرض على GitHub↗400
  • adam-bjtu/openrftA

    ADaM-BJTU/OpenRFT

    0عرض على GitHub↗
    عرض على GitHub↗0
عرض جميع البدائل الـ 30 لـ DeepSeekRL Extended→

الأسئلة الشائعة

ما هي وظيفة brendanhogan/deepseekrl-extended؟

Exploring Applications of GRPO

ما هي الميزات الرئيسية لـ brendanhogan/deepseekrl-extended؟

الميزات الرئيسية لـ brendanhogan/deepseekrl-extended هي: Reasoning Models.

ما هي البدائل مفتوحة المصدر لـ brendanhogan/deepseekrl-extended؟

تشمل البدائل مفتوحة المصدر لـ brendanhogan/deepseekrl-extended: qwenlm/qwen2.5 — Qwen2.5 is a suite of large language model foundation models designed for natural language generation, code… agentica-project/deepscaler. agentica-project/rllm — 🚀 Reinforcement Learning for Language Agents🌟. aidc-ai/marco-o1 — An Open Large Reasoning Model for Real-World Solutions. aliyun/qwen-dianjin. adam-bjtu/openrft.