awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
NovaSky-AI avatar

NovaSky-AI/SkyRL

0
View on GitHub↗
1,611 نجوم·259 تفرعات·Python·apache-2.0·1 مشاهدةdocs.skyrl.ai/docs↗

SkyRL

Features

  • Reinforcement Learning - Full-stack RL library with modular training and inference.
  • Reinforcement Learning Frameworks - Reinforcement learning library for training reasoning models.

سجل النجوم

مخطط تاريخ النجوم لـ novasky-ai/skyrlمخطط تاريخ النجوم لـ novasky-ai/skyrl

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

الأسئلة الشائعة

ما هي الميزات الرئيسية لـ novasky-ai/skyrl؟

الميزات الرئيسية لـ novasky-ai/skyrl هي: Reinforcement Learning, Reinforcement Learning Frameworks.

ما هي البدائل مفتوحة المصدر لـ novasky-ai/skyrl؟

تشمل البدائل مفتوحة المصدر لـ novasky-ai/skyrl: google/dopamine — Dopamine is a reinforcement learning research framework designed for prototyping and testing algorithms across diverse… huggingface/trl — This library provides a comprehensive framework for fine-tuning, aligning, and distilling transformer-based language… alibaba/roll — ROLL is a distributed reinforcement learning framework and model alignment toolkit designed for large language models.… aunum/gold — Reinforcement Learning in Go. inclusionai/areal — AReaL is a system for agent orchestration, distributed model training, and parameter-efficient tuning. It provides a… instadeepai/jumanji — 🕹️ A diverse suite of scalable reinforcement learning environments in JAX.

بدائل مفتوحة المصدر لـ SkyRL

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع SkyRL.
  • google/dopamineالصورة الرمزية لـ google

    google/dopamine

    10,879عرض على GitHub↗

    Dopamine is a reinforcement learning research framework designed for prototyping and testing algorithms across diverse simulated environments. It provides an agent development toolkit that utilizes a flat class hierarchy to facilitate the creation and extension of learning agents. The framework includes a standardization layer via environment wrappers that connect agents to various physics simulations and gaming environments. It also features a high-performance experience replay buffer for storing and sampling transition data to improve training stability, alongside a dedicated hyperparameter

    Jupyter Notebook
    عرض على GitHub↗10,879
  • aunum/goldالصورة الرمزية لـ aunum

    aunum/gold

    351عرض على GitHub↗

    Reinforcement Learning in Go

    Go
    عرض على GitHub↗351
  • alibaba/rollالصورة الرمزية لـ alibaba

    alibaba/ROLL

    2,844عرض على GitHub↗

    ROLL is a distributed reinforcement learning framework and model alignment toolkit designed for large language models. It serves as a scalable training pipeline and GPU cluster manager, providing the infrastructure to align model behavior using reinforcement learning algorithms and preference optimization techniques. The project distinguishes itself through an agentic rollout orchestrator that generates and collects multi-turn interaction trajectories between AI agents and simulated environments. It supports specialized alignment methods including Direct Preference Optimization, reinforcement

    Pythonagenticrlhfrlvr
    عرض على GitHub↗2,844
  • huggingface/trlالصورة الرمزية لـ huggingface

    huggingface/trl

    18,653عرض على GitHub↗

    This library provides a comprehensive framework for fine-tuning, aligning, and distilling transformer-based language models. It serves as a toolkit for adapting models to specialized domains through supervised learning, while offering advanced methodologies to improve output quality and reasoning capabilities. The project distinguishes itself through specialized alignment and optimization techniques, including direct preference optimization and reinforcement learning, which allow models to be tuned against human preferences without complex reward modeling. It further supports training efficie

    Python
    عرض على GitHub↗18,653
عرض جميع البدائل الـ 30 لـ SkyRL→