awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
andri27-ts avatar

andri27-ts/Reinforcement-Learning

0
View on GitHub↗
4,722 स्टार्स·669 फोर्क्स·Jupyter Notebook·MIT·14 व्यूज़andri27-ts.github.io/Reinforcement-Learning↗

Reinforcement Learning

यह प्रोजेक्ट Python में लिखे गए सुदृढीकरण शिक्षण (reinforcement learning) कार्यान्वयनों और शैक्षिक सामग्रियों का एक संग्रह है। यह डीप सुदृढीकरण शिक्षण के माध्यम से नियंत्रण कार्यों को हल करने के लिए न्यूरल नेटवर्क आर्किटेक्चर प्रदान करता है, जो मूल्य-आधारित और नीति-ग्रेडिएंट विधियों तक फैला है।

रिपॉजिटरी में ग्रेडिएंट-आधारित शिक्षण के विकल्प के रूप में विकासवादी रणनीतियों (evolutionary strategies) और आनुवंशिक एल्गोरिदम की एक लाइब्रेरी शामिल है। इसमें आंतरिक सिमुलेशन और ऑफ़लाइन योजना को सक्षम करने के लिए भविष्य के वातावरण राज्यों और पुरस्कारों की भविष्यवाणी करने के लिए एक मॉडल-आधारित सिस्टम भी है।

कोडबेस क्षमताओं की एक विस्तृत श्रृंखला को कवर करता है, जिसमें स्थिर व्यवहार अपडेट के लिए एक्टर-क्रिटिक फ्रेमवर्क और प्रॉक्सिमल पॉलिसी ऑप्टिमाइज़ेशन शामिल है। यह डीप Q-नेटवर्क्स, SARSA वेरिएंट्स, और ड्यूलिंग नेटवर्क आर्किटेक्चर जैसी मूल्य-आधारित शिक्षण तकनीकों को लागू करता है, साथ ही एक्शन चयन और शोर-आधारित अन्वेषण के लिए तंत्र भी शामिल है।

यह प्रोजेक्ट एक कोर्स के रूप में संरचित है, जो डीप सुदृढीकरण शिक्षण और न्यूरल नेटवर्क प्रशिक्षण सिखाने के लिए Python कोड को लेक्चर्स के साथ जोड़ता है।

Features

  • Reinforcement Learning Curricula - Provides a structured curriculum of lectures and Python code for learning deep reinforcement learning.
  • Actor-Critic Architectures - Implements actor-critic architectures that combine policy-learning actors and value-estimating critics to improve training convergence.
  • Evolutionary Strategy Optimization - Includes a library of evolutionary strategies and genetic algorithms as alternatives to gradient-based learning.
  • Deep Q-Learning Implementations - Provides deep Q-learning implementations using neural networks and experience replay to master game environments.
  • Deep Reinforcement Learning Implementations - Provides deep reinforcement learning implementations for solving control tasks using value-based and policy-gradient methods.
  • Evolutionary Strategy Libraries - Ships a dedicated library of evolutionary strategies and genetic algorithms for agent optimization.
  • Policy Gradient Methods - Develops agents for complex control tasks using gradient-based policy optimization and actor-critic architectures.
  • Proximal Policy Optimization - Provides a surrogate objective function optimizer to ensure stable policy updates in continuous action spaces.
  • Model-Based Planning - Features a model-based system for predicting environment states and rewards to enable internal simulation and offline planning.
  • Model-Based RL Systems - Features a model-based RL system that enables internal simulation and offline planning.
  • Policy Gradient Optimizers - Uses gradient-based optimization methods to directly adjust agent behavior within actor-critic architectures.
  • Deep Value-Based Training - Implements value-based learning techniques including deep Q-networks and SARSA variants for discrete environments.
  • State-Action Value Updates - Implements state-action value updates by calculating errors between target rewards and current value predictions.
  • Double DQN Implementations - Implements Double Deep Q-Networks to decouple action selection from value estimation and reduce overestimation bias.
  • Discrete Environment Solvers - Trains agents to find optimal action-value functions specifically for discrete environment settings.
  • Environment Dynamics Modeling - A model-based approach that builds internal representations of the environment to forecast future states and optimize agent actions.
  • Experience Replay Buffers - Provides experience replay buffers to store and sample past interactions, breaking data correlation for stable training.
  • Exploration Strategies - Provides greedy and epsilon-greedy action selection strategies to balance exploration and exploitation.
  • Noise-Based Exploration - Implements noise layers within the network to manage the exploration-exploitation trade-off without epsilon-greedy methods.
  • Genetic Algorithms - Employs genetic algorithms to optimize agent parameters as a scalable alternative to traditional gradient-based methods.
  • Clipped Policy Objectives - Uses a clipped surrogate objective function to ensure stable behavior updates in continuous action spaces.
  • Neural Dynamics Models - Implements neural dynamics models to predict future states and rewards, enabling planning and knowledge transfer.
  • Environment Simulation Planning - Builds internal neural network models to predict future states and rewards for action optimization within simulations.
  • Dueling Network Architectures - Implements neural network architectures that decouple state value estimation from action advantage for improved stability.
  • Sarsa Update Implementations - Implements the Sarsa on-policy temporal difference learning algorithm using deep neural networks.
  • Multi-Step Return Calculations - Provides a forward-view multi-step approach for updating target values to increase learning efficiency.
  • Multi-Step Returns - Implements multi-step temporal difference learning to propagate value information faster through the state space.

स्टार हिस्ट्री

andri27-ts/reinforcement-learning के लिए स्टार हिस्ट्री चार्टandri27-ts/reinforcement-learning के लिए स्टार हिस्ट्री चार्ट

AI सर्च

और अधिक बेहतरीन रिपॉजिटरी खोजें

अपनी ज़रूरत को सरल भाषा में बताएं — AI हजारों क्यूरेटेड ओपन-सोर्स प्रोजेक्ट्स को प्रासंगिकता के आधार पर रैंक करता है।

Start searching with AI

अक्सर पूछे जाने वाले प्रश्न

andri27-ts/reinforcement-learning क्या करता है?

यह प्रोजेक्ट Python में लिखे गए सुदृढीकरण शिक्षण (reinforcement learning) कार्यान्वयनों और शैक्षिक सामग्रियों का एक संग्रह है। यह डीप सुदृढीकरण शिक्षण के माध्यम से नियंत्रण कार्यों को हल करने के लिए न्यूरल नेटवर्क आर्किटेक्चर प्रदान करता है, जो मूल्य-आधारित और नीति-ग्रेडिएंट विधियों तक फैला है।

andri27-ts/reinforcement-learning की मुख्य विशेषताएं क्या हैं?

andri27-ts/reinforcement-learning की मुख्य विशेषताएं हैं: Reinforcement Learning Curricula, Actor-Critic Architectures, Evolutionary Strategy Optimization, Deep Q-Learning Implementations, Deep Reinforcement Learning Implementations, Evolutionary Strategy Libraries, Policy Gradient Methods, Proximal Policy Optimization।

andri27-ts/reinforcement-learning के कुछ ओपन-सोर्स विकल्प क्या हैं?

andri27-ts/reinforcement-learning के ओपन-सोर्स विकल्पों में शामिल हैं: morvanzhou/reinforcement-learning-with-tensorflow — This project is an educational repository of reinforcement learning agents and tutorials implemented using TensorFlow.… packtpublishing/deep-reinforcement-learning-hands-on — This project serves as an educational resource and training framework for developing intelligent agents through deep… ljpzzz/machinelearning — This project is a machine learning implementation library featuring a collection of code examples that implement… yandexdataschool/practical_rl — Practical_RL is a comprehensive educational curriculum and course for learning to design and implement agents that… dennybritz/reinforcement-learning — This repository provides a comprehensive library of reinforcement learning algorithms designed for training autonomous… udacity/deep-reinforcement-learning — This project is a deep reinforcement learning curriculum providing educational materials and implementation exercises…

Reinforcement Learning के ओपन-सोर्स विकल्प

समान ओपन-सोर्स प्रोजेक्ट्स, जो Reinforcement Learning के साथ साझा की गई सुविधाओं के आधार पर रैंक किए गए हैं।
  • morvanzhou/reinforcement-learning-with-tensorflowMorvanZhou का अवतार

    MorvanZhou/Reinforcement-learning-with-tensorflow

    9,464GitHub पर देखें↗

    This project is an educational repository of reinforcement learning agents and tutorials implemented using TensorFlow. It provides a practical codebase for both model-free and model-based learning agents, designed to demonstrate how AI agents learn through trial and error. The collection features detailed implementations of various algorithmic approaches, including Deep Q-Networks and Policy Gradient methods. It specifically covers Actor-Critic architectures for continuous and discrete action spaces, alongside Proximal Policy Optimization and Deep Deterministic Policy Gradients. The framewor

    Pythona3cactor-criticasynchronous-advantage-actor-critic
    GitHub पर देखें↗9,464
  • packtpublishing/deep-reinforcement-learning-hands-onPacktPublishing का अवतार

    PacktPublishing/Deep-Reinforcement-Learning-Hands-On

    3,098GitHub पर देखें↗

    This project serves as an educational resource and training framework for developing intelligent agents through deep reinforcement learning. It provides a collection of practical tutorials and code examples designed to teach the implementation of neural networks for solving complex decision-making tasks. By focusing on hands-on learning, the material guides users through the process of building autonomous systems that improve their performance through trial and error. The framework centers on the integration of standardized simulation environments, allowing agents to interact with diverse tas

    Python
    GitHub पर देखें↗3,098
  • ljpzzz/machinelearningljpzzz का अवतार

    ljpzzz/machinelearning

    8,706GitHub पर देखें↗

    This project is a machine learning implementation library featuring a collection of code examples that implement supervised, unsupervised, and reinforcement learning algorithms from scratch. It provides a comprehensive set of toolkits for core machine learning components, including a natural language processing toolkit, a reinforcement learning framework, and suites for data dimensionality reduction and pattern mining. The library includes specialized implementations for reinforcement learning, such as Q-Learning, Deep Q-Networks, and Actor-Critic agents. The natural language processing capab

    Jupyter Notebookalgorithmsmachinelearningreinforcementlearning
    GitHub पर देखें↗8,706
  • yandexdataschool/practical_rlyandexdataschool का अवतार

    yandexdataschool/Practical_RL

    6,522GitHub पर देखें↗

    Practical_RL is a comprehensive educational curriculum and course for learning to design and implement agents that solve complex decision processes. It provides a structured study program covering the fundamentals of reinforcement learning, from basic trial-and-error behavior to advanced deep reinforcement learning. The project includes specialized guides and frameworks for imitation learning based on expert demonstrations, model-based reinforcement learning using planners, and the training of recurrent neural networks to solve partially observed environments. The materials cover a broad ran

    Jupyter Notebookcourse-materialsdeep-learningdeep-reinforcement-learning
    GitHub पर देखें↗6,522
Reinforcement Learning के सभी 30 विकल्प देखें→