This is a pytorch implementation of SfBC: Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling.
Thanks for checking out the Best-README-Template. If you have a suggestion that would make this better, please fork the repo and create a pull request or simply open an issue with the tag "enhancement". Don't forget to give the project a star! Thanks again! Now go create something AMAZING!…
In this work, we propose Diffusion-QL which utilizes a diffusion model as a highly expressive policy class for behavior cloning and policy regularization. In our approach we learn an action-value function and we add a term maximising action-values to the the training loss of the diffusion model,…
The main features of twitter/diffusion-rl are: Diffusion Reinforcement Learning.
Open-source alternatives to twitter/diffusion-rl include: aaronanima/targf — [Website] [Arxiv]. anuragajay/decision-diffuser. chendrag/sfbc — This is a pytorch implementation of SfBC: Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling. cleandiffuserteam/cleandiffuser — Thanks for checking out the Best-README-Template. If you have a suggestion that would make this better, please fork… jannerm/diffuser — Code for the paper "Planning with Diffusion for Flexible Behavior Synthesis". opendilab/generativerl — English | 简体中文(Simplified Chinese).