How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
DARS: Dynamic Action Re-Sampling to Enhance Coding Agent Performance by Adaptive Tree Traversal
The main features of vaibhavagg303/dars-agent are: Policy Optimization.
Projects with overlapping indexed features include: trademaster-ntu/trademaster — TradeMaster is a reinforcement learning trading framework and algorithmic trading simulator designed for designing and… chenxinan-fdu/polaris — 🌠 A PO st-training recipe for scaling R L on A dvanced R eason I ng model S 🚀. liaomengqi/e3-rl4llms — [25/08/20] : Aceept as EMNLP 2025 Main Conference paper. multimodal-art-projection/treepo — This is the official implementation of TreePO algorithm. netease-youdao/confucius3-math — 💜 Confucius Demo | 🤗 Hugging Face | 🤖 ModelScope |… chanliang/eepo.
TradeMaster is a reinforcement learning trading framework and algorithmic trading simulator designed for designing and testing quantitative trading strategies. The system provides a platform for developing reinforcement learning agents, managing quantitative portfolios, and optimizing trade execution using financial market data. The project features specialized components for multi-modality data preprocessing, a high-fidelity market environment simulation for strategy backtesting, and a quantitative portfolio manager for capital reallocation across multiple assets. It includes a trade executi
🌠 A PO st-training recipe for scaling R L on A dvanced R eason I ng model S 🚀
25/08/20 : Aceept as EMNLP 2025 Main Conference paper