How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
RynnVLA-002: A Unified Vision-Language-Action and World Model
ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning
Official implementation of ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models
[ICRA 2026] VITRA: Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos
The main features of microsoft/vitra are: Embodied Foundation Models.
Projects with overlapping indexed features include: alibaba-damo-academy/worldvla — RynnVLA-002: A Unified Vision-Language-Action and World Model. amap-cvlab/abot-manipulation — ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning. arashakb/actquant — Official implementation of ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models. beingbeyond/being-h — Being-H is BeingBeyond's family of human-centric embodied foundation models. beingbeyond/being-h0 — Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos (ICML 2026). agibottech/genie-envisioner — Join our WeChat Group.