awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

3 مستودعات

Awesome GitHub RepositoriesVision-Language-Action Controllers

A platform that combines vision-language-action models with whole-body controllers to generate coordinated joint commands for humanoid robots from language and image inputs.

Distinct from Humanoid Coordination: Distinct from Humanoid Coordination: focuses on vision-language-action models for control rather than general whole-body coordination systems.

Explore 3 awesome GitHub repositories matching hardware & iot · Vision-Language-Action Controllers. Refine with filters or upvote what's useful.

Awesome Vision-Language-Action Controllers GitHub Repositories

اعثر على أفضل المستودعات باستخدام الذكاء الاصطناعي.سنبحث عن أفضل المستودعات المطابقة باستخدام الذكاء الاصطناعي.
  • nvidia/isaac-gr00tالصورة الرمزية لـ NVIDIA

    NVIDIA/Isaac-GR00T

    6,222عرض على GitHub↗

    Combines vision-language-action models with whole-body controllers to generate coordinated joint commands for humanoid robots.

    Jupyter Notebook
    عرض على GitHub↗6,222
  • openvla/openvlaالصورة الرمزية لـ openvla

    openvla/openvla

    5,305عرض على GitHub↗

    OpenVLA is a vision-language-action model and framework designed for general-purpose robotic manipulation. It provides a robotic policy training framework and a control inference engine that map visual and textual inputs to robotic control actions, enabling zero-shot instruction following on hardware. The project includes a robotics dataset pipeline for standardizing diverse trajectory data and managing dataset mixtures. It supports large-scale model training through distributed GPU compute and sharded data parallelism, alongside parameter-efficient adaptation for fine-tuning models to new ta

    Generates precise robotic control actions by processing combined visual and textual inputs through a trained VLA model.

    Python
    عرض على GitHub↗5,305
  • rlinf/rlinfالصورة الرمزية لـ RLinf

    RLinf/RLinf

    2,502عرض على GitHub↗

    RLinf is a distributed reinforcement learning orchestrator and embodied AI training framework. It provides the infrastructure to train vision-language-action models and robotic policies using a combination of reinforcement learning and supervised fine-tuning. The system is designed for scaling workloads across GPU clusters, managing the placement of actors, rollout workers, and environment components. It features a specialized robotics data collection pipeline for gathering teleoperated demonstrations and simulation trajectories into standardized replay buffers, alongside a hardware interface

    Maps visual observations and language prompts to continuous robotic control commands using VLM backbones.

    Pythonagentic-aiembodied-aireinforcement-learning
    عرض على GitHub↗2,502
  1. Home
  2. Hardware & IoT
  3. Embedded Systems And Robotics
  4. Robotics And Autonomous Systems
  5. Robotics & Drones
  6. Robotics and Control
  7. Humanoid Coordination
  8. Vision-Language-Action Controllers

استكشف الوسوم الفرعية

  • Manipulation Skill TrainingsUses vision-language-action models to enable humanoid robots to perform tasks from language and image inputs. **Distinct from Vision-Language-Action Controllers:** Distinct from Vision-Language-Action Controllers: focuses on training manipulation skills, not general control.