awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoServidor MCPAcerca deCómo clasificamosPrensa
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 repositorios

Awesome GitHub RepositoriesAction Output Models

Processes multimodal inputs including natural language and camera images to generate motor commands for generalized robot manipulation.

Distinct from Vision-Language Models: Distinct from Vision-Language Models: extends vision-language processing to generate motor commands (action output), not just visual and linguistic understanding.

Explore 2 awesome GitHub repositories matching artificial intelligence & ml · Action Output Models. Refine with filters or upvote what's useful.

  1. Home
  2. Artificial Intelligence & ML
  3. Machine Learning
  4. Multimodal Processing Tools
  5. Vision-Language Models
  6. Action Output Models

Awesome Action Output Models GitHub Repositories

Encuentra los mejores repositorios con IA.Buscaremos los repositorios que mejor coincidan usando IA.
  • nvidia/isaac-gr00tAvatar de NVIDIA

    NVIDIA/Isaac-GR00T

    6,222Ver en GitHub↗

    Processes multimodal inputs including natural language and camera images to generate motor commands for generalized robot manipulation.

    Jupyter Notebook
    Ver en GitHub↗6,222
  • openvla/openvlaAvatar de openvla

    openvla/openvla

    5,305Ver en GitHub↗

    OpenVLA is a vision-language-action model and framework designed for general-purpose robotic manipulation. It provides a robotic policy training framework and a control inference engine that map visual and textual inputs to robotic control actions, enabling zero-shot instruction following on hardware. The project includes a robotics dataset pipeline for standardizing diverse trajectory data and managing dataset mixtures. It supports large-scale model training through distributed GPU compute and sharded data parallelism, alongside parameter-efficient adaptation for fine-tuning models to new ta

    Maps visual and textual inputs to tokenized action sequences for general-purpose robotic manipulation.

    Python
    Ver en GitHub↗5,305