awesome-repositories.com
Blog
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectAboutHow we rankPressMCP server
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to microsoft/vitra

Open-source alternatives to VITRA

30 open-source projects similar to microsoft/vitra, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best VITRA alternative.

  • agibottech/genie-envisionerAgibotTech avatar

    AgibotTech/Genie-Envisioner

    550View on GitHub↗

    Join our WeChat Group

    Python
    View on GitHub↗550
  • alibaba-damo-academy/worldvlaalibaba-damo-academy avatar

    alibaba-damo-academy/WorldVLA

    1,074View on GitHub↗

    RynnVLA-002: A Unified Vision-Language-Action and World Model

    Python
    View on GitHub↗1,074
  • amap-cvlab/abot-manipulationamap-cvlab avatar

    amap-cvlab/ABot-Manipulation

    570View on GitHub↗

    ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

    Python
    View on GitHub↗570
  • arashakb/actquantarashakb avatar

    arashakb/ActQuant

    5View on GitHub↗

    Official implementation of ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models

    C++
    View on GitHub↗5
  • beingbeyond/being-hBeingBeyond avatar

    BeingBeyond/Being-H

    1,012View on GitHub↗

    Being-H is BeingBeyond's family of human-centric embodied foundation models.

    Python
    View on GitHub↗1,012
  • beingbeyond/being-h0BeingBeyond avatar

    BeingBeyond/Being-H0

    48View on GitHub↗

    Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos (ICML 2026)

    Python
    View on GitHub↗48

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • buoyancy99/large-video-plannerbuoyancy99 avatar

    buoyancy99/large-video-planner

    250View on GitHub↗

    This repo provides training and inference code for the paper "Large Video Planner Enables Generalizable Robot Control"

    Python
    View on GitHub↗250
  • chowzy069/reconvlaChowzy069 avatar

    Chowzy069/Reconvla

    261View on GitHub↗

    Official implementation of ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver.

    Python
    View on GitHub↗261
  • cladernyjorn/vlm4vlaCladernyJorn avatar

    CladernyJorn/VLM4VLA

    157View on GitHub↗

    Implementation of VLM4VLA

    Python
    View on GitHub↗157
  • declare-lab/nora-1.5declare-lab avatar

    declare-lab/nora-1.5

    106View on GitHub↗

    NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards

    Python
    View on GitHub↗106
  • dexmal/dexboticDexmal avatar

    Dexmal/dexbotic

    1,219View on GitHub↗

    Dexbotic: Open-Source Vision-Language-Action Toolbox

    Python
    View on GitHub↗1,219
  • dreamzero0/dreamzerodreamzero0 avatar

    dreamzero0/dreamzero

    752View on GitHub↗
    Python
    View on GitHub↗752
  • eo-robotics/eo-1E

    eo-robotics/EO-1

    0View on GitHub↗
    View on GitHub↗0
  • facebookresearch/vjepa2facebookresearch avatar

    facebookresearch/vjepa2

    3,021View on GitHub↗

    vjepa2 is a joint-embedding predictive architecture and video self-supervised learning framework. It functions as a visual representation learner and a robotic manipulation model designed to learn representations by predicting future latent states without reconstructing pixels. The system enables the pretraining of video encoders that learn temporally consistent features through masked-token prediction and multi-modal tokenization. It further maps these latent embeddings to specific physical movements via action-conditioned post-training to plan and execute robot arm grasping and picking task

    Python
    View on GitHub↗3,021
  • internrobotics/f1-vlaInternRobotics avatar

    InternRobotics/F1-VLA

    201View on GitHub↗

    F1: A Vision Language Action Model Bridging Understanding and Generation to Actions

    Python
    View on GitHub↗201
  • internrobotics/internvla-a1InternRobotics avatar

    InternRobotics/InternVLA-A1

    414View on GitHub↗

    InternVLA-A1: Unifying Understanding, Generation, and Action for Robotic Manipulation​

    Python
    View on GitHub↗414
  • internrobotics/internvla-m1InternRobotics avatar

    InternRobotics/InternVLA-M1

    416View on GitHub↗

    InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy

    Python
    View on GitHub↗416
  • jiutian-vl/cogvlaJiuTian-VL avatar

    JiuTian-VL/CogVLA

    186View on GitHub↗

    NeurIPS 2025 CogVLA: Cognition-Aligned Vision-Language-Action Models via Instruction-Driven Routing & Sparsification

    Python
    View on GitHub↗186
  • jxbi1010/vla-touchjxbi1010 avatar

    jxbi1010/VLA-Touch

    79View on GitHub↗

    Implementation of RA-L (2026) paper: VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback

    Python
    View on GitHub↗79
  • kahnchana/langtomokahnchana avatar

    kahnchana/LangToMo

    21View on GitHub↗

    WIP Code for LangToMo

    Python
    View on GitHub↗21
  • lukelin-web/voteLukeLIN-web avatar

    LukeLIN-web/VOTE

    26View on GitHub↗

    Vision-Language-Action Optimization with Trajectory Ensemble Voting

    Python
    View on GitHub↗26
  • microsoft/villa-xmicrosoft avatar

    microsoft/villa-x

    204View on GitHub↗

    This is the official repository for villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models.

    Python
    View on GitHub↗204
  • mint-sjtu/evo-vlaMINT-SJTU avatar

    MINT-SJTU/Evo-VLA

    55View on GitHub↗

    Evo-0: Vision-Language-Action Model with Implicit Spatial Understanding.

    View on GitHub↗55
  • nvidia/isaac-gr00tNVIDIA avatar

    NVIDIA/Isaac-GR00T

    6,222View on GitHub↗
    Jupyter Notebook
    View on GitHub↗6,222
  • nvlabs/vla0NVlabs avatar

    NVlabs/vla0

    483View on GitHub↗

    VLA-0: Building State-of-the-Art VLAs with Zero Modification

    Python
    View on GitHub↗483
  • open-gigaai/giga-brain-0open-gigaai avatar

    open-gigaai/giga-brain-0

    2,542View on GitHub↗

    giga-brain-0 is a robot action model framework designed to train and deploy neural networks that map multi-modal sensor data to physical robot control signals. It functions as a robot manipulation controller that processes high-dimensional observations to execute dexterous, long-horizon physical tasks. The project provides a multi-modal robot inference server using a client-server architecture to stream real-time vision and language observations for instant action prediction. It includes an embodiment fine-tuning pipeline to adapt pre-trained base models to specific robot hardware configurati

    Python
    View on GitHub↗2,542
  • opendrivelab/univlaO

    OpenDriveLab/UniVLA

    0View on GitHub↗
    View on GitHub↗0
  • opengalaxea/g0OpenGalaxea avatar

    OpenGalaxea/G0

    0View on GitHub↗

    placeholder for original G0 webpage

    HTML
    View on GitHub↗0
  • openhelix-team/llava-vlaOpenHelix-Team avatar

    OpenHelix-Team/LLaVA-VLA

    196View on GitHub↗

    LLaVA-VLA: A Simple Yet Powerful Vision-Language-Action Model ICRA 2026

    Python
    View on GitHub↗196
  • openhelix-team/spatial-forcingOpenHelix-Team avatar

    OpenHelix-Team/Spatial-Forcing

    249View on GitHub↗

    Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model ICLR2026

    Python
    View on GitHub↗249