awesome-repositories.com
Blog
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetÀ proposNotre méthodologiePresseServeur MCP
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to microsoft/vitra

Open-source alternatives to VITRA

30 open-source projects similar to microsoft/vitra, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best VITRA alternative.

  • agibottech/genie-envisionerAvatar de AgibotTech

    AgibotTech/Genie-Envisioner

    550Voir sur GitHub↗

    Join our WeChat Group

    Python
    Voir sur GitHub↗550
  • alibaba-damo-academy/worldvlaAvatar de alibaba-damo-academy

    alibaba-damo-academy/WorldVLA

    1,074Voir sur GitHub↗

    RynnVLA-002: A Unified Vision-Language-Action and World Model

    Python
    Voir sur GitHub↗1,074
  • amap-cvlab/abot-manipulationAvatar de amap-cvlab

    amap-cvlab/ABot-Manipulation

    570Voir sur GitHub↗

    ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

    Python
    Voir sur GitHub↗570
  • arashakb/actquantAvatar de arashakb

    arashakb/ActQuant

    5Voir sur GitHub↗

    Official implementation of ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models

    C++
    Voir sur GitHub↗5
  • beingbeyond/being-hAvatar de BeingBeyond

    BeingBeyond/Being-H

    1,012Voir sur GitHub↗

    Being-H is BeingBeyond's family of human-centric embodied foundation models.

    Python
    Voir sur GitHub↗1,012
  • beingbeyond/being-h0Avatar de BeingBeyond

    BeingBeyond/Being-H0

    48Voir sur GitHub↗

    Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos (ICML 2026)

    Python
    Voir sur GitHub↗48

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Find more with AI search
  • buoyancy99/large-video-plannerAvatar de buoyancy99

    buoyancy99/large-video-planner

    250Voir sur GitHub↗

    This repo provides training and inference code for the paper "Large Video Planner Enables Generalizable Robot Control"

    Python
    Voir sur GitHub↗250
  • chowzy069/reconvlaAvatar de Chowzy069

    Chowzy069/Reconvla

    261Voir sur GitHub↗

    Official implementation of ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver.

    Python
    Voir sur GitHub↗261
  • cladernyjorn/vlm4vlaAvatar de CladernyJorn

    CladernyJorn/VLM4VLA

    157Voir sur GitHub↗

    Implementation of VLM4VLA

    Python
    Voir sur GitHub↗157
  • declare-lab/nora-1.5Avatar de declare-lab

    declare-lab/nora-1.5

    106Voir sur GitHub↗

    NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards

    Python
    Voir sur GitHub↗106
  • dexmal/dexboticAvatar de Dexmal

    Dexmal/dexbotic

    1,219Voir sur GitHub↗

    Dexbotic: Open-Source Vision-Language-Action Toolbox

    Python
    Voir sur GitHub↗1,219
  • dreamzero0/dreamzeroAvatar de dreamzero0

    dreamzero0/dreamzero

    752Voir sur GitHub↗
    Python
    Voir sur GitHub↗752
  • eo-robotics/eo-1E

    eo-robotics/EO-1

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • facebookresearch/vjepa2Avatar de facebookresearch

    facebookresearch/vjepa2

    3,021Voir sur GitHub↗

    vjepa2 is a joint-embedding predictive architecture and video self-supervised learning framework. It functions as a visual representation learner and a robotic manipulation model designed to learn representations by predicting future latent states without reconstructing pixels. The system enables the pretraining of video encoders that learn temporally consistent features through masked-token prediction and multi-modal tokenization. It further maps these latent embeddings to specific physical movements via action-conditioned post-training to plan and execute robot arm grasping and picking task

    Python
    Voir sur GitHub↗3,021
  • internrobotics/f1-vlaAvatar de InternRobotics

    InternRobotics/F1-VLA

    201Voir sur GitHub↗

    F1: A Vision Language Action Model Bridging Understanding and Generation to Actions

    Python
    Voir sur GitHub↗201
  • internrobotics/internvla-a1Avatar de InternRobotics

    InternRobotics/InternVLA-A1

    414Voir sur GitHub↗

    InternVLA-A1: Unifying Understanding, Generation, and Action for Robotic Manipulation​

    Python
    Voir sur GitHub↗414
  • internrobotics/internvla-m1Avatar de InternRobotics

    InternRobotics/InternVLA-M1

    416Voir sur GitHub↗

    InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy

    Python
    Voir sur GitHub↗416
  • jiutian-vl/cogvlaAvatar de JiuTian-VL

    JiuTian-VL/CogVLA

    186Voir sur GitHub↗

    NeurIPS 2025 CogVLA: Cognition-Aligned Vision-Language-Action Models via Instruction-Driven Routing & Sparsification

    Python
    Voir sur GitHub↗186
  • jxbi1010/vla-touchAvatar de jxbi1010

    jxbi1010/VLA-Touch

    79Voir sur GitHub↗

    Implementation of RA-L (2026) paper: VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback

    Python
    Voir sur GitHub↗79
  • kahnchana/langtomoAvatar de kahnchana

    kahnchana/LangToMo

    21Voir sur GitHub↗

    WIP Code for LangToMo

    Python
    Voir sur GitHub↗21
  • lukelin-web/voteAvatar de LukeLIN-web

    LukeLIN-web/VOTE

    26Voir sur GitHub↗

    Vision-Language-Action Optimization with Trajectory Ensemble Voting

    Python
    Voir sur GitHub↗26
  • microsoft/villa-xAvatar de microsoft

    microsoft/villa-x

    204Voir sur GitHub↗

    This is the official repository for villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models.

    Python
    Voir sur GitHub↗204
  • mint-sjtu/evo-vlaAvatar de MINT-SJTU

    MINT-SJTU/Evo-VLA

    55Voir sur GitHub↗

    Evo-0: Vision-Language-Action Model with Implicit Spatial Understanding.

    Voir sur GitHub↗55
  • nvidia/isaac-gr00tAvatar de NVIDIA

    NVIDIA/Isaac-GR00T

    6,222Voir sur GitHub↗
    Jupyter Notebook
    Voir sur GitHub↗6,222
  • nvlabs/vla0Avatar de NVlabs

    NVlabs/vla0

    483Voir sur GitHub↗

    VLA-0: Building State-of-the-Art VLAs with Zero Modification

    Python
    Voir sur GitHub↗483
  • open-gigaai/giga-brain-0Avatar de open-gigaai

    open-gigaai/giga-brain-0

    2,542Voir sur GitHub↗

    giga-brain-0 is a robot action model framework designed to train and deploy neural networks that map multi-modal sensor data to physical robot control signals. It functions as a robot manipulation controller that processes high-dimensional observations to execute dexterous, long-horizon physical tasks. The project provides a multi-modal robot inference server using a client-server architecture to stream real-time vision and language observations for instant action prediction. It includes an embodiment fine-tuning pipeline to adapt pre-trained base models to specific robot hardware configurati

    Python
    Voir sur GitHub↗2,542
  • opendrivelab/univlaO

    OpenDriveLab/UniVLA

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • opengalaxea/g0Avatar de OpenGalaxea

    OpenGalaxea/G0

    0Voir sur GitHub↗

    placeholder for original G0 webpage

    HTML
    Voir sur GitHub↗0
  • openhelix-team/llava-vlaAvatar de OpenHelix-Team

    OpenHelix-Team/LLaVA-VLA

    196Voir sur GitHub↗

    LLaVA-VLA: A Simple Yet Powerful Vision-Language-Action Model ICRA 2026

    Python
    Voir sur GitHub↗196
  • openhelix-team/spatial-forcingAvatar de OpenHelix-Team

    OpenHelix-Team/Spatial-Forcing

    249Voir sur GitHub↗

    Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model ICLR2026

    Python
    Voir sur GitHub↗249