7 open-source projects similar to dvlab-research/lisa, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.
Official Repo For OMG-LLaVA and OMG-Seg codebase CVPR-24 and NeurIPS-24
Large Language-and-Vision Assistant for Biomedicine, built towards multimodal GPT-4 level capabilities.
ECCV 2024 Best Paper Candidate & TPAMI 2025 PointLLM: Empowering Large Language Models to Understand Point Clouds
The official implement of VITA, VITA15, LongVITA, VITA-Audio, VITA-VLA, and VITA-E.
A Vision-Language Model for Spatial Affordance Prediction in Robotics
Monkey (LMM): Image Resolution and Text Label Are Important Things for Large Multi-modal Models (CVPR 2024 Highlight)
AAAI 2025 DocKylin: A Large Multimodal Model for Visual Document Understanding with Efficient Visual Slimming