How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.
Medical Multimodal LLMs
The repository provides a collection of vision language models, benchmarks, and related applications, released as part of Project MONAI (Medical Open Network for Artificial Intelligence).
【ICML 2025 Spotlight】 Official Repo for Paper ‘’HealthGPT : A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation‘’
This is the code repo for the Med-Flamingo paper.
Large Language-and-Vision Assistant for Biomedicine, built towards multimodal GPT-4 level capabilities.
The main features of microsoft/llava-med are: Medical Multimodal Models, Medical Vision Language Models, Specialized Multimodal Tasks, Pre-training Datasets.
Open-source alternatives to microsoft/llava-med include: taokz/biomedgpt — BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks. snap-stanford/med-flamingo — This is the code repo for the Med-Flamingo paper. dcdmllm/healthgpt — 【ICML 2025 Spotlight】 Official Repo for Paper ‘’HealthGPT : A Medical Large Vision-Language Model for Unifying… freedomintelligence/huatuogpt-vision — Medical Multimodal LLMs. project-monai/vlm-radiology-agent-framework — The repository provides a collection of vision language models, benchmarks, and related applications, released as part… vision-cair/minigpt-med — Open-sourced code of MiniGPT-Med.