How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
Medical Multimodal LLMs
The repository provides a collection of vision language models, benchmarks, and related applications, released as part of Project MONAI (Medical Open Network for Artificial Intelligence).
【ICML 2025 Spotlight】 Official Repo for Paper ‘’HealthGPT : A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation‘’
This is the code repo for the Med-Flamingo paper.
Large Language-and-Vision Assistant for Biomedicine, built towards multimodal GPT-4 level capabilities.
The main features of microsoft/llava-med are: Medical Multimodal Models, Medical Vision Language Models, Specialized Multimodal Tasks, Pre-training Datasets.
Projects with overlapping indexed features include: taokz/biomedgpt — BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks. snap-stanford/med-flamingo — This is the code repo for the Med-Flamingo paper. dcdmllm/healthgpt — 【ICML 2025 Spotlight】 Official Repo for Paper ‘’HealthGPT : A Medical Large Vision-Language Model for Unifying… freedomintelligence/huatuogpt-vision — Medical Multimodal LLMs. project-monai/vlm-radiology-agent-framework — The repository provides a collection of vision language models, benchmarks, and related applications, released as part… vision-cair/minigpt-med — Open-sourced code of MiniGPT-Med.