How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.
Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models
Official implementation of "Visually Dehallucinative Instruction Generation: Know What You Don't Know"
ICLR'24 Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning
NeurIPS 2023 Datasets and Benchmarks Track LAMM: Multi-Modal Large Language Models and Applications as AI Agents
[NeurIPS 2025] The official repository of "Inst-IT: Boosting Multimodal Instance Understanding via Explicit Visual Prompt Instruction Tuning"
The main features of inst-it/inst-it are: Multimodal Benchmarks, Pre-training Datasets.
Open-source alternatives to inst-it/inst-it include: ncsoft/idk — Official implementation of "Visually Dehallucinative Instruction Generation: Know What You Don't Know". openm3d/m3dbench — [ECCV 2024] M3DBench introduces a comprehensive 3D instruction-following dataset with support for interleaved… fuxiaoliu/lrv-instruction — [ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning. hypjudy/sparkles — Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models. openlamm/lamm — [NeurIPS 2023 Datasets and Benchmarks Track] LAMM: Multi-Modal Large Language Models and Applications as AI Agents. open-compass/vlmevalkit — VLMEvalKit is a vision-language model evaluation framework and inference engine designed to run standardized…