awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Back to cmmmu-benchmark/cmmmu

Projects sharing features with CMMMU

30 open-source projects similar to cmmmu-benchmark/cmmmu, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.

  • open-compass/vlmevalkitopen-compass avatar

    open-compass/VLMEvalKit

    3,824View on GitHub↗

    VLMEvalKit is a vision-language model evaluation framework and inference engine designed to run standardized benchmarks and measure model accuracy across diverse visual datasets. It serves as a multimodal model benchmark and performance toolkit for calculating metrics and comparing model responses. The toolkit includes a specialized visual reasoning evaluator that uses adversarial samples to distinguish actual image understanding from reliance on language patterns. It also provides capabilities for image generation evaluation, testing a model's ability to create or modify visuals based on tex

    Pythonchatgptclaudeclip
    View on GitHub↗3,824
  • alenai97/micevalalenai97 avatar

    alenai97/MiCEval

    6View on GitHub↗

    An automatic evaluation framework for Multimodal Chain-of-Thought.

    Python
    View on GitHub↗6
  • bradyfu/video-mmeBradyFU avatar

    BradyFU/Video-MME

    779View on GitHub↗

    ✨✨CVPR 2025 Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    View on GitHub↗779
  • bytedance/lynx-llmbytedance avatar

    bytedance/lynx-llm

    272View on GitHub↗

    paper: https://arxiv.org/abs/2307.02469 page: https://lynx-llm.github.io/

    Pythonresearch
    View on GitHub↗272
  • damo-nlp-sg/m3examDAMO-NLP-SG avatar

    DAMO-NLP-SG/M3Exam

    105View on GitHub↗

    Data and code for paper "M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models"

    Pythonai-educationchatgptevaluation
    View on GitHub↗105

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • dcdmllm/cheetahDCDmllm avatar

    DCDmllm/Cheetah

    354View on GitHub↗

    Fine-tuning Multimodal LLMs to Follow Zero-shot Demonstrative Instructions

    Python
    View on GitHub↗354
  • freedomintelligence/mllm-benchFreedomIntelligence avatar

    FreedomIntelligence/MLLM-Bench

    76View on GitHub↗

    MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria

    Python
    View on GitHub↗76
  • fuxiaoliu/lrv-instructionFuxiaoLiu avatar

    FuxiaoLiu/LRV-Instruction

    297View on GitHub↗

    ICLR'24 Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

    Pythonchatgptevaluationevaluation-metrics
    View on GitHub↗297
  • fuxiaoliu/mmcFuxiaoLiu avatar

    FuxiaoLiu/MMC

    95View on GitHub↗

    NAACL 2024 MMC: Advancing Multimodal Chart Understanding with LLM Instruction Tuning

    Pythonarxivbenchmarkchart
    View on GitHub↗95
  • gzcch/bingogzcch avatar

    gzcch/Bingo

    55View on GitHub↗

    Chenhang Cui, Yiyang Zhou, Xinyu Yang, Shirley Wu, Linjun Zhang, James Zou, Huaxiu Yao *Equal Contribution

    View on GitHub↗55
  • hypjudy/sparklesHYPJUDY avatar

    HYPJUDY/Sparkles

    45View on GitHub↗

    Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models

    Python
    View on GitHub↗45
  • inst-it/inst-itinst-it avatar

    inst-it/inst-it

    40View on GitHub↗

    NeurIPS 2025 The official repository of "Inst-IT: Boosting Multimodal Instance Understanding via Explicit Visual Prompt Instruction Tuning"

    Pythoninstruction-tuninglarge-multimodal-modelsmultimodal
    View on GitHub↗40
  • katha-ai/velocitikatha-ai avatar

    katha-ai/VELOCITI

    8View on GitHub↗

    VELOCITI Benchmark Evaluation and Visualisation Code

    Pythonartificial-intelligenceawesome-listbenchmark
    View on GitHub↗8
  • lerogo/mmgenbenchlerogo avatar

    lerogo/MMGenBench

    119View on GitHub↗

    Official repository of MMGenBench

    Pythonllms-benchmarkingmllmmmgenbench
    View on GitHub↗119
  • lightchen233/m3cotLightChen233 avatar

    LightChen233/M3CoT

    91View on GitHub↗

    @Author: Qiguang Chen @LastEditors: Qiguang Chen @Date: 2024-05-23 20:24:16 @LastEditTime: 2024-05-26 18:09:00 @Description: -->

    Python
    View on GitHub↗91
  • llyx97/tempcompassllyx97 avatar

    llyx97/TempCompass

    132View on GitHub↗

    ACL 2024 Findings "TempCompass: Do Video LLMs Really Understand Videos?", Yuanxin Liu, Shicheng Li, Yi Liu, Yuxiang Wang, Shuhuai Ren, Lei Li, Sishuo Chen, Xu Sun, Lu Hou

    Pythonevaluationtemporal-perceptionvideo-llms
    View on GitHub↗132
  • mathvision-cuhk/mathvisionmathvision-cuhk avatar

    mathvision-cuhk/MathVision

    139View on GitHub↗

    NeurIPS 2024 MATH-Vision dataset and code to measure multimodal mathematical reasoning capabilities.

    Python
    View on GitHub↗139
  • ncsoft/idkncsoft avatar

    ncsoft/idk

    6View on GitHub↗

    Official implementation of "Visually Dehallucinative Instruction Generation: Know What You Don't Know"

    View on GitHub↗6
  • open-compass/mmbenchopen-compass avatar

    open-compass/MMBench

    303View on GitHub↗

    Official Repo of "MMBench: Is Your Multi-modal Model an All-around Player?"

    View on GitHub↗303
  • opengvlab/ask-anythingOpenGVLab avatar

    OpenGVLab/Ask-Anything

    3,341View on GitHub↗

    CVPR2024 HighlightVideoChatGPT ChatGPT with video understanding! And many more supported LMs such as miniGPT4, StableLM, and MOSS.

    Pythonbig-modelcaptioning-videoschat
    View on GitHub↗3,341
  • opengvlab/multi-modality-arenaOpenGVLab avatar

    OpenGVLab/Multi-Modality-Arena

    558View on GitHub↗
    Pythonchatchatbotchatgpt
    View on GitHub↗558
  • openlamm/lammOpenLAMM avatar

    OpenLAMM/LAMM

    317View on GitHub↗

    NeurIPS 2023 Datasets and Benchmarks Track LAMM: Multi-Modal Large Language Models and Applications as AI Agents

    Python
    View on GitHub↗317
  • openm3d/m3dbenchOpenM3D avatar

    OpenM3D/M3DBench

    61View on GitHub↗

    ECCV 2024 M3DBench introduces a comprehensive 3D instruction-following dataset with support for interleaved multi-modal prompts.

    Python3ddatasetinstruction-tuning
    View on GitHub↗61
  • pku-yuangroup/video-benchPKU-YuanGroup avatar

    PKU-YuanGroup/Video-Bench

    140View on GitHub↗

    A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models!

    Pythonbenchmarklarge-language-modelstoolkit
    View on GitHub↗140
  • sail-sg/mmcbenchsail-sg avatar

    sail-sg/MMCBench

    27View on GitHub↗

    Code for the paper Benchmarking Large Multimodal Models against Common Corruptions.

    Python
    View on GitHub↗27
  • tsb0601/mmvptsb0601 avatar

    tsb0601/MMVP

    363View on GitHub↗

    Shengbang Tong, Zhuang Liu, Yuexiang Zhai, Yi Ma, Yann LeCun, Saining Xie

    Python
    View on GitHub↗363
  • wusiwei0410/scimmirWusiwei0410 avatar

    Wusiwei0410/SciMMIR

    25View on GitHub↗

    This is the repo for the paper SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval.

    Python
    View on GitHub↗25
  • x-plug/mplug-owlX-PLUG avatar

    X-PLUG/mPLUG-Owl

    2,542View on GitHub↗

    mPLUG-Owl: The Powerful Multi-modal Large Language Model Family

    Pythonalpacachatbotchatgpt
    View on GitHub↗2,542
  • ys-zong/vl-iclys-zong avatar

    ys-zong/VL-ICL

    69View on GitHub↗

    ICLR 2025 VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

    Python
    View on GitHub↗69
  • ailab-cvc/seed-benchAILab-CVC avatar

    AILab-CVC/SEED-Bench

    364View on GitHub↗

    (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions.

    Python
    View on GitHub↗364