awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
FreedomIntelligence avatar

FreedomIntelligence/MLLM-Bench

0
View on GitHub↗
76 星标·4 分支·Python·8 次浏览

MLLM Bench

MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria

Features

  • Evaluation Benchmarks - Evaluating multimodal models using automated scoring.
  • Multimodal Benchmarks - Evaluation framework using GPT-4V with per-sample criteria.

Star 历史

freedomintelligence/mllm-bench 的 Star 历史图表freedomintelligence/mllm-bench 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

MLLM Bench 的开源替代方案

相似的开源项目,按与 MLLM Bench 的功能重合度排序。
  • bradyfu/video-mmeBradyFU 的头像

    BradyFU/Video-MME

    779在 GitHub 上查看↗

    ✨✨CVPR 2025 Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    在 GitHub 上查看↗779
  • damo-nlp-sg/m3examDAMO-NLP-SG 的头像

    DAMO-NLP-SG/M3Exam

    105在 GitHub 上查看↗

    Data and code for paper "M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models"

    Pythonai-educationchatgptevaluation
    在 GitHub 上查看↗105
  • ailab-cvc/seed-benchAILab-CVC 的头像

    AILab-CVC/SEED-Bench

    364在 GitHub 上查看↗

    (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions.

    Python
    在 GitHub 上查看↗364
  • gzcch/bingogzcch 的头像

    gzcch/Bingo

    55在 GitHub 上查看↗

    Chenhang Cui, Yiyang Zhou, Xinyu Yang, Shirley Wu, Linjun Zhang, James Zou, Huaxiu Yao *Equal Contribution

    在 GitHub 上查看↗55
查看 MLLM Bench 的所有 30 个替代方案→

常见问题解答

freedomintelligence/mllm-bench 是做什么的?

MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria

freedomintelligence/mllm-bench 的主要功能有哪些?

freedomintelligence/mllm-bench 的主要功能包括:Evaluation Benchmarks, Multimodal Benchmarks。

freedomintelligence/mllm-bench 有哪些开源替代品?

freedomintelligence/mllm-bench 的开源替代品包括: damo-nlp-sg/m3exam — Data and code for paper "M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models". hypjudy/sparkles — Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models. bradyfu/video-mme — ✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis. ailab-cvc/seed-bench — (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions. gzcch/bingo — Chenhang Cui, Yiyang Zhou, Xinyu Yang, Shirley Wu, Linjun Zhang, James Zou, Huaxiu Yao *Equal Contribution. lerogo/mmgenbench — Official repository of MMGenBench.