awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
gzcch avatar

gzcch/Bingo

0
View on GitHub↗
55 星标·2 分支·4 次浏览

Bingo

Chenhang Cui, Yiyang Zhou, Xinyu Yang, Shirley Wu, Linjun Zhang, James Zou, Huaxiu Yao *Equal Contribution

Features

  • Evaluation Benchmarks - Analyzes bias and interference challenges in vision models.
  • Multimodal Benchmarks - Benchmark for evaluating hallucination types.

Star 历史

gzcch/bingo 的 Star 历史图表gzcch/bingo 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

Bingo 的开源替代方案

相似的开源项目,按与 Bingo 的功能重合度排序。
  • bradyfu/video-mmeBradyFU 的头像

    BradyFU/Video-MME

    779在 GitHub 上查看↗

    ✨✨CVPR 2025 Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    在 GitHub 上查看↗779
  • damo-nlp-sg/m3examDAMO-NLP-SG 的头像

    DAMO-NLP-SG/M3Exam

    105在 GitHub 上查看↗

    Data and code for paper "M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models"

    Pythonai-educationchatgptevaluation
    在 GitHub 上查看↗105
  • ailab-cvc/seed-benchAILab-CVC 的头像

    AILab-CVC/SEED-Bench

    364在 GitHub 上查看↗

    (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions.

    Python
    在 GitHub 上查看↗364
  • freedomintelligence/mllm-benchFreedomIntelligence 的头像

    FreedomIntelligence/MLLM-Bench

    76在 GitHub 上查看↗

    MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria

    Python
    在 GitHub 上查看↗76
查看 Bingo 的所有 30 个替代方案→

常见问题解答

gzcch/bingo 是做什么的?

Chenhang Cui, Yiyang Zhou, Xinyu Yang, Shirley Wu, Linjun Zhang, James Zou, Huaxiu Yao *Equal Contribution

gzcch/bingo 的主要功能有哪些?

gzcch/bingo 的主要功能包括:Evaluation Benchmarks, Multimodal Benchmarks。

gzcch/bingo 有哪些开源替代品?

gzcch/bingo 的开源替代品包括: damo-nlp-sg/m3exam — Data and code for paper "M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models". hypjudy/sparkles — Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models. bradyfu/video-mme — ✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis. ailab-cvc/seed-bench — (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions. freedomintelligence/mllm-bench — MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria. lerogo/mmgenbench — Official repository of MMGenBench.