awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
DAMO-NLP-SG avatar

DAMO-NLP-SG/M3Exam

0
View on GitHub↗
105 stars·13 forks·Python·10 views

M3Exam

Data and code for paper "M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models"

Features

  • Evaluation Benchmarks - Multilingual and multimodal benchmark for exam-based testing.
  • Multimodal Benchmarks - Multilingual and multilevel benchmark for multimodal models.

Star history

Star history chart for damo-nlp-sg/m3examStar history chart for damo-nlp-sg/m3exam

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does damo-nlp-sg/m3exam do?

Data and code for paper "M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models"

What are the main features of damo-nlp-sg/m3exam?

The main features of damo-nlp-sg/m3exam are: Evaluation Benchmarks, Multimodal Benchmarks.

Which projects share features with damo-nlp-sg/m3exam?

Projects with overlapping indexed features include: hypjudy/sparkles — Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models. gzcch/bingo — Chenhang Cui, Yiyang Zhou, Xinyu Yang, Shirley Wu, Linjun Zhang, James Zou, Huaxiu Yao *Equal Contribution. bradyfu/video-mme — ✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis. ailab-cvc/seed-bench — (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions. freedomintelligence/mllm-bench — MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria. lerogo/mmgenbench — Official repository of MMGenBench.

Projects sharing features with M3Exam

These projects share indexed features with M3Exam. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • bradyfu/video-mmeBradyFU avatar

    BradyFU/Video-MME

    779View on GitHub↗

    ✨✨CVPR 2025 Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    View on GitHub↗779
  • freedomintelligence/mllm-benchFreedomIntelligence avatar

    FreedomIntelligence/MLLM-Bench

    76View on GitHub↗

    MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria

    Python
    View on GitHub↗76
  • ailab-cvc/seed-benchAILab-CVC avatar

    AILab-CVC/SEED-Bench

    364View on GitHub↗

    (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions.

    Python
    View on GitHub↗364
  • gzcch/bingogzcch avatar

    gzcch/Bingo

    55View on GitHub↗

    Chenhang Cui, Yiyang Zhou, Xinyu Yang, Shirley Wu, Linjun Zhang, James Zou, Huaxiu Yao *Equal Contribution

    View on GitHub↗55
Compare all 30 related projects
→