awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
AILab-CVC avatar

AILab-CVC/SEED-Bench

0
View on GitHub↗
364 stars·13 forks·Python·14 views

SEED Bench

(CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions.

Features

  • Evaluation Benchmarks - Benchmarking generative comprehension in multimodal models.
  • Multimodal Benchmarks - Benchmark for evaluating generative comprehension.

Star history

Star history chart for ailab-cvc/seed-benchStar history chart for ailab-cvc/seed-bench

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Projects sharing features with SEED Bench

These projects share indexed features with SEED Bench. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • damo-nlp-sg/m3examDAMO-NLP-SG avatar

    DAMO-NLP-SG/M3Exam

    105View on GitHub↗

    Data and code for paper "M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models"

    Pythonai-educationchatgptevaluation
    View on GitHub↗105
  • freedomintelligence/mllm-benchFreedomIntelligence avatar

    FreedomIntelligence/MLLM-Bench

    76View on GitHub↗

    MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria

    Python
    View on GitHub↗76
  • bradyfu/video-mmeBradyFU avatar

    BradyFU/Video-MME

    779View on GitHub↗

    ✨✨CVPR 2025 Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    View on GitHub↗779
  • gzcch/bingogzcch avatar

    gzcch/Bingo

    55View on GitHub↗

    Chenhang Cui, Yiyang Zhou, Xinyu Yang, Shirley Wu, Linjun Zhang, James Zou, Huaxiu Yao *Equal Contribution

    View on GitHub↗55
Compare all 30 related projects→

Frequently asked questions

What does ailab-cvc/seed-bench do?

(CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions.

What are the main features of ailab-cvc/seed-bench?

The main features of ailab-cvc/seed-bench are: Evaluation Benchmarks, Multimodal Benchmarks.

Which projects share features with ailab-cvc/seed-bench?

Projects with overlapping indexed features include: hypjudy/sparkles — Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models. gzcch/bingo — Chenhang Cui, Yiyang Zhou, Xinyu Yang, Shirley Wu, Linjun Zhang, James Zou, Huaxiu Yao *Equal Contribution. bradyfu/video-mme — ✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis. damo-nlp-sg/m3exam — Data and code for paper "M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models". freedomintelligence/mllm-bench — MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria. lerogo/mmgenbench — Official repository of MMGenBench.