awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
chenllliang avatar

chenllliang/MMEvalPro

0
View on GitHub↗
25 نجوم·3 تفرعات·Python·5 مشاهداتmmevalpro.github.io↗

MMEvalPro

[NAACL 2025] Source code for MMEvalPro, a more trustworthy and efficient benchmark for evaluating LMMs

Features

  • Evaluation Benchmarks - Calibrating benchmarks for trustworthy and efficient evaluation.

سجل النجوم

مخطط تاريخ النجوم لـ chenllliang/mmevalproمخطط تاريخ النجوم لـ chenllliang/mmevalpro

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

الأسئلة الشائعة

ما هي وظيفة chenllliang/mmevalpro؟

[NAACL 2025] Source code for MMEvalPro, a more trustworthy and efficient benchmark for evaluating LMMs

ما هي الميزات الرئيسية لـ chenllliang/mmevalpro؟

الميزات الرئيسية لـ chenllliang/mmevalpro هي: Evaluation Benchmarks.

ما هي البدائل مفتوحة المصدر لـ chenllliang/mmevalpro؟

تشمل البدائل مفتوحة المصدر لـ chenllliang/mmevalpro: pyspur-dev/pyspur. datawhalechina/prompt-engineering-for-developers — This project is a technical curriculum and development guide focused on large language model prompt engineering,… ai45lab/openrt — Open-source red teaming framework for MLLMs with 42+ attack methods. ailab-cvc/seed-bench — (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions. albertwy/gpt-4v-evaluation — Data for evaluating GPT-4V. aifeg/benchlmm — [ECCV 2024] BenchLMM: Benchmarking Cross-style Visual Capability of Large Multimodal Models.

بدائل مفتوحة المصدر لـ MMEvalPro

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع MMEvalPro.
  • pyspur-dev/pyspurالصورة الرمزية لـ PySpur-Dev

    PySpur-Dev/pyspur

    5,677عرض على GitHub↗
    TypeScriptagentagentsai
    عرض على GitHub↗5,677
  • datawhalechina/prompt-engineering-for-developersالصورة الرمزية لـ datawhalechina

    datawhalechina/prompt-engineering-for-developers

    24,267عرض على GitHub↗

    This project is a technical curriculum and development guide focused on large language model prompt engineering, fine-tuning, and the creation of retrieval augmented generation applications. It serves as a comprehensive resource for developers to master crafting precise instructions and textual patterns to improve the quality and predictability of model outputs. The material covers the end-to-end workflow of adapting open-source models to specific datasets and integrating language models with vector databases to generate responses based on private information. It also provides a systematic ap

    Jupyter Notebook
    عرض على GitHub↗24,267
  • ai45lab/openrtالصورة الرمزية لـ AI45Lab

    AI45Lab/OpenRT

    257عرض على GitHub↗

    Open-source red teaming framework for MLLMs with 42+ attack methods

    Python
    عرض على GitHub↗257
  • aifeg/benchlmmالصورة الرمزية لـ AIFEG

    AIFEG/BenchLMM

    86عرض على GitHub↗

    ECCV 2024 BenchLMM: Benchmarking Cross-style Visual Capability of Large Multimodal Models

    Pythonbenchmarkcvdataset
    عرض على GitHub↗86
عرض جميع البدائل الـ 30 لـ MMEvalPro→