awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectServer MCPDespreCum realizăm clasamentulPresă
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Felixgithub2017 avatar

Felixgithub2017/MMCU

0
View on GitHub↗
90 stele·11 fork-uri·Python·3 vizualizări

MMCU

MEASURING MASSIVE MULTITASK CHINESE UNDERSTANDING

Features

  • Evaluation Benchmarks - Benchmark covering medical, legal, psychological, and educational domains.

Istoric stele

Graficul istoricului de stele pentru felixgithub2017/mmcuGraficul istoricului de stele pentru felixgithub2017/mmcu

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Alternative open-source pentru MMCU

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu MMCU.
  • chancefocus/pixiuAvatar chancefocus

    chancefocus/PIXIU

    868Vezi pe GitHub↗

    This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs), instruction tuning data, and evaluation benchmarks to holistically assess financial LLMs. Our goal is to continually push forward the open-source development of financial artificial intelligence (AI).

    Jupyter Notebook
    Vezi pe GitHub↗868
  • coastalcph/lex-glueAvatar coastalcph

    coastalcph/lex-glue

    259Vezi pe GitHub↗

    LexGLUE: A Benchmark Dataset for Legal Language Understanding in English

    Python
    Vezi pe GitHub↗259
  • codefuse-ai/codefuse-devops-evalAvatar codefuse-ai

    codefuse-ai/codefuse-devops-eval

    656Vezi pe GitHub↗

    Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain.

    Python
    Vezi pe GitHub↗656
  • cbluebenchmark/cblueAvatar CBLUEbenchmark

    CBLUEbenchmark/CBLUE

    843Vezi pe GitHub↗

    CBLUE1 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark

    Python
    Vezi pe GitHub↗843
Vezi toate cele 20 alternative pentru MMCU→

Întrebări frecvente

Ce face felixgithub2017/mmcu?

MEASURING MASSIVE MULTITASK CHINESE UNDERSTANDING

Care sunt principalele funcționalități ale felixgithub2017/mmcu?

Principalele funcționalități ale felixgithub2017/mmcu sunt: Evaluation Benchmarks.

Care sunt câteva alternative open-source pentru felixgithub2017/mmcu?

Alternativele open-source pentru felixgithub2017/mmcu includ: chancefocus/pixiu — This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs),… coastalcph/lex-glue — LexGLUE: A Benchmark Dataset for Legal Language Understanding in English. codefuse-ai/codefuse-devops-eval — Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain. dai-shen/laiw — LAiW: A Chinese Legal Large Language Models Benchmark. felixgithub2017/cg-eval — Chinese Generation Evaluation. cbluebenchmark/cblue — [CBLUE1] 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark.