awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektÜber unsRanking-MethodikPresseMCP-Server
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
hkust-nlp avatar

hkust-nlp/ceval

0
View on GitHub↗
1,854 Stars·84 Forks·Python·MIT·2 Aufrufecevalbenchmark.com↗

Ceval

Official github repo for C-Eval, a Chinese evaluation suite for foundation models [NeurIPS 2023]

Features

  • Model Evaluation and Benchmarking - Comprehensive evaluation suite for foundation models in Chinese.
  • Natural Language Processing - Listed in the “Natural Language Processing” section of the FunNLP awesome list.

Star-Verlauf

Star-Verlauf für hkust-nlp/cevalStar-Verlauf für hkust-nlp/ceval

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Häufig gestellte Fragen

Was macht hkust-nlp/ceval?

Official github repo for C-Eval, a Chinese evaluation suite for foundation models [NeurIPS 2023]

Was sind die Hauptfunktionen von hkust-nlp/ceval?

Die Hauptfunktionen von hkust-nlp/ceval sind: Model Evaluation and Benchmarking, Natural Language Processing.

Welche Open-Source-Alternativen gibt es zu hkust-nlp/ceval?

Open-Source-Alternativen zu hkust-nlp/ceval sind unter anderem: ntmc-community/matchzoo — MatchZoo is a deep learning framework designed for building, training, and evaluating neural networks that determine… openlmlab/gaokao-bench — GAOKAO-Bench is an evaluation framework that utilizes GAOKAO questions as a dataset to evaluate large language models. comet-ml/opik — Opik is an observability and evaluation platform designed for generative AI applications and agentic workflows. It… open-compass/opencompass — OpenCompass is an open-source framework for standardized benchmarking of large language models. It provides a… dusty-nv/jetson-inference — jetson-inference is a set of libraries and tools for executing optimized deep learning models on embedded GPU… facebookresearch/parlai — ParlAI is a conversational AI research framework designed for training, evaluating, and sharing dialogue models using…

Open-Source-Alternativen zu Ceval

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit Ceval.
  • ntmc-community/matchzooAvatar von NTMC-Community

    NTMC-Community/MatchZoo

    3,845Auf GitHub ansehen↗

    MatchZoo is a deep learning framework designed for building, training, and evaluating neural networks that determine the relevance and similarity between pairs of textual inputs. It serves as a research platform for neural information retrieval, specifically supporting the development of models for document retrieval, question answering, and ranking tasks. The framework utilizes declarative architecture composition to define complex neural network structures. It includes automated hyper-parameter resolution to populate missing configuration parameters before model compilation and uses callbac

    Python
    Auf GitHub ansehen↗3,845
  • open-compass/opencompassAvatar von open-compass

    open-compass/opencompass

    6,678Auf GitHub ansehen↗

    OpenCompass is an open-source framework for standardized benchmarking of large language models. It provides a configurable evaluation pipeline that supports both objective and subjective assessment, using a dual-engine architecture to handle closed-form answer comparison and open-ended response rating. The framework is designed as a modular platform where datasets, models, and metrics are composed through declarative YAML configuration files. The framework distinguishes itself through its extensible model integration layer, which supports custom models, HuggingFace models, and third-party API

    Pythonbenchmarkchatgptevaluation
    Auf GitHub ansehen↗6,678
  • comet-ml/opikAvatar von comet-ml

    comet-ml/opik

    17,787Auf GitHub ansehen↗

    Opik is an observability and evaluation platform designed for generative AI applications and agentic workflows. It provides a centralized environment for tracing execution flows, managing prompt templates, and monitoring production performance, allowing teams to gain visibility into complex model interactions and tool usage without requiring manual application code changes. The platform distinguishes itself through its integrated approach to the AI development lifecycle, combining distributed trace instrumentation with automated evaluation frameworks. It supports model-as-a-judge scoring, syn

    Pythonevaluationhacktoberfesthacktoberfest2025
    Auf GitHub ansehen↗17,787
  • openlmlab/gaokao-benchAvatar von OpenLMLab

    OpenLMLab/GAOKAO-Bench

    760Auf GitHub ansehen↗

    GAOKAO-Bench is an evaluation framework that utilizes GAOKAO questions as a dataset to evaluate large language models.

    Python
    Auf GitHub ansehen↗760
Alle 30 Alternativen zu Ceval anzeigen→