awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoServidor MCPAcerca deCómo clasificamosPrensa
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
patronus-ai avatar

patronus-ai/financebench

0
View on GitHub↗
328 estrellas·61 forks·Jupyter Notebook·5 vistas

Financebench

Abstract: FinanceBench is a first-of-its-kind test suite for evaluating the performance of LLMs on open book financial question answering (QA). This repository contains an open source sample of 150 annotated examples used in the evaluation and analysis of models assessed in the FinanceBench…

Features

  • Evaluation Benchmarks - Benchmark for open-ended financial Q&A performance.

Historial de estrellas

Gráfico del historial de estrellas de patronus-ai/financebenchGráfico del historial de estrellas de patronus-ai/financebench

Búsqueda con IA

Explora más repositorios increíbles

Describe lo que necesitas en lenguaje sencillo: la IA clasifica miles de proyectos open-source curados por relevancia.

Start searching with AI

Alternativas open-source a Financebench

Proyectos open-source similares, clasificados según cuántas características comparten con Financebench.
  • chancefocus/pixiuAvatar de chancefocus

    chancefocus/PIXIU

    868Ver en GitHub↗

    This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs), instruction tuning data, and evaluation benchmarks to holistically assess financial LLMs. Our goal is to continually push forward the open-source development of financial artificial intelligence (AI).

    Jupyter Notebook
    Ver en GitHub↗868
  • coastalcph/lex-glueAvatar de coastalcph

    coastalcph/lex-glue

    259Ver en GitHub↗

    LexGLUE: A Benchmark Dataset for Legal Language Understanding in English

    Python
    Ver en GitHub↗259
  • codefuse-ai/codefuse-devops-evalAvatar de codefuse-ai

    codefuse-ai/codefuse-devops-eval

    656Ver en GitHub↗

    Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain.

    Python
    Ver en GitHub↗656
  • cbluebenchmark/cblueAvatar de CBLUEbenchmark

    CBLUEbenchmark/CBLUE

    843Ver en GitHub↗

    CBLUE1 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark

    Python
    Ver en GitHub↗843
Ver las 20 alternativas a Financebench→

Preguntas frecuentes

¿Qué hace patronus-ai/financebench?

Abstract: FinanceBench is a first-of-its-kind test suite for evaluating the performance of LLMs on open book financial question answering (QA). This repository contains an open source sample of 150 annotated examples used in the evaluation and analysis of models assessed in the FinanceBench…

¿Cuáles son las características principales de patronus-ai/financebench?

Las características principales de patronus-ai/financebench son: Evaluation Benchmarks.

¿Qué alternativas de código abierto existen para patronus-ai/financebench?

Las alternativas de código abierto para patronus-ai/financebench incluyen: chancefocus/pixiu — This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs),… coastalcph/lex-glue — LexGLUE: A Benchmark Dataset for Legal Language Understanding in English. codefuse-ai/codefuse-devops-eval — Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain. dai-shen/laiw — LAiW: A Chinese Legal Large Language Models Benchmark. felixgithub2017/cg-eval — Chinese Generation Evaluation. cbluebenchmark/cblue — [CBLUE1] 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark.