awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectServer MCPDespreCum realizăm clasamentulPresă
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
patronus-ai avatar

patronus-ai/financebench

0
View on GitHub↗
328 stele·61 fork-uri·Jupyter Notebook·4 vizualizări

Financebench

Abstract: FinanceBench is a first-of-its-kind test suite for evaluating the performance of LLMs on open book financial question answering (QA). This repository contains an open source sample of 150 annotated examples used in the evaluation and analysis of models assessed in the FinanceBench…

Features

  • Evaluation Benchmarks - Benchmark for open-ended financial Q&A performance.

Istoric stele

Graficul istoricului de stele pentru patronus-ai/financebenchGraficul istoricului de stele pentru patronus-ai/financebench

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Alternative open-source pentru Financebench

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu Financebench.
  • chancefocus/pixiuAvatar chancefocus

    chancefocus/PIXIU

    868Vezi pe GitHub↗

    This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs), instruction tuning data, and evaluation benchmarks to holistically assess financial LLMs. Our goal is to continually push forward the open-source development of financial artificial intelligence (AI).

    Jupyter Notebook
    Vezi pe GitHub↗868
  • coastalcph/lex-glueAvatar coastalcph

    coastalcph/lex-glue

    259Vezi pe GitHub↗

    LexGLUE: A Benchmark Dataset for Legal Language Understanding in English

    Python
    Vezi pe GitHub↗259
  • codefuse-ai/codefuse-devops-evalAvatar codefuse-ai

    codefuse-ai/codefuse-devops-eval

    656Vezi pe GitHub↗

    Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain.

    Python
    Vezi pe GitHub↗656
  • cbluebenchmark/cblueAvatar CBLUEbenchmark

    CBLUEbenchmark/CBLUE

    843Vezi pe GitHub↗

    CBLUE1 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark

    Python
    Vezi pe GitHub↗843
Vezi toate cele 20 alternative pentru Financebench→

Întrebări frecvente

Ce face patronus-ai/financebench?

Abstract: FinanceBench is a first-of-its-kind test suite for evaluating the performance of LLMs on open book financial question answering (QA). This repository contains an open source sample of 150 annotated examples used in the evaluation and analysis of models assessed in the FinanceBench…

Care sunt principalele funcționalități ale patronus-ai/financebench?

Principalele funcționalități ale patronus-ai/financebench sunt: Evaluation Benchmarks.

Care sunt câteva alternative open-source pentru patronus-ai/financebench?

Alternativele open-source pentru patronus-ai/financebench includ: chancefocus/pixiu — This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs),… coastalcph/lex-glue — LexGLUE: A Benchmark Dataset for Legal Language Understanding in English. codefuse-ai/codefuse-devops-eval — Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain. dai-shen/laiw — LAiW: A Chinese Legal Large Language Models Benchmark. felixgithub2017/cg-eval — Chinese Generation Evaluation. cbluebenchmark/cblue — [CBLUE1] 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark.