awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
patronus-ai avatar

patronus-ai/financebench

0
View on GitHub↗
328 نجوم·61 تفرعات·Jupyter Notebook·4 مشاهدات

Financebench

Abstract: FinanceBench is a first-of-its-kind test suite for evaluating the performance of LLMs on open book financial question answering (QA). This repository contains an open source sample of 150 annotated examples used in the evaluation and analysis of models assessed in the FinanceBench…

Features

  • Evaluation Benchmarks - Benchmark for open-ended financial Q&A performance.

سجل النجوم

مخطط تاريخ النجوم لـ patronus-ai/financebenchمخطط تاريخ النجوم لـ patronus-ai/financebench

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

بدائل مفتوحة المصدر لـ Financebench

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع Financebench.
  • chancefocus/pixiuالصورة الرمزية لـ chancefocus

    chancefocus/PIXIU

    868عرض على GitHub↗

    This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs), instruction tuning data, and evaluation benchmarks to holistically assess financial LLMs. Our goal is to continually push forward the open-source development of financial artificial intelligence (AI).

    Jupyter Notebook
    عرض على GitHub↗868
  • coastalcph/lex-glueالصورة الرمزية لـ coastalcph

    coastalcph/lex-glue

    259عرض على GitHub↗

    LexGLUE: A Benchmark Dataset for Legal Language Understanding in English

    Python
    عرض على GitHub↗259
  • codefuse-ai/codefuse-devops-evalالصورة الرمزية لـ codefuse-ai

    codefuse-ai/codefuse-devops-eval

    656عرض على GitHub↗

    Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain.

    Python
    عرض على GitHub↗656
  • cbluebenchmark/cblueالصورة الرمزية لـ CBLUEbenchmark

    CBLUEbenchmark/CBLUE

    843عرض على GitHub↗

    CBLUE1 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark

    Python
    عرض على GitHub↗843
عرض جميع البدائل الـ 20 لـ Financebench→

الأسئلة الشائعة

ما هي وظيفة patronus-ai/financebench؟

Abstract: FinanceBench is a first-of-its-kind test suite for evaluating the performance of LLMs on open book financial question answering (QA). This repository contains an open source sample of 150 annotated examples used in the evaluation and analysis of models assessed in the FinanceBench…

ما هي الميزات الرئيسية لـ patronus-ai/financebench؟

الميزات الرئيسية لـ patronus-ai/financebench هي: Evaluation Benchmarks.

ما هي البدائل مفتوحة المصدر لـ patronus-ai/financebench؟

تشمل البدائل مفتوحة المصدر لـ patronus-ai/financebench: chancefocus/pixiu — This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs),… coastalcph/lex-glue — LexGLUE: A Benchmark Dataset for Legal Language Understanding in English. codefuse-ai/codefuse-devops-eval — Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain. dai-shen/laiw — LAiW: A Chinese Legal Large Language Models Benchmark. felixgithub2017/cg-eval — Chinese Generation Evaluation. cbluebenchmark/cblue — [CBLUE1] 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark.