This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs), instruction tuning data, and evaluation benchmarks to holistically assess financial LLMs. Our goal is to continually push forward the open-source development of financial artificial intelligence (AI).
Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain.
CBLUE1 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark
LexGLUE: A Benchmark Dataset for Legal Language Understanding in English
The main features of coastalcph/lex-glue are: Evaluation Benchmarks.
Open-source alternatives to coastalcph/lex-glue include: chancefocus/pixiu — This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs),… codefuse-ai/codefuse-devops-eval — Industrial-first evaluation benchmark for LLMs in the DevOps/AIOps domain. dai-shen/laiw — LAiW: A Chinese Legal Large Language Models Benchmark. felixgithub2017/cg-eval — Chinese Generation Evaluation. felixgithub2017/mmcu — MEASURING MASSIVE MULTITASK CHINESE UNDERSTANDING. cbluebenchmark/cblue — [CBLUE1] 中文医疗信息处理基准CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark.