Cfbenchmark Basic Eval

Evaluates Chinese large language models on financial text processing capabilities, specifically entity recognition, text classification, and content generation within the financial domain. It tests the models' adaptability using zero-shot and few-shot (3 examples) prompting strategies across eight distinct tasks. Use when the user wants to benchmark on CFBenchmark-Basic, or asks about evaluating this task. Reports F1-Score.

qhjqhj00 f29e0a7 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/cfbenchmark-basic-eval commit f29e0a7dda

Frequently asked questions

npx skillmds add qhjqhj00/cfbenchmark-basic-eval