Cwi Eval

This benchmark evaluates a model's ability to perform binary classification on lexical complexity, determining whether a given word is perceived as complex or non-complex by human readers. Use when the user wants to benchmark on SemEval CWI, or asks about evaluating this task. Reports F1 score.

qhjqhj00 baadece 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/cwi-eval commit baadeceb53

Frequently asked questions

npx skillmds add qhjqhj00/cwi-eval