General Nlu Eval

Evaluates whether integrating external knowledge (textual descriptions or embeddings) improves performance on general natural language understanding tasks compared to baseline pre-trained language models. It probes the model's ability to leverage external semantic information to enhance representation learning and decision-making across classification, regression, and sequence labeling benchmarks. Use when the user wants to benchmark on GLUE, Penn Treebank, CoNLL-2003, or asks about evaluating this task. Reports Accuracy / F1 / Pearson correlation / Matthew's correlation.

qhjqhj00 1a4832b 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/general-nlu-eval commit 1a4832bb0f

Frequently asked questions

npx skillmds add qhjqhj00/general-nlu-eval