Ptb Wikitext2 Lm Eval

Evaluates next-token prediction accuracy and long-range dependency modeling in language models, with a specific focus on handling rare and out-of-vocabulary words without expanding vocabulary size. Use when the user wants to benchmark on Penn Treebank, WikiText-2, or asks about evaluating this task. Reports perplexity.

qhjqhj00 4c656fb 2.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/ptb-wikitext2-lm-eval commit 4c656fb136

Frequently asked questions

npx skillmds add qhjqhj00/ptb-wikitext2-lm-eval