Held Out Test Loss

Evaluates language model generalization and overfitting by measuring cross-entropy loss on a held-out test set. It probes how well the model retains predictive performance when trained on repeated or constrained data subsets. Use when the user has predictions and gold and needs to compute held-out test loss.

qhjqhj00 7f50cf5 2.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/held-out-test-loss commit 7f50cf5672

Frequently asked questions

npx skillmds add qhjqhj00/held-out-test-loss