Calibrated Similarity

Measures the semantic novelty of LLM-generated text by quantifying its similarity to the closest segment in the model's pretraining corpus. It probes whether models merely reproduce memorized training data or generalize to produce compositionally distinct outputs. Use when the user has predictions and gold and needs to compute calibrated similarity.

qhjqhj00 6d5d676 3.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/calibrated-similarity commit 6d5d67666c

Frequently asked questions

npx skillmds add qhjqhj00/calibrated-similarity