Pico Lm Perplexity

Compute pico-lm/perplexity via the HuggingFace `evaluate` library. Use when the user has predictions + references and wants the canonical implementation of pico-lm/perplexity.

qhjqhj00 95bfa1f 822 B Updated 3 repo stars

File contents

pico-lm-perplexity

Metric pico-lm/perplexity from the HuggingFace evaluate library.

When to invoke

User asks to compute pico-lm/perplexity or wants HF evaluate's canonical version.

Recipe

import evaluate
metric = evaluate.load("pico-lm/perplexity")
result = metric.compute(predictions=preds, references=refs)
print(result)

Don'ts

  • Don't assume your in-house pico-lm/perplexity matches HF — version conventions vary.
  • Many evaluate metrics have task-specific arguments (average=, lang=, model_type=); read the metric card before reporting numbers.

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/pico-lm-perplexity commit 95bfa1fdd6

Frequently asked questions

npx skillmds add qhjqhj00/pico-lm-perplexity