Guydav Restrictedpython Code Eval

Compute guydav/restrictedpython_code_eval via the HuggingFace `evaluate` library. Use when the user has predictions + references and wants the canonical implementation of guydav/restrictedpython_code_eval.

qhjqhj00 31ea8aa 942 B Updated 3 repo stars

File contents

guydav-restrictedpython-code-eval

Metric guydav/restrictedpython_code_eval from the HuggingFace evaluate library.

When to invoke

User asks to compute guydav/restrictedpython_code_eval or wants HF evaluate's canonical version.

Recipe

import evaluate
metric = evaluate.load("guydav/restrictedpython_code_eval")
result = metric.compute(predictions=preds, references=refs)
print(result)

Don'ts

  • Don't assume your in-house guydav/restrictedpython_code_eval matches HF — version conventions vary.
  • Many evaluate metrics have task-specific arguments (average=, lang=, model_type=); read the metric card before reporting numbers.

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/guydav-restrictedpython-code-eval commit 31ea8aad23

Frequently asked questions

npx skillmds add qhjqhj00/guydav-restrictedpython-code-eval