K4black Codebleu

Compute k4black/codebleu via the HuggingFace `evaluate` library. Use when the user has predictions + references and wants the canonical implementation of k4black/codebleu.

qhjqhj00 903d72c 806 B Updated 3 repo stars

File contents

k4black-codebleu

Metric k4black/codebleu from the HuggingFace evaluate library.

When to invoke

User asks to compute k4black/codebleu or wants HF evaluate's canonical version.

Recipe

import evaluate
metric = evaluate.load("k4black/codebleu")
result = metric.compute(predictions=preds, references=refs)
print(result)

Don'ts

  • Don't assume your in-house k4black/codebleu matches HF — version conventions vary.
  • Many evaluate metrics have task-specific arguments (average=, lang=, model_type=); read the metric card before reporting numbers.

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/k4black-codebleu commit 903d72c698

Frequently asked questions

npx skillmds add qhjqhj00/k4black-codebleu