# K4black Codebleu

> Compute k4black/codebleu via the HuggingFace `evaluate` library. Use when the user has predictions + references and wants the canonical implementation of k4black/codebleu.

- Skill: `qhjqhj00/k4black-codebleu` (Agent Skill)
- Install (CLI): `npx skillmds add qhjqhj00/k4black-codebleu`
- Raw SKILL.md: https://api.skillmd.com/api/skills/qhjqhj00/k4black-codebleu/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: AI & ML
- Author: qhjqhj00 (https://skillmd.com/u/qhjqhj00)
- Updated: 2026-09-08
- Page: https://skillmd.com/skills/qhjqhj00/k4black-codebleu

---


# k4black-codebleu

> Metric `k4black/codebleu` from the HuggingFace `evaluate` library.

## When to invoke

User asks to compute `k4black/codebleu` or wants HF evaluate's canonical version.

## Recipe

```python
import evaluate
metric = evaluate.load("k4black/codebleu")
result = metric.compute(predictions=preds, references=refs)
print(result)
```

## Don'ts

- Don't assume your in-house `k4black/codebleu` matches HF — version conventions vary.
- Many evaluate metrics have task-specific arguments (`average=`, `lang=`, `model_type=`); read the metric card before reporting numbers.

