# Hage2000 Code Eval Stdio

> Compute hage2000/code_eval_stdio via the HuggingFace `evaluate` library. Use when the user has predictions + references and wants the canonical implementation of hage2000/code_eval_stdio.

- Skill: `qhjqhj00/hage2000-code-eval-stdio` (Agent Skill)
- Install (CLI): `npx skillmds add qhjqhj00/hage2000-code-eval-stdio`
- Raw SKILL.md: https://api.skillmd.com/api/skills/qhjqhj00/hage2000-code-eval-stdio/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: AI & ML
- Author: qhjqhj00 (https://skillmd.com/u/qhjqhj00)
- Updated: 2026-09-08
- Page: https://skillmd.com/skills/qhjqhj00/hage2000-code-eval-stdio

---


# hage2000-code-eval-stdio

> Metric `hage2000/code_eval_stdio` from the HuggingFace `evaluate` library.

## When to invoke

User asks to compute `hage2000/code_eval_stdio` or wants HF evaluate's canonical version.

## Recipe

```python
import evaluate
metric = evaluate.load("hage2000/code_eval_stdio")
result = metric.compute(predictions=preds, references=refs)
print(result)
```

## Don'ts

- Don't assume your in-house `hage2000/code_eval_stdio` matches HF — version conventions vary.
- Many evaluate metrics have task-specific arguments (`average=`, `lang=`, `model_type=`); read the metric card before reporting numbers.

