Multilogbench Eval

Evaluates LLMs' ability to generate appropriate logging statements for code callables across six programming languages, testing both snapshot-based code understanding and revision-history-based code evolution contexts. Use when the user wants to benchmark on MultiLogBench, or asks about evaluating this task. Reports exact-match accuracy.

qhjqhj00 5aaa2a7 2.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/multilogbench-eval commit 5aaa2a799a

Frequently asked questions

npx skillmds add qhjqhj00/multilogbench-eval