Md Evalbench Eval

Evaluates large language models on molecular dynamics domain knowledge, LAMMPS scripting syntax comprehension, and automatic generation of executable LAMMPS simulation scripts from natural language instructions. Use when the user wants to benchmark on MD-EvalBench, or asks about evaluating this task. Reports Exec-Success@$k$.

qhjqhj00 037eea5 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/md-evalbench-eval commit 037eea501b

Frequently asked questions

npx skillmds add qhjqhj00/md-evalbench-eval