Model Evaluation Benchmark

Automated reproduction of comprehensive model evaluation benchmarks following the Benchmark Suite V3. Auto-activates for model benchmarking, comparison evaluation, or performance testing between AI models.

rysweet 76a5374 3.4 KB Updated

File contents

rysweet/amplihack-rs/tree/main/amplifier-bundle/skills/model-evaluation-benchmark commit 76a5374055

Frequently asked questions

npx skillmds@latest add rysweet/model-evaluation-benchmark