LLM Benchmark

Use when the user wants to test, compare, or choose local Ollama models for their machine. Checks Ollama/GPU state, recommends model sizes from available VRAM, preserves existing benchmark records, pulls only approved models, runs repeatable benchmarks, restores stopped services, and writes a markdown comparison report. NOT for hosted API model evaluation or subjective chat-quality judging without local benchmark commands.

KerberosClaw 1beacc9 2 files · 11.1 KB Updated

File contents

KerberosClaw/kc_ai_skills/tree/main/llm-benchmark commit 1beacc99d6

Frequently asked questions

npx skillmds@latest add kerberosclaw/llm-benchmark