LLM Evaluation

Implement comprehensive evaluation strategies for LLM applications using automated metrics, human feedback, and benchmarking. Use when testing LLM performance, measuring AI application quality, or establishing evaluation frameworks.

asymmetric-al d29fd28 18.6 KB Updated

File contents

asymmetric-al/compass/tree/main/.agents/skills/llm-evaluation commit d29fd2866a

Frequently asked questions

npx skillmds@latest add asymmetric-al/llm-evaluation