LLM Evaluation Framework

Use when performing llm evaluation framework — provides a structured framework for evaluating large language models (LLMs) for production use. Covers task-specific benchmarking, safety testing, cost analysis, latency measurement, prompt engineering evaluation, and comparison across models to select the optimal LLM for a given use case.

cloudthinker-ai Updated 7 repo stars

File contents

cloudthinker-ai/CloudSkills/tree/main/skills/templates/llm-evaluation-framework commit c6efdd73c7

Frequently asked questions

npx skillmds@latest add cloudthinker-ai/llm-evaluation-framework