Kitaru Replay Experiment

Run and interpret one safe, bounded Kitaru replay comparison against an accepted cohort and exact evaluators. Use when a user wants to replay a cohort, test or compare a model, prompt, system prompt, parameter, agent version, or tool policy, supervise an experiment run, determine whether one candidate helped, or ask for one bounded change worth testing.

zenml-io Updated

File contents

zenml-io/kitaru-skills/tree/main/skills/kitaru-replay-experiment commit 044869ec4f

Frequently asked questions

npx skillmds@latest add zenml-io/kitaru-replay-experiment