Harbor Harness Improvement Loop

Analyze Harbor benchmark jobs and ATIF trajectories, identify evidence-backed weaknesses across LibrAgent builtin tools, prompts, agent guidance, execution policy, and benchmark instrumentation, then design controlled improvements and rerun comparisons. Use when analyzing Harbor or Terminal-Bench results, optimizing the harness, investigating tool-call failures or retries, auditing prompt/context efficiency, or running repeated BM → analysis → improvement cycles.

fritzprix 43f24fd 4 files · 24.9 KB Updated

File contents

fritzprix/libr-agent/tree/main/.agents/skills/harbor-harness-improvement-loop commit 43f24fd228

Frequently asked questions

npx skillmds@latest add fritzprix/harbor-harness-improvement-loop