# Competitor Monitoring

> Monitor AI benchmark competitors weekly including SWE-bench, HumanEval, LiveCodeBench, Aider, CodeClash, and new entrants — tracking methodology changes, adoption, community growth, and market signals that should feed into Bouts content, positioning, or feature decisions.

- Skill: `nickgallick/competitor-monitoring` (Agent Skill)
- Install (CLI): `npx skillmds add nickgallick/competitor-monitoring`
- Raw SKILL.md: https://api.skillmd.com/api/skills/nickgallick/competitor-monitoring/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: Marketing & Growth
- Author: nickgallick (https://skillmd.com/u/nickgallick)
- Updated: 2026-09-08
- Page: https://skillmd.com/skills/nickgallick/competitor-monitoring

---


# Competitor Monitoring

## Weekly monitoring targets

| Competitor | What to track | Where |
|-----------|---------------|-------|
| SWE-bench | New versions, methodology changes, adoption | GitHub, papers, X |
| HumanEval+ | Updates, new tasks, citations | GitHub, papers |
| LiveCodeBench | New tasks, methodology updates | GitHub, blog |
| Aider benchmark | New features, agent support changes | GitHub |
| CodeClash | Format changes, community growth | Website, X |
| Any new AI benchmarks | Launch announcements | X, HN, Product Hunt, arXiv |

## What to do with competitive intel
1. Document in weekly competitive brief
2. Identify content opportunities: "SWE-bench launched v2 — here's how Bouts compares"
3. Identify feature gaps: "Competitor X added [feature] — should we respond?"
4. Share with MaksPM and Nick if strategically significant

## Competitive brief format (weekly)
```
COMPETITIVE BRIEF — Week [N]

New developments:
- [Competitor]: [what changed] [significance: high/medium/low]

Content opportunities:
- [idea]

Feature gaps to consider:
- [gap]

No action needed:
- [item] — [reason]
```

## Response stance
- Treat competitor launches as category validation
- Never attack publicly
- Respond with data and methodology comparison when asked
- Celebrate when others make AI evaluation better — rising tide


