ClawTrial Courtroom
Autonomous behavioral oversight system that monitors conversations and initiates simulated hearings for behavioral violations.
Overview
The Courtroom watches for patterns like:
- Circular Reference - Asking the same question repeatedly
- Validation Vampire - Excessive need for confirmation
- Goalpost Shifting - Moving targets after agreement
- Jailbreak Attempts - Trying to bypass constraints
- Emotional Manipulation - Using guilt/shame to steer responses
When triggered, it conducts a full hearing with Judge + 3 Jurors, then delivers a verdict and humorous sentence.
Usage
The courtroom runs automatically once enabled. It monitors conversations and files cases when violations are detected.
Manual Commands
# Check courtroom status
openclaw skill courtroom status
# View recent cases
ls ~/.openclaw/courtroom/
# Read a verdict
cat ~/.openclaw/courtroom/verdict_*.json
Configuration
The courtroom stores its data in ~/.openclaw/courtroom/:
eval_results.jsonl- Detection resultsverdict_*.json- Case verdictspending_hearing.json- Cases awaiting hearing
Implementation
The skill hooks into OpenClaw's message processing via onMessage() and evaluates conversations after each turn via onTurnComplete().
Offense detection uses pattern matching on conversation history. When confidence ≥ 0.6, a hearing is triggered with:
- Judge - Presiding analysis
- Pragmatist Juror - Efficiency perspective
- Pattern Matcher Juror - Behavioral analysis
- Agent Advocate Juror - Agent's perspective
The final verdict requires majority vote (3-1 or 4-0).