Raven Research Looped Reasoning Eval

Scaffold a research-grade experiment that tests whether a Looped / Recurrent-Depth Transformer (RDT) is better than a parameter-matched vanilla transformer at vulnerability discovery, when both are continued-pretrained on a security corpus. Use when the user says "test the looped-transformer hypothesis", "run the OpenMythos RDT experiment", "compare loop count to trial budget", "does inference-time loops help security reasoning", or asks for a reproducible research pipeline around the Mythos architecture hypothesis. The skill scaffolds the experiment — it does NOT make claims about Mythos itself, and it explicitly bins OpenMythos as a community speculative reconstruction, not Anthropic's actual model.

daemon-blockint-tech Updated

File contents

daemon-blockint-tech/project-raven-d3fend/tree/main/agents/skills/research-looped-reasoning-eval commit fdb2f252e1

Frequently asked questions

npx skillmds@latest add daemon-blockint-tech/raven-research-looped-reasoning-eval