Browsesafe Bench Eval

Evaluates AI browser agents' ability to detect prompt injection attacks embedded in complex, realistic HTML environments. It probes whether models can distinguish malicious intent from benign distractors across diverse attack types, injection strategies, and linguistic styles. Use when the user wants to benchmark on BrowseSafe-Bench, or asks about evaluating this task. Reports balanced accuracy.

qhjqhj00 66e1326 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/browsesafe-bench-eval commit 66e1326b5e

Frequently asked questions

npx skillmds add qhjqhj00/browsesafe-bench-eval