Cipher Crypto Vuln Eval

This benchmark evaluates whether large language models can generate secure cryptographic Python code and avoid common implementation flaws under varying security guidance. It probes the model's ability to follow secure prompting instructions, correctly implement cryptographic primitives, and avoid known anti-patterns like weak hashing or fixed IVs. Use when the user wants to benchmark on CIPHER, or asks about evaluating this task. Reports vulnerability_rates.

qhjqhj00 aeff4c6 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/cipher-crypto-vuln-eval commit aeff4c60a4

Frequently asked questions

npx skillmds add qhjqhj00/cipher-crypto-vuln-eval