Sequential Model Editing Eval

Evaluates the stability and performance of sequential knowledge editing methods on large language models over long horizons. It probes whether editing techniques can maintain factual accuracy, preserve general capabilities, and avoid norm blow-up or catastrophic forgetting across thousands of atomic updates. Use when the user wants to benchmark on CounterFact, ZsRE, WikiBigEdit, GLUE-style tasks (SST, MRPC, RTE, CoLA, MNLI), MMLU, or asks about evaluating this task. Reports Efficacy.

qhjqhj00 ab6a582 4.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/sequential-model-editing-eval commit ab6a582db6

Frequently asked questions

npx skillmds add qhjqhj00/sequential-model-editing-eval