Fine Grained Interpretability Eval

Evaluates the interpretability and faithfulness of neural NLP models by measuring how well token-level saliency rationales align with human-annotated ground truth rationales and model predictions across sentiment analysis, semantic textual similarity, and machine reading comprehension tasks. Use when the user wants to benchmark on Fine-grained Interpretability Benchmark (SA/STS/MRC), or asks about evaluating this task. Reports Token-F1, MAP.

qhjqhj00 bb1239b 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/fine-grained-interpretability-eval commit bb1239b1d1

Frequently asked questions

npx skillmds add qhjqhj00/fine-grained-interpretability-eval