Possible Stories Ifsm Eval

Assesses whether large language models can generate story endings that align with free-form instructions provided alongside a narrative context. It measures both instruction-following accuracy and the model's ability to produce distinct endings for different instructions. Use when the user wants to benchmark on Possible Stories, or asks about evaluating this task. Reports IFSM.

qhjqhj00 1b6ffd3 3.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/possible-stories-ifsm-eval commit 1b6ffd37e3

Frequently asked questions

npx skillmds add qhjqhj00/possible-stories-ifsm-eval