Regen Quality Eval

Evaluates whether increasing the length of a user's purchase history context improves recommendation quality for LLM-based agents. It probes the saturation point of personalization reasoning and the cost-efficiency trade-off of context length. Use when the user wants to benchmark on REGEN, or asks about evaluating this task. Reports quality scores.

qhjqhj00 36965b1 2.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/regen-quality-eval commit 36965b1f45

Frequently asked questions

npx skillmds add qhjqhj00/regen-quality-eval