Hybridrag Bench Eval

Evaluates retrieval-augmented models' ability to perform multi-hop reasoning over hybrid knowledge (unstructured text and knowledge graphs) using time-framed, external scientific literature to prevent parametric memorization. Use when the user wants to benchmark on Arxiv-AI, Arxiv-CY, Arxiv-BIO, or asks about evaluating this task. Reports accuracy.

qhjqhj00 2a781f0 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/hybridrag-bench-eval commit 2a781f0ad7

Frequently asked questions

npx skillmds add qhjqhj00/hybridrag-bench-eval