Triplesumm Video Summarization Eval

Evaluates a model's ability to perform video summarization by predicting frame-level importance scores across visual, textual, and audio modalities. It probes the model's capacity for adaptive multimodal fusion and temporal dependency modeling to identify salient segments in long videos. Use when the user wants to benchmark on MoSu, Mr. HiSum, SumMe, TVSum, or asks about evaluating this task. Reports Kendall’s τ (kTau), Spearman’s ρ (sRho).

qhjqhj00 fbacfb2 4.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/triplesumm-video-summarization-eval commit fbacfb23ec

Frequently asked questions

npx skillmds add qhjqhj00/triplesumm-video-summarization-eval