Hrt Multi Scale Eval

Evaluates a wavelet-inspired multi-resolution transformer across five linguistic granularities, from character morphology to discourse reasoning. It probes the model's capacity for hierarchical composition, long-range dependency modeling up to 16K tokens, and computational efficiency compared to standard transformers. Use when the user wants to benchmark on WikiMorpho, IMDB-BYTE, WordNet Hypernymy (WN-Hyper), SentEval Word Similarity Suite, GLUE Benchmark, SuperGLUE, Long Range Arena (LRA), WikiText-103, DiscoEval Benchmark, NarrativeQA, or asks about evaluating this task. Reports Accuracy, F1-score, Perplexity (PPL), Normalized Efficiency Score (NES).

qhjqhj00 df73b53 5.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/hrt-multi-scale-eval commit df73b53a21

Frequently asked questions

npx skillmds add qhjqhj00/hrt-multi-scale-eval