Bloom Empirical Eval

Evaluates BLOOM model variants against BERT-style and GPT-style baselines across diverse NLP tasks including text classification, question answering, zero/few-shot learning, multilingual transfer, and text generation. Use when the user wants to benchmark on GLUE, SQuAD, XNLI, MARC, Zero/FSL Benchmarks, or asks about evaluating this task. Reports accuracy.

qhjqhj00 b0fa6f5 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/bloom-empirical-eval commit b0fa6f532d

Frequently asked questions

npx skillmds add qhjqhj00/bloom-empirical-eval