LLM Pretrain Finetune Eval

Evaluates memory-efficient gradient compression optimizers against full-rank baselines during LLM pre-training and fine-tuning, measuring final model quality, convergence speed, memory footprint, and training throughput. Use when the user wants to benchmark on C4, MMLU, GLUE, or asks about evaluating this task. Reports Validation PPL, Accuracy.

qhjqhj00 e997925 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/llm-pretrain-finetune-eval commit e997925820

Frequently asked questions

npx skillmds add qhjqhj00/llm-pretrain-finetune-eval