Materials LLM Probe Eval

Evaluates large language models' ability to retrieve materials science knowledge and predict continuous physical properties. It probes the fundamental asymmetry in LLM behavior between symbolic tasks (classification, link prediction) and numerical regression tasks, assessing how fine-tuning affects accuracy and output consistency across modalities. Use when the user wants to benchmark on MatKG, Crystal System Classification, Bandgap Prediction, Dielectric Constant Prediction, or asks about evaluating this task. Reports RMSE, Top-1 accuracy.

qhjqhj00 2adc0eb 3.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/materials-llm-probe-eval commit 2adc0eb6cd

Frequently asked questions

npx skillmds add qhjqhj00/materials-llm-probe-eval