Flip Benchmark Eval

Evaluates the ability of large protein language models to predict protein fitness under constrained, low-data scenarios. It probes mutation-level generalization, overfitting risks, and the impact of model depth and structural information on predictive accuracy across diverse protein families. Use when the user wants to benchmark on FLIP benchmark, or asks about evaluating this task. Reports MSE.

qhjqhj00 e831c28 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/flip-benchmark-eval commit e831c28314

Frequently asked questions

npx skillmds add qhjqhj00/flip-benchmark-eval