Multipl E Low Resource Eval

Evaluates large language models' ability to generate functionally correct code in low-resource programming languages (R and Racket). It probes how well in-context learning strategies and fine-tuning adapt pre-trained models to languages with limited training data and documentation. Use when the user wants to benchmark on MultiPL-E, or asks about evaluating this task. Reports pass@1.

qhjqhj00 cb1c5d4 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/multipl-e-low-resource-eval commit cb1c5d4f8f

Frequently asked questions

npx skillmds add qhjqhj00/multipl-e-low-resource-eval