Ist Unbabel 2022 Qe Eval

Evaluates machine translation quality estimation (QE) by predicting human quality scores at the sentence level and identifying error locations at the word level. It also assesses the model's ability to generate faithful explanations for predicted errors. Use when the user wants to benchmark on IST-Unbabel 2022 QE Shared Task, or asks about evaluating this task. Reports Spearman's rank correlation, Matthew's correlation coefficient (MCC), Recall@K (R@K).

qhjqhj00 d3af055 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/ist-unbabel-2022-qe-eval commit d3af055edd

Frequently asked questions

npx skillmds add qhjqhj00/ist-unbabel-2022-qe-eval