Greek LLM Benchmark Eval

Evaluates open-source (Llama-70b) and closed-source (GPT-4o mini) LLMs across seven distinct NLP tasks in Modern Greek. It probes capabilities in classification, sequence labeling, text generation, and machine translation to assess model performance in a lesser-resourced language setting. Use when the user wants to benchmark on SemEval-2020 Task 12 (OffensEval-2020 Greek), Greek Native Corpus (GNC), Global Voices Greek MT Corpus, Areios Pagos Legal Summarization Corpus, University Help Desk Intent Classification Dataset, Greek NER Annotated Dataset, Greek Treebank (POS), or asks about evaluating this task. Reports macro-F1, BERTScore F1.

qhjqhj00 1fe5784 4.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/greek-llm-benchmark-eval commit 1fe5784828

Frequently asked questions

npx skillmds add qhjqhj00/greek-llm-benchmark-eval