LLM Forecasting Evaluation

Benchmark LLMs on real-world forecasting questions from Metaculus, comparing against human crowds and expert forecasters. Identifies which domains LLMs handle well and where they fall short relative to human intelligence.

adu2021 7fb8c4b 9.1 KB Updated

File contents

adu2021/skillxiv/tree/main/skills/skillxiv-v0.0.2-claude-opus-4.6/llm-forecasting-evaluation commit 7fb8c4bf28

Frequently asked questions

npx skillmds@latest add adu2021/llm-forecasting-evaluation