AI Reliability

Use when evaluating the reliability (score consistency/stability) of an AI/ML selection tool — Concern 6 of Tippins, Oswald & McPhail (2021). Covers reliability as an absolute requirement, what stability means for AI scores, evidence that machine scoring can be as or more reliable than human scoring, the questionable reliability of facial-emotion analysis (including across skin tone, disability, and altered features), and confounds from individual differences in the data generated (e.g., extraversion/verbosity). Triggers: "reliability of an AI assessment", "are the scores stable", "facial emotion recognition reliability", "machine-scored interview reliability", "test-retest for AI hiring", "verbosity confound".

OpenMatter-Network Updated

File contents

OpenMatter-Network/agent-io-skills/tree/main/ai-selection-legal-ethical/skills/ai-reliability commit 5462a05419

Frequently asked questions

npx skillmds@latest add openmatter-network/ai-reliability