Exposing Assumptions in AI Benchmarks through Cognitive Modelling

Jonathan H. Rystr{\o}m; Kenneth C. Enevoldsen

arXiv:2409.16849·cs.AI·September 26, 2024

Exposing Assumptions in AI Benchmarks through Cognitive Modelling

Jonathan H. Rystr{\o}m, Kenneth C. Enevoldsen

PDF

Open Access

TL;DR

This paper introduces a method using cognitive models, specifically Structural Equation Models, to reveal assumptions in AI benchmarks, aiming to improve their validity and guide better dataset development.

Contribution

It presents a novel framework for exposing implicit assumptions in AI benchmarks through explicit cognitive modeling, enhancing theoretical grounding and transparency.

Findings

01

Identifies hidden assumptions in cultural AI benchmarks

02

Demonstrates how cognitive modeling can guide dataset development

03

Provides a framework for more rigorous AI evaluation

Abstract

Cultural AI benchmarks often rely on implicit assumptions about measured constructs, leading to vague formulations with poor validity and unclear interrelations. We propose exposing these assumptions using explicit cognitive models formulated as Structural Equation Models. Using cross-lingual alignment transfer as an example, we show how this approach can answer key research questions and identify missing datasets. This framework grounds benchmark construction theoretically and guides dataset development to improve construct measurement. By embracing transparency, we move towards more rigorous, cumulative AI evaluation science, challenging researchers to critically examine their assessment foundations.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsExplainable Artificial Intelligence (XAI)