1 paper
Jonathan H. Rystrøm, Kenneth C. Enevoldsen
Cultural AI benchmarks often rely on implicit assumptions about measured constructs, leading to vague formulations with poor validity and unclear interrelations. We propose exposin…