2 papers
cs.AI2026
The Illusion of AI Expertise Under Uncertainty: Navigating Elusive Ground Truth via a Probabilistic Paradigm
Aparna Elangovan, Lei Xu, Mahsa Elyasi +8
Benchmarking the capabilities of AI systems, including Large Language Models (LLMs) and Vision Models, typically ignores the impact of uncertainty in the underlying ground truth an…
cs.LG2025
Exploring Human-AI Conceptual Alignment through the Prism of Chess
Semyon Lomasov, Judah Goldfeder, Mehmet Hamza Erol +5
Do AI systems truly understand human concepts or merely mimic surface patterns? We investigate this through chess, where human creativity meets precise strategic concepts. Analyzin…