1 paper
Claas Beger, Ryan Yi, Shuhao Fu +5
OpenAI's o3-preview reasoning model exceeded human accuracy on the ARC-AGI-1 benchmark, but does that mean state-of-the-art models recognize and reason with the abstractions the be…