6 citations · 6 across the 2 of their papers we have counts for
1 paper · 1 filter
Claas Beger, Ryan Yi, Shuhao Fu +5
OpenAI's o3-preview reasoning model exceeded human accuracy on the ARC-AGI-1 benchmark, but does that mean state-of-the-art models recognize and reason with the abstractions the be…