20 citations · 27 across the 5 of their papers we have counts for
1 paper · 2 filters
Claas Beger, Ryan Yi, Shuhao Fu +5
OpenAI's o3-preview reasoning model exceeded human accuracy on the ARC-AGI-1 benchmark, but does that mean state-of-the-art models recognize and reason with the abstractions the be…