2 papers
cs.AI2025
FormulaOne: Measuring the Depth of Algorithmic Reasoning Beyond Competitive Programming
Gal Beniamini, Yuval Dor, Alon Vinnikov +10
Frontier AI models demonstrate formidable breadth of knowledge. But how close are they to true human -- or superhuman -- expertise? Genuine experts can tackle the hardest problems…
cs.SD2025
Summary of the NOTSOFAR-1 Challenge: Highlights and Learnings
Igor Abramovski, Alon Vinnikov, Shalev Shaer +4
The first Natural Office Talkers in Settings of Far-field Audio Recordings (NOTSOFAR-1) Challenge is a pivotal initiative that sets new benchmarks by offering datasets more represe…