Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
PhysElite: How Far Are LLMs from Solving Olympiad-Level Physics Problems?
Ruoran Xu, Wending Gao, Liyunfeng Chen +5
Understanding how (multimodal) large language models perform on physics problems requires benchmarks that reflect the difficulty and breadth of expert-level physical reasoning. Exi…
cs.AI2026
FormalAnalyticGeo: A Neural-Symbolic Based Framework for Multimodal Analytic Geometry Problem Generation
Ruoran Xu, Wending Gao, Xiaoqing Kang +1
Math reasoning has achieved significant progress with the rapid advancement of Multimodal Large Language Models (MLLMs), however analytic geometry remains largely underexplored, pr…