4 papers · 1 filter
CLORE: Content-Level Optimization for Reasoning Efficiency
Yuyang Wu, Qiyao Xue, Guanxing Lu +4
Reinforcement learning post-training has improved the reasoning ability of large language models, but often produces unnecessarily long, repetitive, or semantically opaque reasonin…
Can Agents Price a Reaction? Evaluating LLMs on Chemical Cost Reasoning
Yuyang Wu, Yue Huang, Shuaike Shen +8
Large Language Models (LLMs) have become increasingly capable as tool-using agents, with benchmarks spanning diverse general agentic tasks. Yet rigorous evaluation of scientific to…
Reasoning Path and Latent State Analysis for Multi-view Visual Spatial Reasoning: A Cognitive Science Perspective
Qiyao Xue, Weichen Liu, Shiqi Wang +3
Spatial reasoning is a core aspect of human intelligence that allows perception, inference and planning in 3D environments. However, current vision-language models (VLMs) struggle…
Spatial Reasoning in Multimodal Large Language Models: A Survey of Tasks, Benchmarks and Methods
Weichen Liu, Qiyao Xue, Haoming Wang +3
Spatial reasoning, which requires ability to perceive and manipulate spatial relationships in the 3D world, is a fundamental aspect of human intelligence, yet remains a persistent…