1 citations · 1 across the 4 of their papers we have counts for
5 papers
Milestones over Outcome: Unlocking Geometric Reasoning with Sub-Goal Verifiable Reward
Jianlong Chen, Daocheng Fu, Shengze Xu +6
Multimodal Large Language Models (MLLMs) struggle with complex geometric reasoning, largely because "black box" outcome-based supervision fails to distinguish between lucky guesses…
OS Agents: A Survey on MLLM-based Agents for General Computing Devices Use
Xueyu Hu, Tao Xiong, Biao Yi +26
The dream to create AI assistants as capable and versatile as the fictional J.A.R.V.I.S from Iron Man has long captivated imaginations. With the evolution of (multi-modal) large la…
AB-Cache: Training-Free Acceleration of Diffusion Models via Adams-Bashforth Cached Feature Reuse
Zichao Yu, Zhen Zou, Guojiang Shao +6
Diffusion models have demonstrated remarkable success in generative tasks, yet their iterative denoising process results in slow inference, limiting their practicality. While exist…
DGNN: A Neural PDE Solver Induced by Discontinuous Galerkin Methods
Guanyu Chen, Shengze Xu, Dong Ni +1
We propose a general framework for the Discontinuous Galerkin-induced Neural Network (DGNN), inspired by the Interior Penalty Discontinuous Galerkin Method (IPDGM). In this approac…
An Improved Optimal Proximal Gradient Algorithm for Non-Blind Image Deblurring
Qingsong Wang, Shengze Xu, Xiaojiao Tong +1
Image deblurring remains a central research area within image processing, critical for its role in enhancing image quality and facilitating clearer visual representations across di…