3 papers
cs.SE2026
Teaching LLMs to Learn Tool Trialing and Execution through Environment Interaction
Xingjie Gao, Pengcheng Huang, Zhenghao Liu +6
Equipping Large Language Models (LLMs) with external tools enables them to solve complex real-world problems. However, the robustness of existing methods remains a critical challen…
cs.CL2026
Long-Chain Reasoning Distillation via Adaptive Prefix Alignment
Zhenghao Liu, Zhuoyang Wu, Xinze Li +6
Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, particularly in solving complex mathematical problems. Recent studies show that distilling long re…
cs.AI2025
Empirical Analysis of Decoding Biases in Masked Diffusion Models
Pengcheng Huang, Tianming Liu, Zhenghao Liu +5
Masked diffusion models (MDMs), which leverage bidirectional attention and a denoising process, are narrowing the performance gap with autoregressive models (ARMs). However, their…