1 citations · 1 across the 5 of their papers we have counts for
10 papers
Teaching LLMs to Learn Tool Trialing and Execution through Environment Interaction
Xingjie Gao, Pengcheng Huang, Zhenghao Liu +6
Equipping Large Language Models (LLMs) with external tools enables them to solve complex real-world problems. However, the robustness of existing methods remains a critical challen…
Long-Chain Reasoning Distillation via Adaptive Prefix Alignment
Zhenghao Liu, Zhuoyang Wu, Xinze Li +6
Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, particularly in solving complex mathematical problems. Recent studies show that distilling long re…
Revealing the Attention Floating Mechanism in Masked Diffusion Models
Xin Dai, Pengcheng Huang, Zhenghao Liu +6
Masked diffusion models (MDMs), which leverage bidirectional attention and a denoising process, are narrowing the performance gap with autoregressive models (ARMs). However, their…
Enhancing Long-Chain Reasoning Distillation through Error-Aware Self-Reflection
Zhuoyang Wu, Xinze Li, Zhenghao Liu +7
Large Language Models (LLMs) have exhibited strong reasoning capabilities and achieved remarkable performance in mathematical problem-solving tasks. Recently, distilling reasoning…
LISRec: Modeling User Preferences with Learned Item Shortcuts for Sequential Recommendation
Haidong Xin, Zhenghao Liu, Sen Mei +7
User-item interaction histories are pivotal for sequential recommendation systems but often include noise, such as unintended clicks or actions that fail to reflect genuine user pr…
Enhancing the Patent Matching Capability of Large Language Models via the Memory Graph
Qiushi Xiong, Zhipeng Xu, Zhenghao Liu +6
Intellectual Property (IP) management involves strategically protecting and utilizing intellectual assets to enhance organizational innovation, competitiveness, and value creation.…