6 papers · 1 filter
HippoSpark: An On-Demand Experience System for LLM Reasoning
Jingyao Liu, Danling Meng, Chen Huang +5
Distilling historical trajectories into reusable experience to enhance future problem-solving has become a focal point of recent LLM research. However, existing methods predominant…
Bilevel Optimization of Agent Skills via Monte Carlo Tree Search
Chenyi Huang, Haoting Zhang, Jingxu Xu +2
Agent \texttt{skills} are structured collections of instructions, tools, and supporting resources that help large language model (LLM) agents perform particular classes of tasks. E…
Beyond Prompt: Fine-grained Simulation of Cognitively Impaired Standardized Patients via Stochastic Steering
Weikang Zhang, Zimo Zhu, Zhichuan Yang +3
Simulating Standardized Patients with cognitive impairment offers a scalable and ethical solution for clinical training. However, existing methods rely on discrete prompt engineeri…
Towards Proactive Information Probing: Customer Service Chatbots Harvesting Value from Conversation
Chen Huang, Zitan Jiang, Changyi Zou +2
Customer service chatbots are increasingly expected to serve not merely as reactive support tools for users, but as strategic interfaces for harvesting high-value information and b…
Beyond Solving Math Quiz: Evaluating the Ability of Large Reasoning Models to Ask for Information
Youcheng Huang, Bowen Qin, Chen Huang +3
Large Reasoning Models (LRMs) have demonstrated remarkable problem-solving abilities in mathematics, as evaluated by existing benchmarks exclusively on well-defined problems. Howev…
ELABORATION: A Comprehensive Benchmark on Human-LLM Competitive Programming
Xinwei Yang, Zhaofeng Liu, Chen Huang +4
While recent research increasingly emphasizes the value of human-LLM collaboration in competitive programming and proposes numerous empirical methods, a comprehensive understanding…