activity
20242026
collaborators
Showing cs.AIShow all

6 papers · 1 filter

cs.AI2026

HippoSpark: An On-Demand Experience System for LLM Reasoning

Jingyao Liu, Danling Meng, Chen Huang +5

Distilling historical trajectories into reusable experience to enhance future problem-solving has become a focal point of recent LLM research. However, existing methods predominant…

cs.AI2026

Bilevel Optimization of Agent Skills via Monte Carlo Tree Search

Chenyi Huang, Haoting Zhang, Jingxu Xu +2

Agent \texttt{skills} are structured collections of instructions, tools, and supporting resources that help large language model (LLM) agents perform particular classes of tasks. E…

cs.AI2026

Beyond Prompt: Fine-grained Simulation of Cognitively Impaired Standardized Patients via Stochastic Steering

Weikang Zhang, Zimo Zhu, Zhichuan Yang +3

Simulating Standardized Patients with cognitive impairment offers a scalable and ethical solution for clinical training. However, existing methods rely on discrete prompt engineeri…

cs.AI2026

Towards Proactive Information Probing: Customer Service Chatbots Harvesting Value from Conversation

Chen Huang, Zitan Jiang, Changyi Zou +2

Customer service chatbots are increasingly expected to serve not merely as reactive support tools for users, but as strategic interfaces for harvesting high-value information and b…

cs.AI2025

Beyond Solving Math Quiz: Evaluating the Ability of Large Reasoning Models to Ask for Information

Youcheng Huang, Bowen Qin, Chen Huang +3

Large Reasoning Models (LRMs) have demonstrated remarkable problem-solving abilities in mathematics, as evaluated by existing benchmarks exclusively on well-defined problems. Howev…

cs.AI2025

ELABORATION: A Comprehensive Benchmark on Human-LLM Competitive Programming

Xinwei Yang, Zhaofeng Liu, Chen Huang +4

While recent research increasingly emphasizes the value of human-LLM collaboration in competitive programming and proposes numerous empirical methods, a comprehensive understanding…