2 papers
cs.CL2025
The Valley of Code Reasoning: Scaling Knowledge Distillation of Large Language Models
Muyu He, Muhammad Ali Shafique, Anand Kumar +2
Distilling the thinking traces of a Large Language Model (LLM) with reasoning capabilities into a smaller model has been proven effective. Yet, there is a scarcity of work done on…
cs.AI2025
Impatient Users Confuse AI Agents: High-fidelity Simulations of Human Traits for Testing Agents
Muyu He, Anand Kumar, Tsach Mackey +3
Despite rapid progress in building conversational AI agents, robustness is still largely untested. Small shifts in user behavior, such as being more impatient, incoherent, or skept…