8 papers
CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation
Yaoning Yu, Kai-Min Chang, Ye Yu +3
Online credit card discussions provide a natural setting for studying how consumers communicate about financial products. Simulating these discussions requires more than just gener…
MiroBench: Benchmarking Realism in Agentic Simulation of Real-world Discussions
Yaoning Yu, Ye Yu, Haojing Luo +1
LLM agents are increasingly used to simulate real world interactions, but it remains unclear whether simulated behaviors preserve the content patterns and interaction dynamics of r…
Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation
Ye Yu, Xiaopeng Yuan, Haibo Jin +3
Recent advances in LLM agents enable systems that autonomously refine workflows, accumulate reusable skills, self-train their underlying models, and maintain persistent memory. How…
Now You Hear Me: Audio Narrative Attacks Against Large Audio-Language Models
Ye Yu, Haibo Jin, Yaoning Yu +2
Large audio-language models increasingly operate on raw speech inputs, enabling more seamless integration across domains such as voice assistants, education, and clinical triage. T…
SIPDO: Closed-Loop Prompt Optimization via Synthetic Data Feedback
Yaoning Yu, Ye Yu, Peiyan Zhang +3
Prompt quality plays a critical role in the performance of large language models (LLMs), motivating a growing body of work on prompt optimization. Most existing methods optimize pr…
Large Language Model-based Data Science Agent: A Survey
Ke Chen, Peiran Wang, Yaoning Yu +2
The rapid advancement of Large Language Models (LLMs) has driven novel applications across diverse domains, with LLM-based agents emerging as a crucial area of exploration. This su…