6 papers
TrajWiki: Source-Grounded Memory Trajectories for Long-Horizon Dialogue Agents
Jingyu Sun, Yuyang Xue, Mingyang Li +9
Large language model agents have shown strong capabilities in generating coherent and contextually appropriate responses, yet robust long-horizon dialogue remains limited by the la…
PMMC: Prospective Multimodal Memory Compilation for Long-Term LVLM Agents
Jingyu Sun, Yan Lin, Yuyang Xue +10
Long-term memory is essential for LVLM agents to maintain consistency and integrate information across extended multimodal interactions. Existing agent memory systems, however, oft…
Stress Testing Concept Erasure with Large Language Model Agents
Yuyang Xue, Feng Chen, Zhihua Liu +4
Concept erasure aims to remove semantic concepts from a trained generative model and is increasingly important for responsible AI deployment. However, verifying whether a model has…
Seeing What Is Actually There: PriVE-Bench and PriVE-Tools for Counterfactual Evaluation of Agentic Visual Evidence in VLMs
Jingyu Sun, Jiachen Tu, Yuyang Xue +8
Vision-language models (VLMs) often answer visual questions using learned language and category priors rather than grounding their predictions in the image itself. Counterfactual i…
SWiFT: Soft-Mask Weight Fine-tuning for Bias Mitigation
Junyu Yan, Feng Chen, Yuyang Xue +4
Recent studies have shown that Machine Learning (ML) models can exhibit bias in real-world scenarios, posing significant challenges in ethically sensitive domains such as healthcar…
CRCE: Coreference-Retention Concept Erasure in Text-to-Image Diffusion Models
Yuyang Xue, Edward Moroshko, Feng Chen +3
Text-to-Image diffusion models can produce undesirable content that necessitates concept erasure. However, existing methods struggle with under-erasure, leaving residual traces of…