2 papers
cs.CV2026
Who Drives the Probability Game of VLMs? A Temporal Causal Drive Evaluation Framework
Shuyao Xiao, Shengling Wang, Haoyu Niu +4
Vision-language models (VLMs) are increasingly evaluated on complex image and video understanding tasks, yet conventional metrics primarily assess final-answer quality and reveal l…
cs.CL2026
When Errors Become Memories: Causal Pathway Tracing in Multi-Turn Memory-Augmented LLMs
Shuyao Xiao, Shengling Wang, Xuan Chen +8
Long-term memory enables large language models (LLMs) to preserve and reuse information across interactions, but it can also turn localized errors into persistent risks. Existing w…