3 papers
cs.CV2025
Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models
Jiawei Liang, Jianjie Huang, Xianghao Jiao +3
Multimodal large language models (MLLMs) have achieved strong vision-language performance, yet their token-level visual evidence remains difficult to inspect. Recent logit-lens att…
cs.CV2025
Physical Adversarial Camouflage through Gradient Calibration and Regularization
Jiawei Liang, Siyuan Liang, Jianjie Huang +3
The advancement of deep object detectors has greatly affected safety-critical fields like autonomous driving. However, physical adversarial camouflage poses a significant security…
cs.AI2025
SMA: Who Said That? Auditing Membership Leakage in Semi-Black-box RAG Controlling
Shixuan Sun, Siyuan Liang, Jianjie Huang +2
Retrieval-Augmented Generation (RAG) and its Multimodal Retrieval-Augmented Generation (MRAG) significantly improve the knowledge coverage and contextual understanding of Large Lan…