collaborators

8 papers

cs.CV2026

When Model Priors Conflict with Visual Evidence: Mitigating Commonsense-Driven Hallucinations by Selective Prior Calibration

Kesheng Chen, Yamin Hu, Wenjian Luo

In vision--language models, commonsense-driven hallucination (CDH) occurs when a model's commonsense prior overrides clear visual evidence of an atypical state. For example, a mode…

cs.IR2026

SentAttack: A Sentence-Level Black-Box Adversarial Attack Method for Dense Retrieval Models

Luping Wei, Yamin Hu, Sihan Shang +2

Retrieval-Augmented Generation (RAG) systems typically consist of a dense retrieval (DR) model for initial retrieval and a neural ranking model (NRM) for re-ranking.Existing robust…

cs.LG2026

Rethinking Backdoor Adversarial Unlearning through the Lens of Catastrophic Forgetting in Continual Learning

Zhenqian Zhu, Yamin Hu, Yujiang Liu +5

Existing studies reveal that current backdoor defenses exhibit limited robustness and often fail against specific types of attacks. More concerningly, prevailing safety tuning stra…

cs.CR2026

From Parameters to Feature Space: Task Arithmetic for Backdoor Mitigation in Model Merging

Zhenqian Zhu, Yamin Hu, Yiya Diao +3

Model merging (MM) has gained significant attention as a cost-effective approach to integrate multiple task-specific models into a unified model. However, recent work reveals that…

cs.AI2026

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization

Jiahao Huang, Peilan Xu, Xiaoya Nan +1

Automating operations research (OR) with large language models (LLMs) remains limited by hand-crafted reasoning--execution workflows. Complex OR tasks require adaptive coordination…

cs.CL2026

CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks

Peiqin Lin, Chenyang Lyu, Wenjiang Luo +22

Large language models (LLMs) are now deployed worldwide, inspiring a surge of benchmarks that measure their multilingual and multicultural abilities. However, these benchmarks prio…