13 papers
Yesterday's Shield, Today's Spear: A Self-Evolving Safety Guardrail in Production
Cong Ming, Jingyi Chen, Bin Liu +4
Deployed LLM safety guardrails are predominantly static: trained once and frozen at release, while new jailbreak techniques and previously un-addressed harmful categories emerge wi…
ANFI: Rethinking Neighbor Feature Interaction in Person Re-ID
Xulin Li, Yan Lu, Bin Liu +5
The paper proposes ANFI, an adaptive neighbor feature interaction method for person re-identification that jointly models affinity and discrepancy relations to mitigate the impact…
MFEN:Multi-Frequency Expert Network for Visible-Infrared Person Re-ID
Xulin Li, Yan Lu, Bin Liu +4
Visible-infrared person re-identification (VI-ReID) is challenging due to the large modality discrepancy between visible and infrared images. We contend that this discrepancy is la…
TextFake: Benchmarking AI-Generated Image Detection on Text-Rich Images
Yuning Zhang, Changtao Miao, Mingyu Liao +5
Recent AI-generated image (AIGI) detectors perform well on natural-image benchmarks, but their behavior on text-rich forgeries, such as fabricated screenshots, documents, and news…
Towards Anytime Retrieval: A Benchmark for Anytime Person Re-Identification
Xulin Li, Yan Lu, Bin Liu +6
In real applications, person re-identification (ReID) is expected to retrieve the target person at any time, including both daytime and nighttime, ranging from short-term to long-t…
Learning to Focus and Precise Cropping: A Reinforcement Learning Framework with Information Gaps and Grounding Loss for MLLMs
Xuanpu Zhao, Zhentao Tan, Dianmo Sheng +6
To enhance the perception and reasoning capabilities of multimodal large language models in complex visual scenes, recent research has introduced agent-based workflows. In these wo…