2 papers
cs.CV2026
Concentrate After Imagination: Text-Conditioned Evidence Grounding for Partially Relevant Video Retrieval
Shuaiqi Cheng, Siyu You, Yanbi Wu +3
Partially Relevant Video Retrieval (PRVR) retrieves untrimmed videos when queries describe only short moments. Although recent methods improve local representations, uncertainty mo…
cs.LG2026
Do VLMs Share Safety Neurons Across Modalities?
Jiaxuan Li, Jiahao Zhang, Duc Minh Vo +3
Vision-language models (VLMs) can comply with harmful requests delivered through images, even when their LLM backbones would refuse the same content in text. While prior work chara…