collaborators

5 papers

cs.MM2026

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence

Han Hu, Dongheng Lin, Yuqi Hou +3

Localising multiple sound sources in visual scenes remains a fundamental challenge in multimodal perception due to an inherent circular dependency: separating mixed audio requires…

cs.CV2025

An Exploratory Study on Abstract Images and Visual Representations Learned from Them

Haotian Li, Jianbo Jiao

Imagine living in a world composed solely of primitive shapes, could you still recognise familiar objects? Recent studies have shown that abstract images-constructed by primitive s…

cs.CV2025

SHREC 2025: Protein surface shape retrieval including electrostatic potential

Taher Yacoub, Camille Depenveiller, Atsushi Tatsuma +25

This SHREC 2025 track dedicated to protein surface shape retrieval involved 9 participating teams. We evaluated the performance in retrieval of 15 proposed methods on a large datas…

cs.DB2025

VSAG: An Optimized Search Framework for Graph-based Approximate Nearest Neighbor Search

Xiaoyao Zhong, Haotian Li, Jiabao Jin +11

Approximate nearest neighbor search (ANNS) is a fundamental problem in vector databases and AI infrastructures. Recent graph-based ANNS algorithms have achieved high search accurac…

cs.CV2025

MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models

Xiao Lin, Zhining Liu, Ze Yang +10

Warning: This paper contains examples of harmful language and images. Reader discretion is advised. Recently, vision-language models have demonstrated increasing influence in moral…