collaborators

9 papers

cs.CV2026

AI-generated Images Challenge Visual Trust in High-risk Scenarios

Yi-Zhi Wang, Yichen Xiao, Linan Yue +5

Rapid advances in image generation are eroding the evidentiary value of visual content in settings where authenticity can affect public safety and personal reputation. Yet existing…

cs.CV2026

On Aligning Hierarchical Standardized Embedding for Audio-visual Generalized Zero-shot Learning

Zihan Zhang, Jie Hong, Siyuan Fan +2

Audio-visual Generalized Zero-shot Learning (AV-GZSL) is a challenging task that aims to classify both seen and unseen objects or scenes by integrating data from audio and visual m…

cs.CV2026

EpiAgent: An Agent-Centric System for Ancient Inscription Restoration

Shipeng Zhu, Ang Chen, Na Nie +3

Ancient inscriptions, as repositories of cultural memory, have suffered from centuries of environmental and human-induced degradation. Restoring their intertwined visual and textua…

cs.CV2026

MAIL++: Multi-Modal Bi-directional Agent Layer for Vision-Language Models

Kaixiang Chen, Pengfei Fang, Hui Xue

Adapting large vision-language models (VLMs) such as CLIP to downstream tasks remains challenging, as full fine-tuning is computationally prohibitive and prone to overfitting in lo…

cs.SI2026

IntervenSim: Intervention-Aware Social Network Simulation for Opinion Dynamics

Yunyao Zhang, Zuocheng Ying, Xinglang Zhang +5

LLM-based social network simulation introduces a new computational approach for modeling event evolution in complex online environments. However, existing methods typically simulat…

cs.CV2026

Text-Phase Synergy Network with Dual Priors for Unsupervised Cross-Domain Image Retrieval

Jing Yang, Hui Xue, Shipeng Zhu +1

This paper studies unsupervised cross-domain image retrieval (UCDIR), which aims to retrieve images of the same category across different domains without relying on labeled data. E…