activity
20242026
collaborators

11 papers

cs.CV2026

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

Xuechao Zou, Shun Zhang, Kai Li +6

The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substantial challenges to AI safety. Ho…

cs.CV2026

Toward Stable Semi-Supervised Remote Sensing Segmentation via Co-Guidance and Co-Fusion

Yi Zhou, Xuechao Zou, Shun Zhang +7

Semi-supervised remote sensing (RS) image semantic segmentation offers a promising solution to alleviate the burden of exhaustive annotation, yet it fundamentally struggles with ps…

cs.RO2026

DST-Calib: A Dual-Path, Self-Supervised, Target-Free LiDAR-Camera Extrinsic Calibration Network

Zhiwei Huang, Yanwei Fu, Yi Zhou +3

LiDAR-camera extrinsic calibration is essential for multi-modal data fusion in robotic perception systems. However, existing approaches typically rely on handcrafted calibration ta…

cs.CV2025

Language as an Anchor: Preserving Relative Visual Geometry for Domain Incremental Learning

Shuyi Geng, Tao Zhou, Yi Zhou

A key challenge in Domain Incremental Learning (DIL) is to continually learn under shifting distributions while preserving knowledge from previous domains. Existing methods face a…

cs.CV2025

Curvilinear Structure-preserving Unpaired Cross-domain Medical Image Translation

Zihao Chen, Yi Zhou, Xudong Jiang +4

Unpaired image-to-image translation has emerged as a crucial technique in medical imaging, enabling cross-modality synthesis, domain adaptation, and data augmentation without costl…

cs.AI2025

Bridging the Gap in Ophthalmic AI: MM-Retinal-Reason Dataset and OphthaReason Model toward Dynamic Multimodal Reasoning

Ruiqi Wu, Yuang Yao, Tengfei Ma +6

Multimodal large language models (MLLMs) have recently demonstrated remarkable reasoning abilities with reinforcement learning paradigm. Although several multimodal reasoning model…