collaborators

6 papers

cs.CV2026

UniEmo: Unifying Emotional Understanding and Generation with Learnable Expert Queries

Yijie Zhu, Lingsen Zhang, Zitong Yu +3

Emotional understanding and generation are often treated as separate tasks, yet they are inherently complementary and can mutually enhance each other. In this paper, we propose the…

cs.CV2026

PHASE-Net: Physics-Grounded Harmonic Attention System for Efficient Remote Photoplethysmography Measurement

Bo Zhao, Dan Guo, Junzhe Cao +5

Remote photoplethysmography (rPPG) measurement enables non-contact physiological monitoring but suffers from accuracy degradation under head motion and illumination changes. Existi…

cs.CV2025

MedIQA: A Scalable Foundation Model for Prompt-Driven Medical Image Quality Assessment

Siyi Xun, Yue Sun, Jingkun Chen +5

Rapid advances in medical imaging technology underscore the critical need for precise and automated image quality assessment (IQA) to ensure diagnostic accuracy. Existing medical I…

cs.CV2025

AdaMHF: Adaptive Multimodal Hierarchical Fusion for Survival Prediction

Shuaiyu Zhang, Xun Lin, Rongxiang Zhang +5

The integration of pathologic images and genomic data for survival analysis has gained increasing attention with advances in multimodal learning. However, current methods often ign…

cs.CV2025

TC-GS: Tri-plane based compression for 3D Gaussian Splatting

Taorui Wang, Zitong Yu, Yong Xu

Recently, 3D Gaussian Splatting (3DGS) has emerged as a prominent framework for novel view synthesis, providing high fidelity and rapid rendering speed. However, the substantial da…

cs.CV2025

MoEdit: On Learning Quantity Perception for Multi-object Image Editing

Yanfeng Li, Kahou Chan, Yue Sun +6

Multi-object images are prevalent in various real-world scenarios, including augmented reality, advertisement design, and medical imaging. Efficient and precise editing of these im…