2 citations · 2 across the 3 of their papers we have counts for
4 papers
AlignVid: Training-Free Attention Scaling for Semantic Fidelity in Text-Guided Image-to-Video Generation
Yexin Liu, Wen-Jie Shu, Zile Huang +6
Text-guided image-to-video generation has made substantial progress, yet it still struggles to execute text-specified edits that require substantial changes to a reference image (\…
Incomplete Modality Disentangled Representation for Ophthalmic Disease Grading and Diagnosis
Chengzhi Liu, Zile Huang, Zhe Chen +6
Ophthalmologists typically require multimodal data sources to improve diagnostic accuracy in clinical decisions. However, due to medical device shortages, low-quality data and data…
Better Sampling, towards Better End-to-end Small Object Detection
Zile Huang, Chong Zhang, Mingyu Jin +3
While deep learning-based general object detection has made significant strides in recent years, the effectiveness and efficiency of small object detection remain unsatisfactory. T…
MTSA-SNN: A Multi-modal Time Series Analysis Model Based on Spiking Neural Network
Chengzhi Liu, Zheng Tao, Zihong Luo +1
Time series analysis and modelling constitute a crucial research area. Traditional artificial neural networks struggle with complex, non-stationary time series data due to high com…