5 papers
FlowDC: Flow-Based Decoupling-Decay for Complex Image Editing
Yilei Jiang, Zhen Wang, Yanghao Wang +4
With the surge of pre-trained text-to-image flow matching models, text-based image editing performance has gained remarkable improvement, especially for \underline{simple editing}…
Multimodal Conditional MeshGAN for Personalized Aneurysm Growth Prediction
Long Chen, Ashiv Patel, Mengyun Qiao +8
Personalized, accurate prediction of aortic aneurysm progression is essential for timely intervention but remains challenging due to the need to model both subtle local deformation…
Noise Matters: Optimizing Matching Noise for Diffusion Classifiers
Yanghao Wang, Long Chen
Although today's pretrained discriminative vision-language models (e.g., CLIP) have demonstrated strong perception abilities, such as zero-shot image classification, they also suff…
SleepGMUformer: A gated multimodal temporal neural network for sleep staging
Chenjun Zhao, Xuesen Niu, Xinglin Yu +4
Sleep staging is a key method for assessing sleep quality and diagnosing sleep disorders. However, current deep learning methods face challenges: 1) postfusion techniques ignore th…
MSF: Efficient Diffusion Model Via Multi-Scale Latent Factorize
Haohang Xu, Longyu Chen, Yichen Zhang +2
While diffusion-based generative models have made significant strides in visual content creation, conventional approaches face computational challenges, especially for high-resolut…