2 citations · 2 across the 14 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
SlerpFlow: Spherical Trajectory Correction for Rectified Flow Inversion
Wenbin Duan, Yan Shu, Zhuoyuan Fu +4
Rectified-flow-based diffusion transformers, particularly FLUX, have demonstrated outstanding performance in high-quality image generation. However, achieving fast and accurate inv…
cs.CV2026
Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models
Zhuoyuan Fu, Zeshang Li, Yiqiong Zhang +5
While Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in 2D medical image understanding, their extension to 3D volumetric imaging remains hindered by…
cs.CV2021
GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions
Chenfei Wu, Lun Huang, Qianxi Zhang +5
Generating videos from text is a challenging task due to its high computational requirements for training and infinite possible answers for evaluation. Existing works typically exp…