Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Progressive Representation Learning for Multimodal Sentiment Analysis with Incomplete Modalities
Jindi Bao, Jianjun Qian, Mengkai Yan +1
Multimodal Sentiment Analysis (MSA) seeks to infer human emotions by integrating textual, acoustic, and visual cues. However, existing approaches often rely on all modalities are c…
cs.CV2025
MonoSE(3)-Diffusion: A Monocular SE(3) Diffusion Framework for Robust Camera-to-Robot Pose Estimation
Kangjian Zhu, Haobo Jiang, Yigong Zhang +3
We propose MonoSE(3)-Diffusion, a monocular SE(3) diffusion framework that formulates markerless, image-based robot pose estimation as a conditional denoising diffusion process. Th…
cs.CV2025
Follow Your Motion: A Generic Temporal Consistency Portrait Editing Framework with Trajectory Guidance
Haijie Yang, Zhenyu Zhang, Hao Tang +2
Pre-trained conditional diffusion models have demonstrated remarkable potential in image editing. However, they often face challenges with temporal consistency, particularly in the…