2 citations · 2 across the 9 of their papers we have counts for
12 papers
MonoSE(3)-Diffusion: A Monocular SE(3) Diffusion Framework for Robust Camera-to-Robot Pose Estimation
Kangjian Zhu, Haobo Jiang, Yigong Zhang +3
We propose MonoSE(3)-Diffusion, a monocular SE(3) diffusion framework that formulates markerless, image-based robot pose estimation as a conditional denoising diffusion process. Th…
NAIPv2: Debiased Pairwise Learning for Efficient Paper Quality Estimation
Penghai Zhao, Jinyu Tian, Qinghua Xing +5
The ability to estimate the quality of scientific papers is central to how both humans and AI systems will advance scientific knowledge in the future. However, existing LLM-based e…
AGSwap: Overcoming Category Boundaries in Object Fusion via Adaptive Group Swapping
Zedong Zhang, Ying Tai, Jianjun Qian +2
Fusing cross-category objects to a single coherent object has gained increasing attention in text-to-image (T2I) generation due to its broad applications in virtual reality, digita…
WeatherCycle: Unpaired Multi-Weather Restoration via Color Space Decoupled Cycle Learning
Wenxuan Fang, Jiangwei Weng, Jianjun Qian +2
Unsupervised image restoration under multi-weather conditions remains a fundamental yet underexplored challenge. While existing methods often rely on task-specific physical priors,…
Warm Chat: Diffuse Emotion-aware Interactive Talking Head Avatar with Tree-Structured Guidance
Haijie Yang, Zhenyu Zhang, Hao Tang +2
Generative models have advanced rapidly, enabling impressive talking head generation that brings AI to life. However, most existing methods focus solely on one-way portrait animati…
Dual-Perspective United Transformer for Object Segmentation in Optical Remote Sensing Images
Yanguang Sun, Jiexi Yan, Jianjun Qian +3
Automatically segmenting objects from optical remote sensing images (ORSIs) is an important task. Most existing models are primarily based on either convolutional or Transformer fe…