1 citations · 1 across the 4 of their papers we have counts for
5 papers · 1 filter
ShapeGen: Towards High-Quality 3D Shape Synthesis
Yangguang Li, Xianglong He, Zi-Xin Zou +4
Inspired by generative paradigms in image and video, 3D shape generation has made notable progress, enabling the rapid synthesis of high-fidelity 3D assets from a single image. How…
Cut2Next: Generating Next Shot via In-Context Tuning
Jingwen He, Hongbo Liu, Jiajun Li +4
Effective multi-shot generation demands purposeful, film-like transitions and strict cinematic continuity. Current methods, however, often prioritize basic visual consistency, negl…
TCFormer: Visual Recognition via Token Clustering Transformer
Wang Zeng, Sheng Jin, Lumin Xu +5
Transformers are widely used in computer vision areas and have achieved remarkable success. Most state-of-the-art approaches split images into regular grids and represent each grid…
Hulk: A Universal Knowledge Translator for Human-Centric Tasks
Yizhou Wang, Yixuan Wu, Weizhen He +8
Human-centric perception tasks, e.g., pedestrian detection, skeleton-based action recognition, and pose estimation, have wide industrial applications, such as metaverse and sports…
GUPNet++: Geometry Uncertainty Propagation Network for Monocular 3D Object Detection
Yan Lu, Xinzhu Ma, Lei Yang +6
Geometry plays a significant role in monocular 3D object detection. It can be used to estimate object depth by using the perspective projection between object's physical size and 2…