2 citations · 6 across the 6 of their papers we have counts for
24 papers
DAug: Diffusion-based Channel Augmentation for Radiology Image Retrieval and Classification
Ying Jin, Zhuoran Zhou, Haoquan Fang +1
Medical image understanding requires meticulous examination of fine visual details, with particular regions requiring additional attention. While radiologists build such expertise…
Graph Canvas for Controllable 3D Scene Generation
Libin Liu, Shen Chen, Sen Jia +6
Spatial intelligence is foundational to AI systems that interact with the physical world, particularly in 3D scene generation and spatial comprehension. Current methodologies for 3…
SAMURAI: Adapting Segment Anything Model for Zero-Shot Visual Tracking with Motion-Aware Memory
Cheng-Yen Yang, Hsiang-Wei Huang, Wenhao Chai +2
The Segment Anything Model 2 (SAM 2) has demonstrated strong performance in object segmentation tasks but faces challenges in visual object tracking, particularly when managing cro…
GTA: Global Tracklet Association for Multi-Object Tracking in Sports
Jiacheng Sun, Hsiang-Wei Huang, Cheng-Yen Yang +2
Multi-object tracking in sports scenarios has become one of the focal points in computer vision, experiencing significant advancements through the integration of deep learning tech…
ScalingGaussian: Enhancing 3D Content Creation with Generative Gaussian Splatting
Shen Chen, Jiale Zhou, Zhongyu Jiang +4
The creation of high-quality 3D assets is paramount for applications in digital heritage preservation, entertainment, and robotics. Traditionally, this process necessitates skilled…
Boosting Online 3D Multi-Object Tracking through Camera-Radar Cross Check
Sheng-Yao Kuan, Jen-Hao Cheng, Hsiang-Wei Huang +6
In the domain of autonomous driving, the integration of multi-modal perception techniques based on data from diverse sensors has demonstrated substantial progress. Effectively surp…