22 citations · 91 across the 40 of their papers we have counts for
40 papers · 1 filter
ToaSt: Token Channel Selection and Structured Pruning for Efficient ViT
Hyunchan Moon, Cheonjun Park, Steven L. Waslander
Vision Transformers (ViTs) have achieved remarkable success across various vision tasks, yet their deployment is often hindered by prohibitive computational costs. While structured…
Contour Refinement using Discrete Diffusion in Low Data Regime
Fei Yu Guan, Ian Keefe, Sophie Wilkinson +2
Boundary detection of irregular and translucent objects is an important problem with applications in medical imaging, environmental monitoring and manufacturing, where many of thes…
Complete Gaussian Splats from a Single Image with Denoising Diffusion Models
Ziwei Liao, Mohamed Sayed, Steven L. Waslander +3
Gaussian splatting typically requires dense observations of the scene and can fail to reconstruct occluded and unobserved areas. We propose a latent diffusion model to reconstruct…
Large Self-Supervised Models Bridge the Gap in Domain Adaptive Object Detection
Marc-Antoine Lavoie, Anas Mahmoud, Steven L. Waslander
The current state-of-the-art methods in domain adaptive object detection (DAOD) use Mean Teacher self-labelling, where a teacher model, directly derived as an exponential moving av…
PSA-SSL: Pose and Size-aware Self-Supervised Learning on LiDAR Point Clouds
Barza Nisar, Steven L. Waslander
Self-supervised learning (SSL) on 3D point clouds has the potential to learn feature representations that can transfer to diverse sensors and multiple downstream perception tasks.…
Active 6D Pose Estimation for Textureless Objects using Multi-View RGB Frames
Jun Yang, Wenjie Xue, Sahar Ghavidel +1
Estimating the 6D pose of textureless objects from RGB images is an important problem in robotics. Due to appearance ambiguities, rotational symmetries, and severe occlusions, sing…