6 papers
Efficient RWKV-based Representation Learning for 3D Point Clouds
Yun Liu, Xuefeng Yan, Liangliang Nan +5
The recent receptance weighted key value (RWKV) model combines RNN-style recurrence, offering a linear-complexity alternative to Transformers' quadratic self-attention for modeling…
ClickSeg3D: Few-Click Interactive Segmentation via Semantic Embeddings
Xueyang Kang, Zijian Yu, Kourosh Khoshelham +1
Interactive segmentation allows efficient label generation by leveraging user-provided clicks to progressively refine predictions, which is critical when fully supervised labels ar…
AsyncBEV: Cross-modal Flow Alignment in Asynchronous 3D Object Detection
Shiming Wang, Holger Caesar, Liangliang Nan +1
In autonomous driving, multi-modal perception tasks like 3D object detection typically rely on well-synchronized sensors, both at training and inference. However, despite the use o…
The Overlooked Value of Test-time Reference Sets in Visual Place Recognition
Mubariz Zaffar, Liangliang Nan, Sebastian Scherer +1
Given a query image, Visual Place Recognition (VPR) is the task of retrieving an image of the same place from a reference database with robustness to viewpoint and appearance chang…
NeuSEditor: From Multi-View Images to Text-Guided Neural Surface Edits
Nail Ibrahimli, Julian F. P. Kooij, Liangliang Nan
Implicit surface representations are valued for their compactness and continuity, but they pose significant challenges for editing. Despite recent advancements, existing methods of…
VoteFlow: Enforcing Local Rigidity in Self-Supervised Scene Flow
Yancong Lin, Shiming Wang, Liangliang Nan +2
Scene flow estimation aims to recover per-point motion from two adjacent LiDAR scans. However, in real-world applications such as autonomous driving, points rarely move independent…