7 papers
LXD-SLAM: LiDAR+X Dense SLAM with Configurable Sensor Combinations
Zhong Wang, Lin Zhang, Linfei Li +4
Simultaneous Localization and Mapping (SLAM) is essential for autonomous systems, yet achieving reliable, globally consistent pose estimation and dense mapping in complex environme…
RaysUp: Ultra-light Universal Feature Upsampling via Geometry-Aware Ray Representation
Yuchuan Ding, Linfei Li, Lin Zhang +1
Pre-trained Vision Foundation Models (VFMs) have become central to modern computer vision due to their powerful semantic representations and strong generalization ability. However,…
GS3LAM: Gaussian Semantic Splatting SLAM
Linfei Li, Lin Zhang, Zhong Wang +1
Recently, the multi-modal fusion of RGB, depth, and semantics has shown great potential in dense Simultaneous Localization and Mapping (SLAM). However, a prerequisite for generatin…
RealVLG-R1: A Large-Scale Real-World Visual-Language Grounding Benchmark for Robotic Perception and Manipulation
Linfei Li, Lin Zhang, Ying Shen
Visual-language grounding aims to establish semantic correspondences between natural language and visual entities, enabling models to accurately identify and localize target object…
Representing Sounds as Neural Amplitude Fields: A Benchmark of Coordinate-MLPs and A Fourier Kolmogorov-Arnold Framework
Linfei Li, Lin Zhang, Zhong Wang +3
Although Coordinate-MLP-based implicit neural representations have excelled in representing radiance fields, 3D shapes, and images, their application to audio signals remains under…
SmartSplat: Feature-Smart Gaussians for Scalable Compression of Ultra-High-Resolution Images
Linfei Li, Lin Zhang, Zhong Wang +1
Recent advances in generative AI have accelerated the production of ultra-high-resolution visual content, posing significant challenges for efficient compression and real-time deco…