12 citations · 12 across the 3 of their papers we have counts for
4 papers · 1 filter
RaysUp: Ultra-light Universal Feature Upsampling via Geometry-Aware Ray Representation
Yuchuan Ding, Linfei Li, Lin Zhang +1
Pre-trained Vision Foundation Models (VFMs) have become central to modern computer vision due to their powerful semantic representations and strong generalization ability. However,…
GS3LAM: Gaussian Semantic Splatting SLAM
Linfei Li, Lin Zhang, Zhong Wang +1
Recently, the multi-modal fusion of RGB, depth, and semantics has shown great potential in dense Simultaneous Localization and Mapping (SLAM). However, a prerequisite for generatin…
RealVLG-R1: A Large-Scale Real-World Visual-Language Grounding Benchmark for Robotic Perception and Manipulation
Linfei Li, Lin Zhang, Ying Shen
Visual-language grounding aims to establish semantic correspondences between natural language and visual entities, enabling models to accurately identify and localize target object…
SmartSplat: Feature-Smart Gaussians for Scalable Compression of Ultra-High-Resolution Images
Linfei Li, Lin Zhang, Zhong Wang +1
Recent advances in generative AI have accelerated the production of ultra-high-resolution visual content, posing significant challenges for efficient compression and real-time deco…