activity
20162026
most citedHybVIO: Pushing the Limits of Real-time Visual-inertial Odometry

41 citations · 213 across the 84 of their papers we have counts for

collaborators

127 papers

cs.CV2026

Does Attention-Guided Masking Really Help Object Discovery in Object-Centric Learning?

Youliang Tao, Yanhua Han, Bin Zhao +3

Object-Centric Learning (OCL) aims to decompose images into objects without human annotations. A major family of mainstream methods uses Slot Attention to aggregate image features…

cs.GR2026

TileGS: Tile-Local Depth Binning for Gaussian Splatting Rasterization

Wei Tan, Matias Turkulainen, Lauri Ilola +2

Real-time 3D Gaussian Splatting (3DGS) achieves high rendering quality, but standard rasterization still traverses a globally sorted tile stream that creates long per-tile ranges a…

cs.CV2026

Compressing AI Traffic: Standardized Neural Network Coding of Visual-Token Representations in Split Vision-Language Inference

Reza Heidari, Hamed R. Tavakoli, Juho Kannala

When the visual encoder and the language decoder of a vision-language model (VLM) run on different compute nodes, the intermediate visual-token embeddings become a communicated pay…

cs.CV2026

GeoMix: Descriptor-Free Visual Localization via Global Context and Multi-Detector Training

Yejun Zhang, Xinjue Wang, Zihan Wang +2

Descriptor-free visual localization eliminates high-dimensional descriptor storage, preserves scene privacy, and simplifies map maintenance, yet its accuracy still lags far behind…

cs.CV2026

SplatGuide: Geometric Priors from 3D Gaussians for Pose-Free Novel View Synthesis

Yejun Zhang, Zihan Wang, Xu Ji +8

Generating photorealistic novel views from unposed images requires both 3D geometric understanding and the ability to synthesize unseen content. A natural strategy combines feed-fo…

cs.CV2026

HSA: Hierarchical Slot Attention for Multi-granularity Scene-Decomposition

Neelu Madan, Rongzhen Zhao, Andreas Mogelmose +4

Slot attention is a powerful framework for object-centric learning, decomposing visual scenes into latent slots through iterative competitive attention. However, existing methods s…