2 papers
cs.CV2026
: A "Spot the Difference" Challenge for Large Multimodal Models
Kewei Wei, Bocheng Hu, Jie Cao +13
Modern Large Multimodal Models (LMMs) have demonstrated extraordinary ability in static image and single-state spatial-temporal understanding. However, their capacity to comprehend…
cs.CV2024
Exploring contextual modeling with linear complexity for point cloud segmentation
Yong Xien Chng, Xuchong Qiu, Yizeng Han +3
Point cloud segmentation is an important topic in 3D understanding that has traditionally has been tackled using either the CNN or Transformer. Recently, Mamba has emerged as a pro…