3 papers
cs.CV2024
Adapting Vision-Language Model with Fine-grained Semantics for Open-Vocabulary Segmentation
Yong Xien Chng, Xuchong Qiu, Yizeng Han +3
Despite extensive research, open-vocabulary segmentation methods still struggle to generalize across diverse domains. To reduce the computational cost of adapting Vision-Language M…
cs.CV2024
Exploring contextual modeling with linear complexity for point cloud segmentation
Yong Xien Chng, Xuchong Qiu, Yizeng Han +3
Point cloud segmentation is an important topic in 3D understanding that has traditionally has been tackled using either the CNN or Transformer. Recently, Mamba has emerged as a pro…
cs.RO2024
GOPT: Generalizable Online 3D Bin Packing via Transformer-based Deep Reinforcement Learning
Heng Xiong, Changrong Guo, Jian Peng +5
Robotic object packing has broad practical applications in the logistics and automation industry, often formulated by researchers as the online 3D Bin Packing Problem (3D-BPP). How…