2 papers
cs.CV2024
Adapting Vision-Language Model with Fine-grained Semantics for Open-Vocabulary Segmentation
Yong Xien Chng, Xuchong Qiu, Yizeng Han +3
Despite extensive research, open-vocabulary segmentation methods still struggle to generalize across diverse domains. To reduce the computational cost of adapting Vision-Language M…
cs.RO2024
GOPT: Generalizable Online 3D Bin Packing via Transformer-based Deep Reinforcement Learning
Heng Xiong, Changrong Guo, Jian Peng +5
Robotic object packing has broad practical applications in the logistics and automation industry, often formulated by researchers as the online 3D Bin Packing Problem (3D-BPP). How…