3 papers
cs.CV2026
SpatialMosaic: A Multiview VLM Dataset for Partial Visibility
Kanghee Lee, Jungi Hong, Sion Lee +4
Recent progress in Multimodal Large Language Models (MLLMs) has enabled 3D scene understanding and spatial reasoning directly from multi-view images, without requiring explicit 3D…
cs.CV2025
BUFFER-X: Towards Zero-Shot Point Cloud Registration in Diverse Scenes
Minkyun Seo, Hyungtae Lim, Kanghee Lee +2
Recent advances in deep learning-based point cloud registration have improved generalization, yet most methods still require retraining or manual parameter tuning for each new envi…
cs.CV2024
3D Geometric Shape Assembly via Efficient Point Cloud Matching
Nahyuk Lee, Juhong Min, Junha Lee +4
Learning to assemble geometric shapes into a larger target structure is a pivotal task in various practical applications. In this work, we tackle this problem by establishing local…