3 papers
cs.CV2026
Extending a Large View Synthesis Model for Multi-view Panoptic Segmentation
Kwonyoung Ryu, In-Jae Lee, Jonghyun Jin +3
Large view synthesis models synthesize novel views through cross-view attention without explicit 3D representations, and recent studies have shown that they learn accurate spatial…
cs.CV2026
SpatialMosaic: A Multiview VLM Dataset for Partial Visibility
Kanghee Lee, Jungi Hong, Sion Lee +4
Recent progress in Multimodal Large Language Models (MLLMs) has enabled 3D scene understanding and spatial reasoning directly from multi-view images, without requiring explicit 3D…
cs.CV2025
OpenBox: Annotate Any Bounding Boxes in 3D
In-Jae Lee, Mungyeom Kim, Kwonyoung Ryu +2
Unsupervised and open-vocabulary 3D object detection has recently gained attention, particularly in autonomous driving, where reducing annotation costs and recognizing unseen objec…