3 papers
cs.CV2026
Focus Matters: Phase-Aware Suppression for Hallucination in Vision-Language Models
Sohyeon Kim, Sang Yeon Yoon, Kyeongbo Kong
Large Vision-Language Models (LVLMs) have achieved impressive progress in multimodal reasoning, yet they remain prone to object hallucinations, generating descriptions of objects t…
cs.CV2026
LivingWorld: Interactive 4D World Generation with Environmental Dynamics
Hyeongju Mun, In-Hwan Jin, Sohyeong Kim +1
We introduce LivingWorld, an interactive framework for generating 4D worlds with environmental dynamics from a single image. While recent advances in 3D scene generation enable lar…
cs.CV2024
PIV3CAMS: a multi-camera dataset for multiple computer vision problems and its application to novel view-point synthesis
Sohyeong Kim, Martin Danelljan, Radu Timofte +2
The modern approaches for computer vision tasks significantly rely on machine learning, which requires a large number of quality images. While there is a plethora of image datasets…