From the 1 of 7 linked papers with an AI index.
7 papers
SplatReasoner: Enhancing Embodied Reasoning and Grounding by Novel View Synthesis
Kim Yu-Ji, Dahye Lee, Kim Jun-Seong +6
SplatReasoner integrates query‑conditioned novel view synthesis via 3D Gaussian splatting into vision‑language models, enabling better embodied reasoning and 3D grounding by genera…
SA-ResGS: Self-Augmented Residual 3D Gaussian Splatting for Next Best View Selection
Kim Jun-Seong, Tae-Hyun Oh, Eduardo Pérez-Pellitero +1
We propose Self-Augmented Residual 3D Gaussian Splatting (SA-ResGS), a novel framework to stabilize uncertainty quantification and enhancing uncertainty-aware supervision in next-b…
Factorized Multi-Resolution HashGrid for Efficient Neural Radiance Fields: Execution on Edge-Devices
Kim Jun-Seong, Mingyu Kim, GeonU Kim +2
We introduce Fact-Hash, a novel parameter-encoding method for training on-device neural radiance fields. Neural Radiance Fields (NeRF) have proven pivotal in 3D representations, bu…
HDR-NSFF: High Dynamic Range Neural Scene Flow Fields
Shin Dong-Yeon, Kim Jun-Seong, Kwon Byung-Ki +1
Radiance of real-world scenes typically spans a much wider dynamic range than what standard cameras can capture. While conventional HDR methods merge alternating-exposure frames, t…
Dr. Splat: Directly Referring 3D Gaussian Splatting via Direct Language Embedding Registration
Kim Jun-Seong, GeonU Kim, Kim Yu-Ji +3
We introduce Dr. Splat, a novel approach for open-vocabulary 3D scene understanding leveraging 3D Gaussian Splatting. Unlike existing language-embedded 3DGS methods, which rely on…
SoundBrush: Sound as a Brush for Visual Scene Editing
Kim Sung-Bin, Kim Jun-Seong, Junseok Ko +2
We propose SoundBrush, a model that uses sound as a brush to edit and manipulate visual scenes. We extend the generative capabilities of the Latent Diffusion Model (LDM) to incorpo…