3 citations · 5 across the 11 of their papers we have counts for
8 papers · 1 filter
GAAT: Geometry-Aware Alignment Transformer for Multimodal UAV Perception
Jingpu Yang, Debin Tang, Yilin Sun +4
Unmanned aerial vehicle (UAV) multimodal perception integrates visible (RGB), infrared (IR), synthetic aperture radar (SAR), and depth sensors for scene understanding under diverse…
Chat-Edit-3D++: Interactive 3D and 4D Scene Editing via Large Language Models
Shuangkang Fang, Yufeng Wang, Yi-Hsuan Tsai +4
Recent work on image content manipulation based on vision-language pre-training models has been effectively extended to text-driven 3D scene editing. However, existing schemes for…
WaterClear-GS: Optical-Aware Gaussian Splatting for Underwater Reconstruction and Restoration
Xinrui Zhang, Yufeng Wang, Shuangkang Fang +3
Underwater 3D reconstruction and appearance restoration remain challenging due to the complex optical properties of water, such as wavelength-dependent attenuation and scattering.…
NeRF Is a Valuable Assistant for 3D Gaussian Splatting
Shuangkang Fang, I-Chao Shen, Takeo Igarashi +5
We introduce NeRF-GS, a novel framework that jointly optimizes Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS). This framework leverages the inherent continuous spat…
Chat-Edit-3D: Interactive 3D Scene Editing via Text Prompts
Shuangkang Fang, Yufeng Wang, Yi-Hsuan Tsai +4
Recent work on image content manipulation based on vision-language pre-training models has been effectively extended to text-driven 3D scene editing. However, existing schemes for…
Editing 3D Scenes via Text Prompts without Retraining
Shuangkang Fang, Yufeng Wang, Yi Yang +4
Numerous diffusion models have recently been applied to image synthesis and editing. However, editing 3D scenes is still in its early stages. It poses various challenges, such as t…