14 citations · 29 across the 7 of their papers we have counts for
12 papers
SpatialGrammar: A Domain-Specific Language for LLM-Based 3D Indoor Scene Generation
Song Tang, Kaiyong Zhao, Yuliang Li +5
Automatically generating interactive 3D indoor scenes from natural language is crucial for virtual reality, gaming, and embodied AI. However, existing LLM-based approaches often su…
RA-NeRF: Robust Neural Radiance Field Reconstruction with Accurate Camera Pose Estimation under Complex Trajectories
Qingsong Yan, Qiang Wang, Kaiyong Zhao +4
Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS) have emerged as powerful tools for 3D reconstruction and SLAM tasks. However, their performance depends heavily on ac…
SphereFusion: Efficient Panorama Depth Estimation via Gated Fusion
Qingsong Yan, Qiang Wang, Kaiyong Zhao +4
Due to the rapid development of panorama cameras, the task of estimating panorama depth has attracted significant attention from the computer vision community, especially in applic…
FusionLLM: A Decentralized LLM Training System on Geo-distributed GPUs with Adaptive Compression
Zhenheng Tang, Xueze Kang, Yiming Yin +11
To alleviate hardware scarcity in training large deep neural networks (DNNs), particularly large language models (LLMs), we present FusionLLM, a decentralized training system desig…
Rethinking Disparity: A Depth Range Free Multi-View Stereo Based on Disparity
Qingsong Yan, Qiang Wang, Kaiyong Zhao +3
Existing learning-based multi-view stereo (MVS) methods rely on the depth range to build the 3D cost volume and may fail when the range is too large or unreliable. To address this…
FADNet++: Real-Time and Accurate Disparity Estimation with Configurable Networks
Qiang Wang, Shaohuai Shi, Shizhen Zheng +2
Deep neural networks (DNNs) have achieved great success in the area of computer vision. The disparity estimation problem tends to be addressed by DNNs which achieve much better pre…