6 papers
DreamCS: Geometry-Aware Text-to-3D Generation with Unpaired 3D Reward Supervision
Xiandong Zou, Ruihao Xia, Hongsong Wang +1
While text-to-3D generation has attracted growing interest, existing methods often struggle to produce 3D assets that align well with human preferences. Current preference alignmen…
Temporal-Guided Visual Foundation Models for Event-Based Vision
Ruihao Xia, Junhong Cai, Luziwei Leng +5
Event cameras offer unique advantages for vision tasks in challenging environments, yet processing asynchronous event streams remains an open challenge. While existing methods rely…
Towards Scalable and Consistent 3D Editing
Ruihao Xia, Yang Tang, Pan Zhou
3D editing - the task of locally modifying the geometry or appearance of a 3D asset - has wide applications in immersive content creation, digital entertainment, and AR/VR. However…
DidSee: Diffusion-Based Depth Completion for Material-Agnostic Robotic Perception and Manipulation
Wenzhou Lyu, Jialing Lin, Wenqi Ren +3
Commercial RGB-D cameras often produce noisy, incomplete depth maps for non-Lambertian objects. Traditional depth completion methods struggle to generalize due to the limited diver…
Unsupervised Modality Adaptation with Text-to-Image Diffusion Models for Semantic Segmentation
Ruihao Xia, Yu Liang, Peng-Tao Jiang +4
Despite their success, unsupervised domain adaptation methods for semantic segmentation primarily focus on adaptation between image domains and do not utilize other abundant visual…
Towards Natural Image Matting in the Wild via Real-Scenario Prior
Ruihao Xia, Yu Liang, Peng-Tao Jiang +5
Recent approaches attempt to adapt powerful interactive segmentation models, such as SAM, to interactive matting and fine-tune the models based on synthetic matting datasets. Howev…