62 citations · 94 across the 6 of their papers we have counts for
12 papers
3D-TOGO: Towards Text-Guided Cross-Category 3D Object Generation
Zutao Jiang, Guansong Lu, Xiaodan Liang +4
Text-guided 3D object generation aims to generate 3D objects described by user-defined captions, which paves a flexible way to visualize what we imagined. Although some works have…
Listen As You Wish: Audio based Event Detection via Text-to-Audio Grounding in Smart Cities
Haoyu Tang, Yunxiao Wang, Jihua Zhu +4
With the development of internet of things technologies, tremendous sensor audio data has been produced, which poses great challenges to audio-based event detection in smart cities…
Tensor-based Intrinsic Subspace Representation Learning for Multi-view Clustering
Qinghai Zheng, Yu Zhang, Jihua Zhu +3
As a hot research topic, many multi-view clustering approaches are proposed over the past few years. Nevertheless, most existing algorithms merely take the consensus information am…
Frame-wise Cross-modal Matching for Video Moment Retrieval
Haoyu Tang, Jihua Zhu, Meng Liu +2
Video moment retrieval targets at retrieving a moment in a video for a given language query. The challenges of this task include 1) the requirement of localizing the relevant momen…
Robust Motion Averaging under Maximum Correntropy Criterion
Jihua Zhu, Jie Hu, Huimin Lu +2
Recently, the motion averaging method has been introduced as an effective means to solve the multi-view registration problem. This method aims to recover global motions from a set…
Registration of multi-view point sets under the perspective of expectation-maximization
Jihua Zhu, Jing Zhang, Huimin Lu +1
Registration of multi-view point sets is a prerequisite for 3D model reconstruction. To solve this problem, most of previous approaches either partially explore available informati…