2 papers
cs.CV2024
VividDreamer: Towards High-Fidelity and Efficient Text-to-3D Generation
Zixuan Chen, Ruijie Su, Jiahao Zhu +3
Text-to-3D generation aims to create 3D assets from text-to-image diffusion models. However, existing methods face an inherent bottleneck in generation quality because the widely-u…
cs.CV2023
Tracking Objects and Activities with Attention for Temporal Sentence Grounding
Zeyu Xiong, Daizong Liu, Pan Zhou +1
Temporal sentence grounding (TSG) aims to localize the temporal segment which is semantically aligned with a natural language query in an untrimmed video.Most existing methods extr…