1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Panacea+: Panoramic and Controllable Video Generation for Autonomous Driving
Yuqing Wen, Yucheng Zhao, Yingfei Liu +7
The field of autonomous driving increasingly demands high-quality annotated video training data. In this paper, we propose Panacea+, a powerful and universally applicable framework…
cs.CL2024
How to Understand Named Entities: Using Common Sense for News Captioning
Ning Xu, Yanhui Wang, Tingting Zhang +3
News captioning aims to describe an image with its news article body as input. It greatly relies on a set of detected named entities, including real-world people, organizations, an…
cs.CV2023★ 1 cited
MicroCinema: A Divide-and-Conquer Approach for Text-to-Video Generation
Yanhui Wang, Jianmin Bao, Wenming Weng +12
We present MicroCinema, a straightforward yet effective framework for high-quality and coherent text-to-video generation. Unlike existing approaches that align text prompts with vi…