2 papers
cs.CV2024
DriveScape: Towards High-Resolution Controllable Multi-View Driving Video Generation
Wei Wu, Xi Guo, Weixuan Tang +4
Recent advancements in generative models have provided promising solutions for synthesizing realistic driving videos, which are crucial for training autonomous driving perception m…
cs.CV2024
SGC-VQGAN: Towards Complex Scene Representation via Semantic Guided Clustering Codebook
Chenjing Ding, Chiyu Wang, Boshi Liu +3
Vector quantization (VQ) is a method for deterministically learning features through discrete codebook representations. Recent works have utilized visual tokenizers to discretize v…