10 citations · 12 across the 3 of their papers we have counts for
6 papers · 1 filter
VisualChef: Generating Visual Aids in Cooking via Mask Inpainting
Oleh Kuzyk, Zuoyue Li, Marc Pollefeys +1
Cooking requires not only following instructions but also understanding, executing, and monitoring each step - a process that can be challenging without visual guidance. Although r…
Sat2Scene: 3D Urban Scene Generation from Satellite Images with Diffusion
Zuoyue Li, Zhenqiang Li, Zhaopeng Cui +2
Directly generating scenes from satellite imagery offers exciting possibilities for integration into applications like games and map services. However, challenges arise from signif…
Spatio-Temporal Perturbations for Video Attribution
Zhenqiang Li, Weimin Wang, Zuoyue Li +2
The attribution method provides a direction for interpreting opaque neural networks in a visual way by identifying and visualizing the input regions/pixels that dominate the output…
Sat2Vid: Street-view Panoramic Video Synthesis from a Single Satellite Image
Zuoyue Li, Zhenqiang Li, Zhaopeng Cui +3
We present a novel method for synthesizing both temporally and geometrically consistent street-view panoramic video from a single satellite image and camera trajectory. Existing cr…
Towards Visually Explaining Video Understanding Networks with Perturbation
Zhenqiang Li, Weimin Wang, Zuoyue Li +2
''Making black box models explainable'' is a vital problem that accompanies the development of deep learning networks. For networks taking visual information as input, one basic bu…
Topological Map Extraction from Overhead Images
Zuoyue Li, Jan Dirk Wegner, Aurélien Lucchi
We propose a new approach, named PolyMapper, to circumvent the conventional pixel-wise segmentation of (aerial) images and predict objects in a vector representation directly. Poly…