10 citations · 22 across the 4 of their papers we have counts for
6 papers · 1 filter
Sat2Scene: 3D Urban Scene Generation from Satellite Images with Diffusion
Zuoyue Li, Zhenqiang Li, Zhaopeng Cui +2
Directly generating scenes from satellite imagery offers exciting possibilities for integration into applications like games and map services. However, challenges arise from signif…
Intuitive Surgical SurgToolLoc and SurgVU Challenges Results: 2022-2025
Aneeq Zia, Max Berniker, Rogerio Garcia Nespolo +153
Robotic assisted (RA) surgery promises to transform surgical intervention. Intuitive Surgical is committed to fostering these changes and the machine learning models and algorithms…
Spatio-Temporal Perturbations for Video Attribution
Zhenqiang Li, Weimin Wang, Zuoyue Li +2
The attribution method provides a direction for interpreting opaque neural networks in a visual way by identifying and visualizing the input regions/pixels that dominate the output…
Sat2Vid: Street-view Panoramic Video Synthesis from a Single Satellite Image
Zuoyue Li, Zhenqiang Li, Zhaopeng Cui +3
We present a novel method for synthesizing both temporally and geometrically consistent street-view panoramic video from a single satellite image and camera trajectory. Existing cr…
Manipulation-skill Assessment from Videos with Spatial Attention Network
Zhenqiang Li, Yifei Huang, Minjie Cai +1
Recent advances in computer vision have made it possible to automatically assess from videos the manipulation skills of humans in performing a task, which breeds many important app…
Predicting Gaze in Egocentric Video by Learning Task-dependent Attention Transition
Yifei Huang, Minjie Cai, Zhenqiang Li +1
We present a new computational model for gaze prediction in egocentric videos by exploring patterns in temporal shift of gaze fixations (attention transition) that are dependent on…