3 papers
eess.AS2023
Speech Modeling with a Hierarchical Transformer Dynamical VAE
Xiaoyu Lin, Xiaoyu Bie, Simon Leglaive +2
The dynamical variational autoencoders (DVAEs) are a family of latent-variable deep generative models that extends the VAE to model a sequence of observed data and a corresponding…
cs.CV2022
DSR: Towards Drone Image Super-Resolution
Xiaoyu Lin, Baran Ozaydin, Vidit Vidit +2
Despite achieving remarkable progress in recent years, single-image super-resolution methods are developed with several limitations. Specifically, they are trained on fixed content…
cs.CV2022
Towards Robust Drone Vision in the Wild
Xiaoyu Lin
The past few years have witnessed the burst of drone-based applications where computer vision plays an essential role. However, most public drone-based vision datasets focus on det…