5 papers
On the Sample Efficiency of Inverse Dynamics Models for Semi-Supervised Imitation Learning
Sacha Morin, Moonsub Byeon, Alexia Jolicoeur-Martineau +1
Semi-supervised imitation learning (SSIL) consists in learning a policy from a small dataset of action-labeled trajectories and a much larger dataset of action-free trajectories. S…
One Pass Is Not Enough: Recursive Latent Refinement for Generative Models
Mehdi Esmaeilzadeh, Alexia Jolicoeur-Martineau, Chirag Vashist +1
Despite remarkable progress, image generation is far from solved. The dominant metric, FID, conflates sample fidelity with mode coverage and is close to being saturated. Yet a mode…
Ctrl-Crash: Controllable Diffusion for Realistic Car Crashes
Anthony Gosselin, Ge Ya Luo, Luis Lara +5
Video diffusion techniques have advanced significantly in recent years; however, they struggle to generate realistic imagery of car crashes due to the scarcity of accident events i…
Ctrl-V: Higher Fidelity Video Generation with Bounding-Box Controlled Object Motion
Ge Ya Luo, Zhi Hao Luo, Anthony Gosselin +2
Controllable video generation has attracted significant attention, largely due to advances in video diffusion models. In domains such as autonomous driving, it is essential to deve…
Beyond FVD: Enhanced Evaluation Metrics for Video Generation Quality
Ge Ya Luo, Gian Mario Favero, Zhi Hao Luo +2
The Fréchet Video Distance (FVD) is a widely adopted metric for evaluating video generation distribution quality. However, its effectiveness relies on critical assumptions. Our an…