10 papers
What Images Cannot Say: Language-Guided Olfactory Representation Learning
Eleftherios Tsonis, Xi Wang, Vicky Kalogeiton
Images tell us what a scene looks like, but rarely what it would feel like to be there. While recent datasets pair visual scenes with electronic-nose measurements, aligning smell s…
SF20K Competition 2025: Summary and findings
Ridouane Ghermi, Xi Wang, Vicky Kalogeiton +1
This report presents the results and findings of the first edition of the Short-Films 20K (SF20K) Competition, held in conjunction with the SLoMO Workshop at ICCV 2025. The competi…
Pulp Motion: Framing-aware multimodal camera and human motion generation
Robin Courant, Xi Wang, David Loiseaux +2
Treating human motion and camera trajectory generation separately overlooks a core principle of cinematography: the tight interplay between actor performance and camera work in the…
Soft-Di[M]O: Improving One-Step Discrete Image Generation with Soft Embeddings
Yuanzhi Zhu, Xi Wang, Stéphane Lathuilière +1
One-step generators distilled from Masked Diffusion Models (MDMs) compress multiple sampling steps into a single forward pass, enabling efficient text and image synthesis. However,…
Diffusion Reinforcement Learning via Centered Reward Distillation
Yuanzhi Zhu, Xi Wang, Stéphane Lathuilière +1
Diffusion and flow models achieve State-Of-The-Art (SOTA) generative performance, yet many practically important behaviors such as fine-grained prompt fidelity, compositional corre…
FakeParts: a New Family of AI-Generated DeepFakes
Ziyi Liu, Firas Gabetni, Awais Hussain Sani +5
We introduce FakeParts, a new class of deepfakes characterized by subtle, localized manipulations to specific spatial regions or temporal segments of otherwise authentic videos. Un…