3 papers
cs.CV2026
VIPER: Visual In-Context Physics Reasoning for Physically Plausible Video Generation
Tianxiao Chen, Hanmo Chen, Huajin Chen +3
Modern video generation models can synthesize visually compelling and temporally coherent clips, yet controlling their physical behavior remains difficult with standard text and im…
cs.CV2026
Mobile-VTON: High-Fidelity On-Device Virtual Try-On
Zhenchen Wan, Ce Chen, Runqi Lin +5
Virtual try-on (VTON) has recently achieved impressive visual fidelity, but most existing systems require uploading personal photos to cloud-based GPUs, raising privacy concerns an…
cs.CV2025
MF-VITON: High-Fidelity Mask-Free Virtual Try-On with Minimal Input
Zhenchen Wan, Yanwu xu, Dongting Hu +6
Recent advancements in Virtual Try-On (VITON) have significantly improved image realism and garment detail preservation, driven by powerful text-to-image (T2I) diffusion models. Ho…