3 papers
cs.CV2026
V-RAE: Rethinking Video Latent Spaces for Generation
Minghui Guo, Shengqiong Wu, Hao Fei
Latent video generation relies on autoencoders to define a compact space in which generative models operate. Although video autoencoder architectures have evolved substantially, th…
cs.CV2026
UniM: A Unified Any-to-Any Interleaved Multimodal Benchmark
Yanlin Li, Minghui Guo, Kaiwen Zhang +13
In real-world multimodal applications, systems usually need to comprehend arbitrarily combined and interleaved multimodal inputs from users, while also generating outputs in any in…
physics.app-ph2025
Biomimetic Liquid Metal Cell
Jingyi Li, Mengwen Qiao, Minghui Guo +7
Gallium-based liquid metals, as a broad category of emerging functional materials with unique physical, chemical, and biological properties, offer numerous possibilities for advanc…