3 papers
cs.CV2026
MotionHalluc: Diagnosing Kinematic Hallucinations in Fine-Grained Motion Reasoning
Weile Guo, Shenghong He, Danying Mo +3
Motion instruction generation in cross-video comparison aims to produce corrective feedback that describes the differences between a query and a reference motion. However, existing…
cs.CV2025
The Collapse of Patches
Wei Guo, Shunqi Mao, Zhuonan Liang +2
Observing certain patches in an image reduces the uncertainty of others. Their realization lowers the distribution entropy of each remaining patch feature, analogous to collapsing…
cs.MM2025
Gotta Hear Them All: Towards Sound Source Aware Audio Generation
Wei Guo, Heng Wang, Jianbo Ma +1
Audio synthesis has broad applications in multimedia. Recent advancements have made it possible to generate relevant audios from inputs describing an audio scene, such as images or…