4 papers
Robust Disentangled Counterfactual Learning for Physical Audiovisual Commonsense Reasoning
Mengshi Qi, Changsheng Lv, Huadong Ma
In this paper, we propose a new Robust Disentangled Counterfactual Learning (RDCL) approach for physical audiovisual commonsense reasoning. The task aims to infer objects' physics…
Towards Robust Unsupervised Attention Prediction in Autonomous Driving
Mengshi Qi, Xiaoyang Bi, Pengfei Zhu +1
Robustly predicting attention regions of interest for self-driving systems is crucial for driving safety but presents significant challenges due to the labor-intensive nature of ob…
A New Teacher-Reviewer-Student Framework for Semi-supervised 2D Human Pose Estimation
Wulian Yun, Mengshi Qi, Fei Peng +1
Conventional 2D human pose estimation methods typically require extensive labeled annotations, which are both labor-intensive and expensive. In contrast, semi-supervised 2D human p…
Human Grasp Generation for Rigid and Deformable Objects with Decomposed VQ-VAE
Mengshi Qi, Zhe Zhao, Huadong Ma
Generating realistic human grasps is crucial yet challenging for object manipulation in computer graphics and robotics. Current methods often struggle to generate detailed and real…