Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023
SHARE: Single-view Human Adversarial REconstruction
Shreelekha Revankar, Shijia Liao, Yu Shen +3
The accuracy of 3D Human Pose and Shape reconstruction (HPS) from an image is progressively improving. Yet, no known method is robust across all image distortion. To address issues…
cs.CV2023
ViLA: Efficient Video-Language Alignment for Video Question Answering
Xijun Wang, Junbang Liang, Chun-Kai Wang +4
In this work, we propose an efficient Video-Language Alignment (ViLA) network. Our ViLA model addresses both efficient frame sampling and effective cross-modal alignment in a unifi…