3 papers
cs.CV2024
MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction
Wang Zhao, Jiachen Liu, Sheng Zhang +5
This paper presents a generalizable 3D plane detection and reconstruction framework named MonoPlane. Unlike previous robust estimator-based works (which require multiple images or…
cs.CV2024
ViMo: Generating Motions from Casual Videos
Liangdong Qiu, Chengxing Yu, Yanran Li +6
Although humans have the innate ability to imagine multiple possible actions from videos, it remains an extraordinary challenge for computers due to the intricate camera movements…
cs.LG2024
Integrating Domain Knowledge for handling Limited Data in Offline RL
Briti Gangopadhyay, Zhao Wang, Jia-Fong Yeh +1
With the ability to learn from static datasets, Offline Reinforcement Learning (RL) emerges as a compelling avenue for real-world applications. However, state-of-the-art offline RL…