19 papers
Chehre: An Emoji-Prompted Video Dataset for Perceptually Diverse Facial Expression Recognition
Bita Azari, Zoe Stanley, Avneet Batra +4
Facial expressions are nonverbal social signals used in human interaction, but facial expression recognition datasets often focus on static images, basic emotion categories, or sin…
Function2Scene: 3D Indoor Scene Layout from Functional Specifications
Ruiqi Wang, Qimin Chen, Daniel Ritchie +4
Most text-driven 3D indoor scene synthesis methods generate rooms from object-centric prompts, asking what furniture should be placed rather than how the space is used. Yet in real…
Artiverse: A Diverse and Physically Grounded Dataset for Articulated Objects
Denys Iliash, Jiayi Liu, Egor Fokin +4
We present Artiverse, a diverse and physically grounded dataset of high-quality articulated 3D objects designed for realistic functional modeling and simulation. Artiverse contains…
Functionalization via Structure Completion and Motion Rectification
Mingrui Zhao, Sai Raj Kishore Perla, Kai Wang +8
Acquisition and creation of 3D assets have been largely view- or appearance-driven. As a result, existing digital 3D models often lack the requisite structural components to functi…
Learning to Place Objects with Programs and Iterative Self Training
Adrian Chang, Kai Wang, Yuanbo Li +3
In this work we study indoor scene object placement. Given a 3D indoor scene and an object, the task is to predict placement locations within the scene. Empirical observations of d…
EgoFun3D: Modeling Interactive Objects from Egocentric Videos using Function Templates
Weikun Peng, Denys Iliash, Manolis Savva
We present EgoFun3D, a coordinated task formulation, dataset, and benchmark for modeling interactive 3D objects from egocentric videos. Interactive objects are of high interest for…