activity
20152023
most citedLearning Visual Importance for Graphic Designs and Data Visualizations

168 citations · 376 across the 14 of their papers we have counts for

collaborators
Showing cs.CVShow all

22 papers · 1 filter

cs.CV2023

Language-Guided Music Recommendation for Video via Prompt Analogies

Daniel McKee, Justin Salamon, Josef Sivic +1

We propose a method to recommend music for an input video while allowing a user to guide music selection with free-form natural language. A key challenge of this problem setting is…

cs.CV202228 cited

Monocular Dynamic View Synthesis: A Reality Check

Hang Gao, Ruilong Li, Shubham Tulsiani +2

We study the recent progress on dynamic view synthesis (DVS) from monocular video. Though existing approaches have demonstrated impressive results, we show a discrepancy between th…

cs.CV2022

Neural Volumetric Object Selection

Zhongzheng Ren, Aseem Agarwala, Bryan Russell +2

We introduce an approach for selecting objects in neural volumetric 3D representations, such as multi-plane images (MPI) and neural radiance fields (NeRF). Our approach takes a set…

cs.CV2022

Focal Length and Object Pose Estimation via Render and Compare

Georgy Ponimatkin, Yann Labbé, Bryan Russell +2

We introduce FocalPose, a neural render-and-compare method for jointly estimating the camera-object 6D pose and camera focal length given a single RGB input image depicting a known…

cs.CV2021

Weakly Supervised Human-Object Interaction Detection in Video via Contrastive Spatiotemporal Regions

Shuang Li, Yilun Du, Antonio Torralba +2

We introduce the task of weakly supervised learning for detecting human and object interactions in videos. Our task poses unique challenges as a system does not know what types of…

cs.CV20215 cited

Editing Conditional Radiance Fields

Steven Liu, Xiuming Zhang, Zhoutong Zhang +3

A neural radiance field (NeRF) is a scene model supporting high-quality view synthesis, optimized per scene. In this paper, we explore enabling user editing of a category-level NeR…