2 papers
cs.CV2024
VONet: Unsupervised Video Object Learning With Parallel U-Net Attention and Object-wise Sequential VAE
Haonan Yu, Wei Xu
Unsupervised video object learning seeks to decompose video scenes into structural object representations without any supervision from depth, optical flow, or segmentation. We pres…
cs.CV2014
A Faster Method for Tracking and Scoring Videos Corresponding to Sentences
Haonan Yu, Daniel P. Barrett, Jeffrey Mark Siskind
Prior work presented the sentence tracker, a method for scoring how well a sentence describes a video clip or alternatively how well a video clip depicts a sentence. We present an…