papers
Publications (2)
cs.CV2024
VideoPoet: A Large Language Model for Zero-Shot Video Generation
Dan Kondratyuk, Lijun Yu, Xiuye Gu +28
We present VideoPoet, a language model capable of synthesizing high-quality video, with matching audio, from a large variety of conditioning signals. VideoPoet employs a decoder-on…
cs.CV2023
Learning to Detect Touches on Cluttered Tables
Norberto Adrian Goussies, Kenji Hata, Shruthi Prabhakara +28
We present a novel self-contained camera-projector tabletop system with a lamp form-factor that brings digital intelligence to our tables. We propose a real-time, on-device, learni…