2 papers
cs.CV2024
Contrastive Language Video Time Pre-training
Hengyue Liu, Kyle Min, Hector A. Valdez +1
We introduce LAVITI, a novel approach to learning language, video, and temporal representations in long-form videos via contrastive learning. Different from pre-training on video-t…
cs.CV2023
RepSGG: Novel Representations of Entities and Relationships for Scene Graph Generation
Hengyue Liu, Bir Bhanu
Scene Graph Generation (SGG) has achieved significant progress recently. However, most previous works rely heavily on fixed-size entity representations based on bounding box propos…