activity
20162022
most citedPolysemous Visual-Semantic Embedding for Cross-Modal Retrieval

22 citations · 76 across the 13 of their papers we have counts for

collaborators

24 papers

cs.CV202212 cited

Video Summarization Overview

Mayu Otani, Yale Song, Yang Wang

With the broad growth of video capturing devices and applications on the web, it is more demanding to provide desired video content for users efficiently. Video summarization facil…

cs.CV2022

One Network Doesn't Rule Them All: Moving Beyond Handcrafted Architectures in Self-Supervised Learning

Sharath Girish, Debadeepta Dey, Neel Joshi +5

The current literature on self-supervised learning (SSL) focuses on developing learning objectives to train neural networks more effectively on unlabeled data. The typical developm…

cs.RO2022

COMPASS: Contrastive Multimodal Pretraining for Autonomous Systems

Shuang Ma, Sai Vemprala, Wenshan Wang +4

Learning representations that generalize across tasks and domains is challenging yet necessary for autonomous systems. Although task-driven approaches are appealing, designing mode…

cs.SI2021

On the Virality of Animated GIFs on Tumblr

Yunseok Jang, Yale Song, Gunhee Kim

Animated GIFs are becoming increasingly popular in online communication. People use them to express emotion, share their interests and enhance (or even replace) short-form texting;…

cs.AI20215 cited

CausalCity: Complex Simulations with Agency for Causal Discovery and Reasoning

Daniel McDuff, Yale Song, Jiyoung Lee +7

The ability to perform causal and counterfactual reasoning are central properties of human intelligence. Decision-making systems that can perform these types of reasoning have the…

cs.LG2021

Contrastive Learning of Global-Local Video Representations

Shuang Ma, Zhaoyang Zeng, Daniel McDuff +1

Contrastive learning has delivered impressive results for various tasks in the self-supervised regime. However, existing approaches optimize for learning representations specific t…