activity
20122024
most citedYouTube-8M: A Large-Scale Video Classification Benchmark

923 citations · 987 across the 13 of their papers we have counts for

collaborators

13 papers

cs.IR20241 cited

General Item Representation Learning for Cold-start Content Recommendations

Jooeun Kim, Jinri Kim, Kwangeun Yeo +4

Cold-start item recommendation is a long-standing challenge in recommendation systems. A common remedy is to use a content-based approach, but rich information from raw contents in…

cs.CV2024

Modality-Aware Representation Learning for Zero-shot Sketch-based Image Retrieval

Eunyi Lyou, Doyeon Lee, Jooeun Kim +1

Zero-shot learning offers an efficient solution for a machine learning model to treat unseen categories, avoiding exhaustive data collection. Zero-shot Sketch-based Image Retrieval…

cs.CV2023

Towards Robust and Smooth 3D Multi-Person Pose Estimation from Monocular Videos in the Wild

Sungchan Park, Eunyi You, Inhoe Lee +1

3D pose estimation is an invaluable task in computer vision with various practical applications. Especially, 3D pose estimation for multi-person from a monocular video (3DMPPE) is…

cs.CL20232 cited

Shuffle & Divide: Contrastive Learning for Long Text

Joonseok Lee, Seongho Joe, Kyoungwon Park +4

We propose a self-supervised learning method for long text documents based on contrastive learning. A key to our method is Shuffle and Divide (SaD), a simple text augmentation algo…

cs.CV20231 cited

ContraCluster: Learning to Classify without Labels by Contrastive Self-Supervision and Prototype-Based Semi-Supervision

Seongho Joe, Byoungjip Kim, Hoyoung Kang +5

The recent advances in representation learning inspire us to take on the challenging problem of unsupervised image classification tasks in a principled way. We propose ContraCluste…

cs.CL20233 cited

MAQA: A Multimodal QA Benchmark for Negation

Judith Yue Li, Aren Jansen, Qingqing Huang +3

Multimodal learning can benefit from the representation power of pretrained Large Language Models (LLMs). However, state-of-the-art transformer based LLMs often ignore negations in…