2 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CV2021★ 2 cited
Survey: Transformer based Video-Language Pre-training
Ludan Ruan, Qin Jin
Inspired by the success of transformer-based pre-training methods on natural language tasks and further computer vision tasks, researchers have begun to apply transformer to video…
cs.CV2021★ 1 cited
Team RUC_AIM3 Technical Report at ActivityNet 2021: Entities Object Localization
Ludan Ruan, Jieting Chen, Yuqing Song +2
Entities Object Localization (EOL) aims to evaluate how grounded or faithful a description is, which consists of caption generation and object grounding. Previous works tackle this…
cs.CV2020
YouMakeup VQA Challenge: Towards Fine-grained Action Understanding in Domain-Specific Videos
Shizhe Chen, Weiying Wang, Ludan Ruan +2
The goal of the YouMakeup VQA Challenge 2020 is to provide a common benchmark for fine-grained action understanding in domain-specific videos e.g. makeup instructional videos. We p…