28 citations · 28 across the 3 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2023
MSDA: Monocular Self-supervised Domain Adaptation for 6D Object Pose Estimation
Dingding Cai, Janne Heikkilä, Esa Rahtu
Acquiring labeled 6D poses from real images is an expensive and time-consuming task. Though massive amounts of synthetic RGB images are easy to obtain, the models trained on them s…
cs.CV2016★ 28 cited
Video Summarization using Deep Semantic Features
Mayu Otani, Yuta Nakashima, Esa Rahtu +2
This paper presents a video summarization technique for an Internet video to provide a quick way to overview its content. This is a challenging problem because finding important or…
cs.CV2016
Learning Joint Representations of Videos and Sentences with Web Image Search
Mayu Otani, Yuta Nakashima, Esa Rahtu +2
Our objective is video retrieval based on natural language queries. In addition, we consider the analogous problem of retrieving sentences or generating descriptions given an input…