32 citations · 60 across the 4 of their papers we have counts for
4 papers
ACT2G: Attention-based Contrastive Learning for Text-to-Gesture Generation
Hitoshi Teshima, Naoki Wake, Diego Thomas +3
Recent increase of remote-work, online meeting and tele-operation task makes people find that gesture for avatars and communication robots is more important than we have thought. I…
Not Only Generative Art: Stable Diffusion for Content-Style Disentanglement in Art Analysis
Yankun Wu, Yuta Nakashima, Noa Garcia
The duality of content and style is inherent to the nature of art. For humans, these two elements are clearly different: content refers to the objects and concepts in the piece of…
Video Summarization using Deep Semantic Features
Mayu Otani, Yuta Nakashima, Esa Rahtu +2
This paper presents a video summarization technique for an Internet video to provide a quick way to overview its content. This is a challenging problem because finding important or…
Learning Joint Representations of Videos and Sentences with Web Image Search
Mayu Otani, Yuta Nakashima, Esa Rahtu +2
Our objective is video retrieval based on natural language queries. In addition, we consider the analogous problem of retrieving sentences or generating descriptions given an input…