3 papers
cs.CV2026
M3DDM+: An improved video outpainting by a modified masking strategy
Takuya Murakawa, Takumi Fukuzawa, Ning Ding +1
M3DDM provides a computationally efficient framework for video outpainting via latent diffusion modeling. However, it exhibits significant quality degradation -- manifested as spat…
cs.CV2025
Can masking background and object reduce static bias for zero-shot action recognition?
Takumi Fukuzawa, Kensho Hara, Hirokatsu Kataoka +1
In this paper, we address the issue of static bias in zero-shot action recognition. Action recognition models need to represent the action itself, not the appearance. However, some…
cs.CV2024
Fine-grained length controllable video captioning with ordinal embeddings
Tomoya Nitta, Takumi Fukuzawa, Toru Tamaki
This paper proposes a method for video captioning that controls the length of generated captions. Previous work on length control often had few levels for expressing length. In thi…