4 papers
Audio-Visual Driven Compression for Low-Bitrate Talking Head Videos
Riku Takahashi, Ryugo Morita, Jinjia Zhou
Talking head video compression has advanced with neural rendering and keypoint-based methods, but challenges remain, especially at low bit rates, including handling large head move…
Block based Adaptive Compressive Sensing with Sampling Rate Control
Kosuke Iwama, Ryugo Morita, Jinjia Zhou
Compressive sensing (CS), acquiring and reconstructing signals below the Nyquist rate, has great potential in image and video acquisition to exploit data redundancy and greatly red…
Visual question answering based evaluation metrics for text-to-image generation
Mizuki Miyamoto, Ryugo Morita, Jinjia Zhou
Text-to-image generation and text-guided image manipulation have received considerable attention in the field of image generation tasks. However, the mainstream evaluation methods…
BATINet: Background-Aware Text to Image Synthesis and Manipulation Network
Ryugo Morita, Zhiqiang Zhang, Jinjia Zhou
Background-Induced Text2Image (BIT2I) aims to generate foreground content according to the text on the given background image. Most studies focus on generating high-quality foregro…