61 citations · 80 across the 3 of their papers we have counts for
4 papers
EMoG: Synthesizing Emotive Co-speech 3D Gesture with Diffusion Model
Lianying Yin, Yijun Wang, Tianyu He +5
Although previous co-speech gesture generation methods are able to synthesize motions in line with speech content, it is still not enough to handle diverse and complicated motion d…
Prompt-ICM: A Unified Framework towards Image Coding for Machines with Task-driven Prompts
Ruoyu Feng, Jinming Liu, Xin Jin +3
Image coding for machines (ICM) aims to compress images to support downstream AI analysis instead of human perception. For ICM, developing a unified codec to reduce information red…
Inpaint Anything: Segment Anything Meets Image Inpainting
Tao Yu, Runseng Feng, Ruoyu Feng +4
Modern image inpainting systems, despite the significant progress, often struggle with mask selection and holes filling. Based on Segment-Anything Model (SAM), we make the first at…
Learned Image Compression with Mixed Transformer-CNN Architectures
Jinming Liu, Heming Sun, Jiro Katto
Learned image compression (LIC) methods have exhibited promising progress and superior rate-distortion performance compared with classical image compression standards. Most existin…