activity
20212024
most citedHPPNet: Modeling the Harmonic Structure and Pitch Invariance in Piano Transcription

4 citations · 13 across the 16 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV20231 cited

LiveChat: Video Comment Generation from Audio-Visual Multimodal Contexts

Julien Lalanne, Raphael Bournet, Yi Yu

Live commenting on video, a popular feature of live streaming platforms, enables viewers to engage with the content and share their comments, reactions, opinions, or questions with…

cs.CV2023

Emotionally Enhanced Talking Face Generation

Sahil Goyal, Shagun Uppal, Sarthak Bhagat +3

Several works have developed end-to-end pipelines for generating lip-synced talking faces with various real-world applications, such as teaching and language translation in videos.…

cs.CV20232 cited

Backdoor Attacks Against Deep Image Compression via Adaptive Frequency Trigger

Yi Yu, Yufei Wang, Wenhan Yang +3

Recent deep-learning-based compression methods have achieved superior performance compared with traditional approaches. However, deep learning models have proven to be vulnerable t…

cs.CV2023

Raw Image Reconstruction with Learned Compact Metadata

Yufei Wang, Yi Yu, Wenhan Yang +4

While raw images exhibit advantages over sRGB images (e.g., linearity and fine-grained quantization level), they are not widely used by common users due to the large storage requir…

cs.CV20212 cited

Feature Distillation Interaction Weighting Network for Lightweight Image Super-Resolution

Guangwei Gao, Wenjie Li, Juncheng Li +3

Convolutional neural networks based single-image super-resolution (SISR) has made great progress in recent years. However, it is difficult to apply these methods to real-world scen…