4 citations · 13 across the 16 of their papers we have counts for
5 papers · 1 filter
LiveChat: Video Comment Generation from Audio-Visual Multimodal Contexts
Julien Lalanne, Raphael Bournet, Yi Yu
Live commenting on video, a popular feature of live streaming platforms, enables viewers to engage with the content and share their comments, reactions, opinions, or questions with…
Emotionally Enhanced Talking Face Generation
Sahil Goyal, Shagun Uppal, Sarthak Bhagat +3
Several works have developed end-to-end pipelines for generating lip-synced talking faces with various real-world applications, such as teaching and language translation in videos.…
Backdoor Attacks Against Deep Image Compression via Adaptive Frequency Trigger
Yi Yu, Yufei Wang, Wenhan Yang +3
Recent deep-learning-based compression methods have achieved superior performance compared with traditional approaches. However, deep learning models have proven to be vulnerable t…
Raw Image Reconstruction with Learned Compact Metadata
Yufei Wang, Yi Yu, Wenhan Yang +4
While raw images exhibit advantages over sRGB images (e.g., linearity and fine-grained quantization level), they are not widely used by common users due to the large storage requir…
Feature Distillation Interaction Weighting Network for Lightweight Image Super-Resolution
Guangwei Gao, Wenjie Li, Juncheng Li +3
Convolutional neural networks based single-image super-resolution (SISR) has made great progress in recent years. However, it is difficult to apply these methods to real-world scen…