6 citations · 8 across the 7 of their papers we have counts for
7 papers
Task-Aware Dynamic Transformer for Efficient Arbitrary-Scale Image Super-Resolution
Tianyi Xu, Yiji Zhou, Xiaotao Hu +4
Arbitrary-scale super-resolution (ASSR) aims to learn a single model for image super-resolution at arbitrary magnifying scales. Existing ASSR networks typically comprise an off-the…
Towards Rehearsal-Free Multilingual ASR: A LoRA-based Case Study on Whisper
Tianyi Xu, Kaixun Huang, Pengcheng Guo +4
Pre-trained multilingual speech foundation models, like Whisper, have shown impressive performance across different languages. However, adapting these models to new or specific lan…
Articulated Object Manipulation with Coarse-to-fine Affordance for Mitigating the Effect of Point Cloud Noise
Suhan Ling, Yian Wang, Shiguang Wu +5
3D articulated objects are inherently challenging for manipulation due to the varied geometries and intricate functionalities associated with articulated objects.Point-level afford…
4K-Resolution Photo Exposure Correction at 125 FPS with ~8K Parameters
Yijie Zhou, Chao Li, Jin Liang +3
The illumination of improperly exposed photographs has been widely corrected using deep convolutional neural networks or Transformers. Despite with promising performance, these met…
A First Order Meta Stackelberg Method for Robust Federated Learning
Yunian Pan, Tao Li, Henger Li +3
Previous research has shown that federated learning (FL) systems are exposed to an array of security risks. Despite the proposal of several defensive strategies, they tend to be no…
MAVD: The First Open Large-Scale Mandarin Audio-Visual Dataset with Depth Information
Jianrong Wang, Yuchen Huo, Li Liu +3
Audio-visual speech recognition (AVSR) gains increasing attention from researchers as an important part of human-computer interaction. However, the existing available Mandarin audi…