3 citations · 3 across the 4 of their papers we have counts for
4 papers
Thinking in Directivity: Speech Large Language Model for Multi-Talker Directional Speech Recognition
Jiamin Xie, Ju Lin, Yiteng Huang +8
Recent studies have demonstrated that prompting large language models (LLM) with audio encodings enables effective speech recognition capabilities. However, the ability of Speech L…
Accelerating Diffusion-based Super-Resolution with Dynamic Time-Spatial Sampling
Rui Qin, Qijie Wang, Ming Sun +3
Diffusion models have gained attention for their success in modeling complex distributions, achieving impressive perceptual quality in SR tasks. However, existing diffusion-based S…
A New Dataset and Framework for Real-World Blurred Images Super-Resolution
Rui Qin, Ming Sun, Chao Zhou +1
Recent Blind Image Super-Resolution (BSR) methods have shown proficiency in general images. However, we find that the efficacy of recent methods obviously diminishes when employed…
CasSR: Activating Image Power for Real-World Image Super-Resolution
Haolan Chen, Jinhua Hao, Kai Zhao +4
The objective of image super-resolution is to generate clean and high-resolution images from degraded versions. Recent advancements in diffusion modeling have led to the emergence…