2 papers
cs.SD2023
UniBriVL: Robust Universal Representation and Generation of Audio Driven Diffusion Models
Sen Fang, Bowen Gao, Yangjian Wu +1
Multimodal large models have been recognized for their advantages in various performance and downstream tasks. The development of these models is crucial towards achieving general…
cs.SD2023
Exploring Efficient-Tuned Learning Audio Representation Method from BriVL
Sen Fang, Yangjian Wu, Bowen Gao +2
Recently, researchers have gradually realized that in some cases, the self-supervised pre-training on large-scale Internet data is better than that of high-quality/manually labeled…