1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2023
DiffS2UT: A Semantic Preserving Diffusion Model for Textless Direct Speech-to-Speech Translation
Yongxin Zhu, Zhujin Gao, Xinyuan Zhou +2
While Diffusion Generative Models have achieved great success on image generation tasks, how to efficiently and effectively incorporate them into speech generation especially trans…
cs.AI2023★ 1 cited
Multi-Grained Multimodal Interaction Network for Entity Linking
Pengfei Luo, Tong Xu, Shiwei Wu +3
Multimodal entity linking (MEL) task, which aims at resolving ambiguous mentions to a multimodal knowledge graph, has attracted wide attention in recent years. Though large efforts…