7 citations · 15 across the 3 of their papers we have counts for
3 papers
eess.AS2024★ 6 cited
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
Ye Bai, Jingping Chen, Jitong Chen +52
Modern automatic speech recognition (ASR) model is required to accurately transcribe diverse speech signals (from different domains, languages, accents, etc) given the specific con…
eess.AS2024★ 7 cited
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Philip Anastassiou, Jiawei Chen, Jitong Chen +43
We introduce Seed-TTS, a family of large-scale autoregressive text-to-speech (TTS) models capable of generating speech that is virtually indistinguishable from human speech. Seed-T…
cs.CL2023★ 2 cited
Universal Multi-modal Entity Alignment via Iteratively Fusing Modality Similarity Paths
Bolin Zhu, Xiaoze Liu, Xin Mao +4
The objective of Entity Alignment (EA) is to identify equivalent entity pairs from multiple Knowledge Graphs (KGs) and create a more comprehensive and unified KG. The majority of E…