2 papers
eess.AS2025
PerformSinger: Multimodal Singing Voice Synthesis Leveraging Synchronized Lip Cues from Singing Performance Videos
Ke Gu, Zhicong Wu, Peng Bai +5
Existing singing voice synthesis (SVS) models largely rely on fine-grained, phoneme-level durations, which limits their practical application. These methods overlook the complement…
cs.CV2024
A Cross-Font Image Retrieval Network for Recognizing Undeciphered Oracle Bone Inscriptions
Zhicong Wu, Qifeng Su, Ke Gu +1
Oracle Bone Inscription (OBI) is the earliest mature writing system in China, which represents a crucial stage in the development of hieroglyphs. Nevertheless, the substantial quan…