2 papers
cs.SD2025
Smark: A Watermark for Text-to-Speech Diffusion Models via Discrete Wavelet Transform
Yichuan Zhang, Chengxin Li, Yujie Gu
Text-to-Speech (TTS) diffusion models generate high-quality speech, which raises challenges for the model intellectual property protection and speech tracing for legal use. Audio w…
cs.AI2025
POI-Enhancer: An LLM-based Semantic Enhancement Framework for POI Representation Learning
Jiawei Cheng, Jingyuan Wang, Yichuan Zhang +4
POI representation learning plays a crucial role in handling tasks related to user mobility data. Recent studies have shown that enriching POI representations with multimodal infor…