2 papers
cs.SD2026
PACE: A Playback-Aligned Context Engine for LLM-Based Full-Duplex Voice Dialogue
Shibo Wang, Zicheng Zhang, Libo Wang +1
LLM-based full-duplex voice services allow users to speak while the assistant is responding. Because servers can generate output and advance dialogue state faster than clients can…
cs.CV2025
Facial Attractiveness Prediction in Live Streaming: A New Benchmark and Multi-modal Method
Hui Li, Xiaoyu Ren, Hongjiu Yu +7
Facial attractiveness prediction (FAP) has long been an important computer vision task, which could be widely applied in live streaming for facial retouching, content recommendatio…