3 papers
cs.CV2026
AI-T2I: Aggregating-and-Isolating Cross-Attention to Diffusion Models for Text-to-Image Synthesis
Shipeng Cao, Biao Qian, Haipeng Liu +2
Text-to-image synthesis has made significant progress, benefiting from the strong generative capabilities of diffusion models. However, these models struggle to achieve precise tex…
cs.CL2025
Listening to the Unspoken: Exploring "365" Aspects of Multimodal Interview Performance Assessment
Jia Li, Yang Wang, Wenhao Qian +4
Interview performance assessment is essential for determining candidates' suitability for professional positions. To ensure holistic and fair evaluations, we propose a novel and co…
cs.IR2024
Diffusion-Based Cloud-Edge-Device Collaborative Learning for Next POI Recommendations
Jing Long, Guanhua Ye, Tong Chen +3
The rapid expansion of Location-Based Social Networks (LBSNs) has highlighted the importance of effective next Point-of-Interest (POI) recommendations, which leverage historical ch…