2 papers
cs.LG2025
HoPE: Hybrid of Position Embedding for Long Context Vision-Language Models
Haoran Li, Yingjie Qin, Baoyuan Ou +2
Vision-Language Models (VLMs) have made significant progress in multimodal tasks. However, their performance often deteriorates in long-context scenarios, particularly long videos.…
cs.AI2025
GIST: Cross-Domain Click-Through Rate Prediction via Guided Content-Behavior Distillation
Wei Xu, Haoran Li, Baoyuan Ou +4
Cross-domain Click-Through Rate prediction aims to tackle the data sparsity and the cold start problems in online advertising systems by transferring knowledge from source domains…