2 papers
cs.CL2025
Learning to Focus: Causal Attention Distillation via Gradient-Guided Token Pruning
Yiju Guo, Wenkai Yang, Zexu Sun +3
Large language models (LLMs) have demonstrated significant improvements in contextual understanding. However, their ability to attend to truly critical information during long-cont…
cs.IR2025
Robust Uplift Modeling with Large-Scale Contexts for Real-time Marketing
Zexu Sun, Qiyu Han, Minqin Zhu +3
Improving user engagement and platform revenue is crucial for online marketing platforms. Uplift modeling is proposed to solve this problem, which applies different treatments (e.g…