3 papers
cs.CL2026
ARKD: Adaptive Reinforcement Learning-Guided Bidirectional KL Divergence Distillation for Text Generation
Zilong Liu, Xuewen Zhang, Jinrui Xing +3
Knowledge distillation (KD) is a key technique for compressing Large Language Models (LLMs), yet methods relying on a single KL objective often fail to balance primary distribution…
cs.IR2026
Rethinking Sales Lead Scoring with LLM-based Hierarchical Preference Ranking
Chenyu Zhang, Yiwen Liu, Yin Sun +4
Sales lead conversion in high-stakes domains (e.g., automotive, real estate) differs fundamentally from e-commerce recommendation due to prolonged decision cycles and multi-stage f…
cs.CV2025
MASR: Self-Reflective Reasoning through Multimodal Hierarchical Attention Focusing for Agent-based Video Understanding
Shiwen Cao, Zhaoxing Zhang, Junming Jiao +4
Even in the era of rapid advances in large models, video understanding remains a highly challenging task. Compared to texts or images, videos commonly contain more information with…