2 papers
cs.CV2026
How Many Visual Tokens Do Multimodal Language Models Need? Scaling Visual Token Pruning with F^3A
YiJie Huang, Yiqun Zhang, Zhuoyue Jia +7
Vision-language models improve perception by feeding increasingly long visual token sequences into language backbones, but the resulting inference cost raises a basic scaling quest…
cs.CL2025
Affective Computing in the Era of Large Language Models: A Survey from the NLP Perspective
Yiqun Zhang, Xiaocui Yang, Xingle Xu +8
Affective Computing (AC) integrates computer science, psychology, and cognitive science to enable machines to recognize, interpret, and simulate human emotions across domains such…