2 papers
cs.AI2026
Xiaomi-GUI-0 Technical Report
Wanxia Cao, Chengzhen Duan, Pei Fu +29
Graphical user interface (GUI) agents build on vision-language models to complete user tasks end-to-end in real applications through interface actions such as tapping, swiping, tex…
cs.CL2025
Attention Basin: Why Contextual Position Matters in Large Language Models
Zihao Yi, Delong Zeng, Zhenqing Ling +6
The performance of Large Language Models (LLMs) is significantly sensitive to the contextual position of information in the input. To investigate the mechanism behind this position…