2 papers
cs.CV2026
StepX-Edge: An On-Device UI Vision-Language Model via Architecture-Training-Deployment Co-Design
Yin Wang, Haotian Hu, Jineng Han +4
Deploying a vision-language model with full UI understanding on end devices has long been trapped between accuracy and efficiency: on one side is the accuracy bar for OCR, screen u…
cs.AI2026
SpecPrefetch: Parameter-Efficient Expert Prefetching for Sparse MoE Foundation Models
Jinwei Kong, Runqi Meng, Fanyi Wang +4
Sparse Mixture-of-Experts (MoE) models expand foundation model capacity through conditional expert activation, but their full expert pools remain difficult to deploy under limited…