Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Urban Socio-Semantic Segmentation with Vision-Language Reasoning
Yu Wang, Yi Wang, Rui Dai +4
As hubs of human activity, urban surfaces consist of a wealth of semantic entities. Segmenting these various entities from satellite imagery is crucial for a range of downstream ap…
cs.CV2026
Q-Hawkeye: Reliable Visual Policy Optimization for Image Quality Assessment
Wulin Xie, Rui Dai, Ruidong Ding +4
Image Quality Assessment (IQA) predicts perceptual quality scores consistent with human judgments. Recent RL-based IQA methods built on MLLMs focus on generating visual quality des…
cs.CV2026
Code2World: A GUI World Model via Renderable Code Generation
Yuhao Zheng, Li'an Zhong, Yi Wang +6
Autonomous GUI agents interact with environments by perceiving interfaces and executing actions. As a virtual sandbox, the GUI World model empowers agents with human-like foresight…