2 papers
cs.CV2026
PixelArena: A benchmark for Pixel-Precision Visual Intelligence
Feng Liang, Sizhe Cheng, Chenqi Yi +1
Omni-modal models that have multimodal input and output are emerging. However, benchmarking their multimodal generation, especially in image generation, is challenging due to the s…
cs.CV2025
A Global-Local Cross-Attention Network for Ultra-high Resolution Remote Sensing Image Semantic Segmentation
Chen Yi, Shan LianLei
With the rapid development of ultra-high resolution (UHR) remote sensing technology, the demand for accurate and efficient semantic segmentation has increased significantly. Howeve…