2 papers
cs.CV2026
ComplexityWorld: Benchmarking Vision-Language Models on Verifiable Visual Decision Making
Ningxin Pan, Hanyu Li, Yehui Tang
Vision-language models (VLMs) have made rapid progress in visual perception and increasingly support real-world tasks that depend on images. Many such tasks, however, require more…
cs.CV2025
Seed1.5-VL Technical Report
Dong Guo, Faming Wu, Feida Zhu +194
We present Seed1.5-VL, a vision-language foundation model designed to advance general-purpose multimodal understanding and reasoning. Seed1.5-VL is composed with a 532M-parameter v…