Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Visual Aesthetic Benchmark: Can Frontier Models Judge Beauty?
Yichen Feng, Yuetai Li, Chunjiang Liu +14
Multimodal large language models (MLLMs) are now routinely deployed for visual understanding, generation, and curation. A substantial fraction of these applications require an expl…
cs.CV2025
VisualSphinx: Large-Scale Synthetic Vision Logic Puzzles for RL
Yichen Feng, Zhangchen Xu, Fengqing Jiang +5
Vision language models (VLMs) are expected to perform effective multimodal reasoning and make logically coherent decisions, which is critical to tasks such as diagram understanding…