Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
CAPTURE: A Benchmark and Evaluation for LVLMs in CAPTCHA Resolving
Jianyi Zhang, Ziyin Zhou, Xu Ji +2
Benefiting from strong and efficient multi-modal alignment strategies, Large Visual Language Models (LVLMs) are able to simulate human visual and reasoning capabilities, such as so…
cs.AI2025
Oedipus and the Sphinx: Benchmarking and Improving Visual Language Models for Complex Graphic Reasoning
Jianyi Zhang, Xu Ji, Ziyin Zhou +5
Evaluating the performance of visual language models (VLMs) in graphic reasoning tasks has become an important research topic. However, VLMs still show obvious deficiencies in simu…