Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
ReMI: A Dataset for Reasoning with Multiple Images
Mehran Kazemi, Nishanth Dikkala, Ankit Anand +8
With the continuous advancement of large language models (LLMs), it is essential to create new benchmarks to effectively evaluate their expanding capabilities and identify areas fo…
cs.CV2023
GeomVerse: A Systematic Evaluation of Large Models for Geometric Reasoning
Mehran Kazemi, Hamidreza Alvari, Ankit Anand +3
Large language models have shown impressive results for multi-hop mathematical reasoning when the input question is only textual. Many mathematical reasoning problems, however, con…