3 citations · 6 across the 12 of their papers we have counts for
3 papers · 1 filter
EncQA: Benchmarking Vision-Language Models on Visual Encodings for Charts
Kushin Mukherjee, Donghao Ren, Dominik Moritz +1
Multimodal vision-language models (VLMs) continue to achieve ever-improving scores on chart understanding benchmarks. Yet, we find that this progress does not fully capture the bre…
Large Language Models estimate fine-grained human color-concept associations
Kushin Mukherjee, Timothy T. Rogers, Karen B. Schloss
Concepts, both abstract and concrete, elicit a distribution of association strengths across perceptual color space, which influence aspects of visual cognition ranging from object…
SEVA: Leveraging sketches to evaluate alignment between human and machine visual abstraction
Kushin Mukherjee, Holly Huey, Xuanchen Lu +4
Sketching is a powerful tool for creating abstract images that are sparse but meaningful. Sketch understanding poses fundamental challenges for general-purpose vision algorithms be…