human baseline comparison 1literacy benchmarking 1model evaluation 1multimodal language models 1scientific visualization 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.AI2026
Benchmarking Multimodal Large Language Models for Scientific Visualization Literacy
Patrick Phuoc Do, Chau M. Ta, Chaoli Wang
The paper evaluates six multimodal large language models on a standardized scientific visualization literacy test, comparing their performance to human participants and highlightin…
cs.HC2026
HiLSVA: Design and Evaluation of a Human-in-the-Loop Agentic System for Scientific Visualization
Kuangshi Ai, Patrick Phuoc Do, Chaoli Wang
Large language model (LLM) agents enable natural language interaction for scientific visualization (SciVis). Still, prior systems have essentially prioritized autonomy over human a…
cs.HC2026
SVLAT: Scientific Visualization Literacy Assessment Test
Patrick Phuoc Do, Kaiyuan Tang, Kuangshi Ai +1
Scientific visualization (SciVis) has become an essential means for exploring, understanding, and communicating complex scientific phenomena. However, the field still lacks a valid…