3 papers
cs.HC2025
An Evaluation-Centric Paradigm for Scientific Visualization Agents
Kuangshi Ai, Haichao Miao, Zhimin Li +2
Recent advances in multi-modal large language models (MLLMs) have enabled increasingly sophisticated autonomous visualization agents capable of translating user intentions into dat…
cs.HC2025
ParaView-MCP: An Autonomous Visualization Agent with Direct Tool Use
Shusen Liu, Haichao Miao, Peer-Timo Bremer
While powerful and well-established, tools like ParaView present a steep learning curve that discourages many potential users. This work introduces ParaView-MCP, an autonomous agen…
cs.HC2025
See or Recall: A Sanity Check for the Role of Vision in Solving Visualization Question Answer Tasks with Multimodal LLMs
Zhimin Li, Haichao Miao, Xinyuan Yan +3
Recent developments in multimodal large language models (MLLM) have equipped language models to reason about vision and language jointly. This permits MLLMs to both perceive and an…