7 papers
HierDoc: Hierarchical Page-to-Region Evidence Routing for Long-Document Visual Question Answering
Rongjian Gu, Wengang Zhou, Junyu Xiong +4
Multi-page document visual question answering requires locating sparse evidence at both the page and region levels. Existing approaches typically emphasize one level over the other…
SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents
Kuangshi Ai, Haichao Miao, Kaiyuan Tang +13
Recent advances in large language models (LLMs) have enabled agentic systems to translate natural-language intent into executable scientific visualization (SciVis) tasks. Despite r…
LatentGandr: Visual Exploration of Generative AI Latent Space via Local Embeddings
Mingwei Li, Suyang Li, Daisuke Sakurai +2
Generative AI has demonstrated significant potential in creative design, enabling the rapid generation of visual content and imaginative concepts. Although deep AI models achieve e…
TopoPilot: Reliable Conversational Workflow Automation for Topological Data Analysis and Visualization
Nathaniel Gorski, Shusen Liu, Bei Wang
Recent agentic systems demonstrate that large language models can generate scientific visualizations from natural language. However, reliability remains a major limitation: systems…
Teaching People LLM's Errors and Getting it Right
Nathan Stringham, Fateme Hashemi Chaleshtori, Xinyuan Yan +3
People use large language models (LLMs) when they should not. This is partly because they see LLMs compose poems and answer intricate questions, so they understandably, but incorre…
Visual Exploration of Feature Relationships in Sparse Autoencoders with Curated Concepts
Xinyuan Yan, Shusen Liu, Kowshik Thopalli +1
Sparse autoencoders (SAEs) have emerged as a powerful tool for uncovering interpretable features in large language models (LLMs) through the sparse directions they learn. However,…