6 papers
LiveEvalBench: Toward Open-World Evaluation for Web Generation
Yiyao Wang, Zhen Wen, Yinghao Tang +5
Large language models are increasingly capable of synthesizing executable frontend projects, yet existing benchmarks still treat web generation as a static evaluation problem. We a…
InterDeepResearch: Enabling Human-Agent Collaborative Information Seeking through Interactive Deep Research
Bo Pan, Lunke Pan, Yitao Zhou +4
Deep research systems powered by LLM agents have transformed complex information seeking by automating the iterative retrieval, filtering, and synthesis of insights from massive-sc…
RAGExplorer: A Visual Analytics System for the Comparative Diagnosis of RAG Systems
Haoyu Tian, Yingchaojie Feng, Zhen Wen +3
The advent of Retrieval-Augmented Generation (RAG) has significantly enhanced the ability of Large Language Models (LLMs) to produce factually accurate and up-to-date responses. Ho…
ConceptViz: A Visual Analytics Approach for Exploring Concepts in Large Language Models
Haoxuan Li, Zhen Wen, Qiqi Jiang +7
Large language models (LLMs) have achieved remarkable performance across a wide range of natural language tasks. Understanding how LLMs internally represent knowledge remains a sig…
VIS-Shepherd: Constructing Critic for LLM-based Data Visualization Generation
Bo Pan, Yixiao Fu, Ke Wang +15
Data visualization generation using Large Language Models (LLMs) has shown promising results but often produces suboptimal visualizations that require human intervention for improv…
Exploring Multimodal Prompt for Visualization Authoring with Large Language Models
Zhen Wen, Luoxuan Weng, Yinghao Tang +5
Recent advances in large language models (LLMs) have shown great potential in automating the process of visualization authoring through simple natural language utterances. However,…