Publications (57)
PCLVis: Visual Analytics of Process Communication Latency in Large-Scale Simulation
Chongke Bi, Xin Gao, Baofeng Fu +4
Large-scale simulations on supercomputers have become important tools for users. However, their scalability remains a problem due to the huge communication cost among parallel proc…
OneLabeler: A Flexible System for Building Data Labeling Tools
Yu Zhang, Yun Wang, Haidong Zhang +3
Labeled datasets are essential for supervised machine learning. Various data labeling tools have been built to collect labels in different usage scenarios. However, developing labe…
The Evolving Duet of Two Modalities: A Survey on Integrating Text and Visualization for Data Communication
Xingyu Lan, Xi Li, Yixing Zhang +3
Text plays a fundamental yet understudied role as a narrative device in data visualization. While existing research has extensively explored text as data input and interaction moda…
Monolithic quantum-dot distributed feedback laser array on silicon
Yi Wang, Siming Chen, Ying Yu +12
Electrically-pumped lasers directly grown on silicon are key devices interfacing silicon microelectronics and photonics. We report here, for the first time, an electrically-pumped,…
Intelligent Drill-Down: Large Language Model-Driven Drill-Down Technique for Human-AI Collaborative Visual Exploration
Zhijun Zheng, Tian Qiu, Yuheng Zhao +1
In visual analytics, applying filters to drill-down and extract higher-value insights is a common and important data analysis method. When the drill-down space becomes excessively…
On-Demand Growth of Semiconductor Heterostructures Guided by Physics-Informed Machine Learning
Chao Shen, Yuan Li, Wenkang Zhan +15
Developing tailored semiconductor heterostructures on demand represents a critical capability for addressing the escalating performance demands in electronic and optoelectronic dev…
AI-Press: A Multi-Agent News Generating and Feedback Simulation System Powered by Large Language Models
Xiawei Liu, Shiyue Yang, Xinnong Zhang +6
The rise of various social platforms has transformed journalism. The growing demand for news content has led to the increased use of large language models (LLMs) in news production…
Reciprocal phase transition-enabled electro-optic modulation
Fang Zou, Lei Zou, Ye Tian +15
Electro-optic (EO) modulation is a well-known and essential topic in the field of communications and sensing. Its ultrahigh efficiency is unprecedentedly desired in the current gre…
Conjugate phase-noise cancellation enables submicrometre dual-comb ranging with free-running megahertz-linewidth lasers
Yue You, Zaifan Wu, Yi Zou +9
Frequency-domain dual-comb ranging combines rapid acquisition with interferometric sensitivity, but high-performance implementations often rely on mutually coherent or actively sta…
MMSciBench: Benchmarking Language Models on Chinese Multimodal Scientific Problems
Xinwu Ye, Chengfan Li, Siming Chen +2
Recent advances in large language models (LLMs) and vision-language models (LVLMs) have shown promise across many tasks, yet their scientific reasoning capabilities remain untested…
Unlocking Scientific Concepts: How Effective Are LLM-Generated Analogies for Student Understanding and Classroom Practice?
Zekai Shao, Siyu Yuan, Lin Gao +3
Teaching scientific concepts is essential but challenging, and analogies help students connect new concepts to familiar ideas. Advancements in large language models (LLMs) enable g…
VizBelle: A Design Space of Embellishments for Data Visualization
Qing Chen, Ziyan Liu, Chengwei Wang +4
Visual embellishments, as a form of non-linguistic rhetorical figures, are used to help convey abstract concepts or attract readers' attention. Creating data visualizations with ap…
Highly reliable, ultra-wideband, isolator-free quantum-dot mode-locked frequency combs for optical interconnects beyond 3.2Tb/s
Shujie Pan, Victoria Cao, Yiheng Feng +6
Quantum dot mode-locked laser-based optical frequency combs are emerging as a critical solution for achieving low-cost, high-efficiency, and large-capacity optical interconnects. T…
NetworkNarratives: Data Tours for Visual Network Exploration and Analysis
Wenchao Li, Sarah Schöttler, James Scott-Brown +4
This paper introduces semi-automatic data tours to aid the exploration of complex networks. Exploring networks requires significant effort and expertise and can be time-consuming a…
Do Language Model Agents Align with Humans in Rating Visualizations? An Empirical Study
Zekai Shao, Yi Shan, Yixuan He +6
Large language models encode knowledge in various domains and demonstrate the ability to understand visualizations. They may also capture visualization design knowledge and potenti…
VisTaxa: Developing a Taxonomy of Historical Visualizations
Yu Zhang, Xinyue Chen, Weili Zheng +4
Historical visualizations are a rich resource for visualization research. While taxonomy is commonly used to structure and understand the design space of visualizations, existing t…
SpecEE: Accelerating Large Language Model Inference with Speculative Early Exiting
Jiaming Xu, Jiayi Pan, Yongkang Zhou +5
Early exiting has recently emerged as a promising technique for accelerating large language models (LLMs) by effectively reducing the hardware computation and memory access. In thi…
DTBIA: An Immersive Visual Analytics System for Brain-Inspired Research
Jun-Hsiang Yao, Mingzheng Li, Jiayi Liu +6
The Digital Twin Brain (DTB) is an advanced artificial intelligence framework that integrates spiking neurons to simulate complex cognitive functions and collaborative behaviors. F…
Fine-Tuned Large Language Model for Visualization System: A Study on Self-Regulated Learning in Education
Lin Gao, Jing Lu, Zekai Shao +7
Large Language Models (LLMs) have shown great potential in intelligent visualization systems, especially for domain-specific applications. Integrating LLMs into visualization syste…
ChartAnno: Evaluating MLLMs for Chart Annotation Generation
Zhenghan Chen, Zekai Shao, Lidan Tan +10
Multimodal large language models (MLLMs) have made significant progress in chart understanding, generation, and editing, but their ability to annotate existing charts remains under…
Beyond the Broadcast: Enhancing VR Tennis Broadcasting through Embedded Visualizations and Camera Techniques
Jun-Hsiang Yao, Jielin Feng, Xinfang Tian +3
Virtual Reality (VR) broadcasting has emerged as a promising medium for providing immersive viewing experiences of major sports events such as tennis. However, current VR broadcast…
System Misuse Detection via Informed Behavior Clustering and Modeling
Linara Adilova, Livin Natious, Siming Chen +2
One of the main tasks of cybersecurity is recognizing malicious interactions with an arbitrary system. Currently, the logging information from each interaction can be collected in…
ChartFI: Benchmarking Faithfulness and Insightfulness of Chart Descriptions from Multimodal Large Language Models
Fen Wang, Zekai Shao, Qiman Kang +5
Chart descriptions are essential for accessibility, cross-modal retrieval, and assisting readers in extracting insights from complex visualizations. As multimodal large language mo…
LightVA: Lightweight Visual Analytics with LLM Agent-Based Task Planning and Execution
Yuheng Zhao, Junjie Wang, Linbin Xiang +5
Visual analytics (VA) requires analysts to iteratively propose analysis tasks based on observations and execute tasks by creating visualizations and interactive exploration to gain…
In-situ Self-optimization of Quantum Dot Emission for Lasers by Machine-Learning Assisted Epitaxy
Chao Shen, Wenkang Zhan, Shujie Pan +12
Traditional methods for optimizing light source emissions rely on a time-consuming trial-and-error approach. While in-situ optimization of light source gain media emission during g…
SceneLoom: Communicating Data with Scene Context
Lin Gao, Leixian Shen, Yuheng Zhao +3
In data-driven storytelling contexts such as data journalism and data videos, data visualizations are often presented alongside real-world imagery to support narrative context. How…
LEVA: Using Large Language Models to Enhance Visual Analytics
Yuheng Zhao, Yixing Zhang, Yu Zhang +5
Visual analytics supports data analysis tasks within complex domain problems. However, due to the richness of data types, visual designs, and interaction designs, users need to rec…
CohortVA: A Visual Analytic System for Interactive Exploration of Cohorts based on Historical Data
Wei Zhang, Jason K. Wong, Xumeng Wang +8
In history research, cohort analysis seeks to identify social structures and figure mobilities by studying the group-based behavior of historical figures. Prior works mainly employ…
Exploring Multi-dimensional Data via Subset Embedding
Peng Xie, Wenyuan Tao, Jie Li +2
Multi-dimensional data exploration is a classic research topic in visualization. Most existing approaches are designed for identifying record patterns in dimensional space or subsp…
Creating Emordle: Animating Word Cloud for Emotion Expression
Liwenhan Xie, Xinhuan Shu, Jeon Cheol Su +3
We propose emordle, a conceptual design that animates wordles (compact word clouds) to deliver their emotional context to the audiences. To inform the design, we first reviewed onl…
From Struggle to Success: Context-Aware Guidance for Screen Reader Users in Computer Use
Nan Chen, Jing Lu, Zilong Wang +3
Equal access to digital technologies is critical for education, employment, and social participation. However, mainstream interfaces are visually oriented, creating steep learning…
ElectionSim: Massive Population Election Simulation Powered by Large Language Model Driven Agents
Xinnong Zhang, Jiayu Lin, Libo Sun +10
The massive population election simulation aims to model the preferences of specific groups in particular election scenarios. It has garnered significant attention for its potentia…
A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models
Junjie Ye, Xuanting Chen, Nuo Xu +12
GPT series models, such as GPT-3, CodeX, InstructGPT, ChatGPT, and so on, have gained considerable attention due to their exceptional natural language processing capabilities. Howe…
Narrative Player: Reviving Data Narratives with Visuals
Zekai Shao, Leixian Shen, Haotian Li +4
Data-rich documents are commonly found across various fields such as business, finance, and science. However, a general limitation of these documents for reading is their reliance…
SD2: Slicing and Dicing Scholarly Data for Interactive Evaluation of Academic Performance
Zhichun Guo, Jun Tao, Siming Chen +2
Comprehensively evaluating and comparing researchers' academic performance is complicated due to the intrinsic complexity of scholarly data. Different scholarly evaluation tasks of…
SoMeLVLM: A Large Vision Language Model for Social Media Processing
Xinnong Zhang, Haoyu Kuang, Xinyi Mou +6
The growth of social media, characterized by its multimodal nature, has led to the emergence of diverse phenomena and challenges, which calls for an effective approach to uniformly…
Rethinking Super-Resolution as Text-Guided Details Generation
Chenxi Ma, Bo Yan, Qing Lin +2
Deep neural networks have greatly promoted the performance of single image super-resolution (SISR). Conventional methods still resort to restoring the single high-resolution (HR) s…
Viewpoint Recommendation for Point Cloud Labeling through Interaction Cost Modeling
Yu Zhang, Xinyi Zhao, Chongke Bi +1
Semantic segmentation of 3D point clouds is important for many applications, such as autonomous driving. To train semantic segmentation models, labeled point cloud segmentation dat…
NotebookRAG: Retrieving Multiple Notebooks to Augment the Generation of EDA Notebooks for Crowd-Wisdom
Yi Shan, Yixuan He, Zekai Shao +2
High-quality exploratory data analysis (EDA) is essential in the data science pipeline, but remains highly dependent on analysts' expertise and effort. While recent LLM-based appro…
ChartInsighter: An Approach for Mitigating Hallucination in Time-series Chart Summary Generation with A Benchmark Dataset
Fen Wang, Bomiao Wang, Xueli Shu +4
Effective chart summary can significantly reduce the time and effort decision makers spend interpreting charts, enabling precise and efficient communication of data insights. Previ…
InfoAlign: A Human-AI Co-Creation System for Storytelling with Infographics
Jielin Feng, Xinwu Ye, Qianhui Li +5
Storytelling infographics are a powerful medium for communicating data-driven stories through visual presentation. However, existing authoring tools lack support for maintaining st…
An Indoor Crowd Movement Trajectory Benchmark Dataset
Ying Zhao, Xin Zhao, Siming Chen +2
In recent years, technologies of indoor crowd positioning and movement data analysis have received widespread attention in the fields of reliability management, indoor navigation,…
Enhanced Radiation Hardness of InAs/GaAs Quantum Dot Lasers for Space Communication
Manyang Li, Jianan Duan, Zhiyong Jin +15
Semiconductor lasers have great potential for space laser communication. However, excessive radiation in space can cause laser failure. In principle, quantum dot (QD) lasers are mo…
ProactiveVA: Proactive Visual Analytics with LLM-Based UI Agent
Yuheng Zhao, Xueli Shu, Liwen Fan +3
Visual analytics (VA) is typically applied to complex data, thus requiring complex tools. While visual analytics empowers analysts in data analysis, analysts may get lost in the co…
OldVisOnline: Curating a Dataset of Historical Visualizations
Yu Zhang, Ruike Jiang, Liwenhan Xie +5
With the increasing adoption of digitization, more and more historical visualizations created hundreds of years ago are accessible in digital libraries online. It provides a unique…
SocioVerse: A World Model for Social Simulation Powered by LLM Agents and A Pool of 10 Million Real-World Users
Xinnong Zhang, Jiayu Lin, Xinyi Mou +18
Social simulation is transforming traditional social science research by modeling human behavior through interactions between virtual individuals and their environments. With recen…
Carbon and Silicon, Coexist or Compete? A Survey on Human-AI Interactions in Agent-based Modeling and Simulation
Ziyue Lin, Siqi Shen, Zichen Cheng +2
Recent interest in human-AI interactions in agent-based modeling and simulation (ABMS) has grown rapidly due to the widespread utilization of large language models (LLMs). ABMS is…
Room-temperature continuous-wave Dirac-vortex topological lasers on silicon
Jingwen Ma, Taojie Zhou, Mingchu Tang +9
Robust laser sources are a fundamental building block for contemporary information technologies. Originating from condensed-matter physics, the concept of topology has recently ent…
SimSpark: Interactive Simulation of Social Media Behaviors
Ziyue Lin, Yi Shan, Lin Gao +2
Understanding user behaviors on social media has garnered significant scholarly attention, enhancing our comprehension of how virtual platforms impact society and empowering decisi…
Novelty Detection in Sequential Data by Informed Clustering and Modeling
Linara Adilova, Siming Chen, Michael Kamp
Novelty detection in discrete sequences is a challenging task, since deviations from the process generating the normal data are often small or intentionally hidden. Novelties can b…
Tower of Babel in Cross-Cultural Communication: A Case Study of #Give Me a Chinese Name# Dialogues During the "TikTok Refugees'' Event
Jielin Feng, Zhibo Yang, Jingyi Zhao +4
The sudden influx of "TikTok refugees'' into the Chinese platform RedNote in early 2025 created an unprecedented, large-scale online cross-cultural communication event between the…
VidAnimator: User-Guided Stylized 3D Character Animation from Human Videos
Xinwu Ye, Jun-Hsiang Yao, Jielin Feng +3
With captivating visual effects, stylized 3D character animation has gained widespread use in cinematic production, advertising, social media, and the potential development of virt…
Wakey-Wakey: Animate Text by Mimicking Characters in a GIF
Liwenhan Xie, Zhaoyu Zhou, Kerun Yu +3
With appealing visual effects, kinetic typography (animated text) has prevailed in movies, advertisements, and social media. However, it remains challenging and time-consuming to c…
SynPO: Synergizing Descriptiveness and Preference Optimization for Video Detailed Captioning
Jisheng Dang, Yizhou Zhang, Hao Ye +6
Fine-grained video captioning aims to generate detailed, temporally coherent descriptions of video content. However, existing methods struggle to capture subtle video dynamics and…
SenseWalk: Agent-Based Semantic Trajectory Simulation Powered by Large Language Models in Zoned Environments
Ziyue Lin, Xinhang Xie, Kangyi Wang +1
Semantic trajectory analysis has recently emerged as an approach for modeling human movement by capturing implicit patterns and behaviors through semantic information (e.g., visito…
KinemaFX: A Kinematic-Driven Interactive System for Particle Effect Exploration and Customization
Yifei Zhang, Lin-Ping Yuan, Yuheng Zhao +2
Particle effects are widely used in games and animation to simulate natural phenomena or stylized visual effects. However, creating effect artworks is challenging for non-expert us…
GeoCamera: Telling Stories in Geographic Visualizations with Camera Movements
Wenchao Li, Zhan Wang, Yun Wang +5
In geographic data videos, camera movements are frequently used and combined to present information from multiple perspectives. However, creating and editing camera movements requi…