papers

Publications (57)

cs.CV2025

PCLVis: Visual Analytics of Process Communication Latency in Large-Scale Simulation

Chongke Bi, Xin Gao, Baofeng Fu +4

Large-scale simulations on supercomputers have become important tools for users. However, their scalability remains a problem due to the huge communication cost among parallel proc…

cs.HC2022

OneLabeler: A Flexible System for Building Data Labeling Tools

Yu Zhang, Yun Wang, Haidong Zhang +3

Labeled datasets are essential for supervised machine learning. Various data labeling tools have been built to collect labels in different usage scenarios. However, developing labe…

cs.HC2026

The Evolving Duet of Two Modalities: A Survey on Integrating Text and Visualization for Data Communication

Xingyu Lan, Xi Li, Yixing Zhang +3

Text plays a fundamental yet understudied role as a narrative device in data visualization. While existing research has extensively explored text as data input and interaction moda…

physics.optics2018

Monolithic quantum-dot distributed feedback laser array on silicon

Yi Wang, Siming Chen, Ying Yu +12

Electrically-pumped lasers directly grown on silicon are key devices interfacing silicon microelectronics and photonics. We report here, for the first time, an electrically-pumped,…

cs.HC2026

Intelligent Drill-Down: Large Language Model-Driven Drill-Down Technique for Human-AI Collaborative Visual Exploration

Zhijun Zheng, Tian Qiu, Yuheng Zhao +1

In visual analytics, applying filters to drill-down and extract higher-value insights is a common and important data analysis method. When the drill-down space becomes excessively…

cond-mat.mtrl-sci2025

On-Demand Growth of Semiconductor Heterostructures Guided by Physics-Informed Machine Learning

Chao Shen, Yuan Li, Wenkang Zhan +15

Developing tailored semiconductor heterostructures on demand represents a critical capability for addressing the escalating performance demands in electronic and optoelectronic dev…

cs.CL2024

AI-Press: A Multi-Agent News Generating and Feedback Simulation System Powered by Large Language Models

Xiawei Liu, Shiyue Yang, Xinnong Zhang +6

The rise of various social platforms has transformed journalism. The growing demand for news content has led to the increased use of large language models (LLMs) in news production…

physics.optics2022

Reciprocal phase transition-enabled electro-optic modulation

Fang Zou, Lei Zou, Ye Tian +15

Electro-optic (EO) modulation is a well-known and essential topic in the field of communications and sensing. Its ultrahigh efficiency is unprecedentedly desired in the current gre…

physics.optics2026

Conjugate phase-noise cancellation enables submicrometre dual-comb ranging with free-running megahertz-linewidth lasers

Yue You, Zaifan Wu, Yi Zou +9

Frequency-domain dual-comb ranging combines rapid acquisition with interferometric sensitivity, but high-performance implementations often rely on mutually coherent or actively sta…

cs.LG2025

MMSciBench: Benchmarking Language Models on Chinese Multimodal Scientific Problems

Xinwu Ye, Chengfan Li, Siming Chen +2

Recent advances in large language models (LLMs) and vision-language models (LVLMs) have shown promise across many tasks, yet their scientific reasoning capabilities remain untested…

cs.HC2025

Unlocking Scientific Concepts: How Effective Are LLM-Generated Analogies for Student Understanding and Classroom Practice?

Zekai Shao, Siyu Yuan, Lin Gao +3

Teaching scientific concepts is essential but challenging, and analogies help students connect new concepts to familiar ideas. Advancements in large language models (LLMs) enable g…

cs.HC2022

VizBelle: A Design Space of Embellishments for Data Visualization

Qing Chen, Ziyan Liu, Chengwei Wang +4

Visual embellishments, as a form of non-linguistic rhetorical figures, are used to help convey abstract concepts or attract readers' attention. Creating data visualizations with ap…

cond-mat.mes-hall2025

Highly reliable, ultra-wideband, isolator-free quantum-dot mode-locked frequency combs for optical interconnects beyond 3.2Tb/s

Shujie Pan, Victoria Cao, Yiheng Feng +6

Quantum dot mode-locked laser-based optical frequency combs are emerging as a critical solution for achieving low-cost, high-efficiency, and large-capacity optical interconnects. T…

cs.HC2023

NetworkNarratives: Data Tours for Visual Network Exploration and Analysis

Wenchao Li, Sarah Schöttler, James Scott-Brown +4

This paper introduces semi-automatic data tours to aid the exploration of complex networks. Exploring networks requires significant effort and expertise and can be time-consuming a…

cs.HC2025

Do Language Model Agents Align with Humans in Rating Visualizations? An Empirical Study

Zekai Shao, Yi Shan, Yixuan He +6

Large language models encode knowledge in various domains and demonstrate the ability to understand visualizations. They may also capture visualization design knowledge and potenti…

cs.HC2025

VisTaxa: Developing a Taxonomy of Historical Visualizations

Yu Zhang, Xinyue Chen, Weili Zheng +4

Historical visualizations are a rich resource for visualization research. While taxonomy is commonly used to structure and understand the design space of visualizations, existing t…

cs.DC2025

SpecEE: Accelerating Large Language Model Inference with Speculative Early Exiting

Jiaming Xu, Jiayi Pan, Yongkang Zhou +5

Early exiting has recently emerged as a promising technique for accelerating large language models (LLMs) by effectively reducing the hardware computation and memory access. In thi…

cs.HC2025

DTBIA: An Immersive Visual Analytics System for Brain-Inspired Research

Jun-Hsiang Yao, Mingzheng Li, Jiayi Liu +6

The Digital Twin Brain (DTB) is an advanced artificial intelligence framework that integrates spiking neurons to simulate complex cognitive functions and collaborative behaviors. F…

cs.HC2024

Fine-Tuned Large Language Model for Visualization System: A Study on Self-Regulated Learning in Education

Lin Gao, Jing Lu, Zekai Shao +7

Large Language Models (LLMs) have shown great potential in intelligent visualization systems, especially for domain-specific applications. Integrating LLMs into visualization syste…

cs.AI2026

ChartAnno: Evaluating MLLMs for Chart Annotation Generation

Zhenghan Chen, Zekai Shao, Lidan Tan +10

Multimodal large language models (MLLMs) have made significant progress in chart understanding, generation, and editing, but their ability to annotate existing charts remains under…

cs.HC2025

Beyond the Broadcast: Enhancing VR Tennis Broadcasting through Embedded Visualizations and Camera Techniques

Jun-Hsiang Yao, Jielin Feng, Xinfang Tian +3

Virtual Reality (VR) broadcasting has emerged as a promising medium for providing immersive viewing experiences of major sports events such as tennis. However, current VR broadcast…

cs.CR2019

System Misuse Detection via Informed Behavior Clustering and Modeling

Linara Adilova, Livin Natious, Siming Chen +2

One of the main tasks of cybersecurity is recognizing malicious interactions with an arbitrary system. Currently, the logging information from each interaction can be collected in…

cs.CL2026

ChartFI: Benchmarking Faithfulness and Insightfulness of Chart Descriptions from Multimodal Large Language Models

Fen Wang, Zekai Shao, Qiman Kang +5

Chart descriptions are essential for accessibility, cross-modal retrieval, and assisting readers in extracting insights from complex visualizations. As multimodal large language mo…

cs.HC2025

LightVA: Lightweight Visual Analytics with LLM Agent-Based Task Planning and Execution

Yuheng Zhao, Junjie Wang, Linbin Xiang +5

Visual analytics (VA) requires analysts to iteratively propose analysis tasks based on observations and execute tasks by creating visualizations and interactive exploration to gain…

cond-mat.mes-hall2024

In-situ Self-optimization of Quantum Dot Emission for Lasers by Machine-Learning Assisted Epitaxy

Chao Shen, Wenkang Zhan, Shujie Pan +12

Traditional methods for optimizing light source emissions rely on a time-consuming trial-and-error approach. While in-situ optimization of light source gain media emission during g…

cs.HC2025

SceneLoom: Communicating Data with Scene Context

Lin Gao, Leixian Shen, Yuheng Zhao +3

In data-driven storytelling contexts such as data journalism and data videos, data visualizations are often presented alongside real-world imagery to support narrative context. How…

cs.HC2024

LEVA: Using Large Language Models to Enhance Visual Analytics

Yuheng Zhao, Yixing Zhang, Yu Zhang +5

Visual analytics supports data analysis tasks within complex domain problems. However, due to the richness of data types, visual designs, and interaction designs, users need to rec…

cs.HC2022

CohortVA: A Visual Analytic System for Interactive Exploration of Cohorts based on Historical Data

Wei Zhang, Jason K. Wong, Xumeng Wang +8

In history research, cohort analysis seeks to identify social structures and figure mobilities by studying the group-based behavior of historical figures. Prior works mainly employ…

cs.LG2021

Exploring Multi-dimensional Data via Subset Embedding

Peng Xie, Wenyuan Tao, Jie Li +2

Multi-dimensional data exploration is a classic research topic in visualization. Most existing approaches are designed for identifying record patterns in dimensional space or subsp…

cs.HC2023

Creating Emordle: Animating Word Cloud for Emotion Expression

Liwenhan Xie, Xinhuan Shu, Jeon Cheol Su +3

We propose emordle, a conceptual design that animates wordles (compact word clouds) to deliver their emotional context to the audiences. To inform the design, we first reviewed onl…

cs.HC2026

From Struggle to Success: Context-Aware Guidance for Screen Reader Users in Computer Use

Nan Chen, Jing Lu, Zilong Wang +3

Equal access to digital technologies is critical for education, employment, and social participation. However, mainstream interfaces are visually oriented, creating steep learning…

cs.CL2024

ElectionSim: Massive Population Election Simulation Powered by Large Language Model Driven Agents

Xinnong Zhang, Jiayu Lin, Libo Sun +10

The massive population election simulation aims to model the preferences of specific groups in particular election scenarios. It has garnered significant attention for its potentia…

cs.CL2023

A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models

Junjie Ye, Xuanting Chen, Nuo Xu +12

GPT series models, such as GPT-3, CodeX, InstructGPT, ChatGPT, and so on, have gained considerable attention due to their exceptional natural language processing capabilities. Howe…

cs.HC2025

Narrative Player: Reviving Data Narratives with Visuals

Zekai Shao, Leixian Shen, Haotian Li +4

Data-rich documents are commonly found across various fields such as business, finance, and science. However, a general limitation of these documents for reading is their reliance…

cs.GR2022

SD2: Slicing and Dicing Scholarly Data for Interactive Evaluation of Academic Performance

Zhichun Guo, Jun Tao, Siming Chen +2

Comprehensively evaluating and comparing researchers' academic performance is complicated due to the intrinsic complexity of scholarly data. Different scholarly evaluation tasks of…

cs.CL2024

SoMeLVLM: A Large Vision Language Model for Social Media Processing

Xinnong Zhang, Haoyu Kuang, Xinyi Mou +6

The growth of social media, characterized by its multimodal nature, has led to the emergence of diverse phenomena and challenges, which calls for an effective approach to uniformly…

cs.CV2022

Rethinking Super-Resolution as Text-Guided Details Generation

Chenxi Ma, Bo Yan, Qing Lin +2

Deep neural networks have greatly promoted the performance of single image super-resolution (SISR). Conventional methods still resort to restoring the single high-resolution (HR) s…

cs.HC2026

Viewpoint Recommendation for Point Cloud Labeling through Interaction Cost Modeling

Yu Zhang, Xinyi Zhao, Chongke Bi +1

Semantic segmentation of 3D point clouds is important for many applications, such as autonomous driving. To train semantic segmentation models, labeled point cloud segmentation dat…

cs.HC2026

NotebookRAG: Retrieving Multiple Notebooks to Augment the Generation of EDA Notebooks for Crowd-Wisdom

Yi Shan, Yixuan He, Zekai Shao +2

High-quality exploratory data analysis (EDA) is essential in the data science pipeline, but remains highly dependent on analysts' expertise and effort. While recent LLM-based appro…

cs.CL2025

ChartInsighter: An Approach for Mitigating Hallucination in Time-series Chart Summary Generation with A Benchmark Dataset

Fen Wang, Bomiao Wang, Xueli Shu +4

Effective chart summary can significantly reduce the time and effort decision makers spend interpreting charts, enabling precise and efficient communication of data insights. Previ…

cs.HC2026

InfoAlign: A Human-AI Co-Creation System for Storytelling with Infographics

Jielin Feng, Xinwu Ye, Qianhui Li +5

Storytelling infographics are a powerful medium for communicating data-driven stories through visual presentation. However, existing authoring tools lack support for maintaining st…

cs.OH2021

An Indoor Crowd Movement Trajectory Benchmark Dataset

Ying Zhao, Xin Zhao, Siming Chen +2

In recent years, technologies of indoor crowd positioning and movement data analysis have received widespread attention in the fields of reliability management, indoor navigation,…

physics.app-ph2024

Enhanced Radiation Hardness of InAs/GaAs Quantum Dot Lasers for Space Communication

Manyang Li, Jianan Duan, Zhiyong Jin +15

Semiconductor lasers have great potential for space laser communication. However, excessive radiation in space can cause laser failure. In principle, quantum dot (QD) lasers are mo…

cs.HC2025

ProactiveVA: Proactive Visual Analytics with LLM-Based UI Agent

Yuheng Zhao, Xueli Shu, Liwen Fan +3

Visual analytics (VA) is typically applied to complex data, thus requiring complex tools. While visual analytics empowers analysts in data analysis, analysts may get lost in the co…

cs.HC2023

OldVisOnline: Curating a Dataset of Historical Visualizations

Yu Zhang, Ruike Jiang, Liwenhan Xie +5

With the increasing adoption of digitization, more and more historical visualizations created hundreds of years ago are accessible in digital libraries online. It provides a unique…

cs.CL2025

SocioVerse: A World Model for Social Simulation Powered by LLM Agents and A Pool of 10 Million Real-World Users

Xinnong Zhang, Jiayu Lin, Xinyi Mou +18

Social simulation is transforming traditional social science research by modeling human behavior through interactions between virtual individuals and their environments. With recen…

cs.HC2025

Carbon and Silicon, Coexist or Compete? A Survey on Human-AI Interactions in Agent-based Modeling and Simulation

Ziyue Lin, Siqi Shen, Zichen Cheng +2

Recent interest in human-AI interactions in agent-based modeling and simulation (ABMS) has grown rapidly due to the widespread utilization of large language models (LLMs). ABMS is…

physics.optics2021

Room-temperature continuous-wave Dirac-vortex topological lasers on silicon

Jingwen Ma, Taojie Zhou, Mingchu Tang +9

Robust laser sources are a fundamental building block for contemporary information technologies. Originating from condensed-matter physics, the concept of topology has recently ent…

cs.HC2025

SimSpark: Interactive Simulation of Social Media Behaviors

Ziyue Lin, Yi Shan, Lin Gao +2

Understanding user behaviors on social media has garnered significant scholarly attention, enhancing our comprehension of how virtual platforms impact society and empowering decisi…

cs.LG2023

Novelty Detection in Sequential Data by Informed Clustering and Modeling

Linara Adilova, Siming Chen, Michael Kamp

Novelty detection in discrete sequences is a challenging task, since deviations from the process generating the normal data are often small or intentionally hidden. Novelties can b…

cs.HC2026

Tower of Babel in Cross-Cultural Communication: A Case Study of #Give Me a Chinese Name# Dialogues During the "TikTok Refugees'' Event

Jielin Feng, Zhibo Yang, Jingyi Zhao +4

The sudden influx of "TikTok refugees'' into the Chinese platform RedNote in early 2025 created an unprecedented, large-scale online cross-cultural communication event between the…

cs.HC2025

VidAnimator: User-Guided Stylized 3D Character Animation from Human Videos

Xinwu Ye, Jun-Hsiang Yao, Jielin Feng +3

With captivating visual effects, stylized 3D character animation has gained widespread use in cinematic production, advertising, social media, and the potential development of virt…

cs.HC2023

Wakey-Wakey: Animate Text by Mimicking Characters in a GIF

Liwenhan Xie, Zhaoyu Zhou, Kerun Yu +3

With appealing visual effects, kinetic typography (animated text) has prevailed in movies, advertisements, and social media. However, it remains challenging and time-consuming to c…

cs.AI2026

SynPO: Synergizing Descriptiveness and Preference Optimization for Video Detailed Captioning

Jisheng Dang, Yizhou Zhang, Hao Ye +6

Fine-grained video captioning aims to generate detailed, temporally coherent descriptions of video content. However, existing methods struggle to capture subtle video dynamics and…

cs.HC2026

SenseWalk: Agent-Based Semantic Trajectory Simulation Powered by Large Language Models in Zoned Environments

Ziyue Lin, Xinhang Xie, Kangyi Wang +1

Semantic trajectory analysis has recently emerged as an approach for modeling human movement by capturing implicit patterns and behaviors through semantic information (e.g., visito…

cs.HC2025

KinemaFX: A Kinematic-Driven Interactive System for Particle Effect Exploration and Customization

Yifei Zhang, Lin-Ping Yuan, Yuheng Zhao +2

Particle effects are widely used in games and animation to simulate natural phenomena or stylized visual effects. However, creating effect artworks is challenging for non-expert us…

cs.HC2023

GeoCamera: Telling Stories in Geographic Visualizations with Camera Movements

Wenchao Li, Zhan Wang, Yun Wang +5

In geographic data videos, camera movements are frequently used and combined to present information from multiple perspectives. However, creating and editing camera movements requi…