5 papers
TaskArtisan: Designing Composable Generative Widgets for LLM-Assisted Analysis
Meng Chen, Amy Pavel
People increasingly use chatbots such as ChatGPT for everyday analysis tasks. While chatbots unify many analysis functions (e.g., scripts, visualizations, summaries), long conversa…
DigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use Coaching
Meng Chen, Anya Ji, Tsung-Han Wu +4
Agents are increasingly capable of automating software tasks, but can they teach humans how to use software themselves? We introduce DigitalCoach, a multimodal dataset of 72 human…
VizCrit: Exploring Strategies for Displaying Computational Feedback in a Visual Design Tool
Mingyi Li, Mengyi Chen, Sarah Luo +5
Visual design instructors often provide multi-modal feedback, mixing annotations with text. Prior theory emphasizes the importance of actionable feedback, where "actionability" lie…
Surfacing Variations to Calibrate Perceived Reliability of MLLM-generated Image Descriptions
Meng Chen, Akhil Iyer, Amy Pavel
Multimodal large language models (MLLMs) provide new opportunities for blind and low vision (BLV) people to access visual information in their daily lives. However, these models of…
Lotus: Creating Short Videos From Long Videos With Abstractive and Extractive Summarization
Aadit Barua, Karim Benharrak, Meng Chen +2
Short-form videos are popular on platforms like TikTok and Instagram as they quickly capture viewers' attention. Many creators repurpose their long-form videos to produce short-for…