11 papers
IntentLint: Supporting Intent Scaffolding and Prompt-time Linting in Human-AI Collaborative Data Analysis
Felicia Li Feng, Jian Zhao, Anamaria Crisan
In human-AI collaborative data analysis, as analyses rapidly evolve, the artifacts meant to capture shared understanding often become incomplete or difficult to interpret, leading…
STAGE-Claw: Automated State-based Agent Benchmarking for Realistic Scenarios
Sirui Liang, Bohan Yu, Peiyu Wang +8
Large language models are increasingly used to power personal agents for everyday applications, but evaluating these agents remains a challenge. Existing benchmarks still rely on s…
Chain-of-Thought Hijacking
Jianli Zhao, Tingchen Fu, Rylan Schaeffer +2
Large Reasoning Models (LRMs) improve task performance through extended inference-time reasoning. Although previous studies suggest that longer reasoning should lead to more robust…
ChipLingo: A Systematic Training Framework for Large Language Models in EDA
Lei Li, Xingwen Yu, Jianguo Ni +4
With the rapid advancement of semiconductor technology, Electronic Design Automation (EDA) has become an increasingly knowledge-intensive and document-driven engineering domain. Al…
MindTrellis: Co-Creating Knowledge Structures with AI through Interactive Visual Exploration
Xiang Li, Cara Li, Emily Kuang +2
Knowledge workers face increasing challenges in synthesizing information from multiple documents into structured conceptual understanding. This process is inherently iterative: use…
GameUIAgent: An LLM-Powered Framework for Automated Game UI Design with Structured Intermediate Representation
Wei Zeng, Fengwei An, Zhen Liu +1
Game UI design requires consistent visual assets across rarity tiers yet remains a predominantly manual process. We present GameUIAgent, an LLM-powered agentic framework that trans…