NeuroSync: Intent-Aware Code-Based Problem Solving via Direct LLM Understanding Modification
arXiv:2508.02823 · doi:10.1145/3746059.3747668
Abstract
Conversational LLMs have been widely adopted by domain users with limited programming experience to solve domain problems. However, these users often face misalignment between their intent and generated code, resulting in frustration and rounds of clarification. This work first investigates the cause of this misalignment, which dues to bidirectional ambiguity: both user intents and coding tasks are inherently nonlinear, yet must be expressed and interpreted through linear prompts and code sequences. To address this, we propose direct intent-task matching, a new human-LLM interaction paradigm that externalizes and enables direct manipulation of the LLM understanding, i.e., the coding tasks and their relationships inferred by the LLM prior to code generation. As a proof-of-concept, this paradigm is then implemented in NeuroSync, which employs a knowledge distillation pipeline to extract LLM understanding, user intents, and their mappings, and enhances the alignment by allowing users to intuitively inspect and edit them via visualizations. We evaluate the algorithmic components of NeuroSync via technical experiments, and assess its overall usability and effectiveness via a user study (N=12). The results show that it enhances intent-task alignment, lowers cognitive effort, and improves coding efficiency.
Accepted in UIST 2025
References in corpus (14)
- Knowledge Distillation: A Survey
- Graph of Thoughts: Solving Elaborate Problems with Large Language Models
- The Metacognitive Demands and Opportunities of Generative AI
- Towards Natural Language Interfaces for Data Visualization: A Survey
- Sensecape: Enabling Multilevel Exploration and Sensemaking with Large Language Models
- Luminate: Structured Generation and Exploration of Design Space with Large Language Models for Human-AI Co-Creation
- "What It Wants Me To Say": Bridging the Abstraction Gap Between End-User Programmers and Code-Generating Large Language Models
- DirectGPT: A Direct Manipulation Interface to Interact with Large Language Models
- VISAR: A Human-AI Argumentative Writing Assistant with Visual Programming and Rapid Draft Prototyping
- Scalability of Network Visualisation from a Cognitive Load Perspective
- WaitGPT: Monitoring and Steering Conversational LLM Agent in Data Analysis with On-the-Fly Code Visualization
- Ivie: Lightweight Anchored Explanations of Just-Generated Code
- Validating AI-Generated Code with Live Programming
- NotePlayer: Engaging Jupyter Notebooks for Dynamic Presentation of Analytical Processes