10 papers
Evaluating QAOA expectation values can be as hard as counting optimal solutions
Stuart Hadfield
Evaluating expectation values is a critical task for variational quantum eigensolvers, and for parameterized quantum circuits and other quantum algorithms more generally. We consid…
AgentForge: An Immersive Role-Playing Platform for Learning Agentic Software Engineering
Zihan Fang, Yueke Zhang, Yu Huang
Agentic AI is increasingly used to coordinate planning, implementation, review, and testing in software development, yet it often offers limited transparency into its decisions and…
SCOPE: Leveraging Subgoal Critiques for Code Generation
Yueke Zhang, Yifan Zhang, Zihan Fang +3
Code generation with large language models (LLMs) remains unreliable because generated programs can appear correct while still violating key semantic requirements in the natural la…
DPO-F+: Aligning Code Repair Feedback with Developers' Preferences
Zihan Fang, Yifan Zhang, Yueke Zhang +2
Large Language Models (LLMs) are increasingly used in software engineering tasks, especially code repair. However, developers often struggle to interpret model outputs, limiting ef…
From Conversation to Contribution: Characterizing Coding Agent in Open-Source Software
Zihan Fang, Yueke Zhang, Ningzhi Tang +3
AI coding assistants such as GitHub Copilot and Cursor have evolved from code-suggestion tools into conversational collaborators, enabling vibe-coding workflows in which developers…
EyeMulator: Improving Code Language Models by Mimicking Human Visual Attention
Yifan Zhang, Chen Huang, Yueke Zhang +5
Code Language Models (CodeLLMs) learn token importance from data correlations, whereas human developers attend selectively to semantically salient code. We present EyeMulator, a mo…