works on

From the 1 of 20 linked papers with an AI index.

activity
20242026
collaborators

20 papers

cs.AI2026

AISPA: User-Centric System Prompt Auditing for Large Language Model Applications

Xiangning Lin, Shenzhe Zhu, Shu Yang +23

The paper presents AISPA, a user‑centric framework for auditing the system prompts that guide large language model behavior in commercial AI products, and reports findings from ana…

cs.CL2026

Humanly: A Configurable and Traceable Environment for Human-AI Collaborative Writing

Shenzhe Zhu, Haoqian Zhang, Xu Yang +7

Teachers, conference chairs, and public readers all judge writing from limited evidence, seeing only a finished document and not the process that produced it. Final text alone cann…

cs.AI2026

Interactive Task Alignment as a POMDP

Andy Dai, Zexue He, Zhenyu Zhang +2

Current benchmarks for language models primarily evaluate execution on fully specified tasks. However, real user tasks are often ambiguous. Users arrive with incomplete, explorator…

cs.CL2026

ProACT: Towards Breakdown-Aware Proactive Agent in Multi-User Collaboration

Shu Yang, Difei Xu, Jiaxin Pei +1

Conversational agents are increasingly embedded in human collaborative work, yet they remain fundamentally passive and reactive: they respond to explicit user requests rather than…

cs.SE2026

ChainSWE: Benchmarking Coding Agents on Multi-Bug Software Maintenance

Qirui Jin, Lingching Tung, Kenan Li +13

Language model (LM) agents are increasingly deployed to maintain codebases over extended periods, fixing streams of related defects while carrying context from one fix to the next.…

cs.AI2026

AutoLab: Can Frontier Models Solve Long-Horizon Auto Research and Engineering Tasks?

Zhangchen Xu, Junda Chen, Yue Huang +16

Scientific and engineering progress is fundamentally a long-horizon iterative process: proposing changes, running experiments, measuring outcomes, and continuously refining artifac…