works on

From the 1 of 53 linked papers with an AI index.

activity
20242026
most citedA Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?

1 citations · 3 across the 22 of their papers we have counts for

collaborators
Showing cs.CLShow all

33 papers · 1 filter

cs.CL2026

Humanly: A Configurable and Traceable Environment for Human-AI Collaborative Writing

Shenzhe Zhu, Haoqian Zhang, Xu Yang +7

Teachers, conference chairs, and public readers all judge writing from limited evidence, seeing only a finished document and not the process that produced it. Final text alone cann…

cs.CL2026

ProACT: Towards Breakdown-Aware Proactive Agent in Multi-User Collaboration

Shu Yang, Difei Xu, Jiaxin Pei +1

Conversational agents are increasingly embedded in human collaborative work, yet they remain fundamentally passive and reactive: they respond to explicit user requests rather than…

cs.CL2026

SelfMem: Self-Optimizing Memory for AI Agents

Shu Yang, Junchao Wu, Derek F. Wong +1

While current AI agents support increasingly long context windows, tool use, and skill execution for long-horizon tasks, they still require memory systems to effectively leverage h…

cs.CL2026

AutoMonitor-Bench: Evaluating the Reliability of LLM-Based Misbehavior Monitor

Shu Yang, Jingyu Hu, Tong Li +3

We introduce AutoMonitor-Bench, the first benchmark designed to systematically evaluate the reliability of LLM-based misbehavior monitors across diverse tasks and failure modes. Au…

cs.CL2026

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

Wenrui Zhou, Mohamed Hendy, Shu Yang +5

As video large language models (Video-LLMs) become increasingly integrated into real-world applications that demand grounded multimodal reasoning, ensuring their factual consistenc…

cs.CL20261 cited

A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?

Ada Chen, Yongjiang Wu, Junyuan Zhang +6

Recently, AI-driven interactions with computing devices have advanced from basic prototype tools to sophisticated, LLM-based systems that emulate human-like operations in graphical…