works on

From the 1 of 6 linked papers with an AI index.

collaborators

6 papers

cs.AI2026

Agents Don't Just Agree, They Remember: Benchmarking Persistent Sycophancy in Stateful Personal Agents

Xutao Mao, Liangjie Zhao, Leyao Wang +6

The paper defines persistent sycophancy, where personal agents store user‑provided claims in long‑term memory and later repeat them, and introduces the Personal Agent Sycophancy Be…

cs.CR2026

MemMark: State-Evolution Attribution Watermarking for Agent Long-Term Memory Systems

Haobo Zhang, Xutao Mao, Guangyuan Dong +5

Memory-backed agents need provenance that can survive leaked or migrated snapshots, where logs, visible outputs, and trusted metadata may be absent. We propose MemMark, a state-evo…

cs.CL2025

ParlAI Vote: A Web Platform for Analyzing Gender and Political Bias in Large Language Models

Wenjie Lin, Hange Liu, Yingying Zhuang +5

We present ParlAI Vote, an interactive web platform for exploring European Parliament debates and votes, and for testing LLMs on vote prediction and bias analysis. This web system…

cs.SI2025

MindVote: When AI Meets the Wild West of Social Media Opinion

Xutao Mao, Ezra Xuanru Tao, Leyao Wang

Large Language Models (LLMs) are increasingly used as scalable tools for pilot testing, predicting public opinion distributions before deploying costly surveys. To serve as effecti…

cs.SD2025

Benchmarking Fake Voice Detection in the Fake Voice Generation Arms Race

Xutao Mao, Ke Li, Cameron Baird +2

The rapid advancement of fake voice generation technology has ignited a race with detection systems, creating an urgent need to secure the audio ecosystem. However, existing benchm…

cs.IR2025

Towards Bridging Review Sparsity in Recommendation with Textual Edge Graph Representation

Leyao Wang, Xutao Mao, Xuhui Zhan +5

Textual reviews enrich recommender systems with fine-grained preference signals and enhanced explainability. However, in real-world scenarios, users rarely leave reviews, resulting…