works on

From the 1 of 6 linked papers with an AI index.

collaborators

6 papers

stat.ML2026

Robust Estimation of Sparse Numerical Vectors under Local Differential Privacy

Puning Zhao, Zhikun Zhang, Shaowei Wang +5

The paper proposes a Randomized Projection with Clipping (RPC) method to robustly estimate sparse numerical vectors under local differential privacy, providing theoretical error gu…

cs.AI2026

Executable Agentic Memory for GUI Agent

Zerui Qin, Sheng Yue, Xingyuan Hua +2

Modern GUI agents typically rely on a model-centric and step-wise interaction paradigm, where LLMs must re-interpret the UI and re-decide actions at every screen, which is fragile…

cs.AI2026

Learning to Explore: Scaling Agentic Reasoning via Exploration-Aware Policy Optimization

Xingyuan Hua, Sheng Yue, Ju Ren

Recent advancements in agentic test-time scaling allow models to gather environmental feedback before committing to final actions. A key limitation of existing methods is that they…

cs.LG2026

AdamO: A Collapse-Suppressed Optimizer for Offline RL

Nan Qiao, Sheng Yue, Shuning Wang +1

Offline reinforcement learning (RL) can fail spectacularly when bootstrapped temporal-difference (TD) updates amplify their own errors, driving the critic toward extreme and unusab…

cs.LG2025

FOVA: Offline Federated Reinforcement Learning with Mixed-Quality Data

Nan Qiao, Sheng Yue, Ju Ren +1

Offline Federated Reinforcement Learning (FRL), a marriage of federated learning and offline reinforcement learning, has attracted increasing interest recently. Albeit with some ad…

cs.LG2025

AugFL: Augmenting Federated Learning with Pretrained Models

Sheng Yue, Zerui Qin, Yongheng Deng +3

Federated Learning (FL) has garnered widespread interest in recent years. However, owing to strict privacy policies or limited storage capacities of training participants such as I…