collaborators

11 papers

cs.CV2026

Mitigating Compounding Error via Video Representation Regularization

Taiye Chen, Qi Zhang, Yisen Wang

The paper studies why autoregressive video generation models accumulate errors over time and introduces a lightweight regularization that stabilizes hidden representations, reducin…

cs.AI2026

DWM: Separating World Effects from Actions in Latent World Models

Yi-Ge Zhang, Tianqi Du, Qi Zhang +1

Latent world models underpin much of modern model-based control, yet current action-conditioned formulations supervise the next-latent transition with a single, undifferentiated ta…

cs.AI2026

SAGE: Subgoal-Conditioned Action Generation for Latent World Model Planning

Letian Cheng, Qi Zhang, Yisen Wang

Latent world models have emerged as a powerful planning paradigm by learning action-conditioned predictive dynamics and using them as internal simulators to imagine and evaluate ca…

cs.LG2026

A Generalization Theory for JEPA-Based World Models

Jingyi Cui, Qi Zhang, Hongwei Wen +1

Joint Embedding Predictive Architectures (JEPAs) have recently emerged as a promising paradigm for world modeling by learning predictive dynamics in a latent space rather than gene…

cs.CL2026

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

Honglin Guo, Qi Zhang, Yu Zhang +6

Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grounded reasoning: locating spa…

cs.LG2026

Group-Graph Policy Optimization for Long-Horizon Agentic Reinforcement Learning

Yunan Wang, Minghui Song, Zihan Zhang +6

Group-based Reinforcement Learning (RL) has significantly enhanced Large Language Models (LLMs) in agentic scenarios. To achieve finer-grained policy updates, recent agentic RL fra…