From the 1 of 11 linked papers with an AI index.
11 papers
Mitigating Compounding Error via Video Representation Regularization
Taiye Chen, Qi Zhang, Yisen Wang
The paper studies why autoregressive video generation models accumulate errors over time and introduces a lightweight regularization that stabilizes hidden representations, reducin…
DWM: Separating World Effects from Actions in Latent World Models
Yi-Ge Zhang, Tianqi Du, Qi Zhang +1
Latent world models underpin much of modern model-based control, yet current action-conditioned formulations supervise the next-latent transition with a single, undifferentiated ta…
SAGE: Subgoal-Conditioned Action Generation for Latent World Model Planning
Letian Cheng, Qi Zhang, Yisen Wang
Latent world models have emerged as a powerful planning paradigm by learning action-conditioned predictive dynamics and using them as internal simulators to imagine and evaluate ca…
A Generalization Theory for JEPA-Based World Models
Jingyi Cui, Qi Zhang, Hongwei Wen +1
Joint Embedding Predictive Architectures (JEPAs) have recently emerged as a promising paradigm for world modeling by learning predictive dynamics in a latent space rather than gene…
AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning
Honglin Guo, Qi Zhang, Yu Zhang +6
Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grounded reasoning: locating spa…
Group-Graph Policy Optimization for Long-Horizon Agentic Reinforcement Learning
Yunan Wang, Minghui Song, Zihan Zhang +6
Group-based Reinforcement Learning (RL) has significantly enhanced Large Language Models (LLMs) in agentic scenarios. To achieve finer-grained policy updates, recent agentic RL fra…