From the 1 of 10 linked papers with an AI index.
10 papers
Mitigating Compounding Error via Video Representation Regularization
Taiye Chen, Qi Zhang, Yisen Wang
The paper studies why autoregressive video generation models accumulate errors over time and introduces a lightweight regularization that stabilizes hidden representations, reducin…
DWM: Separating World Effects from Actions in Latent World Models
Yi-Ge Zhang, Tianqi Du, Qi Zhang +1
Latent world models underpin much of modern model-based control, yet current action-conditioned formulations supervise the next-latent transition with a single, undifferentiated ta…
SAGE: Subgoal-Conditioned Action Generation for Latent World Model Planning
Letian Cheng, Qi Zhang, Yisen Wang
Latent world models have emerged as a powerful planning paradigm by learning action-conditioned predictive dynamics and using them as internal simulators to imagine and evaluate ca…
A Generalization Theory for JEPA-Based World Models
Jingyi Cui, Qi Zhang, Hongwei Wen +1
Joint Embedding Predictive Architectures (JEPAs) have recently emerged as a promising paradigm for world modeling by learning predictive dynamics in a latent space rather than gene…
Leveraging Data Symmetries to Select an Optimal Subset of Training Data under Label Noise
Kumar Shubham, Pavan Karjol, Kiran M K +1
The performance of machine learning models often relies on large labeled datasets; however, data collected from diverse sources can contain label noise. Recent work has shown that,…
A Unified Theory of Sparse Dictionary Learning in Mechanistic Interpretability: Piecewise Biconvexity and Spurious Minima
Yiming Tang, Harshvardhan Saini, Zhaoqian Yao +6
As AI models achieve remarkable capabilities across diverse domains, understanding what representations they learn and how they encode concepts has become increasingly important fo…