5 papers
World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems
Runze Li, Hongyin Zhang, Junxi Jin +5
Vision-Language-Action (VLA) models have emerged as a promising paradigm for building embodied agents that ground perception and language into action. However, most existing approa…
NFPO: Stabilized Policy Optimization of Normalizing Flow for Robotic Policy Learning
Diyuan Shi, Yiqi Tang, Zifeng Zhuang +1
Deep Reinforcement Learning (DRL) has experienced significant advancements in recent years and has been widely used in many fields. In DRL-based robotic policy learning, however, c…
Variational empirical Bayes variable selection in high-dimensional logistic regression
Yiqi Tang, Ryan Martin
Logistic regression involving high-dimensional covariates is a practically important problem. Often the goal is variable selection, i.e., determining which few of the many covariat…
On the Convergence of Adaptive Gradient Methods for Nonconvex Optimization
Dongruo Zhou, Jinghui Chen, Yuan Cao +2
Adaptive gradient methods are workhorses in deep learning. However, the convergence guarantees of adaptive gradient methods for nonconvex optimization have not been thoroughly stud…
Empirical Bayes inference in sparse high-dimensional generalized linear models
Yiqi Tang, Ryan Martin
High-dimensional linear models have been widely studied, but the developments in high-dimensional generalized linear models, or GLMs, have been slower. In this paper, we propose an…