language-conditioned control 1quadrotor navigation 1video diffusion 1visual grounding 1world-action model 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.RO2026
AeroAct: Action-Centered World-Action Models for Language-Conditioned Quadrotor Flight
Xinhong Zhang, Qiyuan Zhu, Yubo Huang +8
The paper introduces AeroAct, a world-action model that predicts quadrotor flight actions from egocentric video, proprioceptive data, and language commands, using a video diffusion…
cs.CL2025
HI-TransPA: Hearing Impairments Translation Personal Assistant
Zhiming Ma, Shiyu Gan, Junhao Zhao +10
Hearing-impaired individuals often face significant barriers in daily communication due to the inherent challenges of producing clear speech. To address this, we introduce the Omni…