embodied generation 1embodied video generation 1multimodal foundation models 1multi-view consistency 1robotic scene synthesis 1
From the 1 of 9 linked papers with an AI index.
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Embodied VideoAgent: Persistent Memory from Egocentric Videos and Embodied Sensors Enables Dynamic Scene Understanding
Yue Fan, Xiaojian Ma, Rongpeng Su +4
This paper investigates the problem of understanding dynamic 3D scenes from egocentric observations, a key challenge in robotics and embodied AI. Unlike prior studies that explored…
cs.CV2024
Semantic Gaussians: Open-Vocabulary Scene Understanding with 3D Gaussian Splatting
Jun Guo, Xiaojian Ma, Yue Fan +2
Open-vocabulary 3D scene understanding presents a significant challenge in computer vision, with wide-ranging applications in embodied agents and augmented reality systems. Existin…