15 citations · 23 across the 8 of their papers we have counts for
8 papers
Orchestrating Spatial Semantics via a Zone-Graph Paradigm for Intricate Indoor Scene Generation
Meisheng Zhang, Shizhao Sun, Yang Zhao +3
Autonomous 3D indoor scene synthesis breaks down in non-convex rooms with tightly coupled spatial constraints. Data-driven generators lack topological priors for long-horizon plann…
VideoWeaver: Multimodal Multi-View Video-to-Video Transfer for Embodied Agents
George Eskandar, Fengyi Shen, Mohammad Altillawi +4
Recent progress in video-to-video (V2V) translation has enabled realistic resimulation of embodied AI demonstrations, a capability that allows pretrained robot policies to be trans…
ReMem-VLA: Empowering Vision-Language-Action Model with Memory via Dual-Level Recurrent Queries
Hang Li, Fengyi Shen, Dong Chen +6
Vision-language-action (VLA) models for closed-loop robot control are typically cast under the Markov assumption, making them prone to errors on tasks requiring historical context.…
ConfCtrl: Enabling Precise Camera Control in Video Diffusion via Confidence-Aware Interpolation
Liudi Yang, George Eskandar, Fengyi Shen +5
We address the challenge of novel view synthesis from only two input images under large viewpoint changes. Existing regression-based methods lack the capacity to reconstruct unseen…
CE-NPBG: Connectivity Enhanced Neural Point-Based Graphics for Novel View Synthesis in Autonomous Driving Scenes
Mohammad Altillawi, Fengyi Shen, Liudi Yang +2
Current point-based approaches encounter limitations in scalability and rendering quality when using large 3D point cloud maps because using them directly for novel view synthesis…
ConsistentDreamer: View-Consistent Meshes Through Balanced Multi-View Gaussian Optimization
Onat Şahin, Mohammad Altillawi, George Eskandar +2
Recent advances in diffusion models have significantly improved 3D generation, enabling the use of assets generated from an image for embodied AI simulations. However, the one-to-m…