3 papers
cs.CV2026
PARSE: Part-Aware Relational Spatial Modeling
Yinuo Bai, Peijun Xu, Kuixiang Shao +5
Inter-object relations underpin spatial intelligence, yet existing representations -- linguistic prepositions or object-level scene graphs -- are too coarse to specify which region…
cs.RO2026
AtomicVLA: Unlocking the Potential of Atomic Skill Learning in Robots
Likui Zhang, Tao Tang, Zhihao Zhan +9
Recent advances in Visual-Language-Action (VLA) models have shown promising potential for robotic manipulation tasks. However, real-world robotic tasks often involve long-horizon,…
cs.GR2024
Sophia-in-Audition: Virtual Production with a Robot Performer
Taotao Zhou, Teng Xu, Dong Zhang +5
We present Sophia-in-Audition (SiA), a new frontier in virtual production, by employing the humanoid robot Sophia within an UltraStage environment composed of a controllable lighti…