2 papers
cs.RO2026
ST-: Structured SpatioTemporal VLA for Robotic Manipulation
Chuanhao Ma, Hanyu Zhou, Shihan Peng +3
Vision-language-action (VLA) models have achieved great success on general robotic tasks, but still face challenges in fine-grained spatiotemporal manipulation. Typically, existing…
cs.CV2026
Adapting Depth Anything to Adverse Imaging Conditions with Events
Shihan Peng, Yuyang Xiong, Hanyu Zhou +5
Robust depth estimation under dynamic and adverse lighting conditions is essential for robotic systems. Currently, depth foundation models, such as Depth Anything, achieve great su…