1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.RO2026
ST-: Structured SpatioTemporal VLA for Robotic Manipulation
Chuanhao Ma, Hanyu Zhou, Shihan Peng +3
Vision-language-action (VLA) models have achieved great success on general robotic tasks, but still face challenges in fine-grained spatiotemporal manipulation. Typically, existing…
cs.CV2026
Adapting Depth Anything to Adverse Imaging Conditions with Events
Shihan Peng, Yuyang Xiong, Hanyu Zhou +5
Robust depth estimation under dynamic and adverse lighting conditions is essential for robotic systems. Currently, depth foundation models, such as Depth Anything, achieve great su…
cs.CV2024★ 1 cited
CoSEC: A Coaxial Stereo Event Camera Dataset for Autonomous Driving
Shihan Peng, Hanyu Zhou, Hao Dong +5
Conventional frame camera is the mainstream sensor of the autonomous driving scene perception, while it is limited in adverse conditions, such as low light. Event camera with high…