2 papers
cs.CV2026
SV-WAM: An Efficient Surround-View World-Action Model for End-to-End Autonomous Driving
Jinyang Wang, Shiwei Li, Junjian Wang +12
World models (WMs) have demonstrated strong potential for end-to-end autonomous driving by learning predictive representations of future scene dynamics. However, generating future…
cs.CV2024
Detect an Object At Once without Fine-tuning
Junyu Hao, Jianheng Liu, Yongjia Zhao +5
When presented with one or a few photos of a previously unseen object, humans can instantly recognize it in different scenes. Although the human brain mechanism behind this phenome…