2 papers
cs.CV2026
Two-Stream Interactive Joint Learning of Scene Parsing and Geometric Vision Tasks
Guanfeng Tang, Hongbo Zhao, Ziwei Long +5
Inspired by the human visual system, which operates on two parallel yet interactive streams for contextual and spatial understanding, this article presents Two Interactive Streams…
cs.CV2026
Rebenchmarking Unsupervised Monocular 3D Occupancy Prediction
Zizhan Guo, Yi Feng, Mengtan Zhang +3
Inferring the 3D structure from a single image, particularly in occluded regions, remains a fundamental yet unsolved challenge in vision-centric autonomous driving. Existing unsupe…