2 papers
cs.CV2026
Vision Foundation Models as Generalist Tokenizers for Image Generation
Anlin Zheng, Qi Han, Xin Wen +5
In this work, we explore the largely unexplored direction of building a generalist image tokenizer directly on top of a frozen vision foundation model (VFM). To build this tokenize…
cs.CV2026
Failure Modes for Deep Learning-Based Online Mapping: How to Measure and Address Them
Michael Hubbertz, Qi Han, Tobias Meisen
Deep learning-based online mapping has emerged as a cornerstone of autonomous driving, yet these models frequently fail to generalize beyond familiar environments. We propose a fra…