Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
BabyVLM-V2: Toward Developmentally Grounded Pretraining and Benchmarking of Vision Foundation Models
Shengao Wang, Wenqi Wang, Zecheng Wang +20
Early children's developmental trajectories set up a natural goal for sample-efficient pretraining of vision foundation models. We introduce BabyVLM-V2, a developmentally grounded…
cs.CV2025
Modeling Urban Food Insecurity with Google Street View Images
David Li
Food insecurity is a significant social and public health issue that plagues many urban metropolitan areas around the world. Existing approaches to identifying food insecurity rely…
cs.CV2025
Medical Semantic Segmentation with Diffusion Pretrain
David Li, Anvar Kurmukov, Mikhail Goncharov +2
Recent advances in deep learning have shown that learning robust feature representations is critical for the success of many computer vision tasks, including medical image segmenta…