3 papers
cs.CV2025
Assessing the alignment between infants' visual and linguistic experience using multimodal language models
Alvin Wei Ming Tan, Jane Yang, Tarun Sepuri +6
Figuring out which objects or concepts words refer to is a central language learning challenge for young children. Most models of this process posit that children learn early objec…
cs.CV2025
The BabyView dataset: High-resolution egocentric videos of infants' and young children's everyday experiences
Bria Long, Robert Z. Sparks, Violet Xiang +9
Human children far exceed modern machine learning algorithms in their sample efficiency, achieving high performance in key domains with much less data than current models. This ''d…
cs.CL2024
DevBench: A multimodal developmental benchmark for language learning
Alvin Wei Ming Tan, Sunny Yu, Bria Long +5
How (dis)similar are the learning trajectories of vision-language models and children? Recent modeling work has attempted to understand the gap between models' and humans' data eff…