3 papers
cs.CL2026
Gated Differentiable Working Memory for Long-Context Language Modeling
Lingrui Mei, Shenghua Liu, Yiwei Wang +7
Long contexts challenge transformers: attention scores dilute across thousands of tokens, critical information is often lost in the middle, and models struggle to adapt to novel pa…
cs.CV2025
Assessing the alignment between infants' visual and linguistic experience using multimodal language models
Alvin Wei Ming Tan, Jane Yang, Tarun Sepuri +6
Figuring out which objects or concepts words refer to is a central language learning challenge for young children. Most models of this process posit that children learn early objec…
cs.CV2025
The BabyView dataset: High-resolution egocentric videos of infants' and young children's everyday experiences
Bria Long, Robert Z. Sparks, Violet Xiang +9
Human children far exceed modern machine learning algorithms in their sample efficiency, achieving high performance in key domains with much less data than current models. This ''d…