Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
Does Data Scaling Lead to Visual Compositional Generalization?
Arnas Uselis, Andrea Dittadi, Seong Joon Oh
Compositional understanding is crucial for human intelligence, yet it remains unclear whether contemporary vision models exhibit it. The dominant machine learning paradigm is built…
cs.LG2025
First Hallucination Tokens Are Different from Conditional Ones
Jakob Snel, Seong Joon Oh
Large Language Models (LLMs) hallucinate, and detecting these cases is key to ensuring trust. While many approaches address hallucination detection at the response or span level, r…
cs.LG2025
Intermediate Layer Classifiers for OOD generalization
Arnas Uselis, Seong Joon Oh
Deep classifiers are known to be sensitive to data distribution shifts, primarily due to their reliance on spurious correlations in training data. It has been suggested that these…