2 citations · 3 across the 4 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
Relations, Negations, and Numbers: Looking for Logic in Generative Text-to-Image Models
Colin Conwell, Rupert Tawiah-Quashie, Tomer Ullman
Despite remarkable progress in multi-modal AI research, there is a salient domain in which modern AI continues to lag considerably behind even human children: the reliable deployme…
cs.CV2024★ 1 cited
Using Multimodal Deep Neural Networks to Disentangle Language from Visual Aesthetics
Colin Conwell, Christopher Hamblin, Chelsea Boccagno +4
When we experience a visual stimulus as beautiful, how much of that experience derives from perceptual computations we cannot describe versus conceptual knowledge we can readily tr…