6 papers
Does language help generalization in vision models?
Benjamin Devillers, Bhavin Choksi, Romain Bielawski +1
Vision models trained on multimodal datasets can benefit from the wide availability of large image-caption datasets. A recent model (CLIP) was found to generalize well in zero-shot…
GAttANet: Global attention agreement for convolutional neural networks
Rufin VanRullen, Andrea Alamia
Transformer attention architectures, similar to those developed for natural language processing, have recently proved efficient also in vision, either in conjunction with or as a r…
Predictive coding feedback results in perceived illusory contours in a recurrent neural network
Zhaoyang Pang, Callum Biggs O'May, Bhavin Choksi +1
Modern feedforward convolutional neural networks (CNNs) can now solve some computer vision tasks at super-human levels. However, these networks only roughly mimic human visual perc…
Reconstructing Natural Scenes from fMRI Patterns using BigBiGAN
Milad Mozafari, Leila Reddy, Rufin VanRullen
Decoding and reconstructing images from brain imaging data is a research area of high interest. Recent progress in deep generative neural networks has introduced new opportunities…
Which Neural Network Architecture matches Human Behavior in Artificial Grammar Learning?
Andrea Alamia, Victor Gauducheau, Dimitri Paisios +1
In recent years artificial neural networks achieved performance close to or better than humans in several domains: tasks that were previously human prerogatives, such as language p…
Reconstructing Faces from fMRI Patterns using Deep Generative Neural Networks
Rufin VanRullen, Leila Reddy
While objects from different categories can be reliably decoded from fMRI brain response patterns, it has proved more difficult to distinguish visually similar inputs, such as diff…