5 papers
When Safety Overrides Vision: Exploring Dynamics between Vision Influence and Safety Alignment in Vision-Language Models
Mehak Gupta, Tanmoy Chakraborty
Aligned vision-language models (VLMs) are designed to balance grounded visual reasoning with safe generation behavior. However, we observe a striking phenomenon: under safety-const…
Brain alignment of reasoning and action representations from vision-language and action models during naturalistic gameplay
Subba Reddy Oota, Anant Khandelwal, Khushbu Pahwa +4
Understanding how humans and artificial intelligence systems predict and plan by interacting with their environment is a fundamental challenge at the intersection of neuroscience a…
How does longer temporal context enhance multimodal narrative video processing in the brain?
Prachi Jindal, Anant Khandelwal, Manish Gupta +3
Understanding how humans and artificial intelligence systems process complex narrative videos is a fundamental challenge at the intersection of neuroscience and machine learning. T…
Multilingual Language Models Encode Script Over Linguistic Structure
Aastha A K Verma, Anwoy Chatterjee, Mehak Gupta +1
Multilingual language models (LMs) organize representations for typologically and orthographically diverse languages into a shared parameter space, yet the nature of this internal…
Linguistic properties and model scale in brain encoding: from small to compressed language models
Subba Reddy Oota, Vijay Rowtula, Satya Sai Srinath Namburi +5
Recent work has shown that scaling large language models (LLMs) improves their alignment with human brain activity, yet it remains unclear what drives these gains and which represe…