1 paper · 1 filter
Ramaneswaran Selvakumar, Kaousheik Jayakumar, S Sakshi +3
Audio-Visual Large Language Models (AVLLMs) are emerging as unified interfaces to multimodal perception. We present the first mechanistic interpretability study of AVLLMs, analyzin…