1 paper · 1 filter
Debarpan Bhattacharya, Apoorva Kulkarni, Sriram Ganapathy
The popular success of text-based large language models (LLM) has streamlined the attention of the multimodal community to combine other modalities like vision and audio along with…