1 paper
Khush Attarde, Yusuf Ali, Megha Thukral +3
MLLMs have shown strong zero-shot capabilities across diverse inputs such as across images, video, audio, and text. A crucial, yet underexplored, application of these models lies i…