1 paper · 1 filter
Somnath Kumar, Yash Gadhia, Tanuja Ganu +1
Recent advancements in Multi-modal Large Language Models (MLLMs) have significantly improved their performance in tasks combining vision and language. However, challenges persist i…