1 paper
Joong Ho Choi, Jiayang Zhao, Avani Appalla +3
Deploying large multimodal language models at scale is constrained by token-based inference costs, yet the cost-performance behavior of visual prompting strategies remains poorly c…