1 paper
Run Xu, Lu Li, Rongzhao Zhang +1
Recent multimodal large language models have shown promising ability in generating humorous captions for images, yet they still lack stable control over explicit cultural context,…