2 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 1 cited
TRINS: Towards Multimodal Language Models that Can Read
Ruiyi Zhang, Yanzhe Zhang, Jian Chen +4
Large multimodal language models have shown remarkable proficiency in understanding and editing images. However, a majority of these visually-tuned models struggle to comprehend th…
cs.CL2023★ 2 cited
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints
Albert Lu, Hongxin Zhang, Yanzhe Zhang +2
The limits of open-ended generative models are unclear, yet increasingly important. What causes them to succeed and what causes them to fail? In this paper, we take a prompt-centri…