1 paper · 1 filter
Jake R. Patock, Nicole Catherine Lewis, Kevin McCoy +3
State-of-the-art (SOTA) image and text generation models are multimodal models that have many similarities to large language models (LLMs). Despite achieving strong performances, l…