27 citations · 27 across the 1 of their papers we have counts for
2 papers
cs.CV2023
On the Cultural Gap in Text-to-Image Generation
Bingshuai Liu, Longyue Wang, Chenyang Lyu +4
One challenge in text-to-image (T2I) generation is the inadvertent reflection of culture gaps present in the training data, which signifies the disparity in generated image quality…
cs.CL2023★ 27 cited
Macaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration
Chenyang Lyu, Minghao Wu, Longyue Wang +5
Although instruction-tuned large language models (LLMs) have exhibited remarkable capabilities across various NLP tasks, their effectiveness on other data modalities beyond text ha…