2 papers
cs.CV2022
A Creative Industry Image Generation Dataset Based on Captions
Xiang Yuejia, Lv Chuanhao, Liu Qingdazhu +3
Most image generation methods are difficult to precisely control the properties of the generated images, such as structure, scale, shape, etc., which limits its large-scale applica…
cs.CL2022
On Vision Features in Multimodal Machine Translation
Bei Li, Chuanhao Lv, Zefan Zhou +4
Previous work on multimodal machine translation (MMT) has focused on the way of incorporating vision features into translation but little attention is on the quality of vision mode…