39 citations · 41 across the 3 of their papers we have counts for
3 papers
cs.CL2023
Lost in Translation: When GPT-4V(ision) Can't See Eye to Eye with Text. A Vision-Language-Consistency Analysis of VLLMs and Beyond
Xiang Zhang, Senyu Li, Zijun Wu +1
Recent advancements in multimodal techniques open exciting possibilities for models excelling in diverse tasks involving text, audio, and image processing. Models like GPT-4V, blen…
cs.CV2023★ 2 cited
TTIDA: Controllable Generative Data Augmentation via Text-to-Text and Text-to-Image Models
Yuwei Yin, Jean Kaddour, Xiang Zhang +4
Data augmentation has been established as an efficacious approach to supplement useful information for low-resource datasets. Traditional augmentation techniques such as noise inje…
cs.CL2017★ 39 cited
Which Encoding is the Best for Text Classification in Chinese, English, Japanese and Korean?
Xiang Zhang, Yann LeCun
This article offers an empirical study on the different ways of encoding Chinese, Japanese, Korean (CJK) and English languages for text classification. Different encoding levels ar…