most citedCould We Have Had Better Multilingual LLMs If English Was Not the Central Language?

1 citations · 1 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV2024

Generating Faithful and Salient Text from Multimodal Data

Tahsina Hashem, Weiqing Wang, Derry Tanti Wijaya +2

While large multimodal models (LMMs) have obtained strong performance on many multimodal tasks, they may still hallucinate while generating text. Their performance on detecting sal…

cs.CL20241 cited

Could We Have Had Better Multilingual LLMs If English Was Not the Central Language?

Ryandito Diandaru, Lucky Susanto, Zilu Tang +2

Large Language Models (LLMs) demonstrate strong machine translation capabilities on languages they are trained on. However, the impact of factors beyond training data size on trans…

cs.CL2023

Explain-then-Translate: An Analysis on Improving Program Translation with Self-generated Explanations

Zilu Tang, Mayank Agarwal, Alex Shypula +4

This work explores the use of self-generated natural language explanations as an intermediate step for code-to-code translation with language models. Across three types of explanat…

cs.CL2023

Replicable Benchmarking of Neural Machine Translation (NMT) on Low-Resource Local Languages in Indonesia

Lucky Susanto, Ryandito Diandaru, Adila Krisnadhi +2

Neural machine translation (NMT) for low-resource local languages in Indonesia faces significant challenges, including the need for a representative benchmark and limited data avai…

cs.CY2023

A Novel Method for Analysing Racial Bias: Collection of Person Level References

Muhammed Yusuf Kocyigit, Anietie Andy, Derry Wijaya

Long term exposure to biased content in literature or media can significantly influence people's perceptions of reality, leading to the development of implicit biases that are diff…