5 citations · 9 across the 10 of their papers we have counts for
15 papers
Beyond English-Centric LLMs: What Language Do Multilingual Language Models Think in?
Chengzhi Zhong, Fei Cheng, Qianying Liu +5
In this study, we investigate whether non-English-centric LLMs, despite their strong performance, `think' in their respective dominant language: more precisely, `think' refers to h…
J-CRe3: A Japanese Conversation Dataset for Real-world Reference Resolution
Nobuhiro Ueda, Hideko Habe, Yoko Matsui +5
Understanding expressions that refer to the physical world is crucial for such human-assisting systems in the real world, as robots that must perform actions that are expected by u…
AcTED: Automatic Acquisition of Typical Event Duration for Semi-supervised Temporal Commonsense QA
Felix Virgo, Fei Cheng, Lis Kanashiro Pereira +3
We propose a voting-driven semi-supervised approach to automatically acquire the typical duration of an event and use it as pseudo-labeled data. The human evaluation demonstrates t…
Rapidly Developing High-quality Instruction Data and Evaluation Benchmark for Large Language Models with Minimal Human Effort: A Case Study on Japanese
Yikun Sun, Zhen Wan, Nobuhiro Ueda +4
The creation of instruction data and evaluation benchmarks for serving Large language models often involves enormous human annotation. This issue becomes particularly pronounced wh…
RecMind: Japanese Movie Recommendation Dialogue with Seeker's Internal State
Takashi Kodama, Hirokazu Kiyomaru, Yin Jou Huang +1
Humans pay careful attention to the interlocutor's internal state in dialogues. For example, in recommendation dialogues, we make recommendations while estimating the seeker's inte…
Bilingual Corpus Mining and Multistage Fine-Tuning for Improving Machine Translation of Lecture Transcripts
Haiyue Song, Raj Dabre, Chenhui Chu +2
Lecture transcript translation helps learners understand online courses, however, building a high-quality lecture machine translation system lacks publicly available parallel corpo…