activity
20242026
most citedRobotic Backchanneling in Online Conversation Facilitation: A Cross-Generational Study

3 citations · 13 across the 30 of their papers we have counts for

collaborators

33 papers

cs.MM2026

Multimodal Temporal Modeling for Continuous Group Emotion Recognition in Multi-party Dialogues

Soma Iwata, Koji Inoue, Muyun Wu +3

To realize natural behavior in dialogue agents in multi-party dialogue scenarios, it is important to understand group emotion such as valence and arousal as a whole. Most prior wor…

cs.CL2026

TEIDAN: A Multilingual Multiparty Dialogue Corpus

Taiga Mori, Koji Inoue, Mikey Elmers +2

Multi-party interaction is a central setting for human communication and a necessary target for human-agent interaction systems that must participate in group conversation. Yet ava…

cs.RO2026

Human-robot conversation with multiple participants in noisy public spaces

Divesh Lala, Yogeeswaran Muthukumaran, Vincent Fernandes +10

For noisy real-world environments such as those in open public spaces, spoken dialogue systems for both autonomous robots and avatars should be carefully designed to provide enhanc…

cs.CL2026

CultureConverse: A Multilingual Multi-turn Simulation Harness for Culturally Grounded Assistance in East and Southeast Asia

Bryan Chen Zhengyu Tan, Weihua Zheng, Thong T. Doan +30

Current cultural evaluations for large language models (LLMs) often reduce culture to single-turn factual recall via MCQs, failing to capture a common use case: users seeking pract…

cs.CL2026

MemUse: Moving Memory Evaluation from Direct QA to Natural Integration in Long-Term Human-AI Conversation

Ryuichi Sumida, Koji Inoue, Tatsuya Kawahara

Memory systems for conversational LLMs are conventionally evaluated by direct, fact-seeking questions about prior dialogue (Direct QA): can the model recall fact X from a prior con…

cs.HC2026

Does Listening Matter? Backchanneling and Nodding in AI Clone

Koji Inoue, Kazushi Kato, Tatsuya Kawahara +1

AI clones that imitate a specific person typically reproduce what the person says and how they sound, but not how they listen. We investigate whether adding multimodal listening be…