activity
20192023
most citedMacaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration

27 citations · 29 across the 3 of their papers we have counts for

collaborators

6 papers

cs.CL202327 cited

Macaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration

Chenyang Lyu, Minghao Wu, Longyue Wang +5

Although instruction-tuned large language models (LLMs) have exhibited remarkable capabilities across various NLP tasks, their effectiveness on other data modalities beyond text ha…

cs.CL20221 cited

Robust Task-Oriented Dialogue Generation with Contrastive Pre-training and Adversarial Filtering

Shiquan Yang, Xinting Huang, Jey Han Lau +1

Data artifacts incentivize machine learning models to learn non-transferable generalizations by taking advantage of shortcuts in the data, and there is growing evidence that data a…

cs.CL2020

Generalizable and Explainable Dialogue Generation via Explicit Action Learning

Xinting Huang, Jianzhong Qi, Yu Sun +1

Response generation for task-oriented dialogues implicitly optimizes two objectives at the same time: task completion and language quality. Conditioned response generation serves a…

cs.CL20201 cited

Semi-Supervised Dialogue Policy Learning via Stochastic Reward Estimation

Xinting Huang, Jianzhong Qi, Yu Sun +1

Dialogue policy optimization often obtains feedback until task completion in task-oriented dialogue systems. This is insufficient for training intermediate dialogue turns since sup…

cs.CL2019

MALA: Cross-Domain Dialogue Generation with Action Learning

Xinting Huang, Jianzhong Qi, Yu Sun +1

Response generation for task-oriented dialogues involves two basic components: dialogue planning and surface realization. These two components, however, have a discrepancy in their…

cs.IR2019

CARL: Aggregated Search with Context-Aware Module Embedding Learning

Xinting Huang, Jianzhong Qi, Yu Sun +2

Aggregated search aims to construct search result pages (SERPs) from blue-links and heterogeneous modules (such as news, images, and videos). Existing studies have largely ignored…