2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.IR2022★ 2 cited
Mr. Right: Multimodal Retrieval on Representation of ImaGe witH Text
Cheng-An Hsieh, Cheng-Ping Hsieh, Pu-Jen Cheng
Multimodal learning is a recent challenge that extends unimodal learning by generalizing its domain to diverse modalities, such as texts, images, or speech. This extension requires…
cs.CL2022
An Evaluation Dataset for Legal Word Embedding: A Case Study On Chinese Codex
Chun-Hsien Lin, Pu-Jen Cheng
Word embedding is a modern distributed word representations approach widely used in many natural language processing tasks. Converting the vocabulary in a legal document into a wor…
cs.MM2021
End-to-End Video Question-Answer Generation with Generator-Pretester Network
Hung-Ting Su, Chen-Hsi Chang, Po-Wei Shen +5
We study a novel task, Video Question-Answer Generation (VQAG), for challenging Video Question Answering (Video QA) task in multimedia. Due to expensive data annotation costs, many…