2 citations · 2 across the 4 of their papers we have counts for
4 papers
Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
Youngmin Kim, Jiwan Chung, Jisoo Kim +5
Nonverbal communication is integral to human interaction, with gestures, facial expressions, and body language conveying critical aspects of intent and emotion. However, existing l…
Visually Dehallucinative Instruction Generation: Know What You Don't Know
Sungguk Cha, Jusung Lee, Younghyun Lee +1
"When did the emperor Napoleon invented iPhone?" Such hallucination-inducing question is well known challenge in generative language modeling. In this study, we present an innovati…
Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks
Jusung Lee, Sungguk Cha, Younghyun Lee +1
Having revolutionized natural language processing (NLP) applications, large language models (LLMs) are expanding into the realm of multimodal inputs. Owing to their ability to inte…
Visually Dehallucinative Instruction Generation
Sungguk Cha, Jusung Lee, Younghyun Lee +1
In recent years, synthetic visual instructions by generative language model have demonstrated plausible text generation performance on the visual question-answering tasks. However,…