3 papers
cs.CV2025
Do You Keep an Eye on What I Ask? Mitigating Multimodal Hallucination via Attention-Guided Ensemble Decoding
Yeongjae Cho, Keonwoo Kim, Taebaek Hwang +1
Recent advancements in Large Vision-Language Models (LVLMs) have significantly expanded their utility in tasks like image captioning and visual question answering. However, they st…
cs.CV2025
CL3DOR: Contrastive Learning for 3D Large Multimodal Models via Odds Ratio on High-Resolution Point Clouds
Keonwoo Kim, Yeongjae Cho, Taebaek Hwang +2
Recent research has demonstrated that Large Language Models (LLMs) are not limited to text-only tasks but can also function as multimodal models across various modalities, includin…
cs.CL2024
DEBATE: Devil's Advocate-Based Assessment and Text Evaluation
Alex Kim, Keonwoo Kim, Sangwon Yoon
As natural language generation (NLG) models have become prevalent, systematically assessing the quality of machine-generated texts has become increasingly important. Recent studies…