3 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 2 cited
CODE: Contrasting Self-generated Description to Combat Hallucination in Large Multi-modal Models
Junho Kim, Hyunjun Kim, Yeonju Kim +1
Large Multi-modal Models (LMMs) have recently demonstrated remarkable abilities in visual context understanding and coherent response generation. However, alongside these advanceme…
cs.CV2022★ 3 cited
Speaker-adaptive Lip Reading with User-dependent Padding
Minsu Kim, Hyunjun Kim, Yong Man Ro
Lip reading aims to predict speech based on lip movements alone. As it focuses on visual information to model the speech, its performance is inherently sensitive to personal lip ap…