2 papers
cs.CV2026
First Logit Boosting: Visual Grounding Method to Mitigate Object Hallucination in Large Vision-Language Models
Jiwoo Ha, Jongwoo Baek, Jinhyun So
Recent Large Vision-Language Models (LVLMs) have demonstrated remarkable performance across various multimodal tasks that require understanding both visual and linguistic inputs. H…
cs.LG2026
AROMMA: Unifying Olfactory Embeddings for Single Molecules and Mixtures
Dayoung Kang, JongWon Kim, Jiho Park +3
Public olfaction datasets are small and fragmented across single molecules and mixtures, limiting learning of generalizable odor representations. Recent works either learn single-m…