1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2025
MASH-VLM: Mitigating Action-Scene Hallucination in Video-LLMs through Disentangled Spatial-Temporal Representations
Kyungho Bae, Jinhyung Kim, Sihaeng Lee +3
In this work, we tackle action-scene hallucination in Video Large Language Models (Video-LLMs), where models incorrectly predict actions based on the scene context or scenes based…
cs.LG2024★ 1 cited
EXAONEPath 1.0 Patch-level Foundation Model for Pathology
Juseung Yun, Yi Hu, Jinhyung Kim +2
Recent advancements in digital pathology have led to the development of numerous foundational models that utilize self-supervised learning on patches extracted from gigapixel whole…