2 papers
cs.CV2026
Draft and Refine with Visual Experts
Sungheon Jeong, Ryozo Masukawa, Jihong Park +5
While recent Large Vision-Language Models (LVLMs) exhibit strong multimodal reasoning abilities, they often produce ungrounded or hallucinated responses because they rely too heavi…
cs.CV2025
Uncertainty-Weighted Image-Event Multimodal Fusion for Video Anomaly Detection
Sungheon Jeong, Jihong Park, Mohsen Imani
Most existing video anomaly detectors rely solely on RGB frames, which lack the temporal resolution needed to capture abrupt or transient motion cues, key indicators of anomalous e…