5 papers
K-EXAONE 2.0 Technical Report
Eunbi Choi, Kibong Choi, Sehyun Chun +74
This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research as a step in our effort toward global frontier-scale foundatio…
Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views
Jiho Choi, Seonho Lee, Seojeong Park +1
Multi-view 3D Visual Question Answering (MV3D-VQA) requires integrating partial observations into a coherent 3D scene representation and selecting informative viewpoints for multi-…
PosterForest: Hierarchical Multi-Agent Collaboration for Scientific Poster Generation
Jiho Choi, Seojeong Park, Seongjong Song +1
Automating scientific poster generation requires hierarchical document understanding and coherent content-layout planning. Existing methods often rely on flat summarization or opti…
Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling
Seojeong Park, Jiho Choi, Junyong Kang +3
Recent multimodal large language models have demonstrated strong reasoning ability, yet their reliability as automated evaluators remains limited by a critical weakness: when visua…
MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval
Seojeong Park, Jiho Choi, Kyungjune Baek +1
Video Moment Retrieval (MR) aims to localize moments within a video based on a given natural language query. Given the prevalent use of platforms like YouTube for information retri…