3 papers
cs.CV2026
VinQA: Visual Elements Interleaved Long-form Answer Generation for Real-World Multimodal Document QA
Young Rok Jang, Hyesoo Kong, Kyunghwan An +3
Real-world documents combine text with tables, charts, photographs, and diagrams arranged in diverse layouts, yet existing research on multimodal large language models (MLLMs) for…
cs.IR2026
ARHN: Answer-Centric Relabeling of Hard Negatives with Open-Source LLMs for Dense Retrieval
Hyewon Choi, Jooyoung Choi, Hansol Jang +4
Neural retrievers are often trained on large-scale triplet data comprising a query, a positive passage, and a set of hard negatives. In practice, hard-negative mining can introduce…
cs.CL2025
LGAI-EMBEDDING-Preview Technical Report
Jooyoung Choi, Hyun Kim, Hansol Jang +6
This report presents a unified instruction-based framework for learning generalized text embeddings optimized for both information retrieval (IR) and non-IR tasks. Built upon a dec…