2 papers
cs.CV2025
Progressive Multimodal Search and Reasoning for Knowledge-Intensive Visual Question Answering
Changin Choi, Wonseok Lee, Jungmin Ko +1
Knowledge-intensive visual question answering (VQA) requires external knowledge beyond image content, demanding precise visual grounding and coherent integration of visual and text…
cs.LG2024
Task-Specific Preconditioner for Cross-Domain Few-Shot Learning
Suhyun Kang, Jungwon Park, Wonseok Lee +1
Cross-Domain Few-Shot Learning~(CDFSL) methods typically parameterize models with task-agnostic and task-specific parameters. To adapt task-specific parameters, recent approaches h…