2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 2 cited
Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks
Jusung Lee, Sungguk Cha, Younghyun Lee +1
Having revolutionized natural language processing (NLP) applications, large language models (LLMs) are expanding into the realm of multimodal inputs. Owing to their ability to inte…
cs.CV2021
Zero-Shot Semantic Segmentation via Spatial and Multi-Scale Aware Visual Class Embedding
Sungguk Cha, Yooseung Wang
Fully supervised semantic segmentation technologies bring a paradigm shift in scene understanding. However, the burden of expensive labeling cost remains as a challenge. To solve t…