3 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 3 cited
SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model
Yang Zhan, Zhitong Xiong, Yuan Yuan
Large language models (LLMs) have recently been extended to the vision-language realm, obtaining impressive general multi-modal capabilities. However, the exploration of multi-moda…
cs.CV2023★ 1 cited
Parameter-Efficient Transfer Learning for Remote Sensing Image-Text Retrieval
Yuan Yuan, Yang Zhan, Zhitong Xiong
Vision-and-language pre-training (VLP) models have experienced a surge in popularity recently. By fine-tuning them on specific datasets, significant performance improvements have b…