activity
20172025
most citedAudio Description from Image by Modal Translation Network

24 citations · 40 across the 4 of their papers we have counts for

collaborators

5 papers

eess.IV2025

A Tree-guided CNN for image super-resolution

Chunwei Tian, Mingjian Song, Xiaopeng Fan +3

Deep convolutional neural networks can extract more accurate structural information via deep architectures to obtain good performance in image super-resolution. However, it is not…

cs.CV2024

Efficient Prompt Tuning of Large Vision-Language Model for Fine-Grained Ship Classification

Long Lan, Fengxiang Wang, Xiangtao Zheng +2

Fine-grained ship classification in remote sensing (RS-FGSC) poses a significant challenge due to the high similarity between classes and the limited availability of labeled data,…

cs.CV2022★ 16 cited

Pairwise Comparison Network for Remote Sensing Scene Classification

Zhang Yue, Zheng Xiangtao, Lu Xiaoqiang

Remote sensing scene classification aims to assign a specific semantic label to a remote sensing image. Recently, convolutional neural networks have greatly improved the performanc…

cs.SD2021★ 24 cited

Audio Description from Image by Modal Translation Network

Hailong Ning, Xiangtao Zheng, Yuan Yuan +1

Audio is the main form for the visually impaired to obtain information. In reality, all kinds of visual data always exist, but audio data does not exist in many cases. In order to…

cs.CV2017

Exploring Models and Data for Remote Sensing Image Caption Generation

Xiaoqiang Lu, Binqiang Wang, Xiangtao Zheng +1

Inspired by recent development of artificial satellite, remote sensing images have attracted extensive attention. Recently, noticeable progress has been made in scene classificatio…