7 papers
SeTformer is What You Need for Vision and Language
Pourya Shamsolmoali, Masoumeh Zareapoor, Eric Granger +1
The dot product self-attention (DPSA) is a fundamental component of transformers. However, scaling them to long sequences, like documents or high-resolution images, becomes prohibi…
Efficient Object Detection in Optical Remote Sensing Imagery via Attention-based Feature Distillation
Pourya Shamsolmoali, Jocelyn Chanussot, Huiyu Zhou +1
Efficient object detection methods have recently received great attention in remote sensing. Although deep convolutional networks often have excellent detection accuracy, their dep…
Distance Weighted Trans Network for Image Completion
Pourya Shamsolmoali, Masoumeh Zareapoor, Huiyu Zhou +2
The challenge of image generation has been effectively modeled as a problem of structure priors or transformation. However, existing models have unsatisfactory performance in under…
ClusVPR: Efficient Visual Place Recognition with Clustering-based Weighted Transformer
Yifan Xu, Pourya Shamsolmoali, Jie Yang
Visual place recognition (VPR) is a highly challenging task that has a wide range of applications, including robot navigation and self-driving vehicles. VPR is particularly difficu…
Entropy Transformer Networks: A Learning Approach via Tangent Bundle Data Manifold
Pourya Shamsolmoali, Masoumeh Zareapoor
This paper focuses on an accurate and fast interpolation approach for image transformation employed in the design of CNN architectures. Standard Spatial Transformer Networks (STNs)…
Image Completion via Dual-path Cooperative Filtering
Pourya Shamsolmoali, Masoumeh Zareapoor, Eric Granger
Given the recent advances with image-generating algorithms, deep image completion methods have made significant progress. However, state-of-art methods typically provide poor cross…