collaborators

7 papers

cs.CV2024

SeTformer is What You Need for Vision and Language

Pourya Shamsolmoali, Masoumeh Zareapoor, Eric Granger +1

The dot product self-attention (DPSA) is a fundamental component of transformers. However, scaling them to long sequences, like documents or high-resolution images, becomes prohibi…

cs.CV2023

Efficient Object Detection in Optical Remote Sensing Imagery via Attention-based Feature Distillation

Pourya Shamsolmoali, Jocelyn Chanussot, Huiyu Zhou +1

Efficient object detection methods have recently received great attention in remote sensing. Although deep convolutional networks often have excellent detection accuracy, their dep…

cs.CV2023

Distance Weighted Trans Network for Image Completion

Pourya Shamsolmoali, Masoumeh Zareapoor, Huiyu Zhou +2

The challenge of image generation has been effectively modeled as a problem of structure priors or transformation. However, existing models have unsatisfactory performance in under…

cs.CV2023

ClusVPR: Efficient Visual Place Recognition with Clustering-based Weighted Transformer

Yifan Xu, Pourya Shamsolmoali, Jie Yang

Visual place recognition (VPR) is a highly challenging task that has a wide range of applications, including robot navigation and self-driving vehicles. VPR is particularly difficu…

cs.CV2023

Entropy Transformer Networks: A Learning Approach via Tangent Bundle Data Manifold

Pourya Shamsolmoali, Masoumeh Zareapoor

This paper focuses on an accurate and fast interpolation approach for image transformation employed in the design of CNN architectures. Standard Spatial Transformer Networks (STNs)…

cs.CV2023

Image Completion via Dual-path Cooperative Filtering

Pourya Shamsolmoali, Masoumeh Zareapoor, Eric Granger

Given the recent advances with image-generating algorithms, deep image completion methods have made significant progress. However, state-of-art methods typically provide poor cross…