From the 2 of 6 linked papers with an AI index.
4 papers · 1 filter
VideoSEMA: a scalable and efficient Mamba-like attention for video understanding
Nhat Thanh Tran, Fanghui Xue andShuai Zhang, Fanghui Xue +5
The paper introduces VideoSEMA, a split space‑time attention model for video classification that combines a scalable Mamba‑like spatial attention block with softmax temporal attent…
SEMA: a Scalable and Efficient Mamba like Attention via Token Localization and Averaging
Nhat Thanh Tran, Fanghui Xue, Shuai Zhang +4
The paper introduces SEMA, a new attention mechanism for vision transformers that combines token localization with arithmetic averaging to avoid the dispersion problem of linear at…
USEMA: a Scalable Efficient Mamba Like Attention for Medical Image Segmentation
Elisha Dayag, Nhat Thanh Tran, Jack Xin
Accurate medical image segmentation is an integral part of the medical image analysis pipeline that requires the ability to merge local and global information. While vision transfo…
Deep Image Prior with L0 Gradient Regularizer for Image Smoothing
Nhat Thanh Tran, Kevin Bui, Jack Xin
Image smoothing is a fundamental image processing operation that preserves the underlying structure, such as strong edges and contours, and removes minor details and textures in an…