2 citations · 2 across the 5 of their papers we have counts for
5 papers
MADTP: Multimodal Alignment-Guided Dynamic Token Pruning for Accelerating Vision-Language Transformer
Jianjian Cao, Peng Ye, Shengze Li +4
Vision-Language Transformers (VLTs) have shown great success recently, but are meanwhile accompanied by heavy computation costs, where a major reason can be attributed to the large…
ClipSAM: CLIP and SAM Collaboration for Zero-Shot Anomaly Segmentation
Shengze Li, Jianjian Cao, Peng Ye +3
Recently, foundational models such as CLIP and SAM have shown promising performance for the task of Zero-Shot Anomaly Segmentation (ZSAS). However, either CLIP-based or SAM-based Z…
Collaborative Position Reasoning Network for Referring Image Segmentation
Jianjian Cao, Beiya Dai, Yulin Li +2
Given an image and a natural language expression as input, the goal of referring image segmentation is to segment the foreground masks of the entities referred by the expression. E…
A2S-NAS: Asymmetric Spectral-Spatial Neural Architecture Search For Hyperspectral Image Classification
Lin Zhan, Jiayuan Fan, Peng Ye +1
Existing deep learning-based hyperspectral image (HSI) classification works still suffer from the limitation of the fixed-sized receptive field, leading to difficulties in distinct…
JNDMix: JND-Based Data Augmentation for No-reference Image Quality Assessment
Jiamu Sheng, Jiayuan Fan, Peng Ye +1
Despite substantial progress in no-reference image quality assessment (NR-IQA), previous training models often suffer from over-fitting due to the limited scale of used datasets, r…