4 citations · 13 across the 11 of their papers we have counts for
7 papers · 1 filter
A Training-Free Framework for Video License Plate Tracking and Recognition with Only One-Shot
Haoxuan Ding, Qi Wang, Junyu Gao +1
Traditional license plate detection and recognition models are often trained on closed datasets, limiting their ability to handle the diverse license plate formats across different…
A Bi-Pyramid Multimodal Fusion Method for the Diagnosis of Bipolar Disorders
Guoxin Wang, Sheng Shi, Shan An +5
Previous research on the diagnosis of Bipolar disorder has mainly focused on resting-state functional magnetic resonance imaging. However, their accuracy can not meet the requireme…
SamLP: A Customized Segment Anything Model for License Plate Detection
Haoxuan Ding, Junyu Gao, Yuan Yuan +1
With the emergence of foundation model, this novel paradigm of deep learning has encouraged many powerful achievements in natural language processing and computer vision. There are…
RSSOD-Bench: A large-scale benchmark dataset for Salient Object Detection in Optical Remote Sensing Imagery
Zhitong Xiong, Yanfeng Liu, Qi Wang +1
We present the RSSOD-Bench dataset for salient object detection (SOD) in optical remote sensing imagery. While SOD has achieved success in natural scene images with deep learning,…
Improving Video Retrieval by Adaptive Margin
Feng He, Qi Wang, Zhifan Feng +4
Video retrieval is becoming increasingly important owing to the rapid emergence of videos on the Internet. The dominant paradigm for video retrieval learns video-text representatio…
MAFNet: A Multi-Attention Fusion Network for RGB-T Crowd Counting
Pengyu Chen, Junyu Gao, Yuan Yuan +1
RGB-Thermal (RGB-T) crowd counting is a challenging task, which uses thermal images as complementary information to RGB images to deal with the decreased performance of unimodal RG…