51 citations · 99 across the 12 of their papers we have counts for
15 papers
PF3Det: A Prompted Foundation Feature Assisted Visual LiDAR 3D Detector
Kaidong Li, Tianxiao Zhang, Kuan-Chuan Peng +1
3D object detection is crucial for autonomous driving, leveraging both LiDAR point clouds for precise depth information and camera images for rich semantic information. Therefore,…
Robust 3D Point Clouds Classification based on Declarative Defenders
Kaidong Li, Tianxiao Zhang, Cuncong Zhong +2
3D point cloud classification requires distinct models from 2D image classification due to the divergent characteristics of the respective input data. While 3D point clouds are uns…
Beyond Isolated Heads: Multi-Overlapped-Head Self-Attention for Vision Transformers
Tianxiao Zhang, Bo Luo, Guanghui Wang
Multi-Head Self-Attention (MHSA) is the cornerstone of Vision Transformers, allowing models to capture diverse feature representations by projecting tokens into independent subspac…
Depth-Wise Convolutions in Vision Transformers for Efficient Training on Small Datasets
Tianxiao Zhang, Wenju Xu, Bo Luo +1
The Vision Transformer (ViT) leverages the Transformer's encoder to capture global information by dividing images into patches and achieves superior performance across various comp…
A New Dataset and Comparative Study for Aphid Cluster Detection and Segmentation in Sorghum Fields
Raiyan Rahman, Christopher Indris, Goetz Bramesfeld +8
Aphid infestations are one of the primary causes of extensive damage to wheat and sorghum fields and are one of the most common vectors for plant viruses, resulting in significant…
Aphid Cluster Recognition and Detection in the Wild Using Deep Learning Models
Tianxiao Zhang, Kaidong Li, Xiangyu Chen +7
Aphid infestation poses a significant threat to crop production, rural communities, and global food security. While chemical pest control is crucial for maximizing yields, applying…