6 papers · 1 filter
iStructTab: Structured Feature Sequencing for Multimodal Learning of Image and Tabular Data
Al Zadid Sultan Bin Habib, Md Younus Ahamed, Prashnna Gyawali +2
Multimodal learning of images and tabular data is often impaired by ineffective representations, resulting in redundancy, dispersion, and generalization problems. To tackle this ch…
SemLT3D: Semantic-Guided Expert Distillation for Camera-only Long-Tailed 3D Object Detection
Hao Vo, Khoa Vo, Thinh Phan +5
Camera-only 3D object detection has emerged as a cost-effective and scalable alternative to LiDAR for autonomous driving, yet existing methods primarily prioritize overall performa…
Improving Accuracy and Generalization for Efficient Visual Tracking
Ram Zaveri, Shivang Patel, Yu Gu +1
Efficient visual trackers overfit to their training distributions and lack generalization abilities, resulting in them performing well on their respective in-distribution (ID) test…
Rethinking Self-Supervised Learning Within the Framework of Partial Information Decomposition
Salman Mohamadi, Gianfranco Doretto, Donald A. Adjeroh
Self Supervised learning (SSL) has demonstrated its effectiveness in feature learning from unlabeled data. Regarding this success, there have been some arguments on the role that m…
Direct Coloring for Self-Supervised Enhanced Feature Decoupling
Salman Mohamadi, Gianfranco Doretto, Donald A. Adjeroh
The success of self-supervised learning (SSL) has been the focus of multiple recent theoretical and empirical studies, including the role of data augmentation (in feature decouplin…
FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Generation
Trong Thang Pham, Ngoc-Vuong Ho, Nhat-Tan Bui +8
Developing an interpretable system for generating reports in chest X-ray (CXR) analysis is becoming increasingly crucial in Computer-aided Diagnosis (CAD) systems, enabling radiolo…