9 papers
Foveated Probes Recover Localized Binding Information in Vision Foundation Models
Mateusz Michalkiewicz, Mahsa Baktashmotlagh, Guha Balakrishnan
Frozen vision foundation models are commonly evaluated through a single global image embedding, but this interface can conflate missing information with information lost at readout…
Beyond Instance-Level Self-Supervision in 3D Multi-Modal Medical Imaging
Tan Pan, Shuhao Mei, Yixuan Sun +8
Self-supervised pre-training methods in medical imaging typically treat each individual as an isolated instance, learning representations through augmentation-based objectives or m…
T5-CSBoost: Adversarial Perturbation Resistant LLM Fingerprinting
Gayan K. Kulatilleke, Mahsa Baktashmotlagh, Siamak Layeghy +1
While many AI-generated text (AIGT) detectors achieve strong performance on clean inputs, their accuracy degrades significantly under light paraphrasing, word substitutions, charac…
WisWheat: A Three-Tiered Vision-Language Dataset for Wheat Management
Bowen Yuan, Selena Song, Javier Fernandez +3
Wheat management strategies play a critical role in determining yield. Traditional management decisions often rely on labour-intensive expert inspections, which are expensive, subj…
MOS: Model Synergy for Test-Time Adaptation on LiDAR-Based 3D Object Detection
Zhuoxiao Chen, Junjie Meng, Mahsa Baktashmotlagh +3
LiDAR-based 3D object detection is crucial for various applications but often experiences performance degradation in real-world deployments due to domain shifts. While most studies…
DiPEx: Dispersing Prompt Expansion for Class-Agnostic Object Detection
Jia Syuen Lim, Zhuoxiao Chen, Mahsa Baktashmotlagh +4
Class-agnostic object detection (OD) can be a cornerstone or a bottleneck for many downstream vision tasks. Despite considerable advancements in bottom-up and multi-object discover…