3 papers
cs.LG2026
The Cost of Reasoning: Chain-of-Thought Induces Overconfidence in Vision-Language Models
Robert Welch, Emir Konuk, Kevin Smith
Vision-language models (VLMs) are increasingly deployed in high-stakes settings where reliable uncertainty quantification (UQ) is as important as predictive accuracy. Extended reas…
cs.CV2025
APLA: A Simple Adaptation Method for Vision Transformers
Moein Sorkhei, Emir Konuk, Kevin Smith +1
Existing adaptation techniques typically require architectural modifications or added parameters, leading to high computational costs and complexity. We introduce Attention Project…
cs.CV2025
VORTEX: Challenging CNNs at Texture Recognition by using Vision Transformers with Orderless and Randomized Token Encodings
Leonardo Scabini, Kallil M. Zielinski, Emir Konuk +4
Texture recognition has recently been dominated by ImageNet-pre-trained deep Convolutional Neural Networks (CNNs), with specialized modifications and feature engineering required t…