activity
20242026
collaborators
Showing cs.CVShow all

10 papers · 1 filter

cs.CV2026

GATE: Reliability-Gated Gaussian Evidence Fusion for Training-Free Test-Time Adaptation of Vision-Language Models

Pedram MohajerAnsari, Amir Salarpour, Run Wang +1

Vision-language models such as CLIP and SigLIP provide strong zero-shot recognition, but their predictions can degrade when deployed on target data that differ from the pretraining…

cs.CV2026

Budget-Aware Adaptive Adversarial Patches for Black-Box Object Detection

Pedram MohajerAnsari, Amir Salarpour, David Fernandez +1

Adversarial patches pose a practical threat to modern object detectors. Prior work shows vulnerability, but three gaps limit actionable insight: (i) few \emph{score-based black-box…

cs.CV2026

Sat3R: Satellite DSM Reconstruction via RPC-Aware Depth Fine-tuning

Qiaoyi Yang, Chaoyi Zhou, Xi Liu +9

Accurate Digital Surface Model (DSM) reconstruction from satellite imagery is critical for applications such as disaster response, urban planning, and large-scale geographic mappin…

cs.CV2026

FF3R: Feedforward Feature 3D Reconstruction from Unconstrained views

Chaoyi Zhou, Run Wang, Feng Luo +4

Recent advances in vision foundation models have revolutionized geometry reconstruction and semantic understanding. Yet, most of the existing approaches treat these capabilities in…

cs.CV2026

Comparative Analysis of Patch Attack on VLM-Based Autonomous Driving Architectures

David Fernandez, Pedram MohajerAnsari, Amir Salarpour +3

Vision-language models are emerging for autonomous driving, yet their robustness to physical adversarial attacks remains unexplored. This paper presents a systematic framework for…

cs.CV2026

SLNet: A Super-Lightweight Geometry-Adaptive Network for 3D Point Cloud Recognition

Mohammad Saeid, Amir Salarpour, Pedram MohajerAnsari +1

We present SLNet, a lightweight backbone for 3D point cloud recognition designed to achieve strong performance without the computational cost of many recent attention, graph, and d…