6 papers
Differentiating Through Dual Prices: End-to-End Policy Learning Under Capacity Constraints
Mohammadsaeed Haghi, Mahdi Salmani, Nima Kelidari
Many social services assign scarce resources, such as housing assistance or hospital interventions, to people who arrive one at a time: each arrival must receive a decision immedia…
A Gold-Standard Study of What Makes a Lightweight Game-Playing Agent Strong
Nima Kelidari, Mohammadsaeed Haghi, Mahdi Salmani
Reinforcement learning agents for imperfect-information card games are only as strong as the opponents they train against, and they are hard to grade, since they beat a random oppo…
Early Stopping for Large Reasoning Models via Confidence Dynamics
Parsa Hosseini, Sumit Nawathe, Mahdi Salmani +2
Large reasoning models rely on long chain-of-thought generation to solve complex problems, but extended reasoning often incurs substantial computational cost and can even degrade p…
From Filters to VLMs: Benchmarking Defogging Methods through Object Detection and Segmentation Performance
Ardalan Aryashad, Parsa Razmara, Amin Mahjoub +3
Autonomous driving perception systems are particularly vulnerable in foggy conditions, where light scattering reduces contrast and obscures fine details critical for safe operation…
Sampling and Loss Weights in Multi-Domain Training
Mahdi Salmani, Pratik Worah, Meisam Razaviyayn +1
In the training of large deep neural networks, there is a need for vast amounts of training data. To meet this need, data is collected from multiple domains, such as Wikipedia and…
Rewriting the Budget: A General Framework for Black-Box Attacks Under Cost Asymmetry
Mahdi Salmani, Alireza Abdollahpoorrostam, Seyed-Mohsen Moosavi-Dezfooli
Traditional decision-based black-box adversarial attacks on image classifiers aim to generate adversarial examples by slightly modifying input images while keeping the number of qu…