From the 1 of 10 linked papers with an AI index.
10 papers
Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents
Amirmohammad Farzaneh, Osvaldo Simeone
The paper introduces Think Short, Defer Smart (TSDS), a framework for edge-deployed LLM agents that stops on-device reasoning when actions stabilize and defers uncertain actions to…
Statistically Valid Hyperparameter Selection: From Tuning to Guarantees
Amirmohammad Farzaneh, Osvaldo Simeone
Hyperparameter selection is a critical step in the deployment of modern artificial intelligence systems, given the need to tune degrees of freedom such as inference-time parameters…
Post-Selection Distributional Model Evaluation
Amirmohammad Farzaneh, Osvaldo Simeone
Formal model evaluation methods typically certify that a model satisfies a prescribed target key performance indicator (KPI) level. However, in many applications, the relevant targ…
Synthetic Counterfactual Labels for Efficient Conformal Counterfactual Inference
Amirmohammad Farzaneh, Matteo Zecchin, Osvaldo Simeone
This work addresses the problem of constructing reliable prediction intervals for individual counterfactual outcomes. Existing conformal counterfactual inference (CCI) methods prov…
Optimized Certainty Equivalent Risk-Controlling Prediction Sets
Jiayi Huang, Amirmohammad Farzaneh, Osvaldo Simeone
In safety-critical applications such as medical image segmentation, prediction systems must provide reliability guarantees that extend beyond conventional expected loss control. Wh…
Should I Have Expressed a Different Intent? Counterfactual Generation for LLM-Based Autonomous Control
Amirmohammad Farzaneh, Salvatore D'Oro, Osvaldo Simeone
Large language model (LLM)-powered agents can translate high-level user intents into plans and actions in an environment. Yet after observing an outcome, users may wonder: What if…