works on

From the 1 of 10 linked papers with an AI index.

collaborators

10 papers

stat.ML2026

Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents

Amirmohammad Farzaneh, Osvaldo Simeone

The paper introduces Think Short, Defer Smart (TSDS), a framework for edge-deployed LLM agents that stops on-device reasoning when actions stabilize and defers uncertain actions to…

stat.ML2026

Statistically Valid Hyperparameter Selection: From Tuning to Guarantees

Amirmohammad Farzaneh, Osvaldo Simeone

Hyperparameter selection is a critical step in the deployment of modern artificial intelligence systems, given the need to tune degrees of freedom such as inference-time parameters…

stat.ML2026

Post-Selection Distributional Model Evaluation

Amirmohammad Farzaneh, Osvaldo Simeone

Formal model evaluation methods typically certify that a model satisfies a prescribed target key performance indicator (KPI) level. However, in many applications, the relevant targ…

cs.LG2026

Synthetic Counterfactual Labels for Efficient Conformal Counterfactual Inference

Amirmohammad Farzaneh, Matteo Zecchin, Osvaldo Simeone

This work addresses the problem of constructing reliable prediction intervals for individual counterfactual outcomes. Existing conformal counterfactual inference (CCI) methods prov…

cs.LG2026

Optimized Certainty Equivalent Risk-Controlling Prediction Sets

Jiayi Huang, Amirmohammad Farzaneh, Osvaldo Simeone

In safety-critical applications such as medical image segmentation, prediction systems must provide reliability guarantees that extend beyond conventional expected loss control. Wh…

cs.AI2026

Should I Have Expressed a Different Intent? Counterfactual Generation for LLM-Based Autonomous Control

Amirmohammad Farzaneh, Salvatore D'Oro, Osvaldo Simeone

Large language model (LLM)-powered agents can translate high-level user intents into plans and actions in an environment. Yet after observing an outcome, users may wonder: What if…