From the 1 of 12 linked papers with an AI index.
4 papers · 1 filter
Optimal Stopping of Self-Refining Foundation Models
Kim Hammar, Tansu Alpcan, Emil C. Lupu
Foundation models can improve their outputs through a self-refinement process driven by external feedback. In this process, the model is embedded in an iterative loop where it gene…
Recovery Control in Replicated Systems through Autonomous Multiagent Rollout
Kim Hammar, Yuchao Li
The paper formulates the timing of replica recovery in redundant computing systems as a multi‑agent POMDP and proposes a multi‑agent rollout approach that uses precomputed signalin…
Adaptive Network Security Policies via Belief Aggregation and Rollout
Kim Hammar, Yuchao Li, Tansu Alpcan +2
Evolving security vulnerabilities and shifting operational conditions require frequent updates to network security policies. These updates include adjustments to incident response…
Feature-Based Belief Aggregation for Partially Observable Markov Decision Problems
Yuchao Li, Kim Hammar, Dimitri Bertsekas
We consider a finite-state partially observable Markov decision problem (POMDP) with an infinite horizon and a discounted cost, and we propose a new method for computing a cost fun…