3 papers
math.OC2025
Constrained Average-Reward Intermittently Observable MDPs
Konstantin Avrachenkov, Madhu Dhiman, Veeraruna Kavitha
In Markov Decision Processes (MDPs) with intermittent state information, decision-making becomes challenging due to periods of missing observations. Linear programming (LP) methods…
math.OC2025
Optimal Control with cost: incorporating peak minimization
Madhu Dhiman, Veeraruna Kavitha, Nandyala Hemachandra
Inventory and queueing systems are often designed by controlling weighted combination of some time-averaged performance metrics (like cumulative holding, shortage, server-utilizati…
math.OC2025
Punitive policies to combat misreporting in dynamic supply chains
Madhu Dhiman, Atul Maurya, Veeraruna Kavitha +1
Wholesale price contracts are known to be associated with double marginalization effects, which prevents supply chains from achieving their true market share. In a dynamic setting…