4 papers
Reinforcement Learning for Sequential Solar PV Policy Design under Uncertainty: An Agent-Based Approach
Iias Faiud, Jonaid Shianifar, Michael Schukat +1
Designing effective and fiscally sustainable policies for solar photovoltaic (PV) adoption requires balancing adoption gains against public expenditure under uncertainty and hetero…
Less Traffic, Better Outcomes: Competition-Aware Request Dispatch in Real-Time Ad Exchanges
Jonaid Shianifar, Blaz Mramor, Fangda Zou +5
Real-time bidding (RTB) ad exchanges typically forward nearly all incoming requests to demand-side platforms (DSPs), even though only a small fraction receive bids. This over-distr…
AI World Cup 2026: Benchmarking Large Language Models for End-to-End Football Tournament Prediction
Jonaid Shianifar, Iias Faiud
Large language models (LLMs) are now regularly asked to forecast real-world events, but comparisons are often difficult because models receive different information, use different…
Hindsight Preference Replay Improves Preference-Conditioned Multi-Objective Reinforcement Learning
Jonaid Shianifar, Michael Schukat, Karl Mason
Multi-objective reinforcement learning (MORL) enables agents to optimize vector-valued rewards while respecting user preferences. CAPQL, a preference-conditioned actor-critic metho…