11 citations · 19 across the 3 of their papers we have counts for
5 papers
Sample Complexity and Overparameterization Bounds for Temporal Difference Learning with Neural Network Approximation
Semih Cayci, Siddhartha Satpathi, Niao He +1
In this paper, we study the dynamics of temporal difference learning with neural network-based value function approximation over a general state space, namely, \emph{Neural TD lear…
Group-Fair Online Allocation in Continuous Time
Semih Cayci, Swati Gupta, Atilla Eryilmaz
The theory of discrete-time online learning has been successfully applied in many problems that involve sequential decision-making under uncertainty. However, in many applications…
Continuous-Time Multi-Armed Bandits with Controlled Restarts
Semih Cayci, Atilla Eryilmaz, R. Srikant
Time-constrained decision processes have been ubiquitous in many fundamental applications in physics, biology and computer science. Recently, restart strategies have gained signifi…
Budget-Constrained Bandits over General Cost and Reward Distributions
Semih Cayci, Atilla Eryilmaz, R. Srikant
We consider a budget-constrained bandit problem where each arm pull incurs a random cost, and yields a random reward in return. The objective is to maximize the total expected rewa…
Optimal Learning for Dynamic Coding in Deadline-Constrained Multi-Channel Networks
Semih Cayci, Atilla Eryilmaz
We study the problem of serving randomly arriving and delay-sensitive traffic over a multi-channel communication system with time-varying channel states and unknown statistics. Thi…