8 citations · 9 across the 4 of their papers we have counts for
4 papers
A Case for Validation Buffer in Pessimistic Actor-Critic
Michal Nauman, Mateusz Ostaszewski, Marek Cygan
In this paper, we investigate the issue of error accumulation in critic networks updated via pessimistic temporal difference objectives. We show that the critic approximation error…
Curriculum reinforcement learning for quantum architecture search under hardware errors
Yash J. Patel, Akash Kundu, Mateusz Ostaszewski +3
The key challenge in the noisy intermediate-scale quantum era is finding useful circuits compatible with current device limitations. Variational quantum algorithms (VQAs) offer a p…
On consequences of finetuning on data with highly discriminative features
Wojciech Masarczyk, Tomasz Trzciński, Mateusz Ostaszewski
In the era of transfer learning, training neural networks from scratch is becoming obsolete. Transfer learning leverages prior knowledge for new tasks, conserving computational res…
Reinforcement learning with experience replay and adaptation of action dispersion
Paweł Wawrzyński, Wojciech Masarczyk, Mateusz Ostaszewski
Effective reinforcement learning requires a proper balance of exploration and exploitation defined by the dispersion of action distribution. However, this balance depends on the ta…