◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

D. Šiška

12 papers hereh-index 15604 citations52 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author9
  • last author1

Across the 11 of 12 papers where every author was matched, so the position is known.

fields
  • math.OC6
  • cs.LG3
  • math.PR2
  • q-fin.MF1

identity via Semantic Scholar / OpenAlex

activity
20072026
most citedOn randomized stopping

10 citations · 10 across the 6 of their papers we have counts for

collaborators
Showing 2026Show all

3 papers · 1 filter

math.OC2026

Entropy Regularization Improves Policy Robustness in Continuous-Time Reinforcement Learning

Jialun Cao, Fernando Acero, David Šiška +1

Entropy regularization is widely used in continuous-time reinforcement learning (RL) to reduce sensitivity to environmental perturbations, yet its robustness benefits lack a rigoro…

cs.LG2026

Global linear convergence of entropy-regularized softmax policy gradient beyond tabular MDPs

Ziyue Chen, David Šiška, Lukasz Szpruch

We study the global convergence of policy gradient for infinite-horizon entropy-regularized Markov decision processes (MDPs) with continuous state and action spaces. We consider lo…

cs.LG2026

A note on convergence of Wasserstein policy optimization

David Šiška, Yufei Zhang

Wasserstein Policy Optimization (WPO) is a recently proposed reinforcement learning algorithm that leverages Wasserstein gradient flows to optimize stochastic policies in continuou…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.