◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

D. Šiška

8 papers hereh-index 15604 citations52 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author7

Across the 8 of 8 papers where every author was matched, so the position is known.

fields
  • math.OC5
  • cs.LG3

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.LGShow all

3 papers · 1 filter

cs.LG2026

Global linear convergence of entropy-regularized softmax policy gradient beyond tabular MDPs

Ziyue Chen, David Šiška, Lukasz Szpruch

We study the global convergence of policy gradient for infinite-horizon entropy-regularized Markov decision processes (MDPs) with continuous state and action spaces. We consider lo…

cs.LG2026

A note on convergence of Wasserstein policy optimization

David Šiška, Yufei Zhang

Wasserstein Policy Optimization (WPO) is a recently proposed reinforcement learning algorithm that leverages Wasserstein gradient flows to optimize stochastic policies in continuou…

cs.LG2026

PPO in the Fisher-Rao geometry

Razvan-Andrei Lascu, David Šiška, Łukasz Szpruch

Proximal Policy Optimization (PPO) is widely used in reinforcement learning due to its strong empirical performance, yet it lacks formal guarantees for policy improvement and conve…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.