◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Simon Stepputtis

10 papers hereh-index 161.1k citations45 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author9
  • last author1

Across the 10 of 10 papers where every author was matched, so the position is known.

fields
  • cs.AI3
  • cs.CV3
  • cs.LG3
  • cs.RO1
same name
  • Simon Stepputtis — 15 papers, h 7

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.LGShow all

3 papers · 1 filter

cs.LG2026

Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control

Rohit Kumar Salla, Manoj Saravanan, Simon Stepputtis

Diffusion policies are a powerful policy class for continuous control, but their iterative denoising process creates a substantial computational bottleneck. Reducing this cost requ…

cs.LG2026

Overcoming Valid Action Suppression in Unmasked Policy Gradient Algorithms

Renos Zabounidis, Roy Siegelmann, Mohamad Qadri +3

In reinforcement learning environments with state-dependent action validity, action masking consistently outperforms penalty-based handling of invalid actions, yet existing theory…

cs.LG2026

SCALAR: Learning and Composing Skills through LLM Guided Symbolic Planning and Deep RL Grounding

Renos Zabounidis, Yue Wu, Simon Stepputtis +4

LM-based agents excel when given high-level action APIs but struggle to ground language into low-level control. Prior work has LLMs generate skills or reward functions for RL, but…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.