13 citations · 31 across the 19 of their papers we have counts for
Showing 2023 · cs.LGShow all
2 papers · 2 filters
cs.LG2023★ 1 cited
The Information Pathways Hypothesis: Transformers are Dynamic Self-Ensembles
Md Shamim Hussain, Mohammed J. Zaki, Dharmashankar Subramanian
Transformers use the dense self-attention mechanism which gives a lot of flexibility for long-range connectivity. Over multiple layers of a deep transformer, the number of possible…
cs.LG2023
Probabilistic Constraint for Safety-Critical Reinforcement Learning
Weiqin Chen, Dharmashankar Subramanian, Santiago Paternain
In this paper, we consider the problem of learning safe policies for probabilistic-constrained reinforcement learning (RL). Specifically, a safe policy or controller is one that, w…