10 citations · 13 across the 5 of their papers we have counts for
3 papers · 1 filter
Self-Supervised Contrastive Pre-Training for Multivariate Point Processes
Xiao Shou, Dharmashankar Subramanian, Debarun Bhattacharjya +2
Self-supervision is one of the hallmarks of representation learning in the increasingly popular suite of foundation models including large language models such as BERT and GPT-3, b…
Adaptive Primal-Dual Method for Safe Reinforcement Learning
Weiqin Chen, James Onyejizu, Long Vu +5
Primal-dual methods have a natural application in Safe Reinforcement Learning (SRL), posed as a constrained policy optimization problem. In practice however, applying primal-dual m…
The Information Pathways Hypothesis: Transformers are Dynamic Self-Ensembles
Md Shamim Hussain, Mohammed J. Zaki, Dharmashankar Subramanian
Transformers use the dense self-attention mechanism which gives a lot of flexibility for long-range connectivity. Over multiple layers of a deep transformer, the number of possible…