output
20142026
most citedBootstrap your own latent: A new approach to self-supervised Learning

3.4k citations

Showing 2022Show all

8 papers · 1 filter

cs.LG20221 cited

Understanding Self-Predictive Learning for Reinforcement Learning

Yunhao Tang, Zhaohan Daniel Guo, Pierre Harvey Richemond +13

We study the learning dynamics of self-predictive learning for reinforcement learning, a family of algorithms that learn representations by minimizing the prediction error of their…

stat.ML2022

Optimistic Posterior Sampling for Reinforcement Learning with Few Samples and Tight Guarantees

Daniil Tiapkin, Denis Belomestny, Daniele Calandriello +6

We consider reinforcement learning in an environment modeled by an episodic, finite, stage-dependent Markov decision process of horizon with states, and actions. The pe…

cs.CY2022317 cited

Power to the People? Opportunities and Challenges for Participatory AI

Abeba Birhane, William Isaac, Vinodkumar Prabhakaran +4

Participatory approaches to artificial intelligence (AI) and machine learning (ML) are gaining momentum: the increased attention comes partly with the view that participation opens…

cs.CV20228 cited

AlignSDF: Pose-Aligned Signed Distance Fields for Hand-Object Reconstruction

Zerui Chen, Yana Hasson, Cordelia Schmid +1

Recent work achieved impressive progress towards joint reconstruction of hands and manipulated objects from monocular color images. Existing methods focus on two alternative repres…

cs.LG202217 cited

Subverting machines, fluctuating identities: Re-learning human categorization

Christina Lu, Jackie Kay, Kevin R. McKee

Most machine learning systems that interact with humans construct some notion of a person's "identity," yet the default paradigm in AI research envisions identity with essential at…

cs.LG2022

Marginalized Operators for Off-policy Reinforcement Learning

Yunhao Tang, Mark Rowland, Rémi Munos +1

In this work, we propose marginalized operators, a new class of off-policy evaluation operators for reinforcement learning. Marginalized operators strictly generalize generic multi…