◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yonathan Efroni

24 papers hereh-index 211.8k citations59 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author10
  • middle author10
  • last author3

Across the 23 of 24 papers where every author was matched, so the position is known.

fields
  • cs.LG22
  • cs.AI2
same name
  • Yonathan Efroni — 4 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20182025
most citedAction Robust Reinforcement Learning and Applications in Continuous Control

66 citations · 104 across the 10 of their papers we have counts for

collaborators
Showing 2022Show all

4 papers · 1 filter

cs.LG2022

Tractable Optimality in Episodic Latent MABs

Jeongyeol Kwon, Yonathan Efroni, Constantine Caramanis +1

We consider a multi-armed bandit problem with M latent contexts, where an agent interacts with the environment for an episode of H time steps. Depending on the length of the ep…

cs.LG2022

Reward-Mixing MDPs with a Few Latent Contexts are Learnable

Jeongyeol Kwon, Yonathan Efroni, Constantine Caramanis +1

We consider episodic reinforcement learning in reward-mixing Markov decision processes (RMMDPs): at the beginning of every episode nature randomly picks a latent reward model among…

cs.LG2022★ 1 cited

Provable Reinforcement Learning with a Short-Term Memory

Yonathan Efroni, Chi Jin, Akshay Krishnamurthy +1

Real-world sequential decision making problems commonly involve partial observability, which requires the agent to maintain a memory of history in order to infer the latent states,…

cs.LG2022

Coordinated Attacks against Contextual Bandits: Fundamental Limits and Defense Mechanisms

Jeongyeol Kwon, Yonathan Efroni, Constantine Caramanis +1

Motivated by online recommendation systems, we propose the problem of finding the optimal policy in multitask contextual bandits when a small fraction α<1/2 of tasks (users) are…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.