◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

E. Uchibe

7 papers hereh-index 254.9k citations126 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author6
  • last author1

Across the 7 of 7 papers where every author was matched, so the position is known.

fields
  • cs.LG5
  • cs.RO1
  • stat.ML1

identity via Semantic Scholar / OpenAlex

activity
20172024
most citedUnifying Value Iteration, Advantage Learning, and Dynamic Policy Programming

2 citations · 5 across the 6 of their papers we have counts for

collaborators
Showing 2022Show all

2 papers · 1 filter

cs.LG2022

Enforcing KL Regularization in General Tsallis Entropy Reinforcement Learning via Advantage Learning

Lingwei Zhu, Zheng Chen, Eiji Uchibe +1

Maximum Tsallis entropy (MTE) framework in reinforcement learning has gained popularity recently by virtue of its flexible modeling choices including the widely used Shannon entrop…

cs.LG2022

q-Munchausen Reinforcement Learning

Lingwei Zhu, Zheng Chen, Eiji Uchibe +1

The recently successful Munchausen Reinforcement Learning (M-RL) features implicit Kullback-Leibler (KL) regularization by augmenting the reward function with logarithm of the curr…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.