◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

E. Uchibe

2 papers hereh-index 254.9k citations126 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.LG1
  • stat.ML1

identity via Semantic Scholar / OpenAlex

most citedUnifying Value Iteration, Advantage Learning, and Dynamic Policy Programming

2 citations · 4 across the 2 of their papers we have counts for

collaborators

2 papers

stat.ML2017★ 2 cited

Unifying Value Iteration, Advantage Learning, and Dynamic Policy Programming

Tadashi Kozuno, Eiji Uchibe, Kenji Doya

Approximate dynamic programming algorithms, such as approximate value iteration, have been successfully applied to many complex reinforcement learning tasks, and a better approxima…

cs.LG2017★ 2 cited

Online Meta-learning by Parallel Algorithm Competition

Stefan Elfwing, Eiji Uchibe, Kenji Doya

The efficiency of reinforcement learning algorithms depends critically on a few meta-parameters that modulates the learning updates and the trade-off between exploration and exploi…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.