◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

M. Ziaei

1 paper hereh-index 17941 citations63 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author1

Across the 1 of 1 paper where every author was matched, so the position is known.

fields
  • cs.AI1

identity via Semantic Scholar / OpenAlex

most citedSelf Punishment and Reward Backfill for Deep Q-Learning

8 citations · 8 across the 1 of their papers we have counts for

collaborators

1 paper

cs.AI2020★ 8 cited

Self Punishment and Reward Backfill for Deep Q-Learning

Mohammad Reza Bonyadi, Rui Wang, Maryam Ziaei

Reinforcement learning agents learn by encouraging behaviours which maximize their total reward, usually provided by the environment. In many environments, however, the reward is p…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.