◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Ercüment Ilhan

6 papers hereh-index 592 citations9 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author2

Across the 6 of 6 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • cs.IR1
  • cs.MA1

identity via Semantic Scholar / OpenAlex

activity
20192024
collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2022

Methodical Advice Collection and Reuse in Deep Reinforcement Learning

Sahir, Ercüment İlhan, Srijita Das +1

Reinforcement learning (RL) has shown great success in solving many challenging tasks via use of deep neural networks. Although using deep learning for RL brings immense representa…

cs.LG2021

Action Advising with Advice Imitation in Deep Reinforcement Learning

Ercument Ilhan, Jeremy Gow, Diego Perez-Liebana

Action advising is a peer-to-peer knowledge exchange technique built on the teacher-student paradigm to alleviate the sample inefficiency problem in deep reinforcement learning. Re…

cs.LG2021

Learning on a Budget via Teacher Imitation

Ercument Ilhan, Jeremy Gow, Diego Perez-Liebana

Deep Reinforcement Learning (RL) techniques can benefit greatly from leveraging prior experience, which can be either self-generated or acquired from other entities. Action advisin…

cs.LG2020

Student-Initiated Action Advising via Advice Novelty

Ercument Ilhan, Jeremy Gow, Diego Perez-Liebana

Action advising is a budget-constrained knowledge exchange mechanism between teacher-student peers that can help tackle exploration and sample inefficiency problems in deep reinfor…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.