2 papers
cs.LG2018
Reward learning from human preferences and demonstrations in Atari
Borja Ibarz, Jan Leike, Tobias Pohlen +3
To solve complex real-world problems with reinforcement learning, we cannot rely on manually specified reward functions. Instead, we can have humans communicate an objective to the…
q-bio.NC2018
Growing Critical: Self-Organized Criticality in a Developing Neural System
Felipe Yaroslav Kalle Kossio, Sven Goedeke, Benjamin van den Akker +2
Experiments in various neural systems found avalanches: bursts of activity with characteristics typical for critical dynamics. A possible explanation for their occurrence is an und…