◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Andrei N. Lupu

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author3

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • cs.AI1

identity via Semantic Scholar / OpenAlex

most citedGrounding Aleatoric Uncertainty for Unsupervised Environment Design

3 citations · 3 across the 4 of their papers we have counts for

collaborators

4 papers

cs.LG2024

Behaviour Distillation

Andrei Lupu, Chris Lu, Jarek Liesen +2

Dataset distillation aims to condense large datasets into a small number of synthetic examples that can be used as drop-in replacements when training new models. It has application…

cs.LG2024

Discovering Minimal Reinforcement Learning Environments

Jarek Liesen, Chris Lu, Andrei Lupu +3

Reinforcement learning (RL) agents are commonly trained and evaluated in the same environment. In contrast, humans often train in a specialized environment before being evaluated,…

cs.AI2022

Self-Explaining Deviations for Coordination

Hengyuan Hu, Samuel Sokota, David Wu +4

Fully cooperative, partially observable multi-agent problems are ubiquitous in the real world. In this paper, we focus on a specific subclass of coordination problems in which huma…

cs.LG2022★ 3 cited

Grounding Aleatoric Uncertainty for Unsupervised Environment Design

Minqi Jiang, Michael Dennis, Jack Parker-Holder +5

Adaptive curricula in reinforcement learning (RL) have proven effective for producing policies robust to discrepancies between the train and test environment. Recently, the Unsuper…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.