◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Kaige Yang

5 papers hereh-index 487 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author1
  • first author4

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • cs.IR1

identity via Semantic Scholar / OpenAlex

activity
20182021
most citedDifferentiable Linear Bandit Algorithm

5 citations · 5 across the 3 of their papers we have counts for

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2021

Learn Dynamic-Aware State Embedding for Transfer Learning

Kaige Yang

Transfer reinforcement learning aims to improve the sample efficiency of solving unseen new tasks by leveraging experiences obtained from previous tasks. We consider the setting wh…

cs.LG2020★ 5 cited

Differentiable Linear Bandit Algorithm

Kaige Yang, Laura Toni

Upper Confidence Bound (UCB) is arguably the most commonly used method for linear multi-arm bandit problems. While conceptually and computationally simple, this method highly relie…

cs.LG2019

Laplacian-regularized graph bandits: Algorithms and theoretical analysis

Kaige Yang, Xiaowen Dong, Laura Toni

We consider a stochastic linear bandit problem with multiple users, where the relationship between users is captured by an underlying graph and user preferences are represented as…

cs.LG2019

Error Analysis on Graph Laplacian Regularized Estimator

Kaige Yang, Xiaowen Dong, Laura Toni

We provide a theoretical analysis of the representation learning problem aimed at learning the latent variables (design matrix) Θ of observations Y with the knowledge of the co…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.