◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Randy Jia

3 papers hereh-index 244 citations3 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • stat.ML1
same name
  • Randy Jia — 2 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedPosterior sampling for reinforcement learning: worst-case regret bounds

2 citations · 2 across the 3 of their papers we have counts for

collaborators

3 papers

stat.ML2023

Contextual Bandits for Evaluating and Improving Inventory Control Policies

Dean Foster, Randy Jia, Dhruv Madeka

Solutions to address the periodic review inventory control problem with nonstationary random demand, lost sales, and stochastic vendor lead times typically involve making strong as…

cs.LG2023

Learning an Inventory Control Policy with General Inventory Arrival Dynamics

Sohrab Andaz, Carson Eisenach, Dhruv Madeka +4

In this paper we address the problem of learning and backtesting inventory control policies in the presence of general arrival dynamics -- which we term as a quantity-over-time arr…

cs.LG2017★ 2 cited

Posterior sampling for reinforcement learning: worst-case regret bounds

Shipra Agrawal, Randy Jia

We present an algorithm based on posterior sampling (aka Thompson sampling) that achieves near-optimal worst-case regret bounds when the underlying Markov Decision Process (MDP) is…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.