collaborators

6 papers

astro-ph.GA2026

Magnetic fields in extreme primordial halos: turbulent collapse and implications for early quasar formation

V. B. Díaz, D. R. G. Schleicher, M. A. Latif +1

It is sometimes suggested that the most massive quasars at high redshift may have formed from rare high-sigma peaks in the cosmic density field. We explore here the evolution in a…

cs.LG2026

COOPA: A Modular LLM Agent Architecture for Operations Research Problems

Chuanhao Li, Xiaoan Xu, Dirk Bergemann +3

Operations Research (OR) provides a rigorous framework for high-stakes decision-making, but effective OR modeling requires substantial domain knowledge, mathematical abstraction, a…

cs.LG2026

Decoding Rewards in Competitive Games: Inverse Game Theory with Entropy Regularization

Junyi Liao, Zihan Zhu, Ethan Fang +2

Estimating the unknown reward functions driving agents' behaviors is of central interest in inverse reinforcement learning and game theory. To tackle this problem, we develop a uni…

cs.LG2026

Learning in Context, Guided by Choice: A Reward-Free Paradigm for Reinforcement Learning with Transformers

Juncheng Dong, Bowen He, Moyang Guo +3

In-context reinforcement learning (ICRL) leverages the in-context learning capabilities of transformer models (TMs) to efficiently generalize to unseen sequential decision-making t…

cs.LG2026

In-Context Reinforcement Learning From Suboptimal Historical Data

Juncheng Dong, Moyang Guo, Ethan X. Fang +2

Transformer models have achieved remarkable empirical successes, largely due to their in-context learning capabilities. Inspired by this, we explore training an autoregressive tran…

cs.LG2025

PASTA: A Unified Framework for Offline Assortment Learning

Juncheng Dong, Weibin Mo, Zhengling Qi +3

We study a broad class of assortment optimization problems in an offline and data-driven setting. In such problems, a firm lacks prior knowledge of the underlying choice model, and…