3 papers
cs.AI2026
CausalDS: Benchmarking Causal Reasoning in Data-Science Agents
Andrej Leban, Yuekai Sun
Large language models (LLMs) increasingly act as integrated data-science agents, combining abstract reasoning with advanced tool use. Yet the relevant benchmark landscape largely d…
stat.ML2025
Energy-Tweedie: Score meets Score, Energy meets Energy
Andrej Leban
Denoising and score estimation are classically linked through Tweedie's formula, which relates the posterior mean under Gaussian noise to the Stein score of the noisy marginal. In…
stat.ML2025
Distributional Autoencoders Know the Score
Andrej Leban
The Distributional Principal Autoencoder (DPA) combines distributionally correct reconstruction with principal-component-like interpretability of the encodings. In this work, we pr…