13 papers
TimeRouter: Efficient and Adaptive Routing of Time-Series Foundation Models
Kanghui Ning, Yushan Jiang, Kashif Rasul +3
Time-series foundation models (TSFMs) are increasingly explored as predictive experts within emerging agentic time-series systems. However, TSFMs exhibit heterogeneous inductive bi…
Forecasting with Hyper-Trees
Alexander März, Kashif Rasul
We introduce Hyper-Trees as a novel framework for modeling time series data using gradient boosted trees. Unlike conventional tree-based approaches that forecast time series direct…
SHARP: A Self-Evolving Human-Auditable Rubric Policy for Financial Trading Agents
Xiwen Chen, Wenhui Zhu, Songzhu Zheng +3
Large language models (LLMs) are increasingly deployed for autonomous financial trading, a domain requiring continuous adaptation to noisy, non-stationary markets. Existing self-im…
AlphaLab: Autonomous Multi-Agent Research Across Optimization Domains with Frontier LLMs
Brendan R. Hogan, Xiwen Chen, James T. Wilson +5
We present AlphaLab, an autonomous research harness that leverages frontier LLM agentic capabilities to automate the full experimental cycle in quantitative, computation-intensive…
Proximal Point Nash Learning from Human Feedback
Daniil Tiapkin, Daniele Calandriello, Denis Belomestny +5
Traditional Reinforcement Learning from Human Feedback (RLHF) often relies on reward models, frequently assuming preference structures like the Bradley--Terry model, which may not…
Improving Reasoning for Diffusion Language Models via Group Diffusion Policy Optimization
Kevin Rojas, Jiahe Lin, Kashif Rasul +4
Diffusion language models (DLMs) enable parallel, order-agnostic generation with iterative refinement, offering a flexible alternative to autoregressive large language models (LLMs…