Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Double Preconditioning (DoPr): Optimization for Test-Time Performance, not Validation Loss
Thomas T. Zhang, Alok Shah, Yifei Zhang +3
Many modern applications of deep learning involve training a neural network via a one-step prediction loss (e.g., regression, cross-entropy), but deploy the network by rollin…
cs.LG2024
Regret Analysis of Multi-task Representation Learning for Linear-Quadratic Adaptive Control
Bruce D. Lee, Leonardo F. Toso, Thomas T. Zhang +2
Representation learning is a powerful tool that enables learning over large multitudes of agents or domains by enforcing that all agents operate on a shared set of learned features…