Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
GAIA: Geometry-Adaptive Operator Learning for Forward and Inverse Problems
Meenakshi Krishnan, Pranav Pulijala, Ke Chen +2
Operator learning for partial differential equations (PDEs) on arbitrary geometries builds fast neural surrogates for large-scale simulation. Although recent geometry-adaptive neur…
cs.LG2024
Towards Better Generalization: Weight Decay Induces Low-rank Bias for Neural Networks
Ke Chen, Chugang Yi, Haizhao Yang
We study the implicit bias towards low-rank weight matrices when training neural networks (NN) with Weight Decay (WD). We prove that when a ReLU NN is sufficiently trained with Sto…