3 papers
cs.LG2026
On the Optimizer Dependence of Neural Scaling Laws
Vansh Ramani, Shourya Vir Jain
The scaling exponent in neural scaling laws is commonly treated as a fixed constant set by architecture and data. We present evidence that depends…
cs.LG2026
ART: Adaptive Resampling-based Training for Imbalanced Classification
Arjun Basandrai, Shourya Jain, K. Ilanthenral
Traditional resampling methods for handling class imbalance typically uses fixed distributions, undersampling the majority or oversampling the minority. These static strategies ign…
cs.LG2026
Language Models Entangle Language and Culture
Shourya Jain, Paras Chopra
Users should not be systemically disadvantaged by the language they use for interacting with LLMs; i.e. users across languages should get responses of similar quality irrespective…