deep equilibrium models 1gradient starvation 1implicit layers 1model diagnostics 1physics-informed networks 1reasoning benchmarks 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.AI2026
Numeracy in Large Language Models: Fundamental Limitations and Paths to Improvement
Aoxin Ni
Large language models (LLMs) achieve strong results on mathematical reasoning benchmarks yet remain unreliable on elementary numerical tasks, including magnitude comparison, large-…
cs.LG2026
The Evaluation Protocol Determines the Result: An Independent Reproduction of LeWorldModel on TwoRoom
Joyjeet Singh
LeWorldModel trains a latent world model with a prediction loss and a single anti-collapse regulariser, and reports approximately 87% of goals reached on TwoRoom, its simplest diag…
cs.LG2026
The Equilibrium Is the Initialization: Lazy Identity Collapse in Physics-Structured Deep Equilibrium Reasoning
Joyjeet Singh
Deep equilibrium models promise input-adaptive implicit computation: harder problems should demand more solver iterations, and the solved equilibrium should encode the result of ge…