2 citations · 2 across the 2 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Transmuting prompts into weights
Hanna Mazzawi, Benoit Dherin, Michael Munn +3
A growing body of research has demonstrated that the behavior of large language models can be effectively controlled at inference time by directly modifying their internal states,…
cs.LG2025
Learning by solving differential equations
Benoit Dherin, Michael Munn, Hanna Mazzawi +3
Modern deep learning algorithms use variations of gradient descent as their main learning methods. Gradient descent can be understood as the simplest Ordinary Differential Equation…