1 citations · 1 across the 5 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Why the Third Axis Is Freedom
Michael Timothy Bennett
In generative training, a model produces an output and is penalised for its difference from an example. With one output per comparison, a model that produces one common answer can…
cs.LG2026
Are Flat Minima an Illusion?
Michael Timothy Bennett
Flat minima are an account of why deep networks generalise. However flatness is a matter of form (parameters), while generalisation is of function. The same function can be a resul…