Period ending 2026-09-21
2 new papers
A weekly snapshot of new work published in Multilayer Perceptrons.
Twelve weeks of publication activity for this topic as it is defined today.
Weekly history
What was published in this topic, kept on the site without email delivery.
Period ending 2026-09-21
A weekly snapshot of new work published in Multilayer Perceptrons.
Period ending 2026-09-14
A weekly snapshot of new work published in Multilayer Perceptrons.
Period ending 2026-09-07
A weekly snapshot of new work published in Multilayer Perceptrons.
118 papers
structure matrix" $A(x)$ that we compute for each architecture. Perceptrons, attention layers, mixtures of experts, and convolutions become one model at different $A$. Its training dynamics then close on the order parameter" and, whenever the data matrices share an eigenbasis, reduce to a Lotka--Volterra equation whose modes switch on one after another. The smaller the initial weights, the further apart the switch-on times, and the plateaus appear as a singular limit of a smooth flow; when many modes are unresolved the same events merge into a power law in training time whose exponent the theory predicts. We confirm both numerically across training methods and architectures.