cs.LGSep 23, 2026

NGN: Learning Neural Network Size as a Differentiable Count

Authors: Lixing Li

Organizations: Cornell University Ithaca, NY 14853

Abstract

Neural network size is usually chosen before training, separating architecture selection from weight optimization. We introduce the Neurogenesis Network (NGN), a differentiable parameterization for learning how many ordered structural components a model should use. For each ordered component group, one learnable boundary selects an active prefix while the model parameters are trained. The boundary can grow from a compact initialization and can be deployed by discarding components beyond the learned boundary. Controlled experiments examine convergence of the learned boundary, the performance of deployed prefixes, and comparisons with fixed-size models and alternative approaches to learning capacity. We then apply the same mechanism to MLPs, convolutional and graph networks, Transformers, state-space models, LoRA, and adapters. Across these settings, deploying only the learned prefix usually changes performance little, and the selected architectures perform similarly to fixed models trained at the same size. These results show that structural capacity can be optimized directly as a count.

Figures & tables

Appendix figures & tables13 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. On the Stability of Growth in Structural Plasticity

    May 14, 2026Lute Lillo, Nick CheneyPlasticityGrowth

  2. Every Feedforward Neural Network Definable in an o-Minimal Structure Has Finite Sample Complexity

    May 8, 2026Anastasis Kratsios, Gregory Cousins, Haitz Sáez de Ocáriz Borde +2Feed-ForwardLearnability

  3. Rethinking Neural Width for Alternating Current Optimal Power Flow Proxies

    Jun 2, 2026Dhruvi Khandelwal, Anurag Basistha, Ayushi Jolotia +1Optimal Power FlowDeep Network