Stochastic Gradient Descent

Recent momentum

emerging

0 papers in the last 28 days · 0.0% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this field, kept on the site without email delivery.

Period ending 2026-09-21

21 new papers

A weekly snapshot of new work published in Stochastic Gradient Descent.

Period ending 2026-09-14

17 new papers

A weekly snapshot of new work published in Stochastic Gradient Descent.

Period ending 2026-09-07

12 new papers

A weekly snapshot of new work published in Stochastic Gradient Descent.

Inside this field

Focused directions

541 papers

Latest in Stochastic Gradient Descent

  1. Move on Muon : A Hamiltonian probability gradient flow perspective of Muon optimizer

    May 22, 2026Aratrika Mustafi, Soumya Mukherjee, Bharath K. SriperumbudurMuon OptimizerMuon

  2. A Typed Tensor Language for Federated Learning

    May 20, 2026Theofilos Mailis, Kalliopi-Christina Despotidou, Konstantinos Filippopolitis +6Federated LearningTensor Programs

  3. From SGD to Muon: Adaptive Optimization via Schatten-p Norms

    May 19, 2026Thomas Massena, Corentin Friedrich, Mathieu SerrurierMuon OptimizerOptimizer Design

  4. The Symmetries of Three-Layer ReLU Networks

    May 18, 2026Johanna Marie Gegenfurtner, Moritz Grillo, Guido MontúfarRectified Linear Unit NetworksSymmetry

  5. High-dimensional Limit of SGD for Diagonal Linear Networks

    May 16, 2026Begoña García Malaxechebarría, Courtney Paquette, Maryam Fazel +1Stochastic Gradient DescentStochastic Differential Equations

  6. Inference-Time Machine Unlearning via Gated Activation Redirection

    May 12, 2026Vinícius Conte Turani, Otávio Parraga, João Vitor Boer Abitante +7Training-Free Inference FrameworkInference-Time

  7. Holder Policy Optimisation

    May 12, 2026Yuxiang Chen, Dingli Liang, Yihang Chen +8Group Relative Policy OptimizationGradient Concentration

  8. Sobolev Regularized MMD Gradient Flow

    May 12, 2026Chenyang Tian, Bharath K. Sriperumbudur, Arthur Gretton +1Maximum Mean DiscrepancyWasserstein Gradient Flows