[2503.22478] Almost Bayesian: The Fractal Dynamics of Stochastic Gradient Descent
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status Get status notifications via email or slack
[2503.22478] Almost Bayesian: The Fractal Dynamics of Stochastic Gradient Descent Skip to main content arXiv is now an independent nonprofit! Learn more × Search arXiv Press Enter to search · Advanced search --> Computer Science > Machine Learning arXiv:2503.22478 (cs) [Submitted on 28 Mar 2025 ( v1 ), last revised 16 Mar 2026 (this version, v2)] Title: Almost Bayesian: The Fractal Dynamics of Stochastic Gradient Descent Authors: Max Hennick , Stijn De Baerdemacker View a PDF of the paper titled Almost Bayesian: The Fractal Dynamics of Stochastic Gradient Descent, by Max Hennick a
Explore this link on the map →related reading
- Stochastic gradient descent - Wikipediaen.m.wikipedia.org
- Neural network training makes beautiful fractals | Jascha’s blogsohl-dickstein.github.io
- Bayesian Neural Networkscs.toronto.edu
- Why Does SGD Love Flat Minima? Marginally Better blogrishit-dagli.github.io
- [2101.12176] On the Origin of Implicit Regularization in Stochastic Gradient Descentarxiv.org
- Gregory Gundersengregorygundersen.com
- Jane Street Blog - Does batch size matter?blog.janestreet.com
- Maybe I was too harsh on deep learning theory (three days ago) — LessWronglesswrong.com
- arxiv.org/pdf/1805.08522arxiv.org
- Learning to be Bayesian without Supervisionpapers.nips.cc
- [1512.04202] Preconditioned Stochastic Gradient Descentarxiv.org
- Thoughts on loss landscapes and why deep learning worksberen.io