NumGM project approved
The DFG has approved the 36-month project “Numerically Efficient Learning of Generative Models and Beyond (NumGM).”
We investigate the mathematical foundations of machine learning through the dynamics, geometry and thermodynamics of learning algorithms.
The DFG has approved the 36-month project “Numerically Efficient Learning of Generative Models and Beyond (NumGM).”
The stated application deadline for this NumGM postdoctoral position was 11 September 2026.
Benjamin Gess and Johannes Müller introduce an advective Fisher–Rao metric for optimization on paths of probability measures.
Workshop report · Oberwolfach Reports 23(1), 821–822 (2026). Fluctuating continuum descriptions of stochastic gradient descent.
Workshop report · Oberwolfach Reports 23(1), 698–700 (2026). A large-deviations perspective on spikes and catapult behaviour in stochastic gradient descent.
Deep learning offers a powerful approach to quantum many-body problems via neural network wavefunctions, but their optimization remains a severe bottleneck. Existing optimization methods, including natural gradient descent and stochastic reconfiguration, suffer from spectral gap-dependent convergence that limits their effectiveness on systems…
In this work, we establish the small-noise asymptotic behaviour (namely, the functional law of large numbers and the large deviation principle) for multi-scale McKean–Vlasov diffusions with super-linear kernels. In this setting, the interaction depends on the laws of both the slow component…
Published · Computer Methods in Applied Mechanics and Engineering 462, 119289 (2026). Efficient and robust optimization is essential for neural networks, enabling scientific machine learning models to converge rapidly to very high accuracy — faithfully capturing complex physical behavior governed by differential equations. In…
Large loss spikes in stochastic gradient descent are studied through a rigorous large-deviations analysis for a shallow, fully connected network in the NTK scaling. In contrast to full-batch gradient descent, the catapult phase is shown to split into inflationary and deflationary regimes,…
Physics-Informed Neural Networks (PINNs) are a class of deep learning models aiming to approximate solutions of PDEs by training neural networks to minimize the residual of the equation. Focusing on non-equilibrium fluctuating systems, we propose a physically informed choice of penalization that…
We propose a framework for the design and analysis of optimization algorithms in variational quantum Monte Carlo, drawing on geometric insights into the corresponding function space. The framework translates infinite-dimensional optimization dynamics into tractable parameter-space algorithms through a Galerkin projection onto the…
In this chapter, we utilize dynamical systems to analyze several aspects of machine learning algorithms. As an expository contribution we demonstrate how to re-formulate a wide variety of challenges from deep neural networks, (stochastic) gradient descent, and related topics into dynamical statements.…