
Machine Learning
Main fields of application
We investigate the mathematical foundations of machine learning through the dynamics, geometry and thermodynamics of learning algorithms. Randomness is both a source of fluctuations and a mechanism that can improve stability and exploration.
Current directions include stochastic-gradient dynamics, reinforcement learning, neural PDE solvers, diffusion and generative models, and thermodynamically informed learning. We analyse effective equations and large-deviation behaviour, connect optimization to gradient-flow and information-geometric structures, and develop algorithms that exploit these structures. This creates a two-way exchange: stochastic analysis explains learning at large scale, while questions from machine learning motivate new problems in conservative SPDEs, sampling and numerical analysis.
Current directions
- Stochastic-gradient dynamics
- Generative and diffusion models
- Reinforcement learning
- Scientific and thermodynamically informed learning
- Information geometry and optimization
Related people
Related projects
Selected publications
-
A Dynamical Systems Perspective on the Analysis of Neural Networks
In this chapter, we utilize dynamical systems to analyze several aspects of machine learning algorithms. As an expository contribution we demonstrate how to re-formulate a wide variety of challenges from deep neural networks, (stochastic) gradient descent, and related topics into dynamical statements.…
-
Central Path Proximal Policy Optimization
Published · Exploration in AI Today Workshop at ICML 2025. In constrained Markov decision processes, enforcing constraints during training is often thought of as decreasing the final return. Recently, it was shown that constraints can be incorporated directly into the policy geometry,…
-
Non-Asymptotic Analysis of Projected Gradient Descent for Physics-Informed Neural Networks
Published · Scientific Machine Learning: Emerging Topics, SEMA SIMAI Springer Series (2026). In this work, we provide a non-asymptotic convergence analysis of projected gradient descent for physics-informed neural networks for the Poisson equation. Under suitable assumptions, we show that the optimization error can be…
-
Characterizing Dynamical Stability of Stochastic Gradient Descent in Overparameterized Learning
For overparameterized optimization tasks, such as those found in modern machine learning, global minima are generally not unique. In order to understand generalization in these settings, it is vital to study to which minimum an optimization algorithm converges. The possibility of having…
-
Stochastic Modified Flows for Riemannian Stochastic Gradient Descent
Published · SIAM journal on control and optimization, 62 (2024) 6, pp. 3288-3314. We give quantitative estimates for the rate of convergence of Riemannian stochastic gradient descent (RSGD) to Riemannian gradient flow and to a diffusion process, the so-called Riemannian stochastic modified flow (RSMF). Using…
-
Stochastic Modified Flows, Mean-Field Limits and Dynamics of Stochastic Gradient Descent
Published · Journal of machine learning research, 25 (2024) 30, pp. 1-27. We propose new limiting dynamics for stochastic gradient descent in the small learning rate regime called stochastic modified flows. These SDEs are driven by a cylindrical Brownian motion and improve the so-called…
-
Exponential convergence rates for momentum stochastic gradient descent in the overparametrized setting
Published · Mathematical programming, (2026). We prove explicit bounds on the exponential rate of convergence for the momentum stochastic gradient descent scheme (MSGD) for arbitrary, fixed hyperparameters (learning rate, friction parameter) and its continuous-in-time counterpart in the…
-
Conservative SPDEs as fluctuating mean field limits of stochastic gradient descent
Published · Probability theory and related fields, 192 (2025) 3/4, pp. 1447-1515. The convergence of stochastic interacting particle systems in the mean-field limit to solutions of conservative stochastic partial differential equations is established, with optimal rate of convergence. As a second main result, a…
