SearcharxivSearch

arXiv subjects

Jan Maas

Publications and source records attributed to Jan Maas.

At least 19 recordsLinked to original sources

Learning Discrete Diffusion of Graphs via Free-Energy Gradient Flows

Diffusion-based models on continuous spaces have seen substantial recent progress through the mathematical framework of gradient flows, leveraging the Wasserstein-2 (${W}_2$) metric via the Jordan-Kinderlehrer-Otto (JKO) scheme. Despite the increasing popularity of diffusion models on discrete spaces using continuous-time Markov chains, a parallel theoretical framework based on gradient flows has remained elusive due to intrinsic challenges in translating the ${W}_2$ distance directly into these settings. In this work, we propose the first computational approach addressing these challenges, leveraging an appropriate metric $W_K$ on the simplex of probability distributions, which enables us to interpret widely used discrete diffusion paths, such as the discrete heat equation, as gradient flows of specific free-energy functionals. Through this theoretical insight, we introduce a novel methodology for learning diffusion dynamics over discrete spaces, which recovers the underlying functional directly by leveraging first-order optimality conditions for the JKO scheme. The resulting method optimizes a simple quadratic loss, trains extremely fast, does not require individual sample trajectories, and only needs a numerical preprocessing computing $W_K$-geodesics. We validate our method through extensive numerical experiments on synthetic data, showing that we can recover the underlying functional for a variety of graph classes.

cs.LG

Kinetic Optimal Transport (OTIKIN) -- Part 1: Second-Order Discrepancies Between Probability Measures

This is the first part of a general description in terms of mass transport for time-evolving interacting particles systems, at a mesoscopic level. Beyond kinetic theory, our framework naturally applies in biology, computer vision, and engineering. The central object of our study is a new discrepancy $\mathsf d$ between two probability distributions in position and velocity states, which is reminiscent of the $2$-Wasserstein distance, but of second-order nature. We construct $\mathsf d$ in two steps. First, we optimise over transport plans. The cost function is given by the minimal acceleration between two coupled states on a fixed time horizon $T$. Second, we further optimise over the time horizon $T>0$. We prove the existence of optimal transport plans and maps, and study two time-continuous characterisations of $\mathsf d$. One is given in terms of dynamical transport plans. The other one -- in the spirit of the Benamou--Brenier formula -- is formulated as the minimisation of an action of the acceleration field, constrained by Vlasov's equations. Equivalence of static and dynamical formulations of $\mathsf d$ holds true. While part of this result can be derived from recent, parallel developments in optimal control between measures, we give an original proof relying on two new ingredients: Galilean regularisation of Vlasov's equations and a kinetic Monge--Mather shortening principle. Finally, we establish a first-order differential calculus in the geometry induced by $\mathsf d$, and identify solutions to Vlasov's equations with curves of measures satisfying a certain $\mathsf d$-absolute continuity condition. One consequence is an explicit formula for the $\mathsf d$-derivative of such curves.

math.AP

Stochastic homogenisation of nonlinear minimum-cost flow problems

This paper deals with the large-scale behaviour of nonlinear minimum-cost flow problems on random graphs. In such problems, a random nonlinear cost functional is minimised among all flows (discrete vector-fields) with a prescribed net flux through each vertex. On a stationary random graph embedded in $\mathbb{R}^d$, our main result asserts that these problems converge, in the large-scale limit, to a continuous minimisation problem where an effective cost functional is minimised among all vector fields with prescribed divergence. Our main result is formulated using $\Gamma$-convergence and applies to multi-species problems. The proof employs the blow-up technique by Fonseca and M\"uller in a discrete setting. One of the main challenges overcome is the construction of the homogenised energy density on random graphs without a periodic structure.

math.AP

$L^\infty$-optimal transport of anisotropic log-concave measures and exponential convergence in Fisher's infinitesimal model

We prove upper bounds on the $L^\infty$-Wasserstein distance from optimal transport between strongly log-concave probability densities and log-Lipschitz perturbations. In the simplest setting, such a bound amounts to a transport-information inequality involving the $L^\infty$-Wasserstein metric and the relative $L^\infty$-Fisher information. We show that this inequality can be sharpened significantly in situations where the involved densities are anisotropic. Our proof is based on probabilistic techniques using Langevin dynamics. As an application of these results, we obtain sharp exponential rates of convergence in Fisher's infinitesimal model from quantitative genetics, generalising recent results by Calvez, Poyato, and Santambrogio in dimension 1 to arbitrary dimensions.

math.PR

Improved Convergence of Score-Based Diffusion Models via Prediction-Correction

Score-based generative models (SGMs) are powerful tools to sample from complex data distributions. Their underlying idea is to (i) run a forward process for time $T_1$ by adding noise to the data, (ii) estimate its score function, and (iii) use such estimate to run a reverse process. As the reverse process is initialized with the stationary distribution of the forward one, the existing analysis paradigm requires $T_1\to\infty$. This is however problematic: from a theoretical viewpoint, for a given precision of the score approximation, the convergence guarantee fails as $T_1$ diverges; from a practical viewpoint, a large $T_1$ increases computational costs and leads to error propagation. This paper addresses the issue by considering a version of the popular predictor-corrector scheme: after running the forward process, we first estimate the final distribution via an inexact Langevin dynamics and then revert the process. Our key technical contribution is to provide convergence guarantees which require to run the forward process only for a fixed finite time $T_1$. Our bounds exhibit a mild logarithmic dependence on the input dimension and the subgaussian norm of the target distribution, have minimal assumptions on the data, and require only to control the $L^2$ loss on the score approximation, which is the quantity minimized in practice.

cs.LG

Local Conditions for Global Convergence of Gradient Flows and Proximal Point Sequences in Metric Spaces

This paper deals with local criteria for the convergence to a global minimiser for gradient flow trajectories and their discretisations. To obtain quantitative estimates on the speed of convergence, we consider variations on the classical Kurdyka--{\L}ojasiewicz inequality for a large class of parameter functions. Our assumptions are given in terms of the initial data, without any reference to an equilibrium point. The main results are convergence statements for gradient flow curves and proximal point sequences to a global minimiser, together with sharp quantitative estimates on the speed of convergence. These convergence results apply in the general setting of lower semicontinuous functionals on complete metric spaces, generalising recent results for smooth functionals on $\mathbb{R}^n$. While the non-smooth setting covers very general spaces, it is also useful for (non)-smooth functionals on $\mathbb{R}^n$.

math.OC

Characterisation of gradient flows for a given functional

Let $X$ be a vector field and $Y$ be a co-vector field on a smooth manifold $M$. Does there exist a smooth Riemannian metric $g_{\alpha \beta}$ on $M$ such that $Y_\beta = g_{\alpha \beta} X^\alpha$? The main result of this note gives necessary and sufficient conditions for this to be true. As an application of this result we show that a finite-dimensional ergodic Lindblad equation admits a gradient flow structure for the von Neumann relative entropy if and only if the condition of BKM-detailed balance holds.

math.DG

Homogenisation of dynamical optimal transport on periodic graphs

This paper deals with the large-scale behaviour of dynamical optimal transport on $\mathbb{Z}^d$-periodic graphs with general lower semicontinuous and convex energy densities. Our main contribution is a homogenisation result that describes the effective behaviour of the discrete problems in terms of a continuous optimal transport problem. The effective energy density can be explicitly expressed in terms of a cell formula, which is a finite-dimensional convex programming problem that depends non-trivially on the local geometry of the discrete graph and the discrete energy density. Our homogenisation result is derived from a $\Gamma$-convergence result for action functionals on curves of measures, which we prove under very mild growth conditions on the energy density. We investigate the cell formula in several cases of interest, including finite-volume discretisations of the Wasserstein distance, where non-trivial limiting behaviour occurs.

math.AP

Gradient flow formulation of diffusion equations in the Wasserstein space over a metric graph

This paper contains two contributions in the study of optimal transport on metric graphs. Firstly, we prove a Benamou-Brenier formula for the Wasserstein distance, which establishes the equivalence of static and dynamical optimal transport. Secondly, in the spirit of Jordan-Kinderlehrer-Otto, we show that McKean-Vlasov equations can be formulated as gradient flow of the free energy in the Wasserstein space of probability measures. The proofs of these results are based on careful regularisation arguments to circumvent some of the difficulties arising in metric graphs, namely, branching of geodesics and the failure of semi-convexity of entropy functionals in the Wasserstein space.

math.AP

Evolutionary $\Gamma$-convergence of entropic gradient flow structures for Fokker-Planck equations in multiple dimensions

We consider finite-volume approximations of Fokker-Planck equations on bounded convex domains in $\mathbb{R}^d$ and study the corresponding gradient flow structures. We reprove the convergence of the discrete to continuous Fokker-Planck equation via the method of Evolutionary $\Gamma$-convergence, i.e., we pass to the limit at the level of the gradient flow structures, generalising the one-dimensional result obtained by Disser and Liero. The proof is of variational nature and relies on a Mosco convergence result for functionals in the discrete-to-continuum limit that is of independent interest. Our results apply to arbitrary regular meshes, even though the associated discrete transport distances may fail to converge to the Wasserstein distance in this generality.

math.AP

Trajectorial dissipation and gradient flow for the relative entropy in Markov chains

We study the temporal dissipation of variance and relative entropy for ergodic Markov Chains in continuous time, and compute explicitly the corresponding dissipation rates. These are identified, as is well known, in the case of the variance in terms of an appropriate Hilbertian norm; and in the case of the relative entropy, in terms of a Dirichlet form which morphs into a version of the familiar Fisher information under conditions of detailed balance. Here we obtain trajectorial versions of these results, valid along almost every path of the random motion and most transparent in the backwards direction of time. Martingale arguments and time reversal play crucial roles, as in the recent work of Karatzas, Schachermayer and Tschiderer for conservative diffusions. Extension are developed to general "convex divergences" and to countable state-spaces. The steepest descent and gradient flow properties for the variance, the relative entropy, and appropriate generalizations, are studied along with their respective geometries under conditions of detailed balance, leading to a very direct proof for the HWI inequality of Otto and Villani in the present context.

math.PR

Modeling of chemical reaction systems with detailed balance using gradient structures

We consider various modeling levels for spatially homogeneous chemical reaction systems, namely the chemical master equation, the chemical Langevin dynamics, and the reaction-rate equation. Throughout we restrict our study to the case where the microscopic system satisfies the detailed-balance condition. The latter allows us to enrich the systems with a gradient structure, i.e. the evolution is given by a gradient-flow equation. We present the arising links between the associated gradient structures that are driven by the relative entropy of the detailed-balance steady state. The limit of large volumes is studied in the sense of evolutionary $\Gamma$-convergence of gradient flows. Moreover, we use the gradient structures to derive hybrid models for coupling different modeling levels.

math.AP

Scaling limits of discrete optimal transport

We consider dynamical transport metrics for probability measures on discretisations of a bounded convex domain in $\mathbb{R}^d$. These metrics are natural discrete counterparts to the Kantorovich metric $\mathbb{W}_2$, defined using a Benamou-Brenier type formula. Under mild assumptions we prove an asymptotic upper bound for the discrete transport metric $\mathcal{W}_{\mathcal{T}}$ in terms of $\mathbb{W}_2$, as the size of the mesh $\mathcal{T}$ tends to $0$. However, we show that the corresponding lower bound may fail in general, even on certain one-dimensional and symmetric two-dimensional meshes. In addition, we show that the asymptotic lower bound holds under an isotropy assumption on the mesh, which turns out to be essentially necessary. This assumption is satisfied, e.g., for tilings by convex regular polygons, and it implies Gromov-Hausdorff convergence of the transport metric.

math.AP

Homogenisation of one-dimensional discrete optimal transport

This paper deals with dynamical optimal transport metrics defined by spatial discretisation of the Benamou--Benamou formula for the Kantorovich metric $W_2$. Such metrics appear naturally in discretisations of $W_2$-gradient flow formulations for dissipative PDE. However, it has recently been shown that these metrics do not in general converge to $W_2$, unless strong geometric constraints are imposed on the discrete mesh. In this paper we prove that, in a $1$-dimensional periodic setting, discrete transport metrics converge to a limiting transport metric with a non-trivial effective mobility. This mobility depends sensitively on the geometry of the mesh and on the non-local mobility at the discrete level. Our result quantifies to what extent discrete transport can make use of microstructure in the mesh to reduce the cost of transport.

math.AP

Non-commutative calculus, optimal transport and functional inequalities in dissipative quantum systems

We study dynamical optimal transport metrics between density matrices associated to symmetric Dirichlet forms on finite-dimensional $C^*$-algebras. Our setting covers arbitrary skew-derivations and it provides a unified framework that simultaneously generalizes recently constructed transport metrics for Markov chains, Lindblad equations, and the Fermi Ornstein--Uhlenbeck semigroup. We develop a non-nommutative differential calculus that allows us to obtain non-commutative Ricci curvature bounds, logarithmic Sobolev inequalities, transport-entropy inequalities, and spectral gap estimates.

math.OA

On the geometry of geodesics in discrete optimal transport

We consider the space of probability measures on a discrete set $X$, endowed with a dynamical optimal transport metric. Given two probability measures supported in a subset $Y \subseteq X$, it is natural to ask whether they can be connected by a constant speed geodesic with support in $Y$ at all times. Our main result answers this question affirmatively, under a suitable geometric condition on $Y$ introduced in this paper. The proof relies on an extension result for subsolutions to discrete Hamilton-Jacobi equations, which is of independent interest.

math.MG

Gradient flow and entropy inequalities for quantum Markov semigroups with detailed balance

We study a class of ergodic quantum Markov semigroups on finite-dimensional unital $C^*$-algebras. These semigroups have a unique stationary state $σ$, and we are concerned with those that satisfy a quantum detailed balance condition with respect to $σ$. We show that the evolution on the set of states that is given by such a quantum Markov semigroup is gradient flow for the relative entropy with respect to $σ$ in a particular Riemannian metric on the set of states. This metric is a non-commutative analog of the $2$-Wasserstein metric, and in several interesting cases we are able to show, in analogy with work of Otto on gradient flows with respect to the classical $2$-Wasserstein metric, that the relative entropy is strictly and uniformly convex with respect to the Riemannian metric introduced here. As a consequence, we obtain a number of new inequalities for the decay of relative entropy for ergodic quantum Markov semigroups with detailed balance.

math.OA

Generalized optimal transport with singular sources

We present a generalized optimal transport model in which the mass-preserving constraint for the $L^2$-Wasserstein distance is relaxed by introducing a source term in the continuity equation. The source term is also incorporated in the path energy by means of its squared $L^2$-norm in time of a functional with linear growth in space. This extension of the original transport model enables local density modulation, which is a desirable feature in applications such as image warping and blending. A key advantage of the use of a functional with linear growth in space is that it allows for singular sources and sinks, which can be supported on points or lines. On a technical level, the $L^2$-norm in time ensures a disintegration of the source in time, which we use to obtain the well-posedness of the model and the existence of geodesic paths. Furthermore, a numerical scheme based on the proximal splitting approach (Papadakis et al., 2014) is presented. We compare our model with the corresponding model involving the $L^2(L^2)$-norm of the source, which merges the metamorphosis approach and the optimal transport approaches in imaging. Selected numerical test cases show strikingly different behaviour.

math.NA