Searcharxiv⌕ Search

arXiv subjects

Goncalo dos Reis

Publications and source records attributed to Goncalo dos Reis.

At least 19 recordsLinked to original sources

Numerical approximation of McKean-Vlasov SDEs via stochastic gradient descent

We propose a novel approach to numerically approximate McKean-Vlasov stochastic differential equations (MV-SDE) using stochastic gradient descent (SGD) while avoiding the use of interacting particle systems (IPS) {and the associated simulation costs required to achieve the ``propagation of chaos'' limit}. The SGD technique is deployed to solve a Euclidean minimization problem, obtained by first representing the MV-SDE as a minimization problem over the set of continuous functions of time, and then approximating the domain with a finite-dimensional subspace. Convergence is established by proving certain intermediate stability and moment estimates of the relevant stochastic processes, including the tangent processes. Numerical experiments illustrate the competitive performance of our SGD based method compared to the IPS benchmarks. This work offers a theoretical foundation for using the SGD method in the context of numerical approximation of MV-SDEs, and provides analytical tools to study its stability and convergence.

math.NA↗

Tamed Euler approximation for fully superlinear growth McKean-Vlasov SDE and their particle systems: sharp rates for strong propagation of chaos, convergence and ergodicity

We study McKean--Vlasov Stochastic Differential Equations (MV-SDEs) whose drift and diffusion coefficients are of superlinear growth in \textit{all} their variables thus also superlinear in the measure component (the meaning is specified in the body of the paper). We address the finite and infinite time horizon case. Our contribution is fourfold. (a) We establish well-posedness for this class of equations and the corresponding interacting particle system. (b) We prove two propagation of chaos results with explicit $L^2$-convergence rates: the first, is a general one where the rate degrades as the system's dimension $d$ increases; the second, attains the sharp rate $N^{-1/2}$ (in particle number $N$) uniformly over the dimension $d$ at the cost of a Vlasov kernel structure that is general and of superlinear growth for the measure dependency -- the latter's proof fully avoids the Kantorovich-Rubinstein duality argument. (c) Unlike existing works -- based on semi-implicit schemes or truncated Euler schemes -- we propose a fully explicit tamed Euler scheme that has reduced computational cost (comparatively). The explicit scheme is shown to converge in strong $L^p$-sense with rate $1/2$ (in timestep). (d) Lastly, we establish exponential ergodicity properties and long-time behavior for the MV-SDE, the corresponding interacting particle system, and the tamed scheme. The latter result is, to the best of our knowledge, fully novel.

math.PR↗

Malliavin differentiability of McKean-Vlasov SDEs with common noise

We establish the Malliavin differentiability of McKean-Vlasov stochastic differential equations (MV-SDEs) with common noise under the global Lipschitz assumption in the space variable and the measure variable. Our result gives also meaning to the Malliavin derivative of the conditional law with respect to the common noise. As an application, we derive an integration by parts formula on the Wiener space for the class of common noise MV-SDEs under consideration.

math.PR↗

Many-Server Queueing Systems with Heterogeneous Strategic Servers in Heavy Traffic

In most service systems, the servers are humans who desire to experience a certain level of idleness. In call centers, this manifests itself as the call avoidance behavior, where servers strategically adjust their service rate to strike a balance between the idleness they receive and effort to work harder. Moreover, being humans, each server values this trade-off differently and has different capabilities. Drawing ideas on mean-field games we develop a novel framework relying on measure-valued processes to simultaneously address strategic server behavior and inherent server heterogeneity in service systems. This framework enables us to extend the recent literature on strategic servers in four new directions by: (i) incorporating individual choices of servers, (ii) incorporating individual abilities of servers, (iii) modeling the discomfort experienced by servers due to low levels of idleness, and (iv) considering more general routing policies. Using our framework, we are able to asymptotically characterize asymmetric Nash equilibria for many-server systems with strategic servers. In simpler cases, it has been shown that the purely quality-driven regime is asymptotically optimal. However, we show that if the discomfort increases fast enough as the idleness approaches zero, the quality-and-efficiency-driven regime and other quality driven regimes can be optimal. This is the first time this conclusion appears in the literature.

math.PR↗

Malliavin differentiability of McKean-Vlasov SDEs with locally Lipschitz coefficients

In this short note, we establish Malliavin differentiability of McKean-Vlasov Stochastic Differential Equations (MV-SDEs) with drifts satisfying both a locally Lipschitz and a one-sided Lipschitz assumption, and where the diffusion coefficient is assumed to be uniformly Lipschitz in its variables. As a secondary contribution, we investigate how Malliavin differentiability transfers across the interacting particle system associated with the McKean-Vlasov equation to its limiting equation. This final result requires both spatial and measure differentiability of the coefficients and doubles as a standalone result of independent interest since the study of Malliavin derivatives of weakly interacting particle systems seems novel to the literature. The presentation is didactic and finishes with a discussion on mollification techniques for the Lions derivative.

math.PR↗

An introduction to tensors for path signatures

We present a fit-for-purpose introduction to tensors and their operations. It is envisaged to help the reader become acquainted with its underpinning concepts for the study of path signatures. The text includes exercises, solutions and many intuitive explanations. The material discusses direct sums and tensor products as two possible operations that make the Cartesian product of vectors spaces a vector space. The difference lies in linear Vs. multilinear structures -- the latter being the suitable one to deal with path signatures. The presentation is offered to understand tensors in a deeper sense than just a multidimensional array. The text concludes with the prime example of an algebra in relation to path signatures: the 'tensor algebra'. This manuscript is the extended version (with two extra sections) of a chapter to appear in Open Access in a forthcoming Springer volume ``Signatures Methods in Finance: An Introduction with Computational Applications". The two additional sections here discuss the factoring of tensor product expressions to a minimal number of terms. This problem is relevant for the path signatures theory but not necessary for what is presented in the book. Tensor factorization is an elegant way of becoming familiar with the language of tensors and tensor products. A GitHub repository is attached.

math.HO↗

Wellposedness, exponential ergodicity and numerical approximation of fully super-linear McKean--Vlasov SDEs and associated particle systems

We study a class of McKean--Vlasov Stochastic Differential Equations (MV-SDEs) with drifts and diffusions having super-linear growth in measure and space -- the maps have general polynomial form but also satisfy a certain monotonicity condition. The combination of the drift's super-linear growth in measure (by way of a convolution) and the super-linear growth in space and measure of the diffusion coefficient requires novel technical elements in order to obtain the main results. We establish wellposedness, propagation of chaos (PoC), and under further assumptions on the model parameters, we show an exponential ergodicity property alongside the existence of an invariant distribution. No differentiability or non-degeneracy conditions are required. Further, we present a particle system based Euler-type split-step scheme (SSM) for the simulation of this type of MV-SDEs. The scheme attains, in stepsize, the strong error rate $1/2$ in the non-path-space root-mean-square error metric and we demonstrate the property of mean-square contraction. Our results are illustrated by numerical examples including: estimation of PoC rates across dimensions, preservation of periodic phase-space, and the observation that taming appears to be not a suitable method unless strong dissipativity is present.

math.PR↗

Improved weak convergence for the long time simulation of Mean-field Langevin equations

We study the weak convergence behaviour of the Leimkuhler--Matthews method, a non-Markovian Euler-type scheme with the same computational cost as the Euler scheme, for the approximation of the stationary distribution of a one-dimensional McKean--Vlasov Stochastic Differential Equation (MV-SDE). The particular class under study is known as mean-field (overdamped) Langevin equations (MFL). We provide weak and strong error results for the scheme in both finite and infinite time. We work under a strong convexity assumption. Based on a careful analysis of the variation processes and the Kolmogorov backward equation for the particle system associated with the MV-SDE, we show that the method attains a higher-order approximation accuracy in the long-time limit (of weak order convergence rate $3/2$) than the standard Euler method (of weak order $1$). While we use an interacting particle system (IPS) to approximate the MV-SDE, we show the convergence rate is independent of the dimension of the IPS and this includes establishing uniform-in-time decay estimates for moments of the IPS, the Kolmogorov backward equation and their derivatives. The theoretical findings are supported by numerical tests.

math.NA↗

A Mean Field Ansatz for Zero-Shot Weight Transfer

The pre-training cost of large language models (LLMs) is prohibitive. One cutting-edge approach to reduce the cost is zero-shot weight transfer, also known as model growth for some cases, which magically transfers the weights trained in a small model to a large model. However, there are still some theoretical mysteries behind the weight transfer. In this paper, inspired by prior applications of mean field theory to neural network dynamics, we introduce a mean field ansatz to provide a theoretical explanation for weight transfer. Specifically, we propose the row-column (RC) ansatz under the mean field point of view, which describes the measure structure of the weights in the neural network (NN) and admits a close measure dynamic. Thus, the weights of different sizes NN admit a common distribution under proper assumptions, and weight transfer methods can be viewed as sampling methods. We empirically validate the RC ansatz by exploring simple MLP examples and LLMs such as GPT-3 and Llama-3.1. We show the mean-field point of view is adequate under suitable assumptions which can provide theoretical support for zero-shot weight transfer.

cs.LG↗

High order splitting methods for SDEs satisfying a commutativity condition

In this paper, we introduce a new simple approach to developing and establishing the convergence of splitting methods for a large class of stochastic differential equations (SDEs), including additive, diagonal and scalar noise types. The central idea is to view the splitting method as a replacement of the driving signal of an SDE, namely Brownian motion and time, with a piecewise linear path that yields a sequence of ODEs $-$ which can be discretized to produce a numerical scheme. This new way of understanding splitting methods is inspired by, but does not use, rough path theory. We show that when the driving piecewise linear path matches certain iterated stochastic integrals of Brownian motion, then a high order splitting method can be obtained. We propose a general proof methodology for establishing the strong convergence of these approximations that is akin to the general framework of Milstein and Tretyakov. That is, once local error estimates are obtained for the splitting method, then a global rate of convergence follows. This approach can then be readily applied in future research on SDE splitting methods. By incorporating recently developed approximations for iterated integrals of Brownian motion into these piecewise linear paths, we propose several high order splitting methods for SDEs satisfying a certain commutativity condition. In our experiments, which include the Cox-Ingersoll-Ross model and additive noise SDEs (noisy anharmonic oscillator, stochastic FitzHugh-Nagumo model, underdamped Langevin dynamics), the new splitting methods exhibit convergence rates of $O(h^{3/2})$ and outperform schemes previously proposed in the literature.

math.NA↗

An iterative method for Helmholtz boundary value problems arising in wave propagation

The complex Helmholtz equation $(Δ+ k^2)u=f$ (where $k\in{\mathbb R},u(\cdot),f(\cdot)\in{\mathbb C}$) is a mainstay of computational wave simulation. Despite its apparent simplicity, efficient numerical methods are challenging to design and, in some applications, regarded as an open problem. Two sources of difficulty are the large number of degrees of freedom and the indefiniteness of the matrices arising after discretisation. Seeking to meet them within the novel framework of probabilistic domain decomposition, we set out to rewrite the Helmholtz equation into a form amenable to the Feynman-Kac formula for elliptic boundary value problems. We consider two typical scenarios, the scattering of a plane wave and the propagation inside a cavity, and recast them as a sequence of Poisson equations. By means of stochastic arguments, we find a sufficient and simulatable condition for the convergence of the iterations. Upon discretisation a necessary condition for convergence can be derived by adding up the iterates using the harmonic series for the matrix inverse -- we illustrate the procedure in the case of finite differences. From a practical point of view, our results are ultimately of limited scope. Nonetheless, this unexpected -- even paradoxical -- new direction of attack on the Helmholtz equation proposed by this work offers a fresh perspective on this classical and difficult problem. Our results show that there indeed exists a predictable range $k<k_{max}$ in which this new ansatz works with $k_{max}$ being far below the challenging situation.

math.NA↗

Euler simulation of interacting particle systems and McKean-Vlasov SDEs with fully superlinear growth drifts in space and interaction

We consider in this work the convergence of a split-step Euler type scheme (SSM) for the numerical simulation of interacting particle Stochastic Differential Equation (SDE) systems and McKean-Vlasov Stochastic Differential Equations (MV-SDEs) with full super-linear growth in the spatial and the interaction component in the drift, and non-constant Lipschitz diffusion coefficient. The super-linear growth in the interaction (or measure) component stems from convolution operations with super-linear growth functions allowing in particular application to the granular media equation with multi-well confining potentials. From a methodological point of view, we avoid altogether functional inequality arguments (as we allow for non-constant non-bounded diffusion maps). The scheme attains, in stepsize, a near-optimal classical (path-space) root mean-square error rate of $1/2-\varepsilon$ for $\varepsilon>0$ and an optimal rate $1/2$ in the non-path-space mean-square error metric. Numerical examples illustrate all findings. In particular, the testing raises doubts if taming is a suitable methodology for this type of problem (with convolution terms and non-constant diffusion coefficients).

math.PR↗

Operator-splitting schemes for degenerate, non-local, conservative-dissipative systems

In this paper, we develop a natural operator-splitting variational scheme for a general class of non-local, degenerate conservative-dissipative evolutionary equations. The splitting-scheme consists of two phases: a conservative (transport) phase and a dissipative (diffusion) phase. The first phase is solved exactly using the method of characteristic and DiPerna-Lions theory while the second phase is solved approximately using a JKO-type variational scheme that minimizes an energy functional with respect to a certain Kantorovich optimal transport cost functional. In addition, we also introduce an entropic-regularisation of the scheme. We prove the convergence of both schemes to a weak solution of the evolutionary equation. We illustrate the generality of our work by providing a number of examples, including the kinetic Fokker-Planck equation and the (regularized) Vlasov-Poisson-Fokker-Planck equation.

math.AP↗

A flexible split-step scheme for solving McKean-Vlasov Stochastic Differential Equations

We present an implicit Split-Step explicit Euler type Method (dubbed SSM) for the simulation of McKean-Vlasov Stochastic Differential Equations (MV-SDEs) with drifts of superlinear growth in space, Lipschitz in measure and non-constant Lipschitz diffusion coefficient. The scheme is designed to leverage the structure induced by the interacting particle approximation system, including parallel implementation and the solvability of the implicit equation. The scheme attains the classical $1/2$ root mean square error (rMSE) convergence rate in stepsize and closes the gap left by [18, "Simulation of McKean-Vlasov SDEs with super-linear growth" in IMA Journal of Numerical Analysis, 01 2021. draa099] regarding efficient implicit methods and their convergence rate for this class of McKean-Vlasov SDEs. A sufficient condition for mean-square contractivity of the scheme is presented. Several numerical examples are presented, including a comparative analysis to other known algorithms for this class (Taming and Adaptive time-stepping) across parallel and non-parallel implementations.

math.NA↗

Forward utility and market adjustments in relative investment-consumption games of many players

We study a portfolio management problem featuring many-player and mean field competition, investment and consumption, and relative performance concerns under the forward performance processes (FPP) framework. We focus on agents using power (CRRA) type FPPs for their investment-consumption optimization problem under a common noise Merton market model. We solve both the many-player and mean field game providing closed-form expressions for the solutions where the limit of the former yields the latter. In our case, the FPP framework yields a continuum of solutions for the consumption component as indexed to a market parameter we coin "market-risk relative consumption preference". The parameter permits the agent to set a preference for their consumption going forward in time that, in the competition case, reflects a common market behaviour. We show the FPP framework, under both competition and no-competition, allows the agent to disentangle her risk-tolerance and elasticity of intertemporal substitution (EIS) just like Epstein-Zin preferences under recursive utility framework and unlike the classical utility theory one. This, in turn, allows a finer analysis on the agent's consumption "income" and "substitution" regimes, and, of independent interest, motivates a new strand of economics research on EIS under the FPP framework. We find that competition rescales the agent's perception of consumption in a non-trivial manner. We provide numerical illustrations of our results.

econ.GN↗

Entropic regularisation of non-gradient systems

The theory of Wasserstein gradient flows in the space of probability measures has made an enormous progress over the last twenty years. It constitutes a unified and powerful framework in the study of dissipative partial differential equations (PDEs) providing the means to prove well-posedness, regularity, stability and quantitative convergence to the equilibrium. The recently developed entropic regularisation technique paves the way for fast and efficient numerical methods for solving these gradient flows. However, many PDEs of interest do not have a gradient flow structure and, a priori, the theory is not applicable. In this paper, we develop a time-discrete entropy regularised variational scheme for a general class of such non-gradient PDEs. We prove the convergence of the scheme and illustrate the breadth of the proposed framework with concrete examples including the non-linear kinetic Fokker-Planck (Kramers) equation and a non-linear degenerate diffusion of Kolmogorov type. Numerical simulations are also provided.

math.AP↗

On the relation between Stratonovich and Ito integrals with functional integrands of conditional measure flows

In this small note we explicit the relation between Ito and Stratonovich integrals when conditional measure flow components are present in the integrands. The `correction' term involves Lions-type measure derivatives and clarifies which cross-correlations need to be taken into account. We cast the framework in relation to SDEs of mean-field type depending on conditional flows of measure. The result being trivial under full flows of measure.

math.PR↗