SearcharxivSearch

arXiv subjects

Shukai Du

Publications and source records attributed to Shukai Du.

8 recordsLinked to original sources

Lagrangian Flow Matching: A Least-Action Framework for Principled Path Design

Flow matching trains a neural velocity field by regression against a target velocity associated with a prescribed probability path connecting a simple initial distribution to the data distribution. A central design choice is the path itself. Existing constructions, including rectified and optimal-transport-based paths, transport samples along straight lines between coupled endpoints and thus cover only a narrow class of dynamics. We observe that this corresponds to the simplest case of the least-action principle in classical mechanics, in which the kinetic Lagrangian yields free-particle straight-line trajectories. Building on this observation, we propose Lagrangian flow matching, a physics-based framework in which the probability path and velocity field are determined by minimizing the action of a general Lagrangian subject to the continuity equation and the prescribed endpoints. We show that this dynamic problem admits an equivalent static optimal transport (OT) formulation, yielding a family of simulation-free training objectives that recover OT-based flow matching as the kinetic special case and the trigonometric variance-preserving diffusion path as the harmonic-oscillator case. More general Lagrangians give rise to new probability paths and velocity fields, and numerical experiments show that they induce meaningful changes in the learned dynamics while remaining competitive with existing conditional flow matching models.

cs.LG

Element learning: a systematic approach of accelerating finite element-type methods via machine learning, with applications to radiative transfer

In this paper, we propose a systematic approach for accelerating finite element-type methods by machine learning for the numerical solution of partial differential equations (PDEs). The main idea is to use a neural network to learn the solution map of the PDEs and to do so in an element-wise fashion. This map takes input of the element geometry and the PDEs' parameters on that element, and gives output of two operators -- (1) the in2out operator for inter-element communication, and (2) the in2sol operator (Green's function) for element-wise solution recovery. A significant advantage of this approach is that, once trained, this network can be used for the numerical solution of the PDE for any domain geometry and any parameter distribution without retraining. Also, the training is significantly simpler since it is done on the element level instead on the entire domain. We call this approach element learning. This method is closely related to hybridizbale discontinuous Galerkin (HDG) methods in the sense that the local solvers of HDG are replaced by machine learning approaches. Numerical tests are presented for an example PDE, the radiative transfer equation, in a variety of scenarios with idealized or realistic cloud fields, with smooth or sharp gradient in the cloud boundary transition. Under a fixed accuracy level of $10^{-3}$ in the relative $L^2$ error, and polynomial degree $p=6$ in each element, we observe an approximately 5 to 10 times speed-up by element learning compared to a classical finite element-type method.

math.NA

A note on devising HDG+ projections on polyhedral elements

In this note, we propose a simple way of constructing HDG+ projections on polyhedral elements. The projections enable us to analyze the Lehrenfeld-Schöberl HDG (HDG+) methods in a very concise manner, and make many existing analysis techniques of standard HDG methods reusable for HDG+. The novelty here is an alternative way of constructing the projections without using $M$-decompositions as a middle step. This extends our previous results [S. Du and F.-J. Sayas, SpringerBriefs in Mathematics (2019)] (elliptic problems) and [S. Du and F.-J. Sayas, Math.~Comp.~89~(2020), 1745-1782] (elasticity) to polyhedral meshes.

math.NA

A unified error analysis of HDG methods for the static Maxwell equations

We propose a framework that allows us to analyze different variants of HDG methods for the static Maxwell equations using one simple analysis. It reduces all the work to the construction of projections that best fit the structures of the approximation spaces. As applications, we analyze four variants of HDG methods (denoted by B, H, B+, H+), where two of them are known (variants H, B+) and the other two are new (variants H+, B). Under certain regularity assumption, we show that all the four variants are optimally convergent and that variants B+ and H+ achieve superconvergence without post-processing. For the two known variants, we prove their optimal convergence under weaker requirements of the meshes and the stabilization functions thanks to the new analysis techniques being introduced. For solution with low-regularity, we give an analysis to these methods and investigate the effect of different stabilization functions on the convergence. At the end, we provide numerical experiments to support the analysis.

math.NA

New analytical tools for HDG in elasticity, with applications to elastodynamics

We present some new analytical tools for the error analysis of hybridizable discontinuous Galerkin (HDG) method for linear elasticity. These tools allow us to analyze more variants of HDG method using the projection-based approach, which renders the error analysis simple and concise. The key result is a tailored projection for the Lehrenfeld-Schöberl type HDG (HDG+ for simplicity) methods. By using the projection we recover the error estimates of HDG+ for steady-state and time-harmonic elasticity in a simpler analysis. We also present a semi-discrete (in space) HDG+ method for transient elastic waves and prove it is uniformly-in-time optimal convergent by using the projection-based error analysis. Numerical experiments supporting our analysis are presented at the end.

math.NA

Analysis of models for viscoelastic wave propagation

We consider the problem of waves propagating in a viscoelastic solid. For the material properties of the solid we consider both classical and fractional differentiation in time versions of the Zener, Maxwell, and Voigt models, where the coupling of different models within the same solid are covered as well. Stability of each model is investigated in the Laplace domain, and these are then translated to time-domain estimates. With the use of semigroup theory, some time-domain results are also given which avoid using the Laplace transform and give sharper estimates. We take the time to develop and explain the theory necessary to understand the relation between the equations we solve in the Laplace domain and those in the time-domain which are written using the language of causal tempered distributions. Finally we offer some numerical experiments that highlight some of the differences between the models and how different parameters effect the results.

math.NA

Calculation of Cauchy stress tensor in molecular dynamics system with a generalized Irving-Kirkwood formulism

Irving and Kirkwood formulism (IK formulism) provides a way to compute continuum mechanics quantities at certain location in terms of molecular variables. To make the approach more practical in computer simulation, Hardy proposed to use a spacial kernel function that couples continuum quantities with atomistic information. To reduce irrational fluctuations, Murdoch proposed to use a temporal kernel function to smooth the physical quantities obtained in Hardy's approach. In this paper, we generalize the original IK formulism to systematically incorporate both spacial and temporal average. The Cauchy stress tensor is derived in this generalized IK formulism (g-IK formulism). Analysis is given to illuminate the connection and difference between g-IK formulism and traditional temporal post-process approach. The relationship between Cauchy stress and first Piola-Kirchhoff stress is restudied in the framework of g-IK formulism. Numerical experiments using molecular dynamics are conducted to examine the analysis results.

physics.comp-ph

On the two mutually independent factors that determine the convergence of least-squares projection method

This paper investigates the least-squares projection method for bounded linear operators, which provides a natural regularization scheme by projection for many ill-posed problems. Yet, without additional assumptions, the convergence of this approximation scheme cannot be guaranteed. We reveal that the convergence of least-squares projection method is determined by two independent factors -- the kernel approximability and the offset angle. The kernel approximability is a necessary condition of convergence described with kernel $N(T)$ and its subspaces $N(T){\cap}X_n$, and we give several equivalent characterizations for it (Theorem 1). The offset angle of $X_n$ is defined as the largest canonical angle between space $T^*T(X_n)$ and $T^{\dagger}T(X_n)$ (which are subspaces of $N(T)^\bot$), and it geometrically reflects the rate of convergence (Theorem 2). The paper also presents new observations for the unconvergence examples of Seidman [10, Example 3.1] and Du [2, Example 2.10] under the notions of kernel approximability and offset angle.

math.NA