arXiv Daily Brief

Last updated: 2026-07-22 12:03:19 (UTC+0800)

Papers: 50
Recommended: 2
Recommended 2 papers
All Papers 48 papers (excluding recommended)
#1Reduced Order Modeling of One-Dimensional Conservative PDEs via the Cumulative Distribution Transform
基于累积分布变换的一维守恒型偏微分方程降阶建模
Harbir Antil, Rocío Díaz Martín, Ivan V. Medri, Kristofor E. Pas, Gustavo K. Rohde, Aryan Saxena, Sarswati Shah · 2026-07-19T04:20:24Z
Abstract
We propose a reduced order modeling (ROM) framework for 1D conservative PDEs based on the cumulative distribution transform (CDT). The CDT maps nonnegative, equal-mass states into a Hilbert space in which 1D Wasserstein distances become weighted $L^2$ distances and translations become affine shifts. This makes the transform especially suited for transport-dominated dynamics, where Eulerian linear-subspace ROMs often suffer from slow decay of Kolmogorov widths. We study this phenomenon for scalar conservative dynamics by analyzing the solution manifold in CDT coordinates. For linear transport, the transformed solution manifold is contained in the 2-dimensional space spanned by the transformed initial datum and the constant function, and has zero Kolmogorov $2$-width. For nonlinear hyperbolic conservation laws, we prove two complementary types of estimates: robust $O(n^{-1})$ bounds that rely only on the conservative transport structure and remain meaningful after shock formation, and sharper $O(n^{-2})$ bounds in smooth pre-shock regimes. For conservative advection-diffusion, we show that the CDT trajectory remains within distance $O(\sqrt{DT})$ of the pure-transport plane, and we also obtain sharper $O(D^2T^2)$ estimates under additional regularity or away from initial layers. In both cases, the zero 2-width behavior of linear transport is recovered as the diffusion coefficient tends to zero. Motivated by these estimates, we develop a CDT-POD numerical scheme: snapshots are mapped to CDT space, Proper Orthogonal Decomposition (POD) is performed in transformed coordinates, and the inverse CDT is used to reconstruct physical states. Numerical experiments for several transport-dominated dynamics show that CDT-POD can capture solution manifolds with substantially fewer modes than Eulerian POD.
摘要
我们提出了一种基于累积分布变换(CDT)的一维守恒型PDE降阶建模(ROM)框架。CDT将非负等质量状态映射到Hilbert空间,其中一维Wasserstein距离变为加权L²距离,平移变为仿射偏移。这使得该变换特别适用于以输运为主的动力学,而Eulerian线性子空间ROM在此类问题中常因Kolmogorov宽度衰减缓慢而受限。我们通过分析CDT坐标下的解流形来研究标量守恒动力学的这一现象。对于线性输运,变换后的解流形包含在变换后的初始数据和常值函数张成的二维空间中,且具有零Kolmogorov 2-宽度。对于非线性双曲守恒律,我们证明了两种互补类型的估计:仅依赖于守恒输运结构且在激波形成后仍有意义的鲁棒O(n⁻¹)界,以及在光滑预激波区域更精确的O(n⁻²)界。对于守恒对流扩散,我们证明了CDT轨迹保持在纯输运平面的O(√(DT))距离内,并在额外正则性或远离初始层的情况下获得了更精确的O(D²T²)估计。在两种情况下,线性输运的零2-宽度行为在扩散系数趋于零时恢复。受这些估计启发,我们开发了CDT-POD数值方案:将快照映射到CDT空间,在变换坐标下进行本征正交分解(POD),并使用逆CDT重建物理状态。几个以输运为主的动力学的数值实验表明,CDT-POD能够以显著少于Eulerian POD的模态捕获解流形。
#2A Linear Variable-Step Embedded ETD Scheme with Uniform-in-Time Stability for the 2D Navier--Stokes Equations
二维Navier-Stokes方程具有一致时间稳定性的线性变步长嵌入式ETD格式
Haifeng Wang, Xiaoming Wang, Min Zhang · 2026-07-19T02:43:30Z
Abstract
We propose a linear variable-step exponential time-differencing method for the incompressible Navier--Stokes equations in vorticity--streamfunction formulation on a two-dimensional periodic box. The method consists of a second-order scheme and an embedded first-order variant, yielding a natural mechanism for adaptive time stepping and a posteriori error control. Each time step requires only uniquely solvable linear problems: two heat equation solves, efficiently handled by Fourier methods in the periodic setting, and one linear scalar auxiliary-variable equation, evaluated via Laplace transform and Talbot's numerical inverse transform. The construction combines the ETD framework, a mean-reverting scalar auxiliary variable (mr-SAV), and second-order extrapolation of the nonlinear term. The mean-reverting correction enables long-time stability while preserving full linearity, distinguishing the method from related mr-SAV schemes that require nonlinear algebraic solves. We prove unconditional long-time stability: for uniformly bounded $L^2$ forcing, the discrete vorticity remains bounded in $L^\infty(0,\infty;L^2)$ for all Reynolds numbers and time-step sizes. Numerical experiments
摘要
我们针对二维周期箱上涡量-流函数形式的不可压Navier-Stokes方程提出了一种线性变步长指数时间差分方法。该方法由一个二阶格式和一个嵌入式一阶变体组成,为自适应时间步长和后验误差控制提供了自然机制。每个时间步仅需求解唯一可解的线性问题:两个热方程求解(在周期设置下通过Fourier方法高效处理)和一个线性标量辅助变量方程(通过Laplace变换和Talbot数值逆变换求值)。该构造结合了ETD框架、均值回归标量辅助变量(mr-SAV)和非线性的二阶外推。均值回归修正实现了长期稳定性,同时保持完全线性,区别于需要非线性代数求解的相关mr-SAV方案。我们证明了无条件长期稳定性:对于一致有界的L²强迫项,离散涡量在L^∞(0,∞;L²)中保持有界,与Reynolds数和时间步长无关。数值实验...
#3Field-of-values analysis of augmented Krylov methods for matrix $\varphi$-function actions
矩阵φ-函数作用增广Krylov方法的数值域分析
Xiaobo Liu, Marcel Schweitzer · 2026-07-20T11:03:32Z
Abstract
We revisit established Krylov subspace methods for linear combinations of matrix $\varphi$-function actions from the viewpoint of the block triangular formulation of Al-Mohy and Liu [SIAM J. Sci. Comput., 48 (2026), pp. A726--A747]. In algorithms such as KIOPS [J. Comput. Phys., 372 (2018), pp. 236--255], one uses an augmentation approach based on evaluating the exponential of a slightly larger matrix that contains the operant vectors in its off-diagonal block, and its field of values may therefore grow substantially with these vectors. Typical convergence estimates for Krylov subspace methods result from bounding the error of polynomial approximations for the exponential on the field of values, so that only very pessimistic convergence estimates are available for these methods, in spite of their good practical performance. In contrast, the larger block formulation established by Al-Mohy and Liu involves an operator whose field of values is independent of the operant vectors, leading to more favorable convergence bounds. We work out the details of how these two approaches are connected to each other, which allows us to transfer the convergence bounds from the latter to the former, thus better explaining the observed performance.
摘要
我们从Al-Mohy和Liu的块三角公式角度重新审视了用于矩阵φ-函数作用线性组合的经典Krylov子空间方法。在KIOPS等算法中,使用了一种增广方法,基于对稍大的矩阵(包含操作向量的非对角块)的指数求值,因此其数值域可能随这些向量大幅增长。Krylov子空间方法的典型收敛估计来自于在数值域上对指数多项式逼近误差的界定,因此尽管这些方法在实际中表现良好,但只能得到非常悲观的收敛估计。相比之下,Al-Mohy和Liu建立的大块公式涉及一个算子,其数值域独立于操作向量,从而产生更有利的收敛界。我们详细阐述了这两种方法之间的联系,从而将后者的收敛界转移到前者,更好地解释了观察到的性能。
#4Uniform-in-time rational approximation of the matrix exponential with real poles
矩阵指数实极点有理逼近的一致时间估计
Stefan Güttel, Shuai Shao · 2026-07-20T14:53:45Z
Abstract
We propose two new approaches for constructing families of rational functions with shared real poles that nearly uniformly approximate the functions $\exp(-tz)$ for $z\geq 0$ and $t$ in a positive time interval. The first result concerns the case where all real poles coalesce into a single point. With an appropriate choice of a weight function we are able to derive a closed formula for the asymptotically optimal location of such a pole. We then discuss the more general case where all real poles are distinct. Using Zolotarev's construction of certain optimal rational functions, we present a simple algorithm to derive nearly optimal poles efficiently. We analyze the stability of the numerical evaluation of the resulting rational matrix functions in floating-point arithmetic. By controlling the growth of potential ill-conditioning arising from partial fractions, reliable and highly parallelizable exponential propagators are obtained.
摘要
我们提出了两种新方法,用于构造共享实极点的有理函数族,使其在正时间区间内对函数exp(-tz) (z≥0) 进行近乎一致的逼近。第一个结果涉及所有实极点合并为单点的情况。通过适当选择权函数,我们推导出该极点渐近最优位置的封闭公式。然后讨论所有实极点互不相同的一般情况。利用Zolotarev对某些最优有理函数的构造,我们提出了一种简单算法,可高效地导出近乎最优的极点。我们分析了所得有理矩阵函数在浮点运算中数值求值的稳定性。通过控制由部分分式引起的潜在病态增长,获得了可靠且高度可并行的指数传播子。
#5Space-time tensor-product finite element methods for parabolic problems
抛物问题的时空张量积有限元方法
Richard Löscher, Michael Reichelt, Olaf Steinbach · 2026-07-21T13:48:48Z
Abstract
We study space-time Galerkin--Petrov formulations for parabolic evolution problems and their relation to classical implicit time-stepping schemes. Although such schemes are stable in the usual time-stepping sense, their interpretation as space-time operator equations may lead to conditional stability, with constants depending on the relation between temporal and spatial mesh sizes. We revisit this phenomenon for the continuous Galerkin method of Aziz and Monk, which yields the Crank--Nicolson scheme in the lowest-order case, and provide a detailed space-time error analysis for solutions of both high and low regularity. In particular, the space-time framework allows us to analyze the deteriorated behaviour of classical time-stepping methods for nonsmooth initial data. By applying integration by parts in time, we derive an adjoint space-time formulation that incorporates the initial condition in a natural variational way. In the lowest-order case, this formulation leads to a Rannacher-type smoothing of the initial data. The theoretical results are complemented by numerical experiments.
摘要
我们研究了抛物型演化问题的时空伽辽金-彼得罗夫公式及其与经典隐式时间步进方案的关系。尽管这些方案在通常的时间步进意义下是稳定的,但将其解释为时空算子方程可能导致条件稳定性,其常数取决于时间和空间网格尺寸的关系。我们针对Aziz和Monk的连续伽辽金方法重新审视了这一现象,该方法在最低阶情况下产生克兰克-尼科尔森格式,并提供了高和低正则性解的详细时空误差分析。特别是,时空框架使我们能够分析经典时间步进方法对非光滑初值的退化行为。通过时间分部积分,我们推导出一个伴随时空公式,以自然变分方式纳入初始条件。在最低阶情况下,该公式导致对初值的Rannacher型平滑。数值实验补充了理论结果。
#6Fast reconstruction of tensor tomographic X-ray scattering data for real-time applications
面向实时应用的张量断层扫描X射线散射数据快速重建
André M. Antunes, Daniël M. Pelt, K. Joost Batenburg · 2026-07-21T14:12:27Z
Abstract
X-ray scattering tensor tomography reveals nanoscale structural orientation in 3D, but its reliance on slow iterative reconstruction limits real-time use. We introduce an extension of a direct reconstruction approach for parallel-beam geometries that computes algebraic filters approximating multiple iterative updates with a single filtering and back-projection step. By explicitly separating the tomographic projector from a view-dependent mixing operator, the method generalizes across tensor representations and modes of (small-angle) X-ray scattering measurement. Simulated and experimental results show that the approach approximates iterative reconstructions while reducing computation time by over an order of magnitude, realizing a reconstruction of a 53x53x53x28 tensor volume in 1 second on commercially-available hardware. The resulting speed and stability enable high-throughput analysis and open the door to real-time scattering-based imaging.
摘要
X射线散射张量断层扫描可以揭示三维纳米级结构取向,但其依赖缓慢的迭代重建限制了实时应用。我们介绍了一种对平行束几何直接重建方法的扩展,该方法计算代数滤波器,通过单次滤波和反投影步骤近似多次迭代更新。通过明确分离断层扫描投影仪与视角相关的混合算子,该方法可推广至不同张量表示和(小角)X射线散射测量模式。模拟和实验结果表明,该方法近似迭代重建的同时,将计算时间降低了一个数量级以上,在商用硬件上1秒内实现了53x53x53x28张量体积的重建。由此产生的速度和稳定性使高通量分析成为可能,为实时散射成像打开了大门。
#7Finite element exponential integration for rough solutions of nonlinear wave equations. Part I: Dirichlet boundary conditions on polygonal and polyhedral domains
非线性波动方程粗解有限元指数积分。第一部分:多边形和多面体域上的Dirichlet边界条件
Jiachuan Cao, Benjamin Dörich, Marlis Hochbruck, Buyang Li · 2026-07-17T21:18:53Z
Abstract
We study a fully discrete scheme for nonlinear wave equations on general bounded polygonal/polyhedral domains with initial data $(u^0,v^0)\in H^γ(Ω)\times H^{γ-1}(Ω)$, $0<γ\le 1$, subject to the natural compatibility condition associated with the homogeneous Dirichlet boundary condition. The scheme combines an exponential Euler time integrator with a finite element spatial discretization. In contrast to existing low-regularity error analyses, which are mostly based on Fourier spectral discretizations, our approach applies to general bounded domains and finite element spatial discretizations. We prove rigorous error estimates for low-regularity solutions. The analysis is based on a frequency decomposition of the underlying elliptic operator, used solely as an analytical regularization device and not in the actual implementation, which allows low-regularity techniques to be extended beyond the Fourier spectral framework. The results also indicate that higher-order finite element methods remain advantageous in spatial approximation even for solutions of limited Sobolev regularity. Numerical experiments on different domains and with different polynomial degrees confirm the predicted convergence behavior.
摘要
我们研究了在有界多边形/多面体域上具有初始数据(u⁰,v⁰) ∈ H^γ(Ω)×H^{γ-1}(Ω) (0<γ≤1) 的非线性波动方程的全离散格式,该数据满足齐次Dirichlet边界条件的自然相容性条件。该格式结合了指数Euler时间积分器和有限元空间离散。与现有主要基于Fourier谱离散的低正则性误差分析不同,我们的方法适用于一般有界域和有限元空间离散。我们证明了低正则性解的严格误差估计。分析基于底层椭圆算子的频率分解,仅用作解析正则化工具而非实际实现,使得低正则性技术能够推广到Fourier谱框架之外。结果还表明,即使对于Sobolev正则性有限的解,高阶有限元方法在空间逼近中仍然具有优势。在不同域上具有不同多项式次数的数值实验证实了预期的收敛行为。
#8A Deep Second-Order Stochastic Residual Method for Fully Nonlinear Parabolic PDEs
全非线性抛物型偏微分方程的深度二阶随机残差方法
Zhenhua Zhao, Jihao Long · 2026-07-18T09:43:15Z
Abstract
We introduce the Deep Second-Order Stochastic Residual Method (D2SRM) for high-dimensional, Hessian-dependent fully nonlinear parabolic PDEs. A single scalar space--time network generates derivative-consistent approximations of the solution, gradient, and Hessian, which are trained jointly through second-order Brownian one-step residuals and terminal value and gradient penalties. For globally Lipschitz equations with identity diffusion and sufficiently weak Hessian coupling, we establish well-posedness in a Brownian occupation space and develop a population-level convergence theory. Under additional regularity, an a posteriori estimate bounds the squared full-jet occupation error of any admissible candidate by the time step and its population objective. For approximate population minimizers, the error bound separates time discretization, neural approximation, and population suboptimality; when the latter two terms are $O(h)$, the full-jet occupation norm is $O(h^{1/2})$. Experiments on a 100-dimensional manufactured benchmark compare terminal treatments, probe Hessian couplings inside and outside the proved small-gain range, and show decreasing errors as the time step decreases. The code is available at https://github.com/ZZHPKU/D2SRM.
摘要
我们提出了用于高维、依赖于Hessian的全非线性抛物型偏微分方程的深度二阶随机残差方法(D2SRM)。一个单一的标量时空网络生成解、梯度和Hessian的一致导数近似,并通过二阶布朗一步残差以及终值和梯度惩罚进行联合训练。对于具有单位扩散和足够弱Hessian耦合的全局Lipschitz方程,我们在布朗占据空间中建立了适定性,并发展了总体收敛理论。在额外正则性下,后验估计将任何可接受候选的平方全射占据误差以时间步长及其总体目标为界。对于近似总体极小化器,误差界分离了时间离散化、神经逼近和总体次优性;当后两项为$O(h)$时,全射占据范数为$O(h^{1/2})$。在100维人造基准上的实验比较了终端处理,探测了证明的小增益范围内外的Hessian耦合,并显示了随着时间步长减小误差递减。代码见https://github.com/ZZHPKU/D2SRM。
#9Neural operator preconditioning from mixed dataset for the Helmholtz equations: Application to transcranial ultrasound
基于混合数据集Helmholtz方程的神经算子预处理及其在经颅超声中的应用
Yanfei Xiang · 2026-07-18T17:48:59Z
Abstract
This work develops a neural operator preconditioned subspace method for sequences of linear systems arising from the discretization of the two-dimensional Helmholtz equation in transcranial ultrasound applications. The problem involves strongly heterogeneous, patient-dependent velocity fields that induce severe wave distortion and pose significant challenges for standard iterative solvers. Building on neural network preconditioning framework of Giraud et al. (HAL RR-9593, 2025) and the idealized skull dataset used for the learned optimizer of Stanziola et al. (JCP 441, 2021), neural operator preconditioners are trained on six mixed velocity-source datasets combining randomized source configurations and idealized skull-based velocity fields with random noise. The proposed mixed-dataset strategy aims to improve both computational efficiency and generalization across varying configurations. The neural operator is trained on a coarse grid using a physics-informed loss based on the relative residual of the discrete Helmholtz equation and is incorporated as a nonlinear preconditioner within flexible GMRES (FGMRES). Numerical experiments demonstrate that the resulting hybrid method efficiently solves practical transcranial ultrasound problems on grids 64 times larger than those used during training, whereas both classical GMRES and the learned optimizer fail to converge within comparable computational budgets. Moreover, the proposed method achieves arbitrary solution accuracies and exhibits strong out-of-distribution generalization across diverse source and velocity configurations. This work highlights the importance of dataset design in scientific machine learning and provides a practical framework for integrating matrix-free neural operator preconditioning with Krylov subspace methods for solving practical large-scale Helmholtz problems.
摘要
本文针对经颅超声应用中二维Helmholtz方程离散化产生的序列线性系统,开发了一种神经算子预处理子空间方法。该问题涉及强异质性、患者依赖的速度场,会引起严重的波畸变,对标准迭代求解器构成重大挑战。基于Giraud等人(HAL RR-9593, 2025)的神经网络预处理框架和Stanziola等人(JCP 441, 2021)用于学习优化器的理想化颅骨数据集,神经算子预处理器在六个混合速度-源数据集上训练,这些数据集结合了随机源配置和带有随机噪声的理想化颅骨速度场。提出的混合数据集策略旨在提高计算效率和在不同配置下的泛化能力。神经算子基于离散Helmholtz方程相对残差的物理信息损失在粗网格上训练,并作为非线性预处理器集成到柔性GMRES(FGMRES)中。数值实验表明,所提出的混合方法在比训练时大64倍的网格上有效求解实际经颅超声问题,而经典GMRES和学习优化器在可比计算预算内均无法收敛。此外,所提方法实现了任意求解精度,并在多样化的源和速度配置下表现出强大的分布外泛化能力。这项工作凸显了科学机器学习中数据集设计的重要性,并为将无矩阵神经算子预处理与Krylov子空间方法相结合求解实际大规模Helmholtz问题提供了实用框架。
#10Novel Adaptive Methods for Hyperbolic Conservation Laws Based on New Quasi-Linear Seventh- and Ninth-Order Schemes
基于新型拟线性七阶和九阶格式的双曲守恒律自适应方法
Shaoshuai Chu, Pingyao Feng, Vadim A. Kolotilov, Alexander Kurganov, Vladimir V. Ostapenko · 2026-07-18T18:18:12Z
Abstract
We develop new adaptive numerical schemes for one- and two-dimensional hyperbolic systems of conservation laws. The methodology relies on the use of a smoothness indicator to automatically partition the computational domain into smooth and nonsmooth (``rough``) regions. We then follow the scheme adaption strategy recently introduced in [S. Chu, P. Feng, V. A. Kolotilov, A. Kurganov, and V. V. Ostapenko, Commun. Comput. Phys., accepted], but instead of the quasi-linear (QL) fifth-order finite-difference scheme used there, we employ the new QL seventh- and ninth-order schemes in the smooth regions. A series of numerical experiments for the Euler equations of gas dynamics demonstrates that the new adaptive schemes contain a smaller amount of numerical dissipation and achieve higher resolution compared with their counterpart that uses the QL fifth-order scheme in the smooth areas.
摘要
我们为一维和二维双曲守恒律系统开发了新的自适应数值格式。该方法依赖于使用光滑性指示器自动将计算域划分为光滑和非光滑(“粗糙”)区域。然后我们遵循最近在[S. Chu, P. Feng, V. A. Kolotilov, A. Kurganov, and V. V. Ostapenko, Commun. Comput. Phys., accepted]中引入的格式自适应策略,但在光滑区域中不是使用那里的拟线性五阶有限差分格式,而是采用新的拟线性七阶和九阶格式。针对气体动力学欧拉方程的一系列数值实验表明,新的自适应格式包含更少的数值耗散,并与在光滑区域使用拟线性五阶格式的对应格式相比实现了更高的分辨率。
#11Convergence of Finite Element Methods for Ricci Flow
Ricci流有限元方法的收敛性
Guangwei Gao, Evan S. Gawlik, Buyang Li · 2026-07-19T03:33:00Z
Abstract
The convergence of a finite element discretization for the two-dimensional Ricci flow is proved. In this method, the Ricci flow on a two-dimensional surface is formulated into solution-driven metric evolution, with the metric evolution driven by the Gauss curvature. The Gauss curvature satisfies a parabolic equation that in turn depends on the metric, thereby enhancing the parabolic structure of the problem. The solution-driven metric evolution formulation is discretized by the finite element method, and the convergence of finite element approximations is proved by adapting the matrix-vector formulation developed in the literature initially for studying solution-driven surface evolution in extrinsic curvature flow. In addition to its convergence, the proposed method also preserves important geometric structures of the Ricci flow at the discrete level, such as area conservation and the Gauss-Bonnet theorem. Extensive numerical experiments are presented to demonstrate the convergence of the proposed method as well as the simulation of Ricci flow.
摘要
证明了二维Ricci流的一种有限元离散的收敛性。在该方法中,二维曲面上的Ricci流被表述为解驱动度量演化,其中度量演化由高斯曲率驱动。高斯曲率满足一个抛物型方程,该方程又依赖于度量,从而增强了问题的抛物结构。解驱动度量演化公式通过有限元方法离散,并通过改编文献中最初用于研究外蕴曲率流中解驱动曲面演化的矩阵-向量公式,证明了有限元近似的收敛性。除了收敛性外,所提出的方法还在离散层面上保留了Ricci流的重要几何结构,如面积守恒和Gauss-Bonnet定理。给出了广泛的数值实验,以证明所提出方法的收敛性以及Ricci流的模拟。
#12A Lowest-Order Robust Mixed Finite Element Method with a Third-Order Tensor Variable for Strain Gradient Elasticity
应变梯度弹性最低阶鲁棒混合有限元方法(具有三阶张量变量)
Xuehai Huang, Zheqian Tang · 2026-07-19T08:38:05Z
Abstract
A lowest-order mixed finite element method is developed for the strain gradient elasticity (SGE) model in arbitrary dimensions. We take the physically meaningful third-order double stress tensor $\boldsymbolΦ:=ι^2\operatorname{grad}\boldsymbolσ(\boldsymbol{u})\in\mathbb{S}\otimes\mathbb{R}^d$ as a primary variable and derive a distributional mixed formulation. The double stress is approximated by an $\mathbb{S}\otimes\mathbb{R}^d$-valued extension of the lowest-order Raviart--Thomas element, while the displacement is approximated by the vector-valued linear Crouzeix--Raviart element. Thus, the method avoids both high-degree bubble enrichment and a Nitsche-type treatment of the higher-order boundary condition. We establish parameter-robust discrete stability, an optimal first-order error estimate for fixed parameters, and a complementary parameter-uniform $\mathcal{O}(ι^{1/2}+h)$ error estimate with constants independent of both the size parameter $ι$ and the Lamé coefficient $λ$. In the boundary-layer regime $ι^{1/2}\lesssim h$, the latter retains a first-order convergence rate in $h$. We also develop a local quadratic post-processing and a hybridized formulation. Numerical experiments in two and three dimensions support the theoretical results.
摘要
针对任意维度的应变梯度弹性(SGE)模型,开发了一种最低阶混合有限元方法。我们将物理上有意义的三阶双应力张量Φ:=ι²grad σ(u)∈S⊗R^d作为原始变量,并推导了分布混合公式。双应力通过最低阶Raviart-Thomas单元的S⊗R^d值扩展来近似,而位移通过向量值线性Crouzeix-Raviart单元近似。因此,该方法避免了高阶泡状富集和Nitsche式处理高阶边界条件。我们建立了参数鲁棒的离散稳定性、固定参数下的最优一阶误差估计,以及互补的参数一致O(ι^(1/2)+h)误差估计,常数与尺寸参数ι和Lamé系数λ无关。在边界层区域ι^(1/2)≲h时,后者保持了h的一阶收敛率。我们还开发了局部二次后处理和混合化公式。二维和三维数值实验支持了理论结果。
#13A Rational Discrete Collocation Method for Second Kind Fredholm Equations
第二类弗雷德霍姆方程的有理离散配置方法
Domenico Mezzanotte, Donatella Occorsio, Mario Pezzella, Woula Themistoclakis · 2026-07-19T10:24:38Z
Abstract
In this work we present a novel discrete collocation method for the numerical solution of Fredholm integral equations of the second kind in the space of continuous functions equipped with the uniform norm. The method is based on a rational interpolation scheme recently developed within the general framework of reproducing kernel Hilbert spaces. This rational approximation has no real poles, interpolates the target function at arbitrary Jacobi nodes and exhibits uniformly bounded Lebesgue constants. Moreover, it converges uniformly for all continuous functions at a rate at least equal to that of the best uniform polynomial approximation. These interesting properties are inherited by the resulting numerical method, for which stability, convergence and good conditioning are established under minimal assumptions on the integral kernel. A series of numerical experiments confirm the theoretical findings and indicate that, in the presence of particularly challenging kernels, the proposed approach provides a robust and effective alternative to Nyström-type methods.
摘要
本文提出了一种新的离散配置方法,用于在配备一致范数的连续函数空间中数值求解第二类弗雷德霍姆积分方程。该方法基于最近在再生核希尔伯特空间框架内发展的一种有理插值方案。这种有理近似没有实极点,在任意雅可比节点上插值目标函数,并具有一致有界的勒贝格常数。此外,它对所有连续函数一致收敛,收敛速度至少与最佳一致多项式逼近相同。这些有趣的性质被继承到所得到的数值方法中,在积分核的最小假设下,建立了稳定性、收敛性和良好的条件数。一系列数值实验证实了理论结果,并表明在存在特别具有挑战性的核时,所提出的方法为Nyström型方法提供了一种稳健有效的替代方案。
#14Raviart--Thomas Elements with Geometric Correction for Distorted Quadrilateral Meshes
具有几何校正的Raviart-Thomas单元用于畸变四边形网格
So-Hsiang Chou · 2026-07-19T15:07:46Z
Abstract
Mixed finite element methods based on Raviart--Thomas spaces are widely used for the numerical approximation of second--order elliptic problems in flux form. On quadrilateral meshes, however, the bilinear mapping from the reference element introduces a spatially varying Jacobian, which may violate the inclusion property $\mathrm{div}\,V_h \subset W_h$ for the standard Raviart--Thomas spaces. In this paper we propose a simple modification of the classical Raviart--Thomas elements on quadrilateral meshes. The modification consists of adding geometrically motivated correction terms to the local basis functions in order to compensate for the geometric distortion introduced by the bilinear mapping. The resulting spaces retain the same dimension and degrees of freedom as the classical Raviart--Thomas elements while restoring the compatibility property. We present a general framework for constructing such modified spaces and illustrate the approach by developing modified versions of the lowest order and next--to--lowest order Raviart--Thomas elements. Theoretical analysis establishes optimal approximation properties under the standard shape--regularity assumption for quadrilateral meshes. Numerical experiments on distorted meshes confirm the predicted convergence rates and show that the modified elements yield consistently improved accuracy over the classical Raviart--Thomas formulation as the geometric distortion increases.
摘要
基于Raviart-Thomas空间的混合有限元方法广泛用于通量形式的二阶椭圆问题的数值逼近。然而,在四边形网格上,从参考单元到物理单元的双线性映射引入了空间变化的雅可比,这可能违反标准Raviart-Thomas空间的包含性质$\mathrm{div}\,V_h \subset W_h$。本文提出了四边形网格上经典Raviart-Thomas单元的一个简单修改。修改包括在局部基函数中添加几何动机的校正项,以补偿双线性映射引入的几何畸变。所得空间保留了与经典Raviart-Thomas单元相同的维数和自由度,同时恢复了兼容性。我们提出了构建此类修改空间的一般框架,并通过开发最低阶和次低阶Raviart-Thomas单元的修改版本来说明该方法。理论分析在四边形网格的标准形状正则性假设下建立了最优逼近性质。在畸变网格上的数值实验证实了预期的收敛速度,并表明随着几何畸变的增加,修改后的单元比经典Raviart-Thomas公式始终具有更高的精度。
#15A posteriori error estimates for parabolic PDEs on evolving surfaces
演化曲面上抛物型偏微分方程的后验误差估计
Michael Lantelme · 2026-07-19T21:39:38Z
Abstract
We derive residual-based a posteriori error estimates for parabolic surface PDEs on closed evolving surfaces. The main contribution is to prove efficiency and reliability for the proposed error indicator, which bounds the error quantities globally from above and globally in space and locally in time from below. We extend methods for adaptivity for parabolic PDEs on stationary surfaces to allow for non-trivial coarsening on evolving surfaces. Multiple numerical experiments are given, which illustrate the asymptotic behaviour of the error and effectiveness of the refinement and coarsening.
摘要
我们推导了封闭演化曲面上抛物型曲面偏微分方程的基于残差的后验误差估计。主要贡献是证明了所提出的误差指示器的有效性和可靠性,该指示器从全局上界误差量,并在空间上全局、时间上局部地从下界误差量。我们将适用于稳态曲面上抛物型PDE的自适应方法扩展到允许在演化曲面上进行非平凡粗化。给出了多个数值实验,说明了误差的渐近行为以及细化和粗化的有效性。
#16Finite element exponential integration for rough solutions of nonlinear wave equations. Part II: Dynamic boundary conditions on curved domains
非线性波动方程粗糙解的有限元指数积分方法。第二部分:弯曲域上的动态边界条件
Jiachuan Cao, Benjamin Dörich, Buyang Li · 2026-07-17T21:40:56Z
Abstract
We study nonlinear wave equations with dynamic boundary conditions on smooth bounded domains and analyze a fully discrete approximation in the low-regularity regime. The method combines isoparametric bulk--surface finite elements of degree $k$ with an exponential integrator in time. Assuming only bounded energy of the exact solution, we prove convergence of the displacement--velocity pair in the weak norm $L^2(Ω;Γ)\times H^{-1}(Ω;Γ)$. The scheme achieves first-order convergence in time and spatial convergence of order $h^{2/3}$ for $k=1$ and $h^{(k+2)/(k+3)}$ for $k\ge 2$. In particular, these rates show that higher-order finite elements retain a provable asymptotic advantage even at low regularity. A central difficulty is that the continuous and discrete bulk--surface problems are posed on different geometries and must therefore be compared directly in weak norms. To address this, we develop a weak-norm framework for non-conforming geometries based on lift and adjoint-lift operators, combined with a frequency-decomposition argument. To the best of our knowledge, this is the first fully discrete low-regularity convergence result for nonlinear wave equations with dynamic boundary conditions in a non-conforming bulk--surface finite element setting. Numerical experiments confirm the predicted rates and illustrate the improved efficiency of higher-order methods.
摘要
我们研究了光滑有界域上具有动态边界条件的非线性波动方程,并分析了低正则性区域的全离散逼近。该方法将$k$次等参体-曲面有限元与时间指数积分器结合。假设精确解仅有有界能量,我们证明了位移-速度对在弱范数$L^2(Ω;Γ)\times H^{-1}(Ω;Γ)$中的收敛性。该方案在时间上达到一阶收敛,空间收敛阶对于$k=1$为$h^{2/3}$,对于$k\ge 2$为$h^{(k+2)/(k+3)}$。特别地,这些收敛速率表明,即使在低正则性下,高阶有限元也保留了可证明的渐近优势。一个核心困难是连续和离散的体-曲面问题定义在不同的几何上,因此必须直接在弱范数下进行比较。为解决这一问题,我们基于提升和伴随提升算子,结合频率分解论证,发展了一个非协调几何的弱范数框架。据我们所知,这是第一个针对非协调体-曲面有限元设置中具有动态边界条件的非线性波动方程的全离散低正则性收敛结果。数值实验验证了预测的收敛速率,并说明了高阶方法改进的效率。
#17Weighted Inverse Lax-Wendroff Boundary Treatment of Discontinuous Galerkin Methods for Conservation Laws
守恒律间断伽辽金方法的加权逆Lax-Wendroff边界处理
Yongjie Bi, Yan Jiang, Yong Liu · 2026-07-21T16:32:22Z
Abstract
In this paper, we propose a weighted inverse Lax-Wendroff (WILW) boundary treatment for the discontinuous Galerkin (DG) method on unfitted meshes to efficiently solve hyperbolic conservation laws in complex geometries. The proposed method employs the standard DG scheme for interior cells and reconstructs high-order approximation polynomials via the ILW principle for cut cells near boundaries to impose numerical boundary conditions, effectively eliminating the time-step restriction typically caused by small cut cells. In particular, to address the sensitivity of numerical errors to the geometric size of cut cells in the basic ILW scheme, we raise the reconstruction order at the boundary, ensuring that accuracy becomes independent of the cut-cell size. Furthermore, it incorporates a weighted least-squares reconstruction to reduce the need for complex high-order boundary derivatives during construction. As a result, the method maintains high-order accuracy while significantly improving computational efficiency for multi-dimensional problems. Finally, the stability of the proposed method is theoretically validated through linear stability analysis, and the effectiveness and robustness of the proposed scheme are numerically verified through a series of one-dimensional and two-dimensional numerical experiments for scalar and system equations.
摘要
本文提出了一种适用于非拟合网格上间断伽辽金(DG)方法的加权逆Lax-Wendroff(WILW)边界处理,以高效求解复杂几何中的双曲守恒律。该方法在内部单元使用标准DG格式,并在边界附近切割单元通过ILW原理重构高阶逼近多项式以施加数值边界条件,有效消除通常由小切割单元导致的时间步长限制。特别地,针对基本ILW方案中数值误差对切割单元几何尺寸的敏感性,我们提高了边界处的重构阶数,确保精度独立于切割单元尺寸。此外,它融入了加权最小二乘重构,减少了构造过程中对复杂高阶边界导数的需求。因此,该方法在保持高阶精度的同时,显著提高了多维问题的计算效率。最后,通过线性稳定性分析从理论上验证了所提方法的稳定性,并通过一系列标量和系统方程的一维及二维数值实验验证了方案的有效性和鲁棒性。
#18Nyström Error Beyond $M$-Matrices: A Minimal Diagonally Dominant Obstruction
$M$-矩阵之外的Nyström误差:最小对角占优障碍
Matthew J. Colbrook · 2026-07-21T16:53:51Z
Abstract
We study the nuclear-norm error of a column-selected Nyström approximation to $K=(L+γI)^{-1}$, where $L$ is symmetric diagonally dominant and $γ>0$. Our central question is whether this error has diminishing returns. A Schur-complement identity reduces the question to traces of inverses of principal submatrices. Existing $M$-matrix results settle the case in which $L$ is a symmetric diagonally dominant $M$-matrix (SDDM). However, diagonal dominance alone is not enough: failure occurs already in dimension three. We construct an exact one-parameter SDD family and determine its sharp failure interval. A $2\times2$ identity proves that dimension three is minimal within the SDD class. We then show that failure persists under strict diagonal dominance; with a nonempty selected base set, dimension four is minimal. Finally, we prove invariance under signature switching, derive a three-dimensional formula showing how a signed triangle causes failure, and give an example in which greedy column selection misses the optimal pair. Together, these findings complete the answer to Problem 4.6 in a recent Simons workshop report.
摘要
我们研究了列选择Nyström近似$K=(L+γI)^{-1}$的核范数误差,其中$L$是对角占优对称矩阵,$γ>0$。核心问题是该误差是否具有递减效应。通过Schur补恒等式,问题归结为主子矩阵逆的迹。现有的$M$-矩阵结果解决了$L$为对称对角占优$M$-矩阵(SDDM)的情况。然而,仅对角占优不够:在三维中已经出现失败。我们构造了一个精确的单参数SDD族,并确定了其尖锐失败区间。$2\times2$恒等式证明三维在SDD类中是最小的。我们进一步证明在严格对角占优下失败仍然存在;在非空选定基集下,四维是最小的。最后,我们证明了符号切换下的不变性,推导了一个三维公式展示符号三角形如何导致失败,并给出了一个贪心列选择错过最优对的例子。这些发现共同完善了对近期Simons研讨会报告中问题4.6的回答。
#19Analytic regularity for a fourth-order singularly perturbed boundary balue problem with two small parameters
带有两个小参数的四阶奇异摄动边值问题的解析正则性
I. Sykopetritou, C. Xenophontos · 2026-07-18T11:32:04Z
Abstract
We consider a fourth order singularly perturbed boundary value problem with two small parameters, in one dimension, under the assumption of analytic input data. We show that the solution may be decomposed into a smooth part, two different width boundary layers, and a negligible remainder. We provide estimates for arbitrary order derivatives of each term of the decomposition, which are explicit in the differentiation order and the singular perturbation parameters, and are needed for proving the convergence of high order numerical methods, such as the $p/hp$ versions of the Finite Element Method. We also provide classical differentiability results, which show that the solution will be analytic, if the data are analytic, but negative powers of the singular perturbation parameter(s) show up once we start differentiating.
摘要
我们考虑一个一维四阶奇异摄动边值问题,带有两个小参数,并假设解析输入数据。我们证明解可以分解为一个光滑部分、两个不同宽度的边界层和一个可忽略的余项。我们为分解的每一项提供了任意阶导数的估计,这些估计在微分阶数和奇异摄动参数上是显式的,并且是证明高阶数值方法(如有限元方法的 $p/hp$ 版本)收敛性所必需的。我们还提供了经典的可微性结果,表明如果数据是解析的,解将是解析的,但一旦开始求导,就会出现奇异摄动参数的负幂次。
#20Stability and Robustness Analysis of Regularized Reconstruction Methods for Low-Dose Computed Tomography in Parallel-Beam Geometry
平行束几何低剂量计算机断层扫描正则化重建方法的稳定性和鲁棒性分析
Mohamed Berrada · 2026-07-19T15:30:29Z
Abstract
Low-dose computed tomography (LDCT) reduces radiation exposure but increases the ill-posedness of the reconstruction problem due to noise and sparse data. While regularized methods like Tikhonov and Total Variation (TV) improve image quality, their performance depends heavily on noise characteristics, sampling conditions, and parameter selection. This study presents a systematic stability and robustness analysis of Filtered Back Projection (FBP), Tikhonov regularization, and TV minimization within a 2D parallel-beam CT framework. A unified simulation pipeline based on the Radon transform is developed and evaluated using both the modified Shepp-Logan phantom and a clinical thorax image. Reconstruction behavior is investigated under multiple degradation scenarios involving Gaussian, Poisson, and mixed noise models, across baseline (180 projections) and sparse-view (60 projections) acquisition geometries. To ensure a fair comparison, regularization parameters are optimized for each scenario through an exhaustive SSIM-based grid-search. Quality is assessed via RMSE, PSNR, and SSIM, while robustness is quantified through an empirical Stability Factor S measuring perturbation amplification from measurement to image space. The results show that FBP is highly sensitive to noise and undersampling. Tikhonov regularization improves structural fidelity compared with FBP but remains more sensitive to perturbation than TV. Conversely, TV provides the best compromise between noise suppression, edge preservation, accuracy, and numerical stability. These findings highlight the stability-resolution trade-off in LDCT and demonstrate that the proposed Stability Factor S offers valuable complementary information to conventional metrics.
摘要
低剂量计算机断层扫描(LDCT)减少了辐射暴露,但由于噪声和稀疏数据而增加了重建问题的不适定性。虽然Tikhonov和全变分(TV)等正则化方法改善了图像质量,但其性能很大程度上取决于噪声特性、采样条件和参数选择。本研究在二维平行束CT框架内,对滤波反投影(FBP)、Tikhonov正则化和TV最小化进行了系统的稳定性和鲁棒性分析。基于Radon变换开发了统一的模拟流程,并使用修改的Shepp-Logan体模和临床胸部图像进行评估。在涉及高斯、泊松和混合噪声模型的多种退化场景下,在基线(180个投影)和稀疏视角(60个投影)采集几何中研究了重建行为。为了公平比较,通过基于SSIM的穷举网格搜索为每个场景优化了正则化参数。质量通过RMSE、PSNR和SSIM评估,而鲁棒性通过经验稳定性因子S来量化,该因子测量从测量空间到图像空间的扰动放大。结果表明FBP对噪声和欠采样高度敏感。与FBP相比,Tikhonov正则化改善了结构保真度,但对扰动比TV更敏感。相反,TV在噪声抑制、边缘保持、准确性和数值稳定性之间提供了最佳平衡。这些发现突显了LDCT中的稳定性-分辨率权衡,并表明所提出的稳定性因子S为传统指标提供了有价值的补充信息。
#21Iterated graph Laplacian for image restoration problems
迭代图拉普拉斯算子用于图像复原问题
Stefano Aleotti, Davide Bianchi, Florian Bossmann, Marco Donatelli, Pietro Maurino · 2026-07-19T15:59:30Z
Abstract
We study the graph Laplacian operator as a regularizer in a generalized Tikhonov framework for linear ill-posed problems. The Laplacian is updated iteratively from the current reconstruction, so that progressively sharper structural information about the solution is fed into the regularization term. We introduce three schemes: a standard one that rebuilds the Laplacian from each new iterate; an error-equation scheme that, following the error-based formulation of iterated Tikhonov regularization, builds the Laplacian from an estimate of the reconstruction error rather than of the image itself; and a mixed scheme combining the two. We establish convergence of all three schemes for noisy data under a priori parameter and stopping rules as the noise level tends to zero. Numerical experiments in two-dimensional computed tomography and image deblurring show consistent gains in reconstruction quality and sharper recovery of fine details.
摘要
我们研究了图拉普拉斯算子作为广义Tikhonov框架中的正则化子用于线性不适定问题。拉普拉斯算子从当前重构中迭代更新,从而将逐渐清晰的解的结构信息馈入正则化项。我们引入了三种方案:一种标准方案,从每个新迭代中重建拉普拉斯算子;一种误差方程方案,遵循迭代Tikhonov正则化的基于误差的公式,从重构误差的估计而不是图像本身构建拉普拉斯算子;以及一种结合两者的混合方案。我们建立了所有三种方案在噪声数据下在先验参数和停止规则下的收敛性,当噪声水平趋于零时。在二维计算机断层扫描和图像去模糊中的数值实验显示了一致的重构质量改进和更清晰的细节恢复。
#22FEVessel: Mesh-Independent Analysis of 3D Pressure Vessels with the Label-Free Pretrained Finite Element Method
FEVessel:基于无标签预训练有限元法的3D压力容器网格无关分析
Yipin Sun, Yizheng Wang, Yuzhou Lin, Baiyang Zheng, Xiaoying Zhuang, Timon Rabczuk · 2026-07-19T16:01:34Z
Abstract
Pressure vessel analysis in the chemical, nuclear, and new-energy industries requires solving the same elasticity problem across many materials, geometries, and loads, where mesh quality and repeated solving govern both accuracy and cost. The finite element method (FEM) cannot amortise this repeated cost and fails on degenerate meshes, while the neural operators meant to replace it still need labelled data that FEM must generate. This paper proposes FEVessel, an adaptation of the Pretrained Finite Element Method (PFEM) to three-dimensional (3D) pressure vessels, and validates four capabilities across the two limitations above. FEVessel i) encodes each vessel as a point cloud with coordinate, material, and load channels, ii) pretrains a Transolver operator on the total potential energy instead of FEM labels, and iii) warm-starts iterative solvers with its prediction. A single model generalises across material, geometry, and boundary conditions at a $1.35\%$ relative displacement error, and its $2.07\%$ strain error is about $4.7$ times lower than that of a supervised Fourier neural operator ($9.72\%$), whose structured grid cannot preserve the through-thickness strain. Its warm start cuts algebraic multigrid iterations from $195$ to $18$, a $9.2\times$ end-to-end wall-clock speedup at the $10^{-3}$ engineering tolerance. The model transfers across mesh resolutions without retraining, holding about $3\%$ error at only $30\%$ of the training point density. On inverted and sliver meshes where FEM fails, the error remains below $3.66\%$. To our knowledge, this is the first systematic study of mesh-independent solution on industrially relevant 3D pressure vessels with degenerate meshes. Because training needs no labels, FEVessel works exactly where FEM cannot supply any, removing manual mesh repair from the analysis pipeline.
摘要
化工、核能及新能源行业的压力容器分析需要在多种材料、几何形状和载荷下求解相同的弹性问题,其中网格质量和重复求解同时决定了准确性和成本。有限元法(FEM)无法摊薄这种重复成本,并且在退化网格上失败;而旨在替代它的神经算子仍然需要FEM必须生成的标记数据。本文提出FEVessel,将预训练有限元法(PFEM)适应于三维压力容器,并针对上述两个局限验证了四种能力。FEVessel i) 将每个容器编码为一个包含坐标、材料和载荷通道的点云,ii) 在总势能上预训练一个Transolver算子,而不是FEM标签,以及 iii) 用其预测热启动迭代求解器。单个模型在材料、几何和边界条件上泛化,相对位移误差为1.35%,其2.07%的应变误差比监督傅里叶神经算子(9.72%)低约4.7倍,后者的结构化网格无法保持厚度方向的应变。其热启动将代数多重网格迭代从195次减少到18次,在10^{-3}工程容差下实现了9.2倍的端到端加速。模型无需重新训练即可跨网格分辨率迁移,在仅30%训练点密度下保持约3%的误差。在FEM失败的倒置和狭长网格上,误差仍低于3.66%。据我们所知,这是首次在工业相关三维压力容器上对具有退化网格的网格无关解决方案的系统研究。由于训练不需要标签,FEVessel恰好能在FEM无法提供任何标签的地方工作,从而从分析流程中消除了手动网格修复。
#23A domain decomposition online-learning-enhanced nonlinear elimination preconditioner
一种区域分解在线学习增强的非线性消去预条件子
Pai Zhang, Linyan Gu, Li Luo · 2026-07-20T09:14:24Z
Abstract
Nonlinearly preconditioned inexact Newton methods form an effective class of solvers for large-scale nonlinear algebraic systems arising from the discretization of partial differential equations. A central challenge in nonlinear elimination (NE) preconditioning is the reliable identification of the slowly converging components to be eliminated. Existing selection strategies often rely on problem-specific physical information or user-tuned thresholds applied directly to the raw nonlinear residual, which may contain irregular oscillatory structures near stagnation regions, making the selected bad subset highly sensitive to threshold parameters. In this work, we propose an online-learning-enhanced NE preconditioner that identifies the bad subset from the dominant structure of the nonlinear residual rather than from the raw residual itself. Residual snapshots are collected online during the stagnation phase of the current Newton solve, and an unsupervised extraction model is trained to capture the principal nonlinear imbalance. We consider both a linear extractor based on principal component analysis and nonlinear extractors based on autoencoder neural networks. Moreover, we integrate the approach into a parallel domain decomposition framework, which trains a local extraction model independently on each subdomain. The learned residual reconstruction is then used to define the bad subset and guide the nonlinear elimination process. Numerical experiments on lid-driven cavity flows at Reynolds numbers up to 10,000 show that the proposed method produces more reliable and coherent bad subsets, is robust with respect to both NE and learning parameters, and outperforms the baseline NE preconditioner in terms of the convergence.
摘要
非线性预条件非精确牛顿方法是求解偏微分方程离散化产生的大规模非线性代数系统的一类有效求解器。非线性消去(NE)预条件中的一个核心挑战是可靠地识别要消去的慢收敛分量。现有的选择策略通常依赖于特定问题的物理信息或直接应用于原始非线性残差的用户调整阈值,而原始非线性残差在停滞区域附近可能包含不规则的振荡结构,使得所选不良子集对阈值参数高度敏感。本文中,我们提出了一种在线学习增强的NE预条件子,它从非线性残差的主导结构而不是原始残差本身识别不良子集。在当前牛顿求解的停滞阶段在线收集残差快照,并训练无监督提取模型以捕捉主要非线性不平衡。我们考虑了基于主成分分析的线性提取器和基于自编码器神经网络的非线性提取器。此外,我们将该方法集成到并行区域分解框架中,该框架在每个子域上独立训练局部提取模型。然后使用学习到的残差重构来定义不良子集并指导非线性消去过程。在雷诺数高达10,000的方腔驱动流上的数值实验表明,所提出的方法产生了更可靠和一致的不良子集,对NE和学习参数均具有鲁棒性,并且在收敛性方面优于基线NE预条件子。
#24Cubature from rational approximation
来自有理逼近的数值积分
Gentian Zavalani · 2026-07-20T11:46:25Z
Abstract
We present a numerical construction of cubature rules for area integrals of analytic functions over planar domains with rectifiable Jordan boundary. The starting point is the Cauchy--Green identity. Given a weight $w$, we choose a $\bar\partial$-antiderivative $W$ and reduce the area integral to a contour integral involving the boundary values of $W$. These values are then approximated by a rational function with free poles, computed by the AAA algorithm. The poles inside the domain become cubature nodes, the corresponding residues become weights, and the boundary residual controls the error through an a posteriori estimate, rigorous once the continuous boundary residual is bounded. The same rule admits a dual reading, as the exact integral of a rational interpolant to the integrand, the area analogue of the one-dimensional interpolatory viewpoint. The numerical examples recover the disk mean-value rule and the focal-segment rule of the ellipse to machine precision, reproduce the exact finite quadrature identities of quadrature domains with both separated and confluent nodes, and evaluate logarithmic and Cauchy volume potentials from boundary data alone. The interior poles trace analytic skeletons that we identify tentatively with the mother bodies of potential theory, along with image points that appear without being imposed; for the square the observed convergence is root-exponential.
摘要
我们提出了一个数值构造积分公式的方法,用于计算具有可求长Jordan边界的平面域上解析函数的面积极分。出发点为Cauchy-Green恒等式。给定权重w,我们选择一个$\bar\partial$-原函数W,并将面积分简化为涉及W边界值的围道积分。然后通过AAA算法计算具有自由极点的有理函数来近似这些值。域内的极点成为积分节点,相应的留数成为权重,边界残差通过后验估计控制误差,一旦连续边界残差有界,该估计就是严格的。同一规则具有双重解读,即被积函数的有理插值函数的精确积分,这是一维插值观点的面积模拟。数值示例恢复了圆盘均值规则和椭圆的焦点线段规则,达到机器精度;再现了具有分离和汇合节点的求积域的精确有限求积恒等式;并仅从边界数据评估了对数和Cauchy体积势。内部极点追踪了解析骨架,我们初步将其识别为势论中的母体,以及无需施加就出现的像点;对于正方形,观察到的收敛是根指数级的。
#25Hybrid-Dimensional Biot Problem with an Optimization Based Domain Decomposition Approach
基于优化区域分解的混合维Biot问题
Francesca Marcon, Stefano Scialò · 2026-07-20T12:36:49Z
Abstract
The present work proposes a numerical approach for solving coupled flow and mechanics problems in fractured porous media, represented as mixed-dimensional domains. In this formulation, the elements of the 3D mesh are allowed to arbitrarily intersect the fractures. Displacements are discontinuous across fractures through the use of the eXtended Finite Element Method (XFEM) on the 3D mesh. The mechanical problem is formulated as a saddle-point problem, in which Lagrange multipliers are used to enforce displacement continuity across the fractures. The resulting Lagrange multipliers represent the stress field acting on the fracture surfaces. Likewise, the pressure field is allowed to be discontinuous across fractures through the XFEM formulation on the non-conforming mesh and is computed using an optimization-based domain decomposition strategy specifically designed for mixed-dimensional problems. The fixed-stress splitting scheme is employed to decouple the flow and mechanics subproblems, while the mixed-dimensional pressure problem is solved at each fixed-stress iteration using the Conjugate Gradient (CG) method. The combination of the fixed-stress scheme and the CG solver proves to be highly effective for this class of problems.
摘要
本文提出了一种数值方法,用于求解裂隙多孔介质中耦合流动与力学问题,该介质被表示为混合维区域。在此公式中,3D网格单元允许任意穿过裂隙。通过使用扩展有限元方法(XFEM)在3D网格上,位移在裂隙处是不连续的。力学问题被表述为鞍点问题,其中使用拉格朗日乘子来强制裂隙间的位移连续性。所得的拉格朗日乘子表示作用在裂隙面上的应力场。同样,通过非协调网格上的XFEM公式,压力场在裂隙处允许不连续,并使用专门为混合维问题设计的基于优化的区域分解策略进行计算。采用固定应力分裂方案来解耦流动和力学子问题,而每个固定应力迭代中混合维压力问题使用共轭梯度(CG)方法求解。固定应力方案与CG求解器的组合被证明对此类问题非常有效。
#26Adaptive Mamba Neural Operators
自适应Mamba神经算子
Zeyuan Song, Zheyu Jiang · 2026-07-20T15:12:09Z
Abstract
Accurately solving partial differential equations (PDEs) on arbitrary geometries and a variety of meshes is an important task in science and engineering applications. In this paper, we propose Adaptive Mamba Neural Operators (AMO), which integrates reproducing kernels for state-space models (SSMs) rather than the kernel integral formulation of SSMs. This is achieved by constructing Takenaka-Malmquist systems for the PDEs. AMO offers new representations that align well with the adaptive Fourier decomposition (AFD) theory and can approximate the solution manifold of PDEs on a wide range of geometries and meshes. In several challenging benchmark PDE problems in the fields of fluid physics, solid physics, and finance on point clouds, structured meshes, regular grids, and irregular domains, AMO consistently outperforms state-of-the-art solvers in terms of relative $L^2$ error. Overall, this work presents a new paradigm for designing explainable neural operator frameworks.
摘要
在任意几何形状和各种网格上精确求解偏微分方程(PDE)是科学和工程应用中的重要任务。在本文中,我们提出了自适应Mamba神经算子(AMO),该算子集成了用于状态空间模型(SSM)的再生核,而不是SSM的核积分公式。这是通过为PDE构建Takenaka-Malmquist系统实现的。AMO提供了与自适应Fourier分解(AFD)理论一致的新表示,并且可以在广泛的几何和网格上逼近PDE的解流形。在流体物理、固体物理和金融领域的几个具有挑战性的基准PDE问题中,涉及点云、结构化网格、规则网格和不规则域,AMO在相对L²误差方面始终优于最先进的求解器。总体而言,这项工作为设计可解释的神经算子框架提供了一种新范式。
#27Contraction-Gauge Preconditioning for Quantized Matrix Multiplication
量化矩阵乘法的收缩度量预处理
Piyush Sao, Narasinga Miniskar, Pedro Valero-Lara, Keita Teranishi, Sudip Seal · 2026-07-21T06:09:08Z
Abstract
We study low-precision computation of C=AB with both factors quantized. We derive an exact finite-dimensional identity for the expected squared product error under independent, zero-mean entrywise errors with known variance fields; it holds exactly for non-overloading subtractive dither and for independent stochastic rounding, and we empirically assess deterministic round-to-nearest (RTN). Using the product-preserving equivalence AB=(AT)(T^{-1}B), we formulate contraction-gauge preconditioning: jointly choosing a factor representation and its sharing pattern before quantization. Preconditioning can reduce product error but may require extra transformed, quantized copies of the opposite operand: a shared transform needs one copy, a block-specific transform up to one per block. Within the bounded family of positive diagonal gauges (folds), a geometric program computes a globally optimal shared fold and a linear program decides whether the identity fold is already optimal. For other families we derive computable selection statistics -- tail index for scaling, profile spread for partitioning, coherence and weighted-Gram energy for rotations, slice-energy covariance for hierarchy depth -- with upper bounds for ranking heuristic candidates. Across twelve linear products from a trained three-block image classifier, median within-product rank correlations between dither-model predictions and deterministic-RTN errors are 0.937 at 8 bits and 0.918 at 4 bits. The GP fold cuts held-out product error over the identity fold by 18.0% (8-bit) and 20.5% (4-bit) in geometric mean, beats a SmoothQuant-style grid baseline at both precisions and on ten of twelve products, and lowers composed logit MSE by 15.4% and 26.4%. We thus provide exact stochastic product-error accounting, certified selection within the diagonal family, and a common objective for evaluating reusable transform candidates under RTN.
摘要
我们研究了两个因子都经过量化的低精度计算C=AB。在已知方差场的独立零均值逐项误差下,我们推导了期望平方乘积误差的精确有限维恒等式;该恒等式精确适用于非过载减法抖动和独立随机舍入,并实验评估了确定性最近舍入(RTN)。利用乘积保持等价AB=(AT)(T^{-1}B),我们提出了收缩度量预处理:在量化前联合选择因子表示及其共享模式。预处理可以减少乘积误差,但可能需要额外的变换后量化副本:共享变换需要一个副本,块特定变换每块最多一个副本。在正对角度量(折叠)的有界族内,几何规划计算全局最优共享折叠,线性规划判断恒等折叠是否已最优。对于其他族,我们推导了可计算的选取统计量——缩放的尾指数、分区的剖面扩展、旋转的相干性和加权Gram能量、层次深度的切片能量协方差——并给出了排序启发式候选的上界。在来自训练好的三块图像分类器的十二个线性乘积中,抖动模型预测与确定性RTN误差之间的中位产品内秩相关系数在8位时为0.937,在4位时为0.918。GP折叠相对于恒等折叠将留出乘积误差的几何平均值降低了18.0%(8位)和20.5%(4位),在两种精度下均优于SmoothQuant风格网格基准,并在十二个产品中的十个上表现更好,并将组合logit MSE降低了15.4%和26.4%。因此,我们提供了精确的随机乘积误差核算、对角族内的认证选择,以及评估RTN下可重用变换候选的共同目标。
#28Error Bound and Stability Analysis for a Randomized Singly Diagonally Implicit Runge-Kutta Method
随机单对角隐式Runge-Kutta方法的误差界与稳定性分析
Monika Eisenmann, Marvin Jans, Raphael Kruse, Helmut Podhaisky · 2026-07-21T10:11:16Z
Abstract
A randomized Singly Diagonally Implicit Runge-Kutta (SDIRK) method, based on the randomized trapezoidal rule as the underlying quadrature scheme, is proposed. Every realization of the scheme is an algebraically stable SDIRK method of at least second order. The main result is the proof that the randomized scheme converges with order 2.5 in the root mean square sense under low regularity assumptions. Numerical experiments illustrate the robustness of the new scheme when applied to nonsmooth problems.
摘要
提出了一种基于随机梯形法则作为底层求积方案的随机单对角隐式Runge-Kutta(SDIRK)方法。该方案的每一次实现都是一个至少二阶的代数稳定SDIRK方法。主要结果是证明在低正则性假设下,随机方案在均方意义下以2.5阶收敛。数值实验说明了新方案应用于非光滑问题时的稳健性。
#29Numerical methods for Langevin-type SPDE: an implicit Milstein approach and multilevel Monte Carlo techniques
Langevin型SPDE的数值方法:隐式Milstein方法与多层蒙特卡罗技术
Sascha Portaro, Carlos Vázquez · 2026-07-21T15:20:31Z
Abstract
In this work, we investigate the numerical approximation of degenerate Langevin-type stochastic partial differential equations (SPDEs) in two spatial dimensions. These SPDEs arise in stochastic dynamics and mathematical finance, among other applications. In order to handle the mixed deterministic-stochastic structure of the equation and the degeneracy of the differential operator, we propose a semi-implicit Milstein finite difference scheme for the numerical solution. Through the Fourier analysis of the mean-square stability and convergence, we derive explicit conditions on the coefficients under which the scheme is stable, jointly with explicit convergence rates in terms of the discretization parameters. We further embed the proposed scheme within a Multilevel Monte Carlo (MLMC) framework to reduce the computational cost associated with SPDE simulations, and we derive its theoretical computational complexity. Numerical experiments confirm theoretical convergence rates and show that the MLMC strategy achieves an accuracy comparable to standard Monte Carlo at a fraction of the computational cost, reducing the complexity from $\mathcal{O}(\varepsilon^{-5})$ to $\mathcal{O}(\varepsilon^{-3})$ for a target root-mean-square error $\varepsilon$. These results show that combining semi-implicit Milstein schemes with MLMC techniques provides an effective approach for the numerical simulation of Langevin-type SPDEs.
摘要
本文研究了二维空间中退化Langevin型随机偏微分方程(SPDE)的数值逼近。这些SPDE出现在随机动力学和金融数学等应用中。为了处理方程的混合确定性-随机结构以及微分算子的退化性,我们提出了一种半隐式Milstein有限差分格式用于数值求解。通过均方稳定性和收敛性的傅里叶分析,我们推导了格式稳定的系数显式条件,以及关于离散化参数的显式收敛率。我们进一步将所提格式嵌入多层蒙特卡罗(MLMC)框架,以降低SPDE模拟相关的计算成本,并推导了其理论计算复杂度。数值实验证实了理论收敛率,并表明MLMC策略以一小部分计算成本达到了与标准蒙特卡罗相当的精度,对于目标均方根误差$\varepsilon$,复杂度从$\mathcal{O}(\varepsilon^{-5})$降至$\mathcal{O}(\varepsilon^{-3})$。这些结果表明,半隐式Milstein格式与MLMC技术的结合为Langevin型SPDE的数值模拟提供了一种有效方法。
#30Uniform-in-Time Weak and Ergodic Error Estimates of a Nonlinearity-Explicit Full Discretization for Superlinear SPDEs Driven by Multiplicative Noise
乘性噪声驱动的超线性SPDE非线性显式全离散格式的一致时间弱误差与遍历误差估计
Jingjing Cai, Zhihui Liu, Xiaoming Wu · 2026-07-21T16:19:25Z
Abstract
For a class of superlinear SPDEs driven by multiplicative noise, we prove an (essentially) sharp uniform-in-time (UIT) weak convergence rate for the nonlinearity-explicit Galerkin tamed Euler method (GTEM). Under standard monotonicity assumptions, the proof combines Malliavin calculus with regularity theory for the associated backward Kolmogorov equation (BKE), leading to UIT moment, Hölder, and Malliavin estimates, along with regularity estimates for the BKE solution. These estimates, together with a weak error decomposition and Malliavin integration by parts (IBP) formula, then yield a UIT weak convergence rate $τ^ρ+λ_N^{-ρ}$ for any $ρ\in (0,1)$. Consequently, we obtain a sharp ergodic error estimate between the exact and numerical invariant measures. Numerical experiments support the theory.
摘要
对于一类由乘性噪声驱动的超线性SPDE,我们证明了非线性显式伽辽金驯化欧拉方法(GTEM)的(本质)精确的一致时间(UIT)弱收敛率。在标准单调性假设下,证明结合了Malliavin微积分与相关向后Kolmogorov方程(BKE)的正则性理论,得到了UIT矩估计、Hölder估计和Malliavin估计,以及BKE解的正则性估计。这些估计与弱误差分解和Malliavin分部积分(IBP)公式相结合,进而得出对于任意$ρ\in (0,1)$的UIT弱收敛率$τ^ρ+λ_N^{-ρ}$。因此,我们得到了精确和数值不变测度之间的尖锐遍历误差估计。数值实验支持该理论。
#31Resolution of the ENO-TV conjecture: a parity dichotomy
ENO-TV猜想的解决:奇偶二分法
Zhuoyun Li, Kailiang Wu · 2026-07-21T16:55:27Z
Abstract
We resolve the ENO--TV conjecture, a discrete coercivity problem in compactness theory for entropy-stable approximations of hyperbolic conservation laws. For order-$k$ essentially non-oscillatory (ENO) reconstruction from compactly supported cell averages, it asks whether the nonnegative ENO source times the $(k-1)$st power of the amplitude uniformly controls the $(k+1)$st absolute-jump moment. We prove a parity dichotomy: the estimate holds for odd $k\ge3$ and fails for even $k\ge4$; the known second-order case completes the classification. Localization gives a selection-independent finite-difference functional uniformly comparable to the source and reduces the conjecture to discrete interpolation. For odd orders, summation by parts reveals a hidden square; a discrete Gagliardo--Nirenberg inequality yields coercivity. For even orders, Euler-polynomial blocks from the functional's polynomial kernel yield counterexamples that persist under arbitrarily small perturbations making all affected ENO comparisons strict. We also prove two coercive estimates for every $k\ge2$: control of jumps larger than a fixed fraction of the amplitude and of local blocks modulo sampled polynomials of degree at most $k-2$. Via the Cayley--Sylvester decomposition, we compute the dimensions of homogeneous first-cohomology spaces for the lattice shift on polynomial jump profiles. At fourth order, for a cubic flux and a globally strictly convex entropy, a total-degree-seven component of a reduced entropy-flux mismatch represents a nonzero class on profiles of degree at most two and hence has no translation-invariant finite-stencil $C^7$ local primitive at the zero constant state. Odd-order coercivity persists on globally quasi-uniform meshes, whereas for each $k\ge2$ it fails on a fixed irregular mesh even though every interface contribution remains nonnegative. This failure is due to the mesh geometry.
摘要
我们解决了ENO-TV猜想,这是一个关于双曲守恒律熵稳定逼近的紧性理论中的离散强制性问题。对于从紧支撑单元平均值的k阶基本无振荡(ENO)重构,它询问非负ENO源乘以振幅的(k-1)次幂是否一致地控制(k+1)次绝对跳跃矩。我们证明了奇偶二分法:该估计对奇数k≥3成立,对偶数k≥4不成立;已知的二阶情形完成了分类。局部化给出了一个与源一致可比的与选择无关的有限差分泛函,并将猜想简化为离散插值。对于奇数阶,分部求和揭示了一个隐藏的平方;一个离散的Gagliardo-Nirenberg不等式给出了强制性。对于偶数阶,来自泛函多项式核的欧拉多项式块产生了反例,这些反例在任意小扰动下持续存在,使所有受影响的ENO比较变得严格。我们还证明了每个k≥2的两个强制性估计:控制大于振幅固定分数的跳跃,以及控制局部块模次数最多为k-2的采样多项式。通过Cayley-Sylvester分解,我们计算了多项式跳跃剖面上格点平移的齐次第一上同调空间的维数。在四阶情况下,对于三次通量和全局严格凸熵,简化熵-通量失配的总次数七分量在次数最多为二的剖面上表示非零类,因此在零常数状态下没有平移不变有限模板C^7局部原函数。奇数阶强制性在全局拟均匀网格上持续成立,而对于每个k≥2,它在固定不规则网格上失败,尽管每个界面贡献保持非负。这种失败是由于网格几何形状造成的。
#32Positive-Allocation Companion Predictors for Nonlinear Dynamics and Their Finite-Difference Diagnostics
用于非线性动力学的正分配伴随预测器及其有限差分诊断
Cynthia V. Flores, James E. Pascoe · 2026-07-17T22:09:51Z
Abstract
We introduce a positive-allocation companion construction for Koopman-inspired finite-dimensional prediction of nonlinear dynamical systems. The method determines recurrence coefficients by representing a target observable snapshot as a nonnegative, normalized combination of earlier training snapshots. These coefficients define a companion matrix whose spectral structure is induced by the allocation constraints at the construction stage. We prove that normalized positive allocation places the companion spectrum in the closed unit disk and, because the coefficients sum to one, includes $1$ as an eigenvalue. Additionally, we develop modal and non-modal diagnostics for the resulting model trajectory. When the companion matrix is diagonalizable, the modal representation shows that first and second finite differences act as spectral filters through factors of $λ_\ell-1$ and $(λ_\ell-1)^2$. We also derive $C$-based finite-difference bounds that avoid diagonalization and can be evaluated directly from the training data and companion matrix. Numerical experiments on the FitzHugh--Nagumo and susceptible--infectious--recovered (SIR) models illustrate the behavior of the construction in oscillatory and transient dissipative settings. The examples demonstrate both the interpretability of the companion recurrence and its limitations, particularly when pointwise trajectory agreement degrades while finite-difference and modal diagnostics remain informative.
摘要
我们引入了一种正分配伴随构造,用于Koopman启发的非线性动力学系统的有限维预测。该方法通过将目标可观测快照表示为早期训练快照的非负归一化组合来确定递归系数。这些系数定义了一个伴随矩阵,其谱结构由构造阶段的分配约束诱导。我们证明了归一化正分配将伴随谱置于闭单位圆盘内,并且由于系数之和为1,包含1作为一个特征值。此外,我们为所得模型轨迹开发了模态和非模态诊断。当伴随矩阵可对角化时,模态表示表明,一阶和二阶有限差分通过因子 $\lambda_\ell-1$ 和 $(\lambda_\ell-1)^2$ 充当谱滤波器。我们还推导了基于 $C$ 的有限差分界,避免了对角化,可以直接从训练数据和伴随矩阵中评估。在FitzHugh-Nagumo和易感-感染-恢复(SIR)模型上的数值实验说明了该构造在振荡和瞬态耗散环境中的行为。这些例子展示了伴随递归的可解释性及其局限性,特别是在逐点轨迹一致性下降时,有限差分和模态诊断仍然具有信息性。
#33Augmented Lagrangian preconditioning for a simplified Ericksen--Leslie model of nematic liquid crystals
用于向列液晶简化Ericksen-Leslie模型的增广拉格朗日预处理
Yanying Li, Xu Qian, Jingmin Xia · 2026-07-18T04:10:55Z
Abstract
The numerical solution of the simplified Ericksen--Leslie model for nematic liquid crystals is challenging because the flow and director equations are strongly coupled and because incompressibility and the unit-length condition must be enforced simultaneously. A Lagrange multiplier formulation avoids a small Ginzburg--Landau parameter, but the Newton systems have a double saddle-point structure. We develop an augmented Lagrangian block preconditioner in which both constraints are augmented while their discrete enforcement remains multiplier based. After finite element discretization and backward Euler time integration, the Newton increments are grouped into velocity--director and pressure-multiplier variables. A block-diagonal approximation of the coupled velocity-director block then leads to separate, physically scaled approximations of the pressure and director-multiplier Schur complements. Manufactured-solution tests show the expected spatial accuracy and first-order temporal convergence for the primary variables; the multiplier error reaches a spatial-error floor on the fixed mesh used in the temporal study. In the reported parameter ranges, the outer FGMRES iteration counts are nearly mesh independent, remain stable under time-step and viscosity variation, and improve as the augmentation parameters increase. A smooth benchmark also exhibits monotone decay of the computed total energy.
摘要
简化Ericksen-Leslie向列液晶模型的数值求解具有挑战性,因为流动方程和指向矢方程强耦合,并且必须同时强制执行不可压缩性和单位长度条件。拉格朗日乘子公式避免了小的Ginzburg-Landau参数,但牛顿系统具有双鞍点结构。我们开发了一种增广拉格朗日块预处理方法,其中两个约束都被增广,同时它们的离散强制执行仍然是基于乘子的。在有限元离散化和向后欧拉时间积分之后,牛顿增量被分为速度-指向矢变量和压力-乘子变量。速度-指向矢块的块对角近似导致压力和指向矢-乘子Schur补的独立物理缩放的近似。制造解测试显示了主要变量预期的空间精度和一阶时间收敛性;乘子误差在时间研究中使用的固定网格上达到空间误差下限。在报告的参数范围内,外部FGMRES迭代次数几乎与网格无关,在时间步长和粘度变化下保持稳定,并随着增广参数的增大而改善。一个平滑基准也显示出计算的总能量单调衰减。
#34PolyChopper: a Polyhedron Splitting Scheme
PolyChopper:一种多面体分割方案
Tommaso Sorgente, Fabio Vicini · 2026-07-18T17:04:32Z
Abstract
We introduce and implement a novel scheme, called \textit{PolyChopper}, for splitting a convex polyhedron into a finite set of disjoint convex sub-polyhedra. The algorithm cuts the polyhedron with a plane, which can either be prescribed as an external constraint or automatically determined from the inertia tensor. The resulting intersection vertices are then adjusted to control the complexity and quality of the generated elements, introducing additional ``notch wedges'', tetrahedra and pyramids, according to a user-defined quality parameter. Experimental results demonstrate that the proposed approach is robust and consistently produces sub-polyhedra that are both simpler and of higher quality than the original polyhedron. Furthermore, recursively applying the algorithm generates a hierarchy of polyhedral subdivisions with progressively smaller elements, all guaranteed to remain convex and predominantly tetrahedral.
摘要
我们提出并实现了一种新颖的方案,称为\textit{PolyChopper},用于将凸多面体分割成有限个不相交的凸子多面体。该算法用一个平面切割多面体,该平面可以作为外部约束指定,也可以从惯性张量自动确定。然后调整生成的相交顶点,以控制生成元素的复杂性和质量,根据用户定义的质量参数引入额外的“凹口楔体”、四面体和金字塔。实验结果表明,所提出的方法是鲁棒的,并且一致地生成比原始多面体更简单、质量更高的子多面体。此外,递归应用该算法生成一个多面体细分层次结构,元素逐渐变小,所有元素都保证保持凸性且主要为四面体。
#35Digital Nets on Cubature Nodes: Inheriting Cubature Accuracy on Low-Dimensional Projections
基于求积节点的数字网格:在低维投影上继承求积精度
Takehito Yoshiki · 2026-07-19T05:07:21Z
Abstract
Base-2 digital nets are practical high-dimensional integration rules: the sample budget $N=2^m$ can be chosen independently of the ambient dimension, and the generating matrices provide algebraic control of projections and Walsh-dual weights. They are therefore well suited to problems whose error is governed by weighted or low-dimensional projection structure. However, when one restricts attention to a smooth low-dimensional projected component, a low-dimensional cubature rule with a comparable number of nodes can be substantially more accurate than the projected digital-net points. This raises the question of whether low-dimensional cubature accuracy can be inserted into a high-dimensional digital-net rule without forming the full tensor product. We answer this question by a simple coordinate embedding: read the leading $p$ binary digits of each coordinate as an index into $2^p$ equal-weight cubature nodes, and replace the coordinate by the indexed node. When a projection forms the full $p$-bit grid, the transformed rule coincides on that projection with the corresponding product cubature rule; small projected $t$-values provide sufficient conditions for such full-grid recovery. For general integrands, the error separates into the corresponding product cubature error and a residual digital-net term. Experiments with scrambled Sobol' nets in dimension $50$ illustrate this mechanism and show finite-budget improvements for the smooth low-order and coordinate-decaying test functions considered here.
摘要
基2数字网格是实用的高维积分规则:样本预算 $N=2^m$ 可以独立于环境维数选择,生成矩阵提供了对投影和Walsh对偶权重的代数控制。因此,它们非常适合误差由加权或低维投影结构支配的问题。然而,当仅关注光滑的低维投影分量时,具有相当节点数的低维求积规则可能比投影的数字网格点精确得多。这引发了一个问题:是否可以将低维求积精度插入到高维数字网格规则中,而不形成完整的张量积?我们通过一个简单的坐标嵌入来回答这个问题:将每个坐标的前 $p$ 个二进制数字读入 $2^p$ 个等权求积节点的索引,并将该坐标替换为索引节点。当投影形成完整的 $p$ 位网格时,转换后的规则在该投影上与相应的乘积求积规则一致;小的投影 $t$ 值为这种全网格恢复提供了充分条件。对于一般的被积函数,误差分解为相应的乘积求积误差和残余的数字网格项。在维度 $50$ 上使用加扰Sobol'网格的实验说明了这一机制,并显示了对于此处考虑的平滑低阶和坐标衰减测试函数,在有限预算下的改进。
#36Quantitative Benchmarking of a Split-Field PML FDTD Solver: Slit Diffraction, and Scattering from PEC and Dielectric Cylinders
分裂场PML FDTD求解器的定量基准测试:狭缝衍射以及PEC和介质圆柱的散射
Sabrina Saima, Tasin Intisar · 2026-07-19T17:53:04Z
Abstract
This paper presents a two-dimensional TMz finite-difference time-domain (FDTD) solver based on Yee's scheme for modeling radiation from an infinitely long z-directed line current, with the open region truncated by a Berenger split-field perfectly matched layer (PML). After validating cylindrical-wave propagation and negligible late-time reflections in free space, the solver is applied to three inhomogeneous configurations: (i) diffraction through a one-cell-thick perfectly electrically conducting (PEC) sheet with single and double slits; (ii) scattering from infinitely long PEC cylinders of circular and rectangular cross section; and (iii) scattering from infinitely long dielectric cylinders of varying cross section and permittivity. Beyond qualitative field maps, the diffraction case is characterized quantitatively: a steady-state phasor extracted by a running discrete Fourier transform yields the transmitted intensity, from which the fringe visibility and the far-field pattern are computed and compared against the closed-form Fraunhofer prediction. The single- and double-slit cases are cleanly separated by a visibility that rises from near zero to near unity, and the double-slit interference maxima agree with the grating condition arcsin(m λ_0 / d) to within a fraction of a degree. For dielectric cylinders, the field penetrates the obstacle with the expected reduced internal wavelength λ_0 / \sqrt{ε_r}, and the scattered field strength grows with permittivity contrast. A reference-subtraction method isolates the scattered field throughout. The results confirm that the FDTD-PML framework accurately captures open-region diffraction and geometry- and material-dependent scattering.
摘要
本文提出了一种基于Yee格式的二维TMz时域有限差分(FDTD)求解器,用于模拟无限长z方向线电流的辐射,开放区域由Berenger分裂场完美匹配层(PML)截断。在验证了自由空间中的柱面波传播和可忽略的后期反射后,该求解器被应用于三种非均匀配置:(i)通过单狭缝和双狭缝的一个单元厚完美电导体(PEC)板的衍射;(ii)无限长圆形和矩形截面PEC圆柱的散射;以及(iii)不同截面和介电常数的无限长介质圆柱的散射。除了定性场图外,衍射情况还进行了定量表征:通过运行离散傅里叶变换提取稳态相量,得到透射强度,从中计算条纹可见度和远场图案,并与闭式夫琅禾费预测进行比较。单缝和双缝情况通过可见度从接近零上升到接近1而清晰区分,双缝干涉极大值与光栅条件 arcsin(m λ_0 / d) 的偏差在零点几度以内。对于介质圆柱,场穿透障碍物,具有预期的缩短波长 λ_0 / \sqrt{ε_r},且散射场强度随介电常数对比度增加而增长。一种参考相减方法在整个过程中隔离了散射场。结果证实,FDTD-PML框架准确捕捉了开放区域衍射以及几何和材料相关的散射。
#37FlowSonic: Stable Zero-Shot Music Editing via High-Order Trajectory Integration
FlowSonic:通过高阶轨迹积分实现稳定的零样本音乐编辑
Ali Boudaghi, Hadi Zare · 2026-07-20T04:00:16Z
Abstract
Zero-shot text-guided editing of real-world music recordings requires balancing semantic modification with faithful preservation of the original musical structure. Although recent diffusion transformers trained with rectified flow have achieved remarkable success in text-to-music generation, extending them to edit existing recordings remains challenging because editing requires accurate deterministic inversion, reliable structural preservation, and numerically stable integration throughout the inversion and generation processes. We present FlowSonic, a zero-shot music editing framework built upon a pretrained diffusion transformer trained with rectified flow. FlowSonic first deterministically inverts a real-world recording into the latent space and preserves its musical structure during editing by reusing cross-attention representations extracted during inversion. To improve the numerical reliability of inversion-based editing, we introduce a high-order ODE solver and systematically investigate how different numerical integration schemes influence trajectory stability, structural preservation, and semantic controllability. Comprehensive experiments on timbre-transfer and genre-modification tasks demonstrate that FlowSonic consistently outperforms existing music editing methods across semantic alignment, harmonic preservation, structural consistency, and perceptual audio quality. We further provide geometric and empirical analyses showing how the proposed numerical integration strategy improves latent trajectory stability and leads to more reliable music editing.
摘要
真实音乐录音的零样本文本引导编辑需要在语义修改与忠实保留原始音乐结构之间取得平衡。尽管最近使用整流流训练的扩散变换器在文本到音乐生成方面取得了显著成功,但将其扩展到编辑现有录音仍然具有挑战性,因为编辑需要准确的确定性反演、可靠的结构保留以及在反演和生成过程中的数值稳定积分。我们提出了FlowSonic,一个建立在预训练整流流扩散变换器上的零样本音乐编辑框架。FlowSonic首先将真实录音确定性地反演到潜在空间,并通过重用反演期间提取的交叉注意力表示来在编辑过程中保留其音乐结构。为了提高基于反演的编辑的数值可靠性,我们引入了一个高阶ODE求解器,并系统研究了不同数值积分方案如何影响轨迹稳定性、结构保留和语义可控性。在音色迁移和体裁修改任务上的综合实验表明,FlowSonic在语义对齐、和声保留、结构一致性和感知音频质量方面始终优于现有的音乐编辑方法。我们进一步提供了几何和实证分析,展示了所提出的数值积分策略如何提高潜在轨迹稳定性,从而实现更可靠的音乐编辑。
#38An operator-splitting algorithm for the hypergraph $p$-Laplacian with applications to missing data recovery
超图p-拉普拉斯算子的算子分裂算法及其在缺失数据恢复中的应用
Kehan Shi, Jin Liu, Martin Burger · 2026-07-20T06:51:51Z
Abstract
Hypergraph $p$-Laplacian regularization is a fundamental model in data analysis with successful applications in various tasks. It aims to minimize a nonsmooth and typically large-scale objective function defined as the sum of the $p$-th powers of the Lipschitz regularization over hyperedges. In this paper, we propose an operator-splitting algorithm for the hypergraph $p$-Laplacian that allows us to handle hyperedges separately in a Gauss-Seidel fashion. Each subproblem can be viewed as a generalized graph Lipschitz learning on a hyperedge, for which we introduce an auxiliary variable to overcome the nonsmoothness and solve it with one step of the alternating direction method of multipliers (ADMM). The resulting algorithm performs proximal ADMM updates sequentially over the hyperedges, and its convergence is proven. We test the algorithm on missing data recovery problems, including image sparse inpainting and semi-supervised learning, to demonstrate that it is faster than existing methods.
摘要
超图p-拉普拉斯正则化是数据分析中的一个基本模型,在各种任务中都有成功的应用。它旨在最小化一个非光滑且通常大规模的目标函数,该函数定义为超边上Lipschitz正则化的p次幂之和。在本文中,我们提出了一种超图p-拉普拉斯算子的算子分裂算法,允许我们以高斯-赛德尔方式分别处理超边。每个子问题可以看作超边上的广义图Lipschitz学习,为此我们引入一个辅助变量来克服非光滑性,并用交替方向乘子法(ADMM)的一步求解。所得到的算法在超边上顺序执行近端ADMM更新,并证明了其收敛性。我们在缺失数据恢复问题(包括图像稀疏修复和半监督学习)上测试了该算法,表明它比现有方法更快。
#39Pathwise skew-symmetric discretisation for SDEs with superlinear drift
带有超线性漂移的随机微分方程的路径反对称离散化
Yuga Iguchi, Samuel Livingstone, Giorgos Vasdekis, Rui-Yang Zhang · 2026-07-20T09:24:11Z
Abstract
The skew-symmetric discretisation has recently been proposed as a new robust simulation method for weakly approximating stochastic differential equations (SDEs) with non-globally Lipschitz drift. This work develops a pathwise version of the scheme by representing the noise increment as a skew-normal distribution and coupling it with the driving Brownian increments, thereby enabling its use in the multilevel Monte Carlo (MLMC) framework. Under suitable conditions, we establish strong convergence of order 1/2 in $L^2$. Subsequently, the associated MLMC estimator is shown to have computational complexity $\mathcal{O} \bigl(\varepsilon^{-2} (\log (1/\varepsilon))^2 \bigr)$ to achieve a mean-squared error $\varepsilon^2$. We then analytically compare the proposed scheme with the tamed Euler scheme, another benchmark for robust discretisation. Under a strong inward-drift regime with the current state being far from the stable region, we show that the probability of moving in the wrong direction tends to vanish in the skew-symmetric scheme, whereas the tamed Euler scheme makes such moves with a non-trivial probability. Furthermore, in the MLMC setting, employing a one-dimensional stochastic Ginzburg-Landau model, we specify the range of step sizes for which the asymptotic variance of the coupled level difference obtained via the pathwise skew-symmetric scheme is lower than that obtained via the tamed Euler scheme. Numerical experiments on several model examples support the theoretical rate of strong convergence and demonstrate the stability and effectiveness of the resulting MLMC in the superlinear drift setting.
摘要
反对称离散化最近被提出作为一种新的稳健模拟方法,用于弱逼近具有非全局Lipschitz漂移的随机微分方程(SDE)。本文通过将噪声增量表示为偏正态分布并与驱动布朗增量耦合,开发了该方案的路径版本,从而使其能够在多级蒙特卡罗(MLMC)框架中使用。在适当条件下,我们建立了 $L^2$ 中 1/2 阶的强收敛性。随后,相关的MLMC估计器被证明具有计算复杂度 $\mathcal{O} \bigl(\varepsilon^{-2} (\log (1/\varepsilon))^2 \bigr)$,以达到均方误差 $\varepsilon^2$。然后,我们从解析角度将所提出的方案与驯服Euler方案(另一种稳健离散化的基准)进行了比较。在强内向漂移状态下,当前状态远离稳定区域时,我们表明在反对称方案中,朝错误方向移动的概率趋于零,而驯服Euler方案以非平凡概率进行这样的移动。此外,在MLMC设置中,使用一维随机Ginzburg-Landau模型,我们指定了步长范围,在该范围内通过路径反对称方案获得的耦合水平差异的渐近方差低于通过驯服Euler方案获得的。多个模型算例的数值实验支持了强收敛的理论速率,并展示了所得到的MLMC在超线性漂移设置中的稳定性和有效性。
#40A Direct Approach to Hermite Interpolation
Hermite插值的直接方法
Simon Bossoney, Marc Troyanov · 2026-07-20T14:32:04Z
Abstract
We introduce a family of polynomials satisfying the natural duality relations for Hermite interpolation, analogous to the classical Lagrange interpolation polynomials. They yield an explicit closed formula for the Hermite interpolant with arbitrary multiplicities, without recourse to divided differences, recursive corrections, or auxiliary Bézout identities. We also give a detailed account of Hermite's original approach to the problem, based on an integral formula, which he used both to derive the interpolating polynomial and to estimate the interpolation error for holomorphic data.
摘要
我们引入了一族多项式,满足Hermite插值的自然对偶关系,类似于经典的Lagrange插值多项式。它们给出了任意重数情况下Hermite插值的显式封闭公式,无需借助差商、递归修正或辅助Bézout恒等式。我们还详细阐述了Hermite基于积分公式的原始方法,他利用该公式推导了插值多项式,并估计了全纯数据的插值误差。
#41Krasnosel'skii-Mann iterations beyond asymptotics: a combinatorial analysis
Krasnosel'skii-Mann迭代的超越渐近分析:组合分析
Mario Bravo, Roberto Cominetti · 2026-07-20T16:15:39Z
Abstract
We revisit the classical Krasnosel'skii-Mann fixed point iteration for contractions and nonexpansive maps in general normed spaces. This iteration is ubiquitous across a wide range of areas, including convex optimization, monotone inclusions, Markov decision processes, under-relaxed methods for nonlinear PDEs, and more. Drawing on a remarkable connection with a Markov chain on $\mathbb{Z}^2$, and using counting arguments from enumerative combinatorics of lattice paths, we derive explicit estimates for the distance between iterates, as well as non-asymptotic error bounds for the fixed point residuals. As the contraction parameter approaches one, these bounds smoothly recover the known estimates for nonexpansive maps. Building upon these estimates, we further derive error bounds for inexact Krasnosel'skii-Mann iterations.
摘要
我们重新审视了通用赋范空间中压缩映射和非扩张映射的经典Krasnosel'skii-Mann不动点迭代。该迭代广泛应用于凸优化、单调包含、马尔可夫决策过程、非线性偏微分方程的欠松弛方法等多个领域。利用与$\mathbb{Z}^2$上马尔可夫链的卓越联系,并借助格点路径的枚举组合计数论证,我们推导了迭代之间距离的显式估计,以及不动点残差的非渐近误差界。当压缩参数趋近于1时,这些界平滑地恢复为非扩张映射的已知估计。基于这些估计,我们进一步推导了非精确Krasnosel'skii-Mann迭代的误差界。
#42Feedback Cycles in Exploratory Equilibria
探索性均衡中的反馈循环
Chen-Hung Wu · 2026-07-20T16:20:55Z
Abstract
Entropy regularization smooths equilibrium policies in time-inconsistent stochastic control. At low temperature, the same Gibbs response can strongly amplify errors in learned rewards and dynamics. We show that the derivative of an exploratory equilibrium is governed by a backward Volterra-parabolic resolvent. Along an aligned positive mode, a lower bound has the same exponential order. A block decomposition identifies the source of the amplification: causal paths contribute powers of 1/tau, whereas a positive feedback cycle can produce exponential growth. At fixed temperature, a local equilibrium branch is twice differentiable with respect to finite-dimensional model parameters, which yields a function-valued delta method. A bounded uniformly elliptic diffusion realizes this path-cycle distinction in every finite dimension. Closing one positive cycle changes the root-n linear-response boundary from a power law to order 1/log n; along the cyclic Perron mode, right-endpoint discretization is relatively consistent exactly when N tau^2 -> infinity. An affine model also gives an exact nonlinear transition at the Lambert-W temperature beta T / W(beta T sqrt(n)). Numerical calculations illustrate these rates.
摘要
熵正则化平滑了时间不一致随机控制中的均衡策略。在低温下,相同的吉布斯响应会强烈放大学习奖励和动力学中的误差。我们证明了探索性均衡的导数由后向Volterra-抛物预解式控制。沿对齐的正模,下界具有相同的指数阶。块分解识别了放大的来源:因果路径贡献了1/tau的幂次,而正反馈循环则可能产生指数增长。在固定温度下,局部均衡分支关于有限维模型参数是二次可微的,这产生了函数值delta方法。一个有界一致椭圆扩散在每个有限维中实现了这种路径-循环区分。关闭一个正循环将根n线性响应边界从幂律改变为1/log n阶;沿循环Perron模,右端点离散化在N tau^2 -> ∞时精确相对一致。一个仿射模型也在Lambert-W温度β T / W(β T sqrt(n))处给出了精确的非线性转变。数值计算展示了这些速率。
#43anyakrakusuma: A Python Library for Entropic Schrödinger Bridges on Idealized Geometries
anyakrakusuma:一个用于理想几何上熵Schrödinger桥的Python库
Sandy Hardian Susanto Herho, Dasapta Erwin Irawan, Agus Wahyu Jatmiko, Sito Fossy Biosa, Candrasa Surya Dharma, Edi Riawan, Astyka Pamumpuni, Rendy Dwi Kartiko, Rusmawan Suwarman, Deny Juanda Puradimaja · 2026-07-20T17:23:37Z
Abstract
We present anyakrakusuma, an open-source Python library that solves the discrete static Schrödinger bridge problem, the entropically regularized counterpart of optimal transport, through a log-domain Sinkhorn--Knopp iteration and reconstructs the entropic interpolation between two empirical point clouds. The solver is paired with a diagnostic pipeline that characterizes the optimal coupling and the intermediate distributions through information-theoretic and geometric measures. We exercise the library on four idealized planar cases spanning a circle-to-circle dilation, a spiral-to-mixture fragmentation, a rigid reorientation of two moons, and a Lissajous-to-trefoil deformation. The log-domain formulation is necessary rather than merely convenient at the parameters studied, where the cost-to-regularization ratio reaches four hundred and the Gibbs kernel underflows double precision across most of its range; the iteration nonetheless attains a marginal residual of $10^{-9}$ and unit marginal fidelity in every case. Residual histories decay geometrically over approximately eight decades at per-iteration contraction factors between $0.966$ and $0.976$, which are local rates near the fixed point that lie many orders of magnitude below the worst-case Hilbert-metric bound. The covariance analysis recovers an imposed ninety-degree reorientation to within $0.07^\circ$, roughly forty times smaller than its uncertainty, across a masked interval of near-isotropy on which the principal axis is unobservable. The diagnostics are reported with explicit attention to the regimes in which each is well defined, including the differential entropy, which is meaningful only on the open interpolation interval. The presented cases are constructed rather than measured; quantitative application to empirical point clouds requires further study.
摘要
我们提出了anyakrakusuma,一个开源的Python库,通过对数域Sinkhorn-Knopp迭代求解离散静态Schrödinger桥问题(最优传输的熵正则化对应),并重构两个经验点云之间的熵插值。该求解器配有一个诊断流程,通过信息论和几何度量表征最优耦合和中间分布。我们在四个理想平面案例上测试了该库,包括圆到圆的膨胀、螺旋到混合的碎裂、两个月亮的刚性重定向以及利萨如到三叶结的变形。在所研究的参数下,对数域公式是必要的,而不仅仅是方便的,其中成本与正则化之比达到四百,吉布斯核在大部分范围内下溢双精度;然而,迭代在每个案例中仍达到了$10^{-9}$的边缘残差和单位边缘保真度。残差历史在约八个数量级上几何衰减,每次迭代的收缩因子在0.966到0.976之间,这些是固定点附近的局部速率,比最坏情况下的Hilbert度量界低多个数量级。协方差分析在近各向同性的掩蔽区间内(此时主轴不可观测)恢复了一个施加的九十度重定向,误差在$0.07^\circ$以内,比其不确定性小约四十倍。诊断报告明确关注每个定义良好的区域,包括仅在开放插值区间上有意义的微分熵。呈现的案例是构造的而非测量的;对经验点云的定量应用需要进一步研究。
#44Convergence and almost sure exponential stability of compensated split-step theta scheme for stochastic pantograph models with Poisson random measure
随机比例模型带泊松随机测度的补偿分裂步theta方法的收敛性和几乎必然指数稳定性
Amr Abosenna, Yongchun Zhou, Boping Tian · 2026-07-18T16:07:36Z
Abstract
Recently, stochastic pantograph models have gained an intensive attention and have been used in different fields such as finance, biology, control and stochastic neural networks. It is also more preferable to incorporate jumps during the study of stochastic differential equations. In this paper, stochastic pantograph model with Poisson random measure is studied. The compensated split-step theta technique is applied to the considered model. The numerical scheme exhibits a non divergent attitude and converges to the solution of our model under assumptions addressed later on. Furthermore, the almost sure exponential stability of the numerical scheme is investigated via utilizing the discrete semi-martingale convergence theorem. Finally, theoretical findings are manifested via some numerical examples.
摘要
近年来,随机比例模型引起了广泛关注,并应用于金融、生物学、控制和随机神经网络等不同领域。在随机微分方程的研究中,加入跳跃也是更可取的。本文研究了带泊松随机测度的随机比例模型。将补偿分裂步theta技术应用于所考虑的模型。该数值方案表现出非发散行为,并在后文阐述的假设下收敛到我们的模型解。此外,利用离散半鞅收敛定理研究了数值方案的几乎必然指数稳定性。最后,通过一些数值例子展示了理论结果。
#45Current-Sheet Formation in Electron Magnetohydrodynamics with Split Fractional Dissipation
分裂分数阶耗散电子磁流体动力学中的电流片形成
Ruimeng Hu, Qirui Peng, Xu Yang · 2026-07-21T02:37:09Z
Abstract
Thin current sheets are central small-scale structures in electron magnetohydrodynamics (EMHD), closely associated with energy dissipation and fast magnetic reconnection at electron scales. We study their formation numerically in a $2\frac{1}{2}$-dimensional EMHD system on a periodic domain with split fractional dissipation, where the magnetic potential and the vertical magnetic component are damped separately. The local theory is governed by a symmetric combined damping balance, but the numerical onset of small-scale growth need not follow this symmetry. A scaling analysis identifies the out-of-plane current as the primary concentration observable, since it is regularized only through the magnetic-potential equation. Using a validated Fourier pseudospectral exponential time-differencing solver with resolution-controlled diagnostics, we find a clear decay/concentration dichotomy. The onset boundary is markedly asymmetric: current-sheet formation appears to be controlled mainly by damping of the magnetic potential, rather than by the combined damping strength. The analyticity strip collapses to the grid scale, the concentration sharpens under grid refinement, and the observed growth is consistent with an energy-critical self-similar rate, with exponent near three. These experiments indicate that magnetic-potential damping is the apparent binding constraint for current-sheet concentration, refining the symmetric sum picture.
摘要
薄电流片是电子磁流体动力学(EMHD)中的中心小尺度结构,与电子尺度上的能量耗散和快速磁重联密切相关。我们在具有分裂分数阶耗散的周期域上的二维半EMHD系统中数值研究它们的形成,其中磁势和垂直磁场分量分别被阻尼。局部理论由对称的组合阻尼平衡控制,但小尺度增长的数值开始不一定遵循这种对称性。尺度分析识别出面外电流为主要集中可观测量,因为它仅通过磁势方程被正则化。使用经过验证的傅里叶伪谱指数时间差分求解器和分辨率控制的诊断,我们发现了一个清晰的衰减/集中二分法。开始边界明显不对称:电流片形成似乎主要由磁势阻尼控制,而不是组合阻尼强度。解析性条塌缩到网格尺度,浓缩在网格细化下加剧,观察到的增长与能量临界自相似速率一致,指数接近三。这些实验表明,磁势阻尼是电流片浓缩的明显约束条件,细化了对称和图像。
#46A Second-Moment Theory for Floating-Point Reduction Trees
浮点归约树的二阶矩理论
Piyush Sao, Narasinga Miniskar, Pedro Valero-Lara, Keita Teranishi, Sudip Seal · 2026-07-21T06:27:43Z
Abstract
Summation error depends on partial-sum order, which standard worst-case bounds omit. To capture this dependence, we derive an exact mean-square error (MSE) recurrence for a binary reduction tree T under conditionally unbiased rounding. With unit roundoff u, the constant-nu model sets the local variance at pre-rounding value x to nu u^2 x^2. Its leading tree-dependent cost for the input vector p is p^T K_T p, where the common-ancestor kernel K_T counts the internal ancestors shared by each pair of leaves. For i.i.d. inputs of mean mu and variance tau^2, this expected cost is tau^2 Lambda_1(T) + mu^2 Lambda_2(T), where Lambda_1 is total leaf depth and Lambda_2 sums squared internal-subtree sizes; Lambda_1 governs centered inputs, while Lambda_2 captures nonzero means. We use these statistics to characterize optimal tree topologies and schedules. Balanced and sequential trees attain the centered extrema. For k inputs, optimal two-stage sequential blocking yields root-mean-square (RMS) error scaling as k^{3/4}. For fixed-stage hierarchies, geometric schedules are optimal for centered inputs, whereas the optimal noncentered stage exponents halve successively. For independent centered inputs with unequal variances, Huffman coding minimizes variance-weighted depth over free leaf assignments. We extend the kernel to matrix multiplication through operand Gram matrices. We then test the approximation under round-to-nearest using exact residuals. Across binary64, binary32, and software-emulated binary16 and bfloat16, the model recovers the ordering among tree topologies; K_T tracks AR(1) partial-sum costs. For GEMM, independently calibrated predictions differ from measurements by at most 3% on the tested grid. A reduction tree extracted from an array library predicts the measured RMS scaling. However, stagnation and bias in positive low-precision sums limit the model's applicability.
摘要
求和误差取决于部分和顺序,标准最坏情况界限忽略了这一点。为了捕捉这种依赖性,我们推导了在条件无偏舍入下二元归约树T的精确均方误差(MSE)递推关系。以单位舍入u为单位,常数nu模型将预舍入值x处的局部方差设置为nu u^2 x^2。其对输入向量p的主要树相关代价为p^T K_T p,其中共同祖先核K_T计算每对叶子共享的内部祖先数量。对于均值为mu、方差为tau^2的独立同分布输入,期望代价为tau^2 Λ_1(T) + mu^2 Λ_2(T),其中Λ_1是总叶子深度,Λ_2是内部子树大小的平方和;Λ_1控制中心化输入,而Λ_2捕获非零均值。我们使用这些统计量来表征最优树拓扑和调度。平衡树和顺序树达到中心极值。对于k个输入,最优两阶段顺序分块产生均方根(RMS)误差缩放为k^{3/4}。对于固定阶段层次结构,几何调度对中心化输入是最优的,而非中心化输入的最优阶段指数依次减半。对于具有不等方差的独立中心化输入,哈夫曼编码在自由叶子分配上最小化方差加权深度。我们通过操作数格拉姆矩阵将核扩展到矩阵乘法。然后我们使用精确残差在最近舍入下测试近似。在binary64、binary32和软件模拟的binary16和bfloat16中,该模型恢复了树拓扑之间的顺序;K_T跟踪AR(1)部分和成本。对于GEMM,独立校准的预测与测量值在测试网格上相差不超过3%。从数组库中提取的归约树预测了测量的RMS缩放。然而,正低精度求和中的停滞和偏差限制了模型的适用性。
#47Boundary-Adapted PINNs for Elliptic Dirichlet Problems: $H^2(Ω)$ A Priori Error Bounds with Application to Mean Escape Time Computation
适应边界的PINN用于椭圆Dirichlet问题:H^2(Ω)先验误差界及其在平均逃逸时间计算中的应用
Nathanael Tepakbong, Jun Fan, Xiang Zhou, Ding-Xuan Zhou · 2026-07-21T15:03:42Z
Abstract
Motivated by the numerical computation of the Mean Escape Time (MET) $τ:Ω\to\mathbb{R}$ of a stochastic process from a bounded domain $Ω\subseteq\mathbb{R}^d$, we study elliptic Dirichlet boundary value problems (BVPs) using boundary-enforced Physics-Informed Neural Networks (PINNs), in which the Dirichlet condition is imposed exactly by multiplying the network output with a predefined distance-to-boundary approximation $ρ$. Combining approximation-theoretic and statistical-learning arguments for Rectified Quadratic Unit (ReQU) and hyperbolic tangent (tanh) networks, we derive a priori error bounds that make explicit the dependence on $ρ$. In particular, we show that exact boundary enforcement alone is not enough for $H^2(Ω)$ error bounds, and that a sufficient and essentially necessary condition is for $ρ$ to be a smooth distance approximation $\textit{normalized to first order}$, of the kind constructed in arXiv:2104.08426 [math.NA]. We thereby identify this subclass of $\textit{boundary-adapted}$ PINNs as the appropriate neural network ansatz for solving Dirichlet BVPs. Numerical experiments support the theory, showing that appropriate choices of $ρ$ improve accuracy and convergence, while poorly chosen distance functions can substantially degrade the solution. Our proof also yields new VC-dimension bounds for hypothesis spaces of higher-order derivatives of ReQU and tanh networks, together with new approximation bounds for shallow ReQU networks in higher-order Sobolev norms, all of which are of important independent interest.
摘要
受随机过程从有界域Ω⊆ℝ^d的平均逃逸时间τ:Ω→ℝ的数值计算启发,我们研究了使用边界强制的物理信息神经网络(PINN)的椭圆Dirichlet边值问题,其中通过将网络输出乘以预定义的距离边界近似ρ来精确施加Dirichlet条件。结合分段线性二次单位(ReQU)和双曲正切(tanh)网络的逼近理论和统计学习论证,我们推导了先验误差界,明确了对ρ的依赖性。特别地,我们表明仅精确边界执行不足以获得H^2(Ω)误差界,充分且本质必要的条件是ρ是一阶归一化的光滑距离近似,类似于arXiv:2104.08426 [math.NA]中构造的类型。因此,我们确定这个适应边界的PINN子类是求解Dirichlet边值问题的适当神经网络ansatz。数值实验支持该理论,表明适当选择ρ可提高精度和收敛性,而选择不当的距离函数会显著降低解的质量。我们的证明还得到了ReQU和tanh网络高阶导数假设空间的新VC维界限,以及浅层ReQU网络在高阶Sobolev范数中的新逼近界限,所有这些都具有重要的独立意义。
#481-Lipschitz Neural Networks on Hadamard Manifolds
Hadamard流形上的1-Lipschitz神经网络
Davide Murari, Marta Ghirardelli, Ben Adcock, Elena Celledoni, Brynjulf Owren, Carola-Bibiane Schönlieb · 2026-07-21T17:54:34Z
Abstract
Controlling the Lipschitz constant of a neural network is a standard way to promote robustness and stability. Most existing constraining strategies are designed for Euclidean spaces. In this work, we construct and analyze a class of 1-Lipschitz neural networks on Hadamard manifolds. Our layers are of gradient-descent type, $1$-Lipschitz, and quasi-$α$-firmly nonexpansive. The core building blocks of the proposed architecture are Busemann functions, and we exploit the properties of Busemann gradient flows to design $1$-Lipschitz geometry-preserving layers. We provide explicit constructions and examples for hyperbolic manifolds and the manifold of symmetric positive definite (SPD) matrices. We test the proposed architecture in two numerical experiments: robust classification on the Poincaré disk and masked-Wishart covariance reconstruction. On the Poincaré disk, the proposed networks yield robust classifiers under hyperbolic perturbations. On the SPD manifold, we train SPD-valued denoisers and adopt them as a Plug-and-Play prior for a masked-Wishart covariance reconstruction problem. We show improved results from the nonexpansive denoiser over static, data-only, and Log-Euclidean denoising baselines, and empirically test its convergence properties.
摘要
控制神经网络的Lipschitz常数是促进鲁棒性和稳定性的标准方法。大多数现有的约束策略是针对欧几里得空间设计的。在这项工作中,我们构造并分析了一类在Hadamard流形上的1-Lipschitz神经网络。我们的层是梯度下降类型、1-Lipschitz和准α-牢固非扩张的。所提出架构的核心构建块是Busemann函数,我们利用Busemann梯度流的性质来设计1-Lipschitz保几何层。我们为双曲流形和对称正定(SPD)矩阵流形提供了显式构造和示例。我们在两个数值实验中测试了所提出的架构:Poincaré盘上的鲁棒分类和掩码Wishart协方差重建。在Poincaré盘上,所提出的网络在双曲扰动下产生鲁棒分类器。在SPD流形上,我们训练SPD值去噪器,并将其作为即插即用先验用于掩码Wishart协方差重建问题。我们展示了非扩张去噪器相比静态、仅数据和Log-Euclidean去噪基线的改进结果,并经验性地测试了其收敛性质。