arXiv Daily Brief

Last updated: 2026-07-24 12:02:14 (UTC+0800)

Papers: 51
Recommended: 3
Recommended 3 papers
All Papers 48 papers (excluding recommended)
#1A hierarchical sparse-grid particle method for the Vlasov--Poisson system
用于Vlasov-Poisson系统的分层稀疏网格粒子方法
Fabrice Deluzet, Clément Guillet, Jacek Narski · 2026-07-22T08:34:33Z
Abstract
We introduce a hierarchical sparse-grid (HSG) particle method for the numerical solution of the Vlasov--Poisson system. Sparse-grid PIC methods have so far been formulated within finite-difference frameworks, most notably through the sparse-grid combination technique (SGCT), which ties them to tensor-product Cartesian grids and globally defined component grids. This paper brings sparse-grid particle methods into the Galerkin setting: the field equation is solved in variational form on a hierarchical sparse-grid space spanned by B-splines of arbitrary degree, and the traditional charge deposition step is replaced by a direct Galerkin projection of the raw Monte Carlo density estimator onto this space. Beyond preserving the mesh-complexity and noise-reduction benefits of sparse grids, this reformulation substantially extends the versatility of the approach, opening the way to spatial adaptivity and to non-rectangular geometries, which are notoriously difficult to accommodate within the SGCT framework. We carry out a probabilistic error analysis decomposing the numerical error into a grid-based bias and a statistical noise component. Under mixed-derivative regularity assumptions on the particle distribution, the bias of the charge density in the $\mathrm{L}^2$-norm is shown to scale as $\mathcal{O}(h^{p+1}|\log h|^{d-1})$, where $p$ is the B-spline degree and $d$ the spatial dimension, and the statistical error in the $\mathrm{L}^1$-norm as $\mathcal{O}(|\log h|^{(d-1)/2}(Nh)^{-1/2})$, matching the accuracy of high-order SGCT-PIC methods. Corresponding bounds are derived for the electric field. The theoretical estimates are validated on classical kinetic plasma benchmarks, including configurations with limited regularity and strong anisotropies that are known to be challenging for sparse-PIC approximations.
摘要
我们介绍了一种用于Vlasov-Poisson系统数值求解的分层稀疏网格(HSG)粒子方法。迄今为止,稀疏网格PIC方法已在有限差分框架下建立,最著名的是通过稀疏网格组合技术(SGCT),该方法将其限制在张量积笛卡尔网格和全局定义的组件网格上。本文将稀疏网格粒子方法引入Galerkin框架:场方程在由任意次B样条张成的分层稀疏网格空间上以变分形式求解,并且传统的电荷沉积步骤被原始蒙特卡洛密度估计器在该空间上的直接Galerkin投影所取代。除了保留稀疏网格在网格复杂度和降噪方面的优势,这种重新表述大大扩展了该方法的通用性,为空间自适应性和非矩形几何区域开辟了道路,而这些在SGCT框架内是出了名的难以处理。我们进行了概率误差分析,将数值误差分解为基于网格的偏差和统计噪声分量。在粒子分布具有混合导数正则性的假设下,电荷密度的$\mathrm{L}^2$范数偏差呈$\mathcal{O}(h^{p+1}|\log h|^{d-1})$量级,其中$p$是B样条次数,$d$是空间维度;$\mathrm{L}^1$范数的统计误差为$\mathcal{O}(|\log h|^{(d-1)/2}(Nh)^{-1/2})$,与高阶SGCT-PIC方法的精度相匹配。电场也推导出了相应的界。这些理论估计在经典动力学等离子体基准测试中得到了验证,包括那些正则性有限和强各向异性(已知对稀疏PIC逼近具有挑战性)的配置。
#2Hierarchical Log-Gaussian Relaxation on a Fixed D3Q125 Velocity Set
固定D3Q125速度集上的分层对数高斯松弛
Bjørn Wu · 2026-07-23T02:13:37Z
Abstract
We develop a hierarchical order-resolved relaxation model for a fixed D3Q125 discrete-velocity kinetic formulation. Conventional adaptive collision models often use one scalar rarefaction or nonequilibrium indicator for all retained moment orders, thereby coupling distinct kinetic sectors. Here, a shared macroscopic-gradient background is combined separately with second-, third-, and fourth-order thermodynamic nonequilibrium indicators to define effective measures K2, K3, and K4, each driving its own log-Gaussian relaxation spectrum. Pure-order perturbation tests verify selective activation, with nonmatching sectors remaining at roundoff level. Homogeneous mixed-order, amplitude, and composition tests show lower residual nonequilibrium than a common-sensor model while preserving positive populations. In a smooth periodic compression wave at the stated reference discretization and in the TNE-only sensor limit, the peak total nonequilibrium intensity is reduced by 6.565%, with reductions throughout the domain and in all retained moment sectors. Additional timestep, transport-discretization, relaxation-spectrum, uniform-boost, long-time, and shear-wave studies show that the sign of the hierarchical correction is robust over the tested configurations, while its magnitude depends on timestep, transport scheme, sensor frame, and relaxation spectrum. The periodic benchmarks preserve the principal global invariants to floating-point accuracy and remain positive. These results establish the mechanism and numerical behavior of order-resolved activation on a fixed velocity set; independent kinetic-reference validation is still required before claiming universal accuracy improvement.
摘要
我们针对固定的D3Q125离散速度动力学公式,开发了一种分层阶次分辨的松弛模型。传统的自适应碰撞模型通常对所有保留的矩阶使用单一的标量稀薄度或非平衡指示器,从而耦合了不同的动力学扇区。这里,一个共享的宏观梯度背景分别与二阶、三阶和四阶热力学非平衡指示器结合,定义了有效度量K2、K3和K4,每个度量驱动其自身的对数高斯松弛谱。纯阶次扰动测试验证了选择性激活,不匹配的扇区保持在舍入误差水平。均匀混合阶次、幅度和组成测试显示,与公共传感器模型相比,残留非平衡更低,同时保持群体为正。在指定参考离散化和仅TNE传感器极限下的光滑周期压缩波中,峰值总非平衡强度降低了6.565%,并在整个域和所有保留矩扇区中均有降低。额外的时间步长、输运离散化、松弛谱、均匀加速、长时间和剪切波研究表明,分层校正的符号在测试配置中保持稳健,而其大小取决于时间步长、输运方案、传感器框架和松弛谱。周期基准测试将主要全局不变量保持到浮点精度,并保持为正。这些结果确立了固定速度集上阶次分辨激活的机制和数值行为;在声称通用精度改进之前,仍需要独立的动力学参考验证。
#3Relative entropy analysis of the Jin-Xin model: Theory and Numerics
Jin-Xin模型的相对熵分析:理论与数值方法
Christina Mahmoud · 2026-07-22T08:09:25Z
Abstract
This work studies the linear Jin-Xin relaxation model by means of the relative entropy method, both at the continuous and fully discrete levels. Under the strict subcharacteristic condition, we obtain an O($ε$) estimate for the convergence toward the equilibrium transport equation. We then develop discrete relative entropy estimates for a Lie splitting scheme and an asymptotic preserving staggered scheme. A fixed-time analysis further explains the O($ε$^2 ) behavior observed in the strongly stiff regime and the transition occurring when $ε$ becomes comparable to the time step.
摘要
本文利用相对熵方法研究了线性Jin-Xin松弛模型,包括连续层面和全离散层面。在严格的子特征条件下,我们得到了收敛到平衡输运方程的$\mathcal{O}(ε)$估计。然后,我们为Lie分裂格式和渐近保持交错格式建立了离散相对熵估计。固定时间分析进一步解释了强刚性区域中观察到的$\mathcal{O}(ε^2)$行为以及当$ε$与时间步长相当时发生的转变。
#4Fast reconstruction of tensor tomographic X-ray scattering data for real-time applications
面向实时应用的张量断层扫描X射线散射数据快速重建
André M. Antunes, Daniël M. Pelt, K. Joost Batenburg · 2026-07-21T14:12:27Z
Abstract
X-ray scattering tensor tomography reveals nanoscale structural orientation in 3D, but its reliance on slow iterative reconstruction limits real-time use. We introduce an extension of a direct reconstruction approach for parallel-beam geometries that computes algebraic filters approximating multiple iterative updates with a single filtering and back-projection step. By explicitly separating the tomographic projector from a view-dependent mixing operator, the method generalizes across tensor representations and modes of (small-angle) X-ray scattering measurement. Simulated and experimental results show that the approach approximates iterative reconstructions while reducing computation time by over an order of magnitude, realizing a reconstruction of a 53x53x53x28 tensor volume in 1 second on commercially-available hardware. The resulting speed and stability enable high-throughput analysis and open the door to real-time scattering-based imaging.
摘要
X射线散射张量断层扫描可以揭示三维纳米级结构取向,但其依赖缓慢的迭代重建限制了实时应用。我们介绍了一种对平行束几何直接重建方法的扩展,该方法计算代数滤波器,通过单次滤波和反投影步骤近似多次迭代更新。通过明确分离断层扫描投影仪与视角相关的混合算子,该方法可推广至不同张量表示和(小角)X射线散射测量模式。模拟和实验结果表明,该方法近似迭代重建的同时,将计算时间降低了一个数量级以上,在商用硬件上1秒内实现了53x53x53x28张量体积的重建。由此产生的速度和稳定性使高通量分析成为可能,为实时散射成像打开了大门。
#5Bilinear Systems with Quadratic Outputs: $\mathcal{H}_2$ Analysis, Optimality Conditions for Model Reduction, and Algorithmic Solutions
二次输出的双线性系统:$\mathcal{H}_2$分析、模型降阶的最优性条件及算法解
Heike Faßbender, Serkan Gugercin, Till Peters · 2026-07-21T18:44:26Z
Abstract
Bilinear systems with quadratic outputs (BQO) have recently emerged as an important system class, arising naturally in applications where both the dynamics and the quantities of interest depend nonlinearly on the state. Despite the growing interest in this class of systems, a systematic $\mathcal{H}_2$ framework for BQO systems has been lacking. In this paper, we develop such a framework by establishing an $\mathcal{H}_2$ inner product and norm for BQO systems, deriving output bounds in terms of the $\mathcal{H}_2$ norm, and obtaining first-order optimality conditions for $\mathcal{H}_2$ optimal model reduction. Building on these theoretical foundations, we propose an algorithm that computes a reduced BQO system satisfying these optimality conditions, and thus generalizing existing $\mathcal{H}_2$ optimal methods for bilinear and linear quadratic output systems. The effectiveness of the proposed framework is demonstrated on two numerical examples.
摘要
具有二次输出的双线性系统(BQO)最近已成为一种重要的系统类别,自然出现在动力学和感兴趣的量都非线性地依赖于状态的应用中。尽管对该类系统的兴趣日益增长,但一直缺乏针对BQO系统的系统化$\mathcal{H}_2$框架。在本文中,我们通过建立BQO系统的$\mathcal{H}_2$内积和范数,推导出基于$\mathcal{H}_2$范数的输出界,并得到$\mathcal{H}_2$最优模型降阶的一阶最优性条件,从而发展了这样一个框架。在这些理论基础之上,我们提出了一种算法,该算法计算满足这些最优性条件的简化BQO系统,从而推广了现有的针对双线性和线性二次输出系统的$\mathcal{H}_2$最优方法。通过两个数值示例证明了所提出框架的有效性。
#6Robust Hierarchical Matrix Compression of Acoustic Volume and Boundary Integral Operators
声学体积和边界积分算子的鲁棒分层矩阵压缩
Alberto Almuna-Morales, Danilo Aballay, Ignacio Labarca-Figueroa, Elwin van 't Wout · 2026-07-21T18:25:10Z
Abstract
Discretizing integral formulations of the Helmholtz equation yields dense linear systems. Hence, simulating acoustic models at larger scales or higher frequencies is typically constrained by memory capacity. Fast algorithms, such as hierarchical matrix compression, reduce the memory footprint substantially while controlling the approximation error in matrix-vector multiplications. However, the commonly used Adaptive Cross Approximation suffers from early-convergence problems, where the iterative construction of low-rank decompositions stops before reaching the targeted error tolerance. This failure arises when the error estimator does not capture significant components of the matrix structure under partial pivoting. This manuscript proposes a new diagonal convergence criterion, additional matrix elements for the pivoting strategy, an extended admissibility condition, and a sustained convergence check to improve the robustness of hierarchical matrix compression. These modifications improve compression reliability without increasing memory. We tested our compression strategy on various discretized volume and boundary integral operators. The computational results show that our approach successfully compresses all benchmark matrices within predefined tolerances, thereby resolving the early-convergence issues encountered in standard algorithms. This robust matrix compression was achieved at the same memory footprint as alternative compression strategies. Furthermore, a complexity analysis shows log-linear memory scaling with mesh refinement at constant frequency. Finally, we successfully applied our robust matrix compression algorithm to a coupled system of volume and boundary integral operators that models transcranial ultrasound propagation. This confirms the feasibility of our robust algorithm to accelerate large-scale simulations with high-resolution meshes in a biomedical application.
摘要
离散Helmholtz方程的积分形式会得到稠密线性系统。因此,在更大尺度或更高频率下模拟声学模型通常受到内存容量的限制。快速算法,如分层矩阵压缩,在控制矩阵向量乘法中的近似误差的同时,显著减少了内存占用。然而,常用的自适应交叉逼近会遇到早期收敛问题,即在达到目标误差容限之前,低秩分解的迭代构建就停止了。当误差估计器无法在部分主元消去下捕捉到矩阵结构的重要分量时,就会发生这种失败。本文提出了一种新的对角收敛准则、用于主元选择策略的附加矩阵元素、扩展的可接受性条件以及持续收敛检查,以提高分层矩阵压缩的鲁棒性。这些修改在不增加内存的情况下提高了压缩可靠性。我们在各种离散化的体积和边界积分算子上测试了我们的压缩策略。计算结果表明,我们的方法成功地将所有基准矩阵压缩在预定义容限内,从而解决了标准算法中遇到的早期收敛问题。这种鲁棒矩阵压缩是在与替代压缩策略相同的内存占用下实现的。此外,复杂度分析显示,在恒定频率下,随着网格细化,内存呈对数线性缩放。最后,我们成功地将我们的鲁棒矩阵压缩算法应用于一个耦合的体积和边界积分算子系统,该系统模拟经颅超声传播。这证实了我们的鲁棒算法在生物医学应用中加速具有高分辨率网格的大规模仿真的可行性。
#7Space-time tensor-product finite element methods for parabolic problems
抛物问题的时空张量积有限元方法
Richard Löscher, Michael Reichelt, Olaf Steinbach · 2026-07-21T13:48:48Z
Abstract
We study space-time Galerkin--Petrov formulations for parabolic evolution problems and their relation to classical implicit time-stepping schemes. Although such schemes are stable in the usual time-stepping sense, their interpretation as space-time operator equations may lead to conditional stability, with constants depending on the relation between temporal and spatial mesh sizes. We revisit this phenomenon for the continuous Galerkin method of Aziz and Monk, which yields the Crank--Nicolson scheme in the lowest-order case, and provide a detailed space-time error analysis for solutions of both high and low regularity. In particular, the space-time framework allows us to analyze the deteriorated behaviour of classical time-stepping methods for nonsmooth initial data. By applying integration by parts in time, we derive an adjoint space-time formulation that incorporates the initial condition in a natural variational way. In the lowest-order case, this formulation leads to a Rannacher-type smoothing of the initial data. The theoretical results are complemented by numerical experiments.
摘要
我们研究了抛物型演化问题的时空伽辽金-彼得罗夫公式及其与经典隐式时间步进方案的关系。尽管这些方案在通常的时间步进意义下是稳定的,但将其解释为时空算子方程可能导致条件稳定性,其常数取决于时间和空间网格尺寸的关系。我们针对Aziz和Monk的连续伽辽金方法重新审视了这一现象,该方法在最低阶情况下产生克兰克-尼科尔森格式,并提供了高和低正则性解的详细时空误差分析。特别是,时空框架使我们能够分析经典时间步进方法对非光滑初值的退化行为。通过时间分部积分,我们推导出一个伴随时空公式,以自然变分方式纳入初始条件。在最低阶情况下,该公式导致对初值的Rannacher型平滑。数值实验补充了理论结果。
#8High-Order Exponential Integrators with Improved Uniform Accuracy for Charged-Particle in a Perpendicular Strong Magnetic Field
用于垂直强磁场中带电粒子的具有改进一致精度的高阶指数积分器
Zhihao Qi, Weibing Deng · 2026-07-23T09:27:47Z
Abstract
This paper considers a class of charged-particle dynamics problems in which the particle is subjected to a magnetic force, with a magnetic flux density inversely proportional to a small parameter $0<\varepsilon\ll 1$, and a nonlinear electric force. The resulting highly oscillatory behavior poses significant challenges for numerical computation. To enhance the performance of exponential integrators (EIs), this paper employs a technique that linearizes the ordinary differential equation through a dimension-raising approach. Based on this technique, a new family of EIs is developed that achieves arbitrarily high order. For short-time simulations on the interval $[0,T]$, it is rigorously proved that the proposed method--which employs auxiliary polynomials of degree $k$ and a time step $Δt$--satisfies two distinct error bounds: $O(\varepsilon Δt^{k+1})$ and $O(\varepsilon^{k+2})$. The latter bound guarantees that the algorithm stays accurate even when the step size is of order $O(1)$. Furthermore, when a large step size $\varepsilon^{-1}Δt$ is used to simulate the long-term dynamics over $[0,\varepsilon^{-1}T]$, the numerical scheme attains a uniform convergence rate of $O(Δt^{k+1})$. Several numerical experiments confirm these theoretical results.
摘要
本文考虑一类带电粒子动力学问题,其中粒子受到磁力作用,磁通密度与一个小参数$0<\varepsilon\ll 1$成反比,以及非线性电力。由此产生的高度振荡行为给数值计算带来了重大挑战。为了提升指数积分器(EI)的性能,本文采用了一种通过维数提升方法线性化常微分方程的技术。基于该技术,开发了一系列新的EI,可实现任意高阶。对于区间$[0,T]$上的短时模拟,严格证明了所提出的方法(使用$k$次辅助多项式和时间步长$\Delta t$)满足两个不同的误差界:$O(\varepsilon \Delta t^{k+1})$和$O(\varepsilon^{k+2})$。后者保证了即使在步长量级为$O(1)$时算法仍保持精确。此外,当使用大步长$\varepsilon^{-1}\Delta t$模拟$[0,\varepsilon^{-1}T]$上的长期动力学时,数值格式达到了$O(\Delta t^{k+1})$的一致收敛速度。数值实验证实了这些理论结果。
#9yancc: A GPU-accelerated, differentiable solver for neoclassical transport in tokamaks and stellarators
yancc:用于托卡马克和仿星器中新经典输运的GPU加速可微求解器
Rory Conlin, Matt Landreman · 2026-07-23T02:33:16Z
Abstract
We present yancc, a new GPU-accelerated solver for the drift kinetic equation that computes neoclassical transport fluxes, flows, and currents in tokamaks and stellarators. The drift kinetic equation is challenging to solve numerically due to strong advection-dominance, recirculating flows, internal boundary layers, severe anisotropy, and high dimensionality. The code solves both the full four-dimensional drift kinetic equation (retaining speed-dependent collisions, energy scattering, and full interspecies coupling), and the reduced monoenergetic form. The discretization combines a Maxwell polynomial collocation grid in speed with finite differences in pitch angle and the flux surface coordinates, using a modified upwind stencil designed to improve diagonal dominance for multigrid efficiency. The resulting linear system is solved with a multigrid-preconditioned Krylov method. Built in JAX, yancc is fully differentiable, enabling gradient-based optimization and adjoint sensitivity analysis. Benchmarks against MONKES and SFINCS show agreement within 1\% across a range of collisionalities, geometries, and multi-species configurations. yancc achieves roughly an order of magnitude speedup over SFINCS on a per-scan basis while using an order of magnitude less memory, with runtime remaining nearly flat across the full range of collisionality. The combination of speed, low memory footprint, and differentiability makes yancc well suited for integration into stellarator optimization workflows, uncertainty quantification, and profile prediction.
摘要
我们介绍yancc,一种新的GPU加速漂移动力学方程求解器,用于计算托卡马克和仿星器中的新经典输运通量、流和电流。漂移动力学方程由于强对流主导、再循环流、内部边界层、严重各向异性和高维度而难以数值求解。该代码求解完整的四维漂移动力学方程(保留速度依赖碰撞、能量散射和完整种类间耦合)以及简化的单能形式。离散化结合了速度上的Maxwell多项式配置网格、俯仰角和磁面坐标上的有限差分,使用改进的上风模板以提高对角占优性,从而提升多重网格效率。得到的线性系统通过多重网格预条件Krylov方法求解。yancc基于JAX构建,完全可微,支持基于梯度的优化和伴随敏感性分析。与MONKES和SFINCS的基准测试表明,在不同碰撞率、几何和多种类配置下,一致性在1%以内。yancc在每次扫描基础上比SFINCS快大约一个数量级,同时内存使用量少一个数量级,且运行时间在全部碰撞率范围内几乎保持平坦。速度、低内存占用和可微性的结合使yancc非常适合集成到仿星器优化工作流、不确定性量化和轮廓预测中。
#10Structure-Preserving Spectral Dynamic Programming on Compact Lie Groups
紧李群上的保结构谱动态规划
Shanqing Liu, Yang Qi · 2026-07-23T02:23:57Z
Abstract
We study spectral approximations of the dynamic programming semigroup for finite-horizon optimal control on a connected compact Lie group $G$, and of the associated first-order Hamilton-Jacobi-Bellman equation. The Bellman operator is monotone and non-expansive in the supremum norm, while the Peter-Weyl decomposition of $L^{2}(G)$, on which every Fourier method on $G$ rests, is orthogonal, and the mismatch is quantitative. The natural sup-norm error recursion of the Galerkin iteration is amplified at every step by the Lebesgue constant of the spectral projection, which grows logarithmically on $S^{1}$ and polynomially on compact Lie groups of rank one, including $\mathrm{SO}(3)$, and in computation the iteration violates elementary bounds within a few steps. We restore the dynamic programming structure at the discrete level by replacing the orthogonal projection with spectral filters of Markov type. An auxiliary heat-kernel/vanishing-viscosity scheme yields qualitative sup-norm convergence for Lipschitz data. The main result is a Fejér-type filter on $G$, finite-rank, positivity preserving and non-expansive, together with a convergence theorem at the rate $O(\sqrtδ+\sqrt{ε+1/(δN^2)})$ for Lipschitz data, where $δ$ is the time step, $N$ the spectral resolution and $ε$ the viscosity. The viscosity may be zero, and the coupling $δ=N^{-1}$ then gives the rate $N^{-1/2}$. The proof interprets the filter as a small random perturbation of the controlled dynamics, requires neither a priori regularity of the value function nor a consistency argument in the viscosity sense for the filtering step, and extends to a fully discrete realization based on positive cubature, with exact Wigner transport on $\mathrm{SO}(3)$. Numerical experiments confirm the predicted rates and filter bias and quantify the frame dependence of two chart-based baselines.
摘要
我们研究连通紧李群$G$上有限时域最优控制的动态规划半群的谱逼近,以及相关的一阶Hamilton-Jacobi-Bellman方程。Bellman算子在最大模范数下是单调和非膨胀的,而$L^{2}(G)$的Peter-Weyl分解(所有傅里叶方法在$G$上的基础)是正交的,这种不匹配是定量的。Galerkin迭代的自然上确界误差递推在每一步被谱投影的勒贝格常数放大,该常数在$S^{1}$上对数增长,在一阶紧李群(包括$\mathrm{SO}(3)$)上多项式增长,并且在计算中迭代在几步内违反基本界。我们通过用Markov型谱滤波器替换正交投影来恢复离散层面的动态规划结构。一个辅助的热核/消失粘性方案为Lipschitz数据提供了定性的上确界收敛。主要结果是在$G$上的Fejér型滤波器,具有有限秩、保正性和非膨胀性,以及Lipschitz数据下的收敛定理,收敛率为$O(\sqrtδ+\sqrt{ε+1/(δN^2)})$,其中$δ$是时间步长,$N$是谱分辨率,$ε$是粘性。粘性可以为零,此时耦合$δ=N^{-1}$给出$N^{-1/2}$的收敛率。证明将该滤波器解释为受控动力学的小随机扰动,不需要价值函数的先验正则性或滤波步骤的粘性意义下的一致性论证,并扩展到基于正立方体的全离散实现,在$\mathrm{SO}(3)$上具有精确的Wigner输运。数值实验证实了预测的收敛率和滤波器偏差,并量化了两个基于图基线的框架依赖性。
#11Flux-Corrected Diagonal Frog: second order and positivity at all time steps
通量校正对角蛙跳:所有时间步长上的二阶和正性
Andrey Itkin · 2026-07-22T17:53:40Z
Abstract
By Godunov's theorem, linear second-order finite-difference schemes for the Fokker-Planck equation cannot preserve positivity. The Diagonal Frog (DF) framework previously bypassed this barrier using eventual positivity, but required a strict minimum time step. This paper resolves the small-step limitation using a nonlinear extension of the DF solvers. We split the second-order directional operator into a monotone M-matrix core and an antidiffusive flux correction. A Zalesak-type limiter is then applied iteratively within the implicit banded solve. The resulting Flux-Corrected DF (FCDF) schemes (variants A and B) are unconditionally positive across all time steps. Because the limiter acts on fluxes rather than point values, these schemes conserve discrete mass exactly and maintain second-order accuracy. Crucially, the limiter activates only within unresolved layers. This ensures the global $L_1$ convergence remains second-order uniformly in the cell Péclet number, avoiding the first-order degradation seen in the Chang-Cooper scheme. The method's Picard iteration is contractive under a purely convective step restriction. To support arbitrary step sizes, we develop an active-set reformulation. This solves the system using a semismooth Newton iteration, where computational cost scales only with the number of nodes where positivity binds. Finally, we introduce a defect-corrected time stepping approach that restores second-order time accuracy. Numerical experiments on Ornstein-Uhlenbeck and advection-dominated benchmarks confirm our claims.
摘要
根据Godunov定理,用于Fokker-Planck方程的线性二阶有限差分格式无法保持正性。对角蛙跳(DF)框架先前通过最终正性绕过了这一障碍,但需要严格的最小时间步长。本文通过DF求解器的非线性扩展解决了小步长限制。我们将二阶方向算子分裂为一个单调M矩阵核心和一个反扩散通量校正。然后,在隐式带状求解中迭代应用Zalesak型限制器。由此产生的通量校正DF(FCDF)格式(变体A和B)在所有时间步长上无条件保持正性。由于限制器作用于通量而非点值,这些格式精确保持离散质量,并保持二阶精度。关键的是,限制器仅在未解析层内激活。这确保了全局$L_1$收敛性在单元Péclet数上保持二阶均匀,避免了Chang-Cooper方案中看到的一阶退化。该方法的Picard迭代在纯粹对流步长限制下是压缩的。为支持任意步长,我们开发了一个活动集重新表述。这使用半光滑牛顿迭代求解系统,其中计算成本仅与正性约束的节点数成比例。最后,我们引入了一种缺陷校正时间推进方法,恢复了二阶时间精度。关于Ornstein-Uhlenbeck和对流主导基准测试的数值实验证实了我们的论断。
#12A Third-Order Maximum-Principle-Preserving CWENO Scheme for Two-Dimensional Nonlocal Conservation Laws
二维非局部守恒律的三阶最大原理保持CWENO格式
Anika Beckers, Jan Friedrich · 2026-07-22T09:29:40Z
Abstract
We present a third-order finite volume central WENO scheme for systems of nonlocal conservation laws in two spatial dimensions. The CWENO reconstruction of the conservative variable provides polynomials that can be evaluated in the entire domain, which is of advantage when approximating the nonlocal terms. Moreover, this method can be augmented with a limiter that preserves the maximum-principle and especially positivity of the solution.
摘要
我们提出了一种用于二维空间中非局部守恒律系统的三阶有限体积中心WENO格式。守恒变量的CWENO重构提供了可在整个域上求值的多项式,这在逼近非局部项时具有优势。此外,该方法可以增加一个限制器,以保持最大原理,特别是解的正性。
#13A Hyper-Reduced Neural Network-Augmented Semi-smooth Newton Method for Nonlinear Parametric Variational Inequalities
非线性参数变分不等式的超缩减神经网络增强半光滑牛顿方法
Sofiane Ezzehi, Virginie Ehrlacher, Guillaume Enchéry, Thibault Faney · 2026-07-23T12:04:26Z
Abstract
We propose a model order reduction framework for nonlinear parametrized variational inequalities arising in computational mechanics. The high-dimensional model is written in mixed primal-dual form with projection-based complementarity conditions, leading to nonlinear nonsmooth algebraic systems solved by a semi-smooth Newton method in primal-dual form. On this basis, reduced models are constructed by proper orthogonal decomposition (POD) of both primal and dual solution snapshots, and the resulting reduced systems are solved by semi-smooth Newton iterations in the reduced space. To address cases where low-dimensional linear spaces provide limited approximation efficiency, we introduce a neural-network-augmented reduced model. Two feedforward networks learn corrections in the truncated POD coordinates of the primal and dual variables, defining a nonlinear manifold approximation that is embedded directly in the semi-smooth Newton iterations. The online cost associated with high-dimensional residual evaluations is reduced through hyper-reduction, using a sparse cubature approach based on greedy nonnegative least squares. Particular attention is paid to the interaction between hyper-reduction and the learned nonlinear manifold. The proposed methodology is assessed on two nonlinear variational inequalities with distinct sources of nonlinearity: a two-dimensional obstacle problem with a cubic nonlinearity in the state equation, and a three-dimensional frictional contact problem in which the Coulomb law induces a nonlinear projection in the constraint equation. Numerical results compare the high-dimensional model, the linear reduced model, the neural-network-augmented reduced model, and their hyper-reduced variants, demonstrating accurate approximations with substantial reductions in online computational cost.
摘要
我们提出了一种用于计算力学中非线性参数变分不等式的模型降阶框架。高维模型以混合原-对偶形式写出,带有基于投影的互补条件,导致通过原-对偶形式的半光滑牛顿法求解的非线性非光滑代数系统。在此基础上,通过原变量和对偶变量解快照的POD构建降阶模型,并在降阶空间中通过半光滑牛顿迭代求解得到的降阶系统。为解决低维线性空间逼近效率有限的情况,我们引入了一种神经网络增强的降阶模型。两个前馈网络学习原变量和对偶变量截断POD坐标中的修正,定义了直接嵌入半光滑牛顿迭代的非线性流形逼近。通过基于贪婪非负最小二乘的稀疏求积方法,降低了与高维残差评估相关的在线成本。特别关注超缩减与学习到的非线性流形之间的相互作用。所提出的方法在两个具有不同非线性源的非线性变分不等式上进行了评估:一个是在状态方程中具有三次非线性的二维障碍问题,另一个是库仑摩擦定律在约束方程中引入非线性投影的三维摩擦接触问题。数值结果比较了高维模型、线性降阶模型、神经网络增强降阶模型及其超缩减变体,展示了精确的逼近以及在线计算成本的大幅降低。
#14A Structure-Adaptive Random Feature Method for High-Dimensional Elliptic PDEs
一种用于高维椭圆PDE的结构自适应随机特征方法
Jiale Linghu, Hao Dong, Yangshuai Wang · 2026-07-22T06:10:49Z
Abstract
Random-feature methods reduce high-dimensional elliptic PDE collocation to linear coefficient problems, but full-dimensional trial spaces overlook lower-dimensional structure. We introduce the Hierarchical Analysis-of-Variance Random Feature Method (HA-RFM), which selects coordinate blocks using closed Sobol indices of the PDE residual, identifies oblique low-rank features from fitted-predictor gradients, and couples all retained features in one regularized least-squares solve. Under structural and stability hypotheses, we establish an $L^2$ error bound that links solution and residual truncation to finite-width approximation and regularized finite-sample fitting, and we derive guarantees for width and structure recovery. The resulting width is polynomial in the dimension at fixed interaction order, with dimension-independent higher-order contributions under uniform structural control. Residual screening achieves exact recovery of the prescribed three-pair support, while fitted-predictor gradients recover oblique directions through dimension $50$. In random-ridge tests, less than $1\%$ additional width reduces errors by factors of $14$-$39$ over coordinate blocks and $34$-$100$ over equal-width full-dimensional RFM. Semilinear computations extend HA-RFM through dimension $100$, while dense and distributed interactions delineate the coordinate families required for broader structure.
摘要
随机特征方法将高维椭圆PDE配置简化为线性系数问题,但全维试验空间忽略了低维结构。我们引入了分层方差分析随机特征方法(HA-RFM),该方法使用PDE残差的闭Sobol指标选择坐标块,从拟合预测器梯度中识别倾斜低秩特征,并将所有保留的特征耦合在一个正则化最小二乘求解中。在结构和稳定性假设下,我们建立了一个$L^2$误差界,该界将解和残差截断与有限宽度近似和正则化有限样本拟合联系起来,并且我们推导了宽度和结构恢复的保证。在固定交互阶数下,所得宽度是维度的多项式,在均匀结构控制下,高阶贡献与维度无关。残差筛选实现了对规定的三对支撑的精确恢复,而拟合预测器梯度通过维度50恢复了倾斜方向。在随机岭测试中,与坐标块相比,不到$1\%$的额外宽度将误差降低了14-39倍,与等宽度全维RFM相比降低了34-100倍。半线性计算将HA-RFM扩展到维度100,而稠密和分布式交互则描绘了更广泛结构所需的坐标族。
#15librla: Randomized Linear Algebra Library
librla:随机线性代数库
Adrianna Gillman, Zydrunas Gimbutas · 2026-07-22T21:13:55Z
Abstract
The library \texttt{librla} is a randomized linear algebra library that is specifically designed for the intermediate-sized matrices (of dimension up to roughly 10,000) that arise in applications such as reduced order modeling, fast direct solvers, least squares solves and, in some settings, data compression. \texttt{librla} is the first software package that is both stable and efficient in several high-level languages: MATLAB, Python and Julia. It also provides increased functionality over existing software. Specifically, it allows the user to choose to create a factorization based on a fixed rank or a desired tolerance. The factorization options include QR, SVD and the interpolative decomposition. Additionally, the factorization can be generated either with access to the matrix or access to a matrix-vector multiplication routine. Numerical results compare the Python implementation with the available PyTorch and SciPy randomized factorizations. Performance of \texttt{librla} in the three languages is comparable.
摘要
库\texttt{librla}是一个随机线性代数库,专门为中等规模矩阵(维数大约达到10,000)设计,这些矩阵出现在降阶建模、快速直接求解器、最小二乘求解以及某些情况下的数据压缩等应用中。\texttt{librla}是第一个在多种高级语言(MATLAB、Python和Julia)中既稳定又高效的软件包。它还提供了比现有软件更多的功能。具体来说,它允许用户选择基于固定秩或期望容差创建分解。分解选项包括QR、SVD和插值分解。此外,分解可以通过访问矩阵或访问矩阵-向量乘法的例程来生成。数值结果将Python实现与可用的PyTorch和SciPy随机分解进行比较。\texttt{librla}在三种语言中的性能相当。
#16A stability-preserving polytopal discontinuous Galerkin method for the Fisher-Kolmogorov model with applications to neurodegenerative diseases
一种保稳定的多面体间断Galerkin方法用于Fisher-Kolmogorov模型及其在神经退行性疾病中的应用
Paola Francesca Antonietti, Francesca Bonizzoni, Mattia Corti, Nicola De March, Salvatore Di Noto, Francesco Regazzoni · 2026-07-23T09:36:34Z
Abstract
The Fisher-Kolmogorov model is one of the most widely used models in the study of neurodegenerative diseases, owing to its simple structure as a nonlinear reaction-diffusion equation. In particular, it is commonly employed to describe proteinopathies such as Alzheimer's and Parkinson's diseases. Under suitable assumptions, non-negativity of the solution is guaranteed at the continuous level, which is physically relevant since the solution represents a relative concentration. However, this property is not generally preserved at the discrete level, potentially leading to unphysical and unstable numerical approximations. In this work, we analyze a modified version of the Fisher-Kolmogorov model that stabilizes the dynamics around the unstable equilibrium $c=0$. For the spatial discretization, we adopt a discontinuous Galerkin method on polygonal and polyhedral meshes, coupled with the Crank-Nicolson scheme for time integration. We derive stability and $a$-priori error estimates for the semi-discrete problem. The theoretical findings are supported by numerical experiments, including convergence studies in both two and three dimensions. Finally, we validate the model through simulations of $α$-synuclein diffusion in a two-dimensional agglomerated brain section, demonstrating the high-order accuracy and robustness of the proposed method.
摘要
Fisher-Kolmogorov模型是神经退行性疾病研究中最广泛使用的模型之一,因为其结构简单,是一个非线性反应扩散方程。它特别常用于描述阿尔茨海默病和帕金森病等蛋白质病变。在适当假设下,解的保非负性在连续层面得到保证,这在物理上是相关的,因为解代表相对浓度。然而,该性质在离散层面通常不保持,可能导致非物理和不稳定的数值逼近。在这项工作中,我们分析了Fisher-Kolmogorov模型的一个修正版本,该版本在不稳定平衡点$c=0$附近稳定动力学。对于空间离散,我们采用了多边形和多面体网格上的间断Galerkin方法,并结合Crank-Nicolson格式进行时间积分。我们推导了半离散问题的稳定性和先验误差估计。数值实验支持了理论发现,包括二维和三维的收敛性研究。最后,我们通过模拟二维聚集脑切片中$α$-突触核蛋白的扩散来验证模型,展示了所提方法的高阶精度和鲁棒性。
#17Jordan algebras, hemiplex numbers, and the Cholesky decomposition of arbitrary symmetric matrices
约当代数、半复数和任意对称矩阵的Cholesky分解
Alan Edelman, Timothy E. Holy · 2026-07-23T14:46:17Z
Abstract
Positive-semidefinite matrices are most efficiently factored using the Cholesky decomposition. For indefinite matrices, the Cholesky factorization does not exist, and the alternatives face greater challenges in achieving numeric stability and preservation of banded structure. Here we pursue an analogy between the requirement for positive-semidefinite matrices and the solution of the quadratic equation x^2 = c for c <= 0. It is shown that a non-associative algebra, called the hemiplex numbers, allows the Cholesky factorization to be computed for arbitrary symmetric matrices. Crucially, the hemiplex Cholesky factorization does not require pivoting for its existence or stability, allowing it to preserve banded structure. For singular matrices it produces a parametrization of the null space, and provides opportunity for truncation of nearly-null directions in a manner similar to common usage of the singular value decomposition. The hemiplex Cholesky factorization may be a practically useful addition to the tools for solving symmetric linear equations.
摘要
半正定矩阵最有效地使用Cholesky分解进行分解。对于不定矩阵,Cholesky分解不存在,而替代方法在数值稳定性和带状结构保持方面面临更大的挑战。这里我们类比二次方程$x^2 = c$在$c leq 0$时解的要求,研究半正定矩阵的要求。我们证明了一种非结合代数(称为半复数)允许对任意对称矩阵计算Cholesky分解。关键在于,半复数Cholesky分解的存在或稳定性不需要选主元,因此可以保持带状结构。对于奇异矩阵,它产生了零空间的一个参数化,并提供了类似于奇异值分解常用方式的截断近零方向的机会。半复数Cholesky分解可能是求解对称线性方程工具的一个实用补充。
#183D Uncertainty Quantification for the Photo-Acoustic Tomography
光声断层成像的三维不确定性量化
Babak Maboudi Afkham, Amal Mohammed A Alghami, Hassan Yazdanian, Tanja Tarvainen · 2026-07-23T14:38:00Z
Abstract
Photoacoustic tomography (PAT) is a promising modality for high-resolution biomedical imaging, motivating the need for reliable uncertainty quantification (UQ) of reconstructed images. Bayesian approaches provide a rigorous framework for UQ but remain computationally challenging for realistic three-dimensional PAT and are sensitive to numerical approximations in the governing wave equation. We develop a finite-element Bayesian UQ framework for PAT that accommodates complex computational domains and detector geometries while enabling large-scale three-dimensional inference. The proposed methodology reformulates the randomize-then-optimize (RTO) sampling strategy as a matrix-free algorithm that generates independent posterior samples using only forward and adjoint wave propagations. Particular attention is given to constructing an adjoint discretization that forms an exact transpose pair with the discrete forward operator while remaining consistent with the continuous PAT adjoint, enabling efficient least-squares solvers within the sampling procedure. We investigate the influence of temporal discretization, artificial boundary conditions, and adjoint consistency on posterior uncertainty and identify discretization strategies that avoid numerical artifacts. The framework is validated against exact posterior statistics, existing Bayesian PAT methods, and Hamiltonian Monte Carlo using the No-U-Turn Sampler (NUTS), and is demonstrated on a three-dimensional problem with approximately $2\times 10^5$ unknowns on a general finite-element domain. To the best of our knowledge, this is the first large-scale Bayesian PAT study on general three-dimensional finite-element geometries, and the methodology extends naturally to a broad class of linear PDE-constrained inverse problems.
摘要
光声断层成像(PAT)是一种有前景的高分辨率生物医学成像模态,因此需要对重建图像进行可靠的不确定性量化(UQ)。贝叶斯方法为UQ提供了严格的框架,但对于实际的三维PAT在计算上仍然具有挑战性,并且对控制波动方程的数值近似敏感。我们为PAT开发了一种有限元贝叶斯UQ框架,该框架能够处理复杂的计算域和探测器几何形状,同时实现大规模三维推断。所提出的方法将随机-然后-优化(RTO)采样策略重新表述为一种无矩阵算法,仅使用正向和伴随波传播即可生成独立的后验样本。特别关注构建伴随离散化,使其与离散正算子形成精确转置对,同时与连续PAT伴随保持一致,从而在采样过程中实现高效的最小二乘求解器。我们研究了时间离散化、人工边界条件和伴随一致性对后验不确定性的影响,并确定了避免数值伪影的离散策略。该框架针对精确后验统计量、现有贝叶斯PAT方法和使用No-U-Turn采样器的哈密顿蒙特卡洛进行了验证,并在一个一般有限元域上的三维问题中进行了演示,该问题约有$2\times 10^5$个未知量。据我们所知,这是首个在一般三维有限元几何上进行的大规模贝叶斯PAT研究,该方法自然扩展到更广泛的线性PDE约束反问题。
#19An unfitted boundary algebraic equation method with Calderón preconditioning for 2D Stokes flow in irregular geometry
一种带有Calderón预处理的非拟合边界代数方程方法用于不规则几何中的二维Stokes流动
Wenjun Ying, Qing Xia · 2026-07-23T13:20:27Z
Abstract
We present an unfitted boundary algebraic equation method for the two-dimensional exterior/interior Stokes equations on a staggered MAC grid. By constructing an explicit free-space pair of velocity and pressure lattice Green's functions (LGFs) from free-space Laplace LGFs, we represent homogeneous fields using sources supported exclusively on thin staggered boundary layers. This formulation imposes physical Dirichlet data at cut points via local interpolation, while sampled-normal rank updates remove hydrostatic null modes associated with single or multiple obstacles. The workflow parallels that of classical boundary integral formulations and requires no artificial boundary conditions for exterior flows, but follows a discretize-then-represent route and does not require singular/near-singular quadrature. The resulting dense boundary system is solved via GMRES, utilizing a componentwise discrete Calderón preconditioner built from the scalar Laplace kernel and padded FFTs for fast volume convolutions. Extensive numerical validation, including multiply connected domains, narrow gaps, and Moffatt eddies, confirms discrete incompressibility to solver accuracy and recovers the expected Moffatt eddy scaling. We achieve second-order velocity and pressure convergence and bound maximum discrete divergence within numerical accuracy. The discrete Calderón preconditioner reduces the condition number by orders of magnitude and yields nearly mesh-independent conditioning in exterior configurations, while remaining effective---though more demanding---for narrow-gap and fine-grid interior problems.
摘要
我们提出了一种用于交错MAC网格上二维外部/内部Stokes方程的非拟合边界代数方程方法。通过从自由空间Laplace LGFs构造显式的自由空间速度-压力对晶格格林函数(LGFs),我们使用仅支持在薄交错边界层上的源来表示均匀场。该公式通过局部插值在切割点施加物理Dirichlet数据,而采样法向秩更新消除了与单个或多个障碍物相关的静水零模态。工作流程类似于经典的边界积分公式,不需要外部流动的人工边界条件,但遵循先离散后表示的路径,并且不需要奇异/近奇异求积。得到的稠密边界系统通过GMRES求解,利用基于标量Laplace核和填充快速傅里叶变换进行快速体积卷积的分量离散Calderón预条件子。大量的数值验证,包括多连通域、窄间隙和Moffatt涡,证实了离散不可压性达到求解器精度,并恢复了预期的Moffatt涡标度。我们实现了二阶速度和压力收敛,并将最大离散散度限制在数值精度范围内。离散Calderón预条件子将条件数降低了几个数量级,并在外部配置中产生几乎与网格无关的条件数,而对于窄间隙和细网格内部问题,虽然要求更高,但仍然有效。
#20On the Approximation of the Unitary Operator Group Associated with a Rotation Matrix and Its Applications to Abstract Hyperbolic Equations
关于旋转矩阵相关幺正算子群的逼近及其在抽象双曲方程中的应用
Jemal Rogava, Zurab Vashakidze · 2026-07-23T13:15:07Z
Abstract
The solution of the Cauchy problem for homogeneous abstract hyperbolic equations, together with its derivative, admits a vector representation in terms of a unitary operator group associated with a rotation matrix. A rational approximation of this unitary group is constructed and shown to possess optimal fourth-order convergence. The order of convergence is determined in accordance with the smoothness scale. Based on this rational approximation, a two-layer semi-discrete scheme is constructed for the approximate solution of Cauchy problems for nonhomogeneous abstract hyperbolic equations in both the linear and semi-linear settings. The convergence properties of the scheme are examined in relation to the regularity of the solution.
摘要
齐次抽象双曲方程柯西问题的解及其导数可以通过与旋转矩阵相关的幺正算子群进行向量表示。构造了该幺正群的有理逼近,并证明其具有最优四阶收敛性。收敛阶根据光滑度尺度确定。基于该有理逼近,构建了一个两层半离散格式,用于近似求解线性和半线性非齐次抽象双曲方程的柯西问题。研究了该格式的收敛性质与解的正则性之间的关系。
#21Adaptive Non-Linear Partition of Unity Methods for Scattered Data Interpolation with Discontinuities
适应非线性的单位分解方法用于含间断散乱数据插值
Adeeba Haider, Roberto Cavoretto, Juan Ruiz-Álvarez, Dionisio F. Yáñez · 2026-07-23T13:07:22Z
Abstract
Scattered data approximation with discontinuities is challenging due to the Gibbs phenomenon, which significantly reduces accuracy near interfaces. The recently introduced Non-Linear Partition of Unity Method (NL-PUM) addresses this by combining Radial Basis Function (RBF) interpolation with a non-linear Weighted Essentially Non-Oscillatory (WENO) strategy. While effective, NL-PUM's performance relies heavily on two fixed hyperparameters: the RBF shape parameter and the patch radius. This work extends NL-PUM by adapting both hyperparameters locally using Leave-One-Out Cross-Validation (LOOCV) minimized via Global Optimization with Optimistic Improvement (GOOI). Our main innovation is a smoothness indicator linking a discontinuity-aware shrinkage process to LOOCV-based radius selection: patches in smooth regions remain unchanged, while those near discontinuities automatically shrink to avoid crossing the interface. The resulting method, LOOCV-NL-PUM-GOOI, requires no prior knowledge of interface geometry and introduces no extra cost beyond standard adaptive shape parameter selection. Numerical experiments on synthetic test functions with jump discontinuities and a real-data application to Norwegian Fjords elevation data confirm that this approach substantially reduces approximation error near discontinuities while preserving full accuracy in smooth regions.
摘要
含间断的散乱数据逼近由于吉布斯现象而具有挑战性,该现象显著降低了界面附近的精度。最近提出的非线性单位分解法(NL-PUM)通过结合径向基函数(RBF)插值和非线性加权本质无振荡(WENO)策略解决了这个问题。虽然有效,但NL-PUM的性能严重依赖于两个固定的超参数:RBF形状参数和补丁半径。本文通过使用留一交叉验证(LOOCV)并通过乐观改进全局优化(GOOI)最小化,局部调整这两个超参数,从而扩展了NL-PUM。我们的主要创新是一个光滑度指示器,将感知间断的收缩过程与基于LOOCV的半径选择联系起来:光滑区域的补丁保持不变,而间断附近的补丁自动收缩以避免穿过界面。由此产生的方法LOOCV-NL-PUM-GOOI不需要先验的界面几何知识,并且除了标准的自适应形状参数选择外不引入额外成本。在具有跳跃间断的合成测试函数上的数值实验以及挪威峡湾高程数据的实际应用证实,该方法显著减少了间断附近的逼近误差,同时在光滑区域保持了完整的精度。
#22A Mixed Discrete Cosserat Rod Formulation
一种混合离散Cosserat杆公式
Tianxiang Dai, Marco Herrmann, Jonas Breuling, Remco I. Leine, Simon R. Eugster · 2026-07-23T10:53:30Z
Abstract
In this communication we propose a discrete Cosserat rod formulation in which a slender elastic rod is represented as a chain of rigid bodies (nodes) coupled by compliant elastic forces and moments acting between adjacent node pairs. Discrete dilatation, shear, torsion and curvature strain measures are evaluated from the relative kinematics of each node pair, while the constitutive behavior is expressed in compliance form through independent stress degrees of freedom. We show that the resulting model arises rigorously from a mixed Petrov--Galerkin Cosserat rod finite element formulation (FEM) at linear kinematic interpolation order when the internal virtual work is integrated by the midpoint rule and the external and inertial contributions by the trapezoidal rule. The proposed formulation inherits the robustness and the absence of locking from the underlying mixed FEM while simultaneously exposing a two-node coupling structure that mirrors discrete rod models from the computer graphics community. This is in sharp contrast to the dense coupling of strain-parameterized reduced-order models often used in soft robotic applications. Three numerical examples involving piecewise-varying cross sections, tendon-driven actuation under different spatial discretizations, and coupled longitudinal-torsional dynamics confirm the accuracy, robustness, and convergence behavior of the presented approach.
摘要
在本文中,我们提出了一种离散的Cosserat杆公式,其中细长弹性杆被表示为一系列刚体(节点),通过相邻节点对之间的柔性弹性力和力矩耦合。离散的膨胀、剪切、扭转和曲率应变度量由每个节点对的相对运动学评估,而本构行为通过独立的应力自由度以柔度形式表达。我们证明,当内部虚功由中点积分、外部和惯性贡献由梯形积分时,所得到的模型严格来自线性运动学插值阶的混合Petrov-Galerkin Cosserat杆有限元公式。所提出的公式继承了底层混合FEM的鲁棒性和无锁特性,同时暴露了与计算机图形学社区中的离散杆模型相似的两节点耦合结构。这与软体机器人应用中常用的应变参数化降阶模型的密集耦合形成鲜明对比。三个数值例子涉及分段变化的横截面、不同空间离散化下的肌腱驱动以及耦合的纵向-扭转动力学,证实了所提出方法的准确性、鲁棒性和收敛行为。
#23A Damped SWIFT Method for European Option Pricing: Coefficients Decay, Truncation, and Error Analysis
用于欧式期权定价的阻尼SWIFT方法:系数衰减、截断和误差分析
Davide Trevisani, José Germán López Salas, Chiheb Ben Hammouda, Cornelis W. Oosterlee · 2026-07-23T09:38:32Z
Abstract
We introduce a damped variant of the Shannon Wavelet Inverse Fourier Technique (SWIFT) for pricing European options when the characteristic function of the underlying model is available. The key idea is to apply an exponential damping transformation to the payoff, which enables the direct computation of Fourier coefficients in the frequency domain without introducing an additional physical-domain truncation parameter. We provide a rigorous analysis of the decay of these coefficients by exploiting the singularity structure of the associated Fourier transforms. For light-tailed models, we obtain Gaussian-type decay estimates, while for semi-heavy and heavy-tailed models whose singularities are poles, algebraic branch points, or logarithmic branch points, we derive exponential decay bounds with explicit polynomial prefactors. The resulting sharp bounds make it possible to truncate the Fourier series without relying on the cumulants of the underlying density, which are often unavailable or difficult to compute in practice. We further derive an error decomposition separating projection, truncation, and quadrature errors, and translate the analysis into practical rules for selecting the damping parameter, resolution level, and truncation range. Numerical experiments demonstrate that the proposed approach consistently improves the accuracy of the original SWIFT method while requiring a significantly smaller number of Fourier coefficients and remaining stable in cases where the undamped method deteriorates.
摘要
我们介绍了一种阻尼变体的香农小波逆傅里叶技术(SWIFT),用于在基础模型的特征函数可用时对欧式期权进行定价。关键思想是对收益应用指数阻尼变换,从而能够在频域中直接计算傅里叶系数,而无需引入额外的物理域截断参数。我们通过利用相关傅里叶变换的奇异性结构,对这些系数的衰减进行了严格分析。对于轻尾模型,我们获得了高斯型衰减估计;而对于奇点为极点、代数分支点或对数分支点的半重尾和重尾模型,我们推导了具有显式多项式前因子的指数衰减界。由此得到的尖锐界使得无需依赖基础密度的累积量(实践中通常不可用或难以计算)即可截断傅里叶级数。我们进一步推导了将投影误差、截断误差和求积误差分离的误差分解,并将分析转化为选择阻尼参数、分辨率水平和截断范围的实际规则。数值实验表明,所提出的方法在需要显著更少的傅里叶系数的同时,持续提高了原始SWIFT方法的精度,并且在无阻尼方法退化的情形下保持稳定。
#24A virtual element method for Darcy's problem coupled with a heat equation
一种用于Darcy问题与热方程耦合的虚拟元方法
Danilo Amigo, Felipe Lepe · 2026-07-22T22:08:57Z
Abstract
In two dimensions, we develop a virtual element method to solve the Darcy's problem coupled with a nonlinear heat equation. This coupled model may allow for thermal diffusion and viscosity as a function of temperature. Under standard discretization assumptions and appropriate assumptions on the data, we prove the well posedness of the proposed numerical scheme. We also derive optimal error estimates under suitable regularity assumptions for the solution and appropriate assumptions on the data. We conclude with a series of numerical tests performed on different families of meshes that complement the theoretical findings.
摘要
在二维中,我们开发了一种虚拟元方法来求解Darcy问题与非线性热方程的耦合。该耦合模型允许热扩散和粘度作为温度的函数。在标准离散化假设和数据适当假设下,我们证明了所提议数值格式的适定性。我们还在解的正则性假设和数据适当假设下推导了最优误差估计。最后,我们在不同系列的网格上进行了一系列数值测试,对理论结果进行了补充。
#25Convergence and mixed-precision preconditioning for the naive Jacobi eigenvalue algorithm
朴素雅可比特征值算法的收敛性与混合精度预处理
Erna Begovic, Marija Miloloza Pandur, Ana Perkovic · 2026-07-22T22:02:04Z
Abstract
The paper studies a Jacobi-type method for the eigenvalue problem of general complex matrices with simple eigenvalues. The method applies elementary triangular similarity transformations in order to annihilate selected off-diagonal elements and, when convergent, produces highly accurate eigenvalues. We give a new proof of its asymptotic quadratic convergence and derive an explicit, verifiable bound that describes the region in which this convergence is guaranteed. To make the method applicable well beyond matrices already close to the diagonal form, we introduce a preconditioning strategy. We use two types of preconditioners, both based on theoretical convergence results. The preconditioner is computed at lower precision to reduce computational cost, the associated similarity transformation is applied either at working or at higher precision, to preserve spectral information, while the main algorithm performs at working precision. Numerical experiments demonstrate that the resulting algorithm is robust and produces very accurate eigenvalues.
摘要
本文研究了一种用于具有简单特征值的一般复矩阵特征值问题的雅可比型方法。该方法应用初等三角相似变换以消去选定的非对角元素,并且在收敛时产生高精度特征值。我们给出了其渐近二次收敛性的新证明,并推导了一个明确的、可验证的界,描述了保证收敛的区域。为了使该方法适用于远未接近对角形式的矩阵,我们引入了一种预处理策略。我们使用了两种类型的预处理器,均基于理论收敛结果。预处理器以较低精度计算以减少计算成本,相关的相似变换以工作精度或更高精度应用以保留谱信息,而主算法以工作精度执行。数值实验表明,所得算法具有鲁棒性,并能产生非常精确的特征值。
#26PG-KINN: A Physics-Informed Petrov-Galerkin Kolmogorov-Arnold Network for Solving Forward and Inverse PDEs
PG-KINN:一种用于求解正问题和反问题的物理信息Petrov-Galerkin Kolmogorov-Arnold网络
Amirhossein Sadr, Nima Soltani, Vahideh Moghtadaiee, Aida Pakniyat, Dara Rahmati, Saeid Gorgin · 2026-07-22T17:08:29Z
Abstract
Physics-informed learning of partial differential equations (PDEs) has been dominated by multilayer perceptrons (MLPs), whose spectral bias and dense parameterization limit both accuracy and interpretability. Kolmogorov Arnold Networks (KANs) mitigate these limitations because their learnable spline activations are structurally aligned with the piecewise-polynomial bases of classical discretizations. However, the way a PDE is cast into a loss functional is as decisive as the choice of approximator: strong-form residual minimization requires high-order derivatives and heavily weighted losses, the energy (Bubnov-Galerkin) form is restricted to self-adjoint operators and, as we show, collapses to a trivial solution for parameter-identification problems, and boundary integral forms require a known fundamental solution. We propose PG-KINN, a physics-informed KAN built on a Petrov-Galerkin formulation in which the trial space is a KAN and the test space is an independent, compactly supported, piecewise-polynomial space evaluated with Gauss-Legendre quadrature. Integration by parts lowers the differentiation order while retaining applicability to general non-self-adjoint, nonlinear, and inverse problems; the localized test functions turn the global residual into a set of element-wise weak residuals with favorable conditioning. On a suite of benchmarks spanning crack singularities, stress concentration, Neo-Hookean hyperelasticity, inverse parameter identification in heterogeneous media, and complex geometries, PG-KINN consistently outperforms legacy MLP baselines and state-of-the-art KAN-based strong/energy/inverse formulations (PIKAN). These results position the Petrov-Galerkin coupling of KAN trial spaces and polynomial test spaces as a robust and accurate route for AI-based computational mechanics.
摘要
偏微分方程(PDE)的物理信息学习一直由多层感知器(MLP)主导,其频谱偏差和密集参数化限制了准确性和可解释性。Kolmogorov Arnold网络(KAN)缓解了这些限制,因为其可学习的样条激活函数在结构上与经典离散化的分段多项式基一致。然而,PDE如何被转化为损失函数与逼近器的选择同样关键:强形式残差最小化需要高阶导数和重加权损失;能量(Bubnov-Galerkin)形式限于自伴算子,并且正如我们所示,对于参数识别问题退化为平凡解;而边界积分形式需要已知的基本解。我们提出了PG-KINN,一种基于Petrov-Galerkin公式的物理信息KAN,其中试验空间是KAN,测试空间是一个独立的、紧支撑的、分段多项式空间,并使用Gauss-Legendre求积进行评估。分部积分降低了微分阶数,同时保持了对一般非自伴、非线性和反问题的适用性;局部测试函数将全局残差转化为一组具有良好条件数的单元弱残差。在一系列覆盖裂纹奇异性、应力集中、Neo-Hookean超弹性、异质介质中逆参数识别以及复杂几何形状的基准测试中,PG-KINN始终优于传统的MLP基线和最先进的基于KAN的强/能量/逆公式(PIKAN)。这些结果将KAN试验空间与多项式测试空间的Petrov-Galerkin耦合定位为基于AI的计算力学中一种稳健且准确的途径。
#27Shape optimization for thermoelasticity with temperature-dependent material parameters
温度相关材料参数的热弹性形状优化
Marc Dambrine, Helmut Harbrecht, Viacheslav Karnaev · 2026-07-22T16:18:54Z
Abstract
We consider the numerical solution of shape optimization problems for thermoelasticity with temperature-dependent material parameters. We show the existence of the shape derivative and derive an expression for generic functionals of domain integral type. Numerical results are presented for two settings: minimization of the compliance under a volume constraint, and minimization of the volume under a constraint on the $L^2$-norm of the von Mises stress. We use the finite element method for solving the underlying boundary value problems and the level set method for the representation of the actual domain.
摘要
我们考虑具有温度相关材料参数的热弹性形状优化问题的数值解。我们证明了形状导数的存在性,并推导了域积分型通用泛函的表达式。给出了两种设置的数值结果:在体积约束下最小化柔度,以及在冯·米塞斯应力$L^2$范数约束下最小化体积。我们使用有限元方法求解底层边值问题,并使用水平集方法表示实际域。
#28Wavenumber-explicit stability and preasymptotic error analysis of UPML finite element method for obstacle scattering problems
障碍散射问题UPML有限元方法的波数显式稳定性与预渐近误差分析
Yuhao Wang, Weiying Zheng · 2026-07-22T16:18:44Z
Abstract
This paper develops a wavenumber-explicit stability and preasymptotic error analysis for finite element approximation of the two-dimensional Helmholtz scattering problem with a uniaxial perfectly matched layer (UPML) truncation. The analysis is based on direct estimates of the stretched Green kernel associated with the Cartesian complex coordinate transformation. We establish explicit stability for the truncated UPML problem. In particular, we prove that the inf-sup constant of the truncated UPML formulation is $μ_L = O(k^{-1})$ in a natural $k$-weighted $H^1$ norm. As a consequence, we also obtain an exponential decay estimate for the PML truncation error. Based on this stability estimate, we formulate a linear continuous interior penalty finite element method (CIP-FEM) on the truncated UPML domain. A key ingredient of the analysis is to establish a piecewise $H^{1+s}$-regularity result (for any $0<s<1$), which accounts for coefficient jumps across PML interfaces and singularities induced by Cartesian corner geometries. This fractional regularity is sufficient to derive wavenumber-explicit preasymptotic error estimates. Numerical experiments confirm the theoretical predictions and illustrate the effectiveness of the UPML CIP-FEM in the high-frequency regime.
摘要
本文发展了二维亥姆霍兹散射问题采用单轴完美匹配层截断的有限元近似的波数显式稳定性和预渐近误差分析。该分析基于与笛卡尔复坐标变换相关的拉伸格林核的直接估计。我们建立了截断UPMP问题的显式稳定性。特别地,我们证明了在自然k加权H¹范数下,截断UPML公式的inf-sup常数为μ_L = O(k^{-1})。由此,我们还得到了PML截断误差的指数衰减估计。基于该稳定性估计,我们在截断UPML域上构造了一种线性连续内罚有限元方法。分析的一个关键要素是建立分片H^{1+s}正则性结果(对任意0<s<1),该结果考虑了跨PML界面的系数跳跃和由笛卡尔角几何引起的奇异性。这种分数正则性足以推导波数显式的预渐近误差估计。数值实验验证了理论预测,并说明了UPML CIP-FEM在高频区域的有效性。
#29Mixed finite element discretization of intrinsic geometrically exact beams for explicit multibody dynamics
显式多体动力学中内蕴几何精确梁的混合有限元离散
Andrea Brugnoli, Philipp L. Kinon, Francesco Sanfedino, Peter Betsch, Olivier A. Bauchau · 2026-07-22T15:05:41Z
Abstract
The Reissner-Simo and Hodges models are two equivalent continuous descriptions of finite-strain beam dynamics. The Reissner-Simo formulation uses displacements and rotations, while the Hodges formulation is intrinsic and avoids both variables. Although equivalent in theory, the two approaches behave differently after discretization and offer distinct numerical advantages. In this work, we develop a structure-preserving discretization of the intrinsic formulation. Because the intrinsic equations involve linear differential operators, both kinematic and dynamic boundary conditions can be imposed naturally using mixed finite elements. The resulting formulation also enables multibody systems to be assembled without algebraic constraints, avoiding the stiff differential-algebraic equations typically introduced by kinematic constraints. We demonstrate the approach on different examples, also showing that closed kinematic loops can be modeled without algebraic constraints. The resulting interconnected systems retain a port-Hamiltonian structure,with all nonlinearities confined to the interconnection operator. This structure allows exact energy preservation when combined with implicit midpoint time integration. Furthermore the scheme appear to require less Newton iterations compared to existing energy preserving scheme.
摘要
Reissner-Simo和Hodges模型是有限应变梁动力学的两种等效连续描述。Reissner-Simo公式使用位移和旋转,而Hodges公式是内蕴的,避免了这两个变量。尽管在理论上等价,但这两种方法在离散化后的行为不同,并具有不同的数值优势。在这项工作中,我们发展了内蕴公式的结构保持离散化。由于内蕴方程涉及线性微分算子,动力学和运动学边界条件都可以通过混合有限元自然地施加。由此产生的公式还允许无代数约束地组装多体系统,避免了通常由运动学约束引入的刚性微分代数方程。我们在不同的例子中展示了该方法,还表明封闭运动学回路可以在没有代数约束的情况下建模。所得互连系统保留了端口-哈密顿结构,所有非线性都局限于互连算子。这种结构在与隐式中点时间积分结合时允许精确的能量保持。此外,该方案似乎比现有的能量保持方案需要更少的牛顿迭代。
#30Hard Guarantees at a Measured Price: Entropy-Stable Learned Finite Volumes for Compressible Flow
以可衡量的代价获得硬保证:可压缩流动的熵稳定学习有限体积法
Denis Gueyffier · 2026-07-22T14:05:43Z
Abstract
Learned solvers for compressible flow are usually compared to classical methods at equal mesh resolution rather than at equal computational cost, and they typically offer no guarantee that their solutions remain physically admissible. We present a learned finite volume scheme for the two-dimensional Euler equations on unstructured meshes, admissible by construction and with an entropy-stable interior flux. We evaluate it under protocols fixed before any computation: frozen thresholds, falsification clauses, negative controls, a factor decomposition of the learned components, and an iso-cost comparison against the refined classical baseline. The decomposition produced the central result: the guarantee machinery alone, with both learned heads switched off (the unlearned skeleton), is the strongest scheme at equal mesh on every periodic case. At equal wall-clock cost the picture inverts into a map. Learning pays robustly only on the wall case whose boundary-condition type it never saw (10.8%). Its periodic gains flip sign with the evaluation draw (+10% on one held-out case, -12% on the hardest). The skeleton is the only method whose iso-cost gain never changes sign, at a measured overhead of 1.74x per step. The guaranteed variant completes 36 of 36 rollouts, Mach extrapolation and unseen wall included, with zero negativity events. We fix the guaranteed scheme's one remaining out-of-distribution weakness, Mach extrapolation, at inference time: with scale-invariant network inputs, a specific-entropy floor, and no retraining, the corrected arm overtakes the unconstrained arm on one Mach case, cuts its deficit on the other by a third, passes the skeleton on the unseen wall, and keeps the guarantee. A spatial gate closes the loop: activating the heads only near the walls beats both the skeleton and the corrected arm, and transfers unchanged to a second wall geometry.
摘要
可压缩流动的学习求解器通常与经典方法在相同网格分辨率下进行比较,而不是在相同计算成本下,并且它们通常不能保证其解保持物理可接受性。我们提出了一种用于非结构化网格上二维Euler方程的学习有限体积格式,该格式通过构造具有可接受性,并具有熵稳定的内部通量。我们在任何计算之前固定的协议下对其进行评估:冻结阈值、证伪子句、阴性对照、学习组件的因子分解,以及针对细化经典基线的等成本比较。分解产生了中心结果:仅凭保证机制,两个学习头部都关闭(未学习的骨架),在每个周期案例的相同网格上是最强的格式。在相同的墙钟时间成本下,情况反转为一幅地图。学习仅在其从未见过的边界条件类型的壁面案例上稳健地付出代价(10.8%)。其周期增益随评估图翻转(在一个保留案例上+10%,在最难的案例上-12%)。骨架是唯一等成本增益从不改变符号的方法,其测量开销为每步1.74倍。受保证的变体完成了36个中的36个展开,包括马赫外推和未见壁面,零负性事件。我们在推理时修复了受保证格式唯一剩余的分布外弱点——马赫外推:通过尺度不变的网络输入、特定的熵下限,并且无需重新训练,修正后的分支在一个马赫案例上超过了无约束分支,在另一个案例上将其缺陷减少了三分之一,在未见壁面上超过了骨架,并保持了保证。一个空间门关闭了循环:仅在壁面附近激活头部,既优于骨架又优于修正分支,并且不变地转移到第二个壁面几何。
#31Boundary-preserving Lamperti--Itô--Taylor approximations for some stochastic differential equations
某类随机微分方程的保边界Lamperti–Itô–Taylor逼近
Johan Ulander · 2026-07-22T13:53:41Z
Abstract
In this work, we propose high-order boundary-preserving numerical schemes for the strong approximation for some scalar stochastic differential equations with invariant domains being open and bounded intervals. The proposed methods involve using the Lamperti transform to map the SDE to another SDE with additive noise with a trivial invariant domain. Then, by imposing regularity assumptions on the original coefficient functions, we can guarantee that the drift coefficient function of the transformed SDE is regular, and known high-order schemes can be used to achieve the desired convergence order. We confirm the theoretical results with numerical experiments.
摘要
在这项工作中,我们针对具有开放有界区间不变域的一些标量随机微分方程,提出了高阶保边界数值格式用于强逼近。所提方法涉及使用Lamperti变换将随机微分方程映射为另一个具有平凡不变域的加性噪声随机微分方程。然后,通过对原始系数函数施加正则性假设,我们可以保证变换后随机微分方程的漂移系数函数是正则的,并且可以使用已知的高阶格式来实现所需的收敛阶。我们通过数值实验确认了理论结果。
#32An Optimization Approach to Weight Collocation for Scattered Spherical Data
散乱球面数据的加权配置优化方法
Congpei An, Xiannan Hu, Xiaoming Yuan · 2026-07-22T12:35:55Z
Abstract
We introduce an optimization approach for constructing spherical quadrature rules on arbitrarily scattered data. Rather than designing node placements, the new approach focuses on optimally computing the weights for fixed configurations. Motivated by Pólya's necessary and sufficient conditions for quadrature convergence in 1933, we argue that pursuing weight positivity and high algebraic exactness for scattered data approximation is not necessary. To align the quadrature design with the underlying theory of approximation, we construct convex optimization models with suitable objective functionals by examining the accuracy of numerical integration with reproducing kernels of Sobolev spaces and the performance of hyperinterpolation with Marcinkiewicz-Zygmund (MZ) inequalities. The resulting optimization models encode the spatial distribution of the scattered sites and the analytic properties of the target function spaces. The proposed approach enables the derivation of rigorous theoretical stability bounds, and the resulting quadrature weights are efficiently computable by modern convex optimization techniques. Numerical results are reported to demonstrate the performance of the optimization approach for fundamental approximation tasks such as numerical integration and hyperinterpolation for scattered spherical data.
摘要
我们提出了一种在任意散乱数据上构造球面求积规则的优化方法。新方法不设计节点布置,而是专注于为固定配置最优地计算权重。受Pólya在1933年提出的求积收敛的充要条件的启发,我们论证了追求权重正性和高代数精确度对于散乱数据逼近并不是必要的。为了使求积设计与逼近理论的基础保持一致,我们通过检查使用Sobolev空间再生核的数值积分精度以及使用Marcinkiewicz-Zygmund不等式的超插值性能,构建了具有适当目标泛函的凸优化模型。由此产生的优化模型编码了散乱点的空间分布和目标函数空间的分析性质。所提出的方法能够推导严格的稳定性理论界,并且求积权重可以通过现代凸优化技术高效计算。数值结果报告了优化方法在基本逼近任务(如散乱球面数据的数值积分和超插值)中的性能。
#33An efficient Galerkin method for high-frequency scattering problems using Wilson bases
使用Wilson基的高效伽辽金方法求解高频散射问题
T. Chaumont-Frelet, M. Ingremeau, F. Proust · 2026-07-22T09:26:59Z
Abstract
We propose a new Galerkin discretization scheme for wave scattering problems that is based on microlocalised basis functions. We show that the proposed method can be made uniformly accurate for large wavenumbers $k$ with a number of degrees of freedom only scaling as $k^{d-1/2}$, while leading to an essentially sparse linear system. In contrast, finite element methods are known to require a number of degrees of freedom scaling at least as $k^d$ to achieve the same property. A similar method based on a Gabor frame was previously introduced by two of the authors, but it was suffering from severe conditioning issues. In the present work, by replacing the Gabor frame by a Wilson basis, we completely alleviate this problem. We rigorously establish error estimates and condition number bounds for the proposed method, and we provide one-dimensional numerical examples illustrating our theoretical findings.
摘要
我们提出了一种基于微局部基函数的波散射问题新的伽辽金离散格式。我们证明,所提方法对于大波数k可以做到一致精确,自由度数量仅按k^{d-1/2}缩放,同时导致本质上稀疏的线性系统。相比之下,已知有限元方法需要自由度数量至少按k^d缩放才能达到相同的性质。之前两位作者曾引入一种基于Gabor框架的类似方法,但该方法存在严重的条件数问题。在本工作中,通过用Wilson基替代Gabor框架,我们完全缓解了这个问题。我们严格建立了所提方法的误差估计和条件数界,并提供了二维数值例子说明我们的理论发现。
#34Multi-domain FEM-BEM coupling with several impenetrable obstacles
含多个不可穿透障碍物的多域FEM-BEM耦合
Antonin Boisneault, Marcella Bonazzoli, Xavier Claeys, Pierre Marchand · 2026-07-22T08:30:33Z
Abstract
In a recent paper we have analyzed a new formulation of the coupling of finite and boundary element methods (FEM-BEM) for Helmholtz problems, involving several heterogeneous bounded subdomains and one homogeneous unbounded subdomain. This formulation, called Generalized Optimized Schwarz Method (GOSM), is substructured, that is, its unknowns are associated with the subdomains interfaces. To derive the GOSM, the first step was to prove that a solution to the Helmholtz problem satisfies a specific multi-domain variational formulation, which involves one operator for each subdomain, each operator being independent of the others. In the present contribution, we design a variational formulation of that type for a more general geometrical and material configuration: several heterogeneous bounded subdomains, impenetrable obstacles and homogeneous subdomains are allowed. Note that, like in our recent paper, we assume that only one subdomain is unbounded, and that its boundary is bounded. The domain partition can have cross-points, that is, points where at least three subdomains are adjacent. We also prove that a solution to the initial Helmholtz problem can be recovered from a solution to the multi-domain variational formulation. This shows that the GOSM is a general and flexible framework to model acoustic wave propagation, as it can handle both multi-domain FEM-BEM coupling and (weakly imposed) boundary conditions on several obstacles.
摘要
在最近的一篇论文中,我们分析了一种新的有限元和边界元方法耦合公式,用于亥姆霍兹问题,涉及多个非均匀有界子域和一个均匀无界子域。这种公式称为广义优化Schwarz方法,是子结构化的,即其未知量与子域界面相关。为了推导GOSM,第一步是证明亥姆霍兹问题的解满足特定的多域变分公式,该公式对每个子域涉及一个算子,每个算子独立于其他算子。在本贡献中,我们为更一般的几何和材料配置设计了这种类型的变分公式:允许存在多个非均匀有界子域、不可穿透障碍物和均匀子域。注意,与我们的近期论文一样,我们假设只有一个子域是无界的,且其边界是有界的。域划分可以具有交叉点,即至少三个子域相邻的点。我们还证明了可以从多域变分公式的解恢复原始亥姆霍兹问题的解。这表明GOSM是模拟声波传播的一个通用且灵活的框架,因为它可以处理多域FEM-BEM耦合和多个障碍物上的弱施加边界条件。
#35A phase-field neural solver for moving contact line problems with dynamic boundary conditions
具有动态边界条件的移动接触线问题的相场神经求解器
Ziyan Chen, Jinpeng Zhang, Pai Zhang, Li Luo · 2026-07-22T02:39:58Z
Abstract
Phase-field models based on the Cahn--Hilliard equation coupled with dynamic boundary conditions provide a thermodynamically consistent framework for moving contact line (MCL) problems. Although physics-informed neural networks (PINNs) offer a mesh-free approach for solving partial differential equations, their direct application to MCL problems remains challenging due to long-time error accumulation, sharp interfacial profiles, localized contact line dynamics, and complex contact angle evolution. In this work, we propose MCL-PINNs, a specialized phase-field neural solver designed for MCL problems with dynamic boundary conditions. The method is built on a discrete-time formulation and incorporates several key techniques, including a multi-network time-marching scheme, a relaxed distribution constraint on the neural network outputs, variable scaling for sharply varying solution features, adaptive loss weighting, adaptive collocation sampling with interface extraction, and, when applicable, symmetry preservation through neural network inputs. These techniques improve the capability of the neural solver in resolving sharp interfacial profiles and contact line motion. The proposed method is validated through three numerical examples involving droplet coalescence, shear-induced droplet deformation, and dynamic wetting in a heterogeneous channel. The numerical results show that MCL-PINNs significantly improve prediction accuracy and robustness compared with standard PINNs formulations, enabling reliable resolution of complex interfacial evolution and moving contact line dynamics.
摘要
基于Cahn-Hilliard方程与动态边界条件耦合的相场模型为移动接触线问题提供了一个热力学一致的框架。尽管物理信息神经网络提供了求解偏微分方程的无网格方法,但由于长时间误差积累、尖锐界面轮廓、局部接触线动力学和复杂的接触角演化,将其直接应用于MCL问题仍然具有挑战性。在这项工作中,我们提出了MCL-PINNs,一种专门设计用于具有动态边界条件的MCL问题的相场神经求解器。该方法基于离散时间公式,并包含几项关键技术,包括多网络时间推进方案、神经网络输出的松弛分布约束、陡变解特征的变量缩放、自适应损失加权、带界面提取的自适应配置点采样,以及在适用时通过神经网络输入保持对称性。这些技术提高了神经求解器解析尖锐界面轮廓和接触线运动的能力。通过三个数值例子验证了所提方法:液滴聚并、剪切诱导液滴变形和异质通道中的动态润湿。数值结果表明,与标准PINNs公式相比,MCL-PINNs显著提高了预测精度和鲁棒性,从而能够可靠地解析复杂的界面演化和移动接触线动力学。
#36A Splitting Architecture for Exact Reduced Coulomb Friction
精确约化库仑摩擦的分裂架构
Hongcheng Song, Ye Fan, Uri M. Ascher, Dinesh K. Pai · 2026-07-21T22:00:12Z
Abstract
Existing approaches to frictional contact dynamics typically either modify the Coulomb law to improve numerical robustness or solve the exact law in a fully coupled monolithic form. However, in its reduced form, exact Coulomb friction can be written as a cone complementarity problem with an augmented velocity, which reveals a natural split between a cone-constrained linear response and a scalar non-associated coupling induced by tangential velocity. We exploit this structure in the solver design. Our method uses an outer iteration to update the non-associated coupling explicitly, and an inner solve for a strongly convex cone-constrained quadratic program. This separation also makes the inner solver modular, so different numerical schemes can be used without changing the outer iteration. We evaluate the method on rigid-body benchmarks with stick-slip transitions and frictional stacking, and show that it reproduces exact Coulomb complementarity without smoothing or relaxing the friction law.
摘要
现有的摩擦接触动力学方法通常要么修改库仑定律以改善数值鲁棒性,要么以完全耦合的整体形式求解精确定律。然而,在其约化形式中,精确库仑摩擦可以写成一个带有增广速度的锥互补问题,这揭示了锥约束线性响应和由切向速度引起的标量非关联耦合之间的自然分裂。我们在求解器设计中利用了这一结构。我们的方法使用外部迭代显式更新非关联耦合,内部求解强凸锥约束二次规划。这种分离也使内部求解器模块化,因此可以在不改变外部迭代的情况下使用不同的数值方案。我们在具有粘滑转换和摩擦堆积的刚体基准上评估了该方法,并表明它再现了精确的库仑互补性,而无需平滑或松弛摩擦定律。
#371-Lipschitz Neural Networks on Hadamard Manifolds
Hadamard流形上的1-Lipschitz神经网络
Davide Murari, Marta Ghirardelli, Ben Adcock, Elena Celledoni, Brynjulf Owren, Carola-Bibiane Schönlieb · 2026-07-21T17:54:34Z
Abstract
Controlling the Lipschitz constant of a neural network is a standard way to promote robustness and stability. Most existing constraining strategies are designed for Euclidean spaces. In this work, we construct and analyze a class of 1-Lipschitz neural networks on Hadamard manifolds. Our layers are of gradient-descent type, $1$-Lipschitz, and quasi-$α$-firmly nonexpansive. The core building blocks of the proposed architecture are Busemann functions, and we exploit the properties of Busemann gradient flows to design $1$-Lipschitz geometry-preserving layers. We provide explicit constructions and examples for hyperbolic manifolds and the manifold of symmetric positive definite (SPD) matrices. We test the proposed architecture in two numerical experiments: robust classification on the Poincaré disk and masked-Wishart covariance reconstruction. On the Poincaré disk, the proposed networks yield robust classifiers under hyperbolic perturbations. On the SPD manifold, we train SPD-valued denoisers and adopt them as a Plug-and-Play prior for a masked-Wishart covariance reconstruction problem. We show improved results from the nonexpansive denoiser over static, data-only, and Log-Euclidean denoising baselines, and empirically test its convergence properties.
摘要
控制神经网络的Lipschitz常数是促进鲁棒性和稳定性的标准方法。大多数现有的约束策略是针对欧几里得空间设计的。在这项工作中,我们构造并分析了一类在Hadamard流形上的1-Lipschitz神经网络。我们的层是梯度下降类型、1-Lipschitz和准α-牢固非扩张的。所提出架构的核心构建块是Busemann函数,我们利用Busemann梯度流的性质来设计1-Lipschitz保几何层。我们为双曲流形和对称正定(SPD)矩阵流形提供了显式构造和示例。我们在两个数值实验中测试了所提出的架构:Poincaré盘上的鲁棒分类和掩码Wishart协方差重建。在Poincaré盘上,所提出的网络在双曲扰动下产生鲁棒分类器。在SPD流形上,我们训练SPD值去噪器,并将其作为即插即用先验用于掩码Wishart协方差重建问题。我们展示了非扩张去噪器相比静态、仅数据和Log-Euclidean去噪基线的改进结果,并经验性地测试了其收敛性质。
#38Resolution of the ENO-TV conjecture: a parity dichotomy
ENO-TV猜想的解决:奇偶二分法
Zhuoyun Li, Kailiang Wu · 2026-07-21T16:55:27Z
Abstract
We resolve the ENO--TV conjecture, a discrete coercivity problem in compactness theory for entropy-stable approximations of hyperbolic conservation laws. For order-$k$ essentially non-oscillatory (ENO) reconstruction from compactly supported cell averages, it asks whether the nonnegative ENO source times the $(k-1)$st power of the amplitude uniformly controls the $(k+1)$st absolute-jump moment. We prove a parity dichotomy: the estimate holds for odd $k\ge3$ and fails for even $k\ge4$; the known second-order case completes the classification. Localization gives a selection-independent finite-difference functional uniformly comparable to the source and reduces the conjecture to discrete interpolation. For odd orders, summation by parts reveals a hidden square; a discrete Gagliardo--Nirenberg inequality yields coercivity. For even orders, Euler-polynomial blocks from the functional's polynomial kernel yield counterexamples that persist under arbitrarily small perturbations making all affected ENO comparisons strict. We also prove two coercive estimates for every $k\ge2$: control of jumps larger than a fixed fraction of the amplitude and of local blocks modulo sampled polynomials of degree at most $k-2$. Via the Cayley--Sylvester decomposition, we compute the dimensions of homogeneous first-cohomology spaces for the lattice shift on polynomial jump profiles. At fourth order, for a cubic flux and a globally strictly convex entropy, a total-degree-seven component of a reduced entropy-flux mismatch represents a nonzero class on profiles of degree at most two and hence has no translation-invariant finite-stencil $C^7$ local primitive at the zero constant state. Odd-order coercivity persists on globally quasi-uniform meshes, whereas for each $k\ge2$ it fails on a fixed irregular mesh even though every interface contribution remains nonnegative. This failure is due to the mesh geometry.
摘要
我们解决了ENO-TV猜想,这是一个关于双曲守恒律熵稳定逼近的紧性理论中的离散强制性问题。对于从紧支撑单元平均值的k阶基本无振荡(ENO)重构,它询问非负ENO源乘以振幅的(k-1)次幂是否一致地控制(k+1)次绝对跳跃矩。我们证明了奇偶二分法:该估计对奇数k≥3成立,对偶数k≥4不成立;已知的二阶情形完成了分类。局部化给出了一个与源一致可比的与选择无关的有限差分泛函,并将猜想简化为离散插值。对于奇数阶,分部求和揭示了一个隐藏的平方;一个离散的Gagliardo-Nirenberg不等式给出了强制性。对于偶数阶,来自泛函多项式核的欧拉多项式块产生了反例,这些反例在任意小扰动下持续存在,使所有受影响的ENO比较变得严格。我们还证明了每个k≥2的两个强制性估计:控制大于振幅固定分数的跳跃,以及控制局部块模次数最多为k-2的采样多项式。通过Cayley-Sylvester分解,我们计算了多项式跳跃剖面上格点平移的齐次第一上同调空间的维数。在四阶情况下,对于三次通量和全局严格凸熵,简化熵-通量失配的总次数七分量在次数最多为二的剖面上表示非零类,因此在零常数状态下没有平移不变有限模板C^7局部原函数。奇数阶强制性在全局拟均匀网格上持续成立,而对于每个k≥2,它在固定不规则网格上失败,尽管每个界面贡献保持非负。这种失败是由于网格几何形状造成的。
#39Nyström Error Beyond $M$-Matrices: A Minimal Diagonally Dominant Obstruction
$M$-矩阵之外的Nyström误差:最小对角占优障碍
Matthew J. Colbrook · 2026-07-21T16:53:51Z
Abstract
We study the nuclear-norm error of a column-selected Nyström approximation to $K=(L+γI)^{-1}$, where $L$ is symmetric diagonally dominant and $γ>0$. Our central question is whether this error has diminishing returns. A Schur-complement identity reduces the question to traces of inverses of principal submatrices. Existing $M$-matrix results settle the case in which $L$ is a symmetric diagonally dominant $M$-matrix (SDDM). However, diagonal dominance alone is not enough: failure occurs already in dimension three. We construct an exact one-parameter SDD family and determine its sharp failure interval. A $2\times2$ identity proves that dimension three is minimal within the SDD class. We then show that failure persists under strict diagonal dominance; with a nonempty selected base set, dimension four is minimal. Finally, we prove invariance under signature switching, derive a three-dimensional formula showing how a signed triangle causes failure, and give an example in which greedy column selection misses the optimal pair. Together, these findings complete the answer to Problem 4.6 in a recent Simons workshop report.
摘要
我们研究了列选择Nyström近似$K=(L+γI)^{-1}$的核范数误差,其中$L$是对角占优对称矩阵,$γ>0$。核心问题是该误差是否具有递减效应。通过Schur补恒等式,问题归结为主子矩阵逆的迹。现有的$M$-矩阵结果解决了$L$为对称对角占优$M$-矩阵(SDDM)的情况。然而,仅对角占优不够:在三维中已经出现失败。我们构造了一个精确的单参数SDD族,并确定了其尖锐失败区间。$2\times2$恒等式证明三维在SDD类中是最小的。我们进一步证明在严格对角占优下失败仍然存在;在非空选定基集下,四维是最小的。最后,我们证明了符号切换下的不变性,推导了一个三维公式展示符号三角形如何导致失败,并给出了一个贪心列选择错过最优对的例子。这些发现共同完善了对近期Simons研讨会报告中问题4.6的回答。
#40Weighted Inverse Lax-Wendroff Boundary Treatment of Discontinuous Galerkin Methods for Conservation Laws
守恒律间断伽辽金方法的加权逆Lax-Wendroff边界处理
Yongjie Bi, Yan Jiang, Yong Liu · 2026-07-21T16:32:22Z
Abstract
In this paper, we propose a weighted inverse Lax-Wendroff (WILW) boundary treatment for the discontinuous Galerkin (DG) method on unfitted meshes to efficiently solve hyperbolic conservation laws in complex geometries. The proposed method employs the standard DG scheme for interior cells and reconstructs high-order approximation polynomials via the ILW principle for cut cells near boundaries to impose numerical boundary conditions, effectively eliminating the time-step restriction typically caused by small cut cells. In particular, to address the sensitivity of numerical errors to the geometric size of cut cells in the basic ILW scheme, we raise the reconstruction order at the boundary, ensuring that accuracy becomes independent of the cut-cell size. Furthermore, it incorporates a weighted least-squares reconstruction to reduce the need for complex high-order boundary derivatives during construction. As a result, the method maintains high-order accuracy while significantly improving computational efficiency for multi-dimensional problems. Finally, the stability of the proposed method is theoretically validated through linear stability analysis, and the effectiveness and robustness of the proposed scheme are numerically verified through a series of one-dimensional and two-dimensional numerical experiments for scalar and system equations.
摘要
本文提出了一种适用于非拟合网格上间断伽辽金(DG)方法的加权逆Lax-Wendroff(WILW)边界处理,以高效求解复杂几何中的双曲守恒律。该方法在内部单元使用标准DG格式,并在边界附近切割单元通过ILW原理重构高阶逼近多项式以施加数值边界条件,有效消除通常由小切割单元导致的时间步长限制。特别地,针对基本ILW方案中数值误差对切割单元几何尺寸的敏感性,我们提高了边界处的重构阶数,确保精度独立于切割单元尺寸。此外,它融入了加权最小二乘重构,减少了构造过程中对复杂高阶边界导数的需求。因此,该方法在保持高阶精度的同时,显著提高了多维问题的计算效率。最后,通过线性稳定性分析从理论上验证了所提方法的稳定性,并通过一系列标量和系统方程的一维及二维数值实验验证了方案的有效性和鲁棒性。
#41Uniform-in-Time Weak and Ergodic Error Estimates of a Nonlinearity-Explicit Full Discretization for Superlinear SPDEs Driven by Multiplicative Noise
乘性噪声驱动的超线性SPDE非线性显式全离散格式的一致时间弱误差与遍历误差估计
Jingjing Cai, Zhihui Liu, Xiaoming Wu · 2026-07-21T16:19:25Z
Abstract
For a class of superlinear SPDEs driven by multiplicative noise, we prove an (essentially) sharp uniform-in-time (UIT) weak convergence rate for the nonlinearity-explicit Galerkin tamed Euler method (GTEM). Under standard monotonicity assumptions, the proof combines Malliavin calculus with regularity theory for the associated backward Kolmogorov equation (BKE), leading to UIT moment, Hölder, and Malliavin estimates, along with regularity estimates for the BKE solution. These estimates, together with a weak error decomposition and Malliavin integration by parts (IBP) formula, then yield a UIT weak convergence rate $τ^ρ+λ_N^{-ρ}$ for any $ρ\in (0,1)$. Consequently, we obtain a sharp ergodic error estimate between the exact and numerical invariant measures. Numerical experiments support the theory.
摘要
对于一类由乘性噪声驱动的超线性SPDE,我们证明了非线性显式伽辽金驯化欧拉方法(GTEM)的(本质)精确的一致时间(UIT)弱收敛率。在标准单调性假设下,证明结合了Malliavin微积分与相关向后Kolmogorov方程(BKE)的正则性理论,得到了UIT矩估计、Hölder估计和Malliavin估计,以及BKE解的正则性估计。这些估计与弱误差分解和Malliavin分部积分(IBP)公式相结合,进而得出对于任意$ρ\in (0,1)$的UIT弱收敛率$τ^ρ+λ_N^{-ρ}$。因此,我们得到了精确和数值不变测度之间的尖锐遍历误差估计。数值实验支持该理论。
#42Numerical methods for Langevin-type SPDE: an implicit Milstein approach and multilevel Monte Carlo techniques
Langevin型SPDE的数值方法:隐式Milstein方法与多层蒙特卡罗技术
Sascha Portaro, Carlos Vázquez · 2026-07-21T15:20:31Z
Abstract
In this work, we investigate the numerical approximation of degenerate Langevin-type stochastic partial differential equations (SPDEs) in two spatial dimensions. These SPDEs arise in stochastic dynamics and mathematical finance, among other applications. In order to handle the mixed deterministic-stochastic structure of the equation and the degeneracy of the differential operator, we propose a semi-implicit Milstein finite difference scheme for the numerical solution. Through the Fourier analysis of the mean-square stability and convergence, we derive explicit conditions on the coefficients under which the scheme is stable, jointly with explicit convergence rates in terms of the discretization parameters. We further embed the proposed scheme within a Multilevel Monte Carlo (MLMC) framework to reduce the computational cost associated with SPDE simulations, and we derive its theoretical computational complexity. Numerical experiments confirm theoretical convergence rates and show that the MLMC strategy achieves an accuracy comparable to standard Monte Carlo at a fraction of the computational cost, reducing the complexity from $\mathcal{O}(\varepsilon^{-5})$ to $\mathcal{O}(\varepsilon^{-3})$ for a target root-mean-square error $\varepsilon$. These results show that combining semi-implicit Milstein schemes with MLMC techniques provides an effective approach for the numerical simulation of Langevin-type SPDEs.
摘要
本文研究了二维空间中退化Langevin型随机偏微分方程(SPDE)的数值逼近。这些SPDE出现在随机动力学和金融数学等应用中。为了处理方程的混合确定性-随机结构以及微分算子的退化性,我们提出了一种半隐式Milstein有限差分格式用于数值求解。通过均方稳定性和收敛性的傅里叶分析,我们推导了格式稳定的系数显式条件,以及关于离散化参数的显式收敛率。我们进一步将所提格式嵌入多层蒙特卡罗(MLMC)框架,以降低SPDE模拟相关的计算成本,并推导了其理论计算复杂度。数值实验证实了理论收敛率,并表明MLMC策略以一小部分计算成本达到了与标准蒙特卡罗相当的精度,对于目标均方根误差$\varepsilon$,复杂度从$\mathcal{O}(\varepsilon^{-5})$降至$\mathcal{O}(\varepsilon^{-3})$。这些结果表明,半隐式Milstein格式与MLMC技术的结合为Langevin型SPDE的数值模拟提供了一种有效方法。
#43Boundary-Adapted PINNs for Elliptic Dirichlet Problems: $H^2(Ω)$ A Priori Error Bounds with Application to Mean Escape Time Computation
适应边界的PINN用于椭圆Dirichlet问题:H^2(Ω)先验误差界及其在平均逃逸时间计算中的应用
Nathanael Tepakbong, Jun Fan, Xiang Zhou, Ding-Xuan Zhou · 2026-07-21T15:03:42Z
Abstract
Motivated by the numerical computation of the Mean Escape Time (MET) $τ:Ω\to\mathbb{R}$ of a stochastic process from a bounded domain $Ω\subseteq\mathbb{R}^d$, we study elliptic Dirichlet boundary value problems (BVPs) using boundary-enforced Physics-Informed Neural Networks (PINNs), in which the Dirichlet condition is imposed exactly by multiplying the network output with a predefined distance-to-boundary approximation $ρ$. Combining approximation-theoretic and statistical-learning arguments for Rectified Quadratic Unit (ReQU) and hyperbolic tangent (tanh) networks, we derive a priori error bounds that make explicit the dependence on $ρ$. In particular, we show that exact boundary enforcement alone is not enough for $H^2(Ω)$ error bounds, and that a sufficient and essentially necessary condition is for $ρ$ to be a smooth distance approximation $\textit{normalized to first order}$, of the kind constructed in arXiv:2104.08426 [math.NA]. We thereby identify this subclass of $\textit{boundary-adapted}$ PINNs as the appropriate neural network ansatz for solving Dirichlet BVPs. Numerical experiments support the theory, showing that appropriate choices of $ρ$ improve accuracy and convergence, while poorly chosen distance functions can substantially degrade the solution. Our proof also yields new VC-dimension bounds for hypothesis spaces of higher-order derivatives of ReQU and tanh networks, together with new approximation bounds for shallow ReQU networks in higher-order Sobolev norms, all of which are of important independent interest.
摘要
受随机过程从有界域Ω⊆ℝ^d的平均逃逸时间τ:Ω→ℝ的数值计算启发,我们研究了使用边界强制的物理信息神经网络(PINN)的椭圆Dirichlet边值问题,其中通过将网络输出乘以预定义的距离边界近似ρ来精确施加Dirichlet条件。结合分段线性二次单位(ReQU)和双曲正切(tanh)网络的逼近理论和统计学习论证,我们推导了先验误差界,明确了对ρ的依赖性。特别地,我们表明仅精确边界执行不足以获得H^2(Ω)误差界,充分且本质必要的条件是ρ是一阶归一化的光滑距离近似,类似于arXiv:2104.08426 [math.NA]中构造的类型。因此,我们确定这个适应边界的PINN子类是求解Dirichlet边值问题的适当神经网络ansatz。数值实验支持该理论,表明适当选择ρ可提高精度和收敛性,而选择不当的距离函数会显著降低解的质量。我们的证明还得到了ReQU和tanh网络高阶导数假设空间的新VC维界限,以及浅层ReQU网络在高阶Sobolev范数中的新逼近界限,所有这些都具有重要的独立意义。
#44Error Bound and Stability Analysis for a Randomized Singly Diagonally Implicit Runge-Kutta Method
随机单对角隐式Runge-Kutta方法的误差界与稳定性分析
Monika Eisenmann, Marvin Jans, Raphael Kruse, Helmut Podhaisky · 2026-07-21T10:11:16Z
Abstract
A randomized Singly Diagonally Implicit Runge-Kutta (SDIRK) method, based on the randomized trapezoidal rule as the underlying quadrature scheme, is proposed. Every realization of the scheme is an algebraically stable SDIRK method of at least second order. The main result is the proof that the randomized scheme converges with order 2.5 in the root mean square sense under low regularity assumptions. Numerical experiments illustrate the robustness of the new scheme when applied to nonsmooth problems.
摘要
提出了一种基于随机梯形法则作为底层求积方案的随机单对角隐式Runge-Kutta(SDIRK)方法。该方案的每一次实现都是一个至少二阶的代数稳定SDIRK方法。主要结果是证明在低正则性假设下,随机方案在均方意义下以2.5阶收敛。数值实验说明了新方案应用于非光滑问题时的稳健性。
#45A Second-Moment Theory for Floating-Point Reduction Trees
浮点归约树的二阶矩理论
Piyush Sao, Narasinga Miniskar, Pedro Valero-Lara, Keita Teranishi, Sudip Seal · 2026-07-21T06:27:43Z
Abstract
Summation error depends on partial-sum order, which standard worst-case bounds omit. To capture this dependence, we derive an exact mean-square error (MSE) recurrence for a binary reduction tree T under conditionally unbiased rounding. With unit roundoff u, the constant-nu model sets the local variance at pre-rounding value x to nu u^2 x^2. Its leading tree-dependent cost for the input vector p is p^T K_T p, where the common-ancestor kernel K_T counts the internal ancestors shared by each pair of leaves. For i.i.d. inputs of mean mu and variance tau^2, this expected cost is tau^2 Lambda_1(T) + mu^2 Lambda_2(T), where Lambda_1 is total leaf depth and Lambda_2 sums squared internal-subtree sizes; Lambda_1 governs centered inputs, while Lambda_2 captures nonzero means. We use these statistics to characterize optimal tree topologies and schedules. Balanced and sequential trees attain the centered extrema. For k inputs, optimal two-stage sequential blocking yields root-mean-square (RMS) error scaling as k^{3/4}. For fixed-stage hierarchies, geometric schedules are optimal for centered inputs, whereas the optimal noncentered stage exponents halve successively. For independent centered inputs with unequal variances, Huffman coding minimizes variance-weighted depth over free leaf assignments. We extend the kernel to matrix multiplication through operand Gram matrices. We then test the approximation under round-to-nearest using exact residuals. Across binary64, binary32, and software-emulated binary16 and bfloat16, the model recovers the ordering among tree topologies; K_T tracks AR(1) partial-sum costs. For GEMM, independently calibrated predictions differ from measurements by at most 3% on the tested grid. A reduction tree extracted from an array library predicts the measured RMS scaling. However, stagnation and bias in positive low-precision sums limit the model's applicability.
摘要
求和误差取决于部分和顺序,标准最坏情况界限忽略了这一点。为了捕捉这种依赖性,我们推导了在条件无偏舍入下二元归约树T的精确均方误差(MSE)递推关系。以单位舍入u为单位,常数nu模型将预舍入值x处的局部方差设置为nu u^2 x^2。其对输入向量p的主要树相关代价为p^T K_T p,其中共同祖先核K_T计算每对叶子共享的内部祖先数量。对于均值为mu、方差为tau^2的独立同分布输入,期望代价为tau^2 Λ_1(T) + mu^2 Λ_2(T),其中Λ_1是总叶子深度,Λ_2是内部子树大小的平方和;Λ_1控制中心化输入,而Λ_2捕获非零均值。我们使用这些统计量来表征最优树拓扑和调度。平衡树和顺序树达到中心极值。对于k个输入,最优两阶段顺序分块产生均方根(RMS)误差缩放为k^{3/4}。对于固定阶段层次结构,几何调度对中心化输入是最优的,而非中心化输入的最优阶段指数依次减半。对于具有不等方差的独立中心化输入,哈夫曼编码在自由叶子分配上最小化方差加权深度。我们通过操作数格拉姆矩阵将核扩展到矩阵乘法。然后我们使用精确残差在最近舍入下测试近似。在binary64、binary32和软件模拟的binary16和bfloat16中,该模型恢复了树拓扑之间的顺序;K_T跟踪AR(1)部分和成本。对于GEMM,独立校准的预测与测量值在测试网格上相差不超过3%。从数组库中提取的归约树预测了测量的RMS缩放。然而,正低精度求和中的停滞和偏差限制了模型的适用性。
#46Contraction-Gauge Preconditioning for Quantized Matrix Multiplication
量化矩阵乘法的收缩度量预处理
Piyush Sao, Narasinga Miniskar, Pedro Valero-Lara, Keita Teranishi, Sudip Seal · 2026-07-21T06:09:08Z
Abstract
We study low-precision computation of C=AB with both factors quantized. We derive an exact finite-dimensional identity for the expected squared product error under independent, zero-mean entrywise errors with known variance fields; it holds exactly for non-overloading subtractive dither and for independent stochastic rounding, and we empirically assess deterministic round-to-nearest (RTN). Using the product-preserving equivalence AB=(AT)(T^{-1}B), we formulate contraction-gauge preconditioning: jointly choosing a factor representation and its sharing pattern before quantization. Preconditioning can reduce product error but may require extra transformed, quantized copies of the opposite operand: a shared transform needs one copy, a block-specific transform up to one per block. Within the bounded family of positive diagonal gauges (folds), a geometric program computes a globally optimal shared fold and a linear program decides whether the identity fold is already optimal. For other families we derive computable selection statistics -- tail index for scaling, profile spread for partitioning, coherence and weighted-Gram energy for rotations, slice-energy covariance for hierarchy depth -- with upper bounds for ranking heuristic candidates. Across twelve linear products from a trained three-block image classifier, median within-product rank correlations between dither-model predictions and deterministic-RTN errors are 0.937 at 8 bits and 0.918 at 4 bits. The GP fold cuts held-out product error over the identity fold by 18.0% (8-bit) and 20.5% (4-bit) in geometric mean, beats a SmoothQuant-style grid baseline at both precisions and on ten of twelve products, and lowers composed logit MSE by 15.4% and 26.4%. We thus provide exact stochastic product-error accounting, certified selection within the diagonal family, and a common objective for evaluating reusable transform candidates under RTN.
摘要
我们研究了两个因子都经过量化的低精度计算C=AB。在已知方差场的独立零均值逐项误差下,我们推导了期望平方乘积误差的精确有限维恒等式;该恒等式精确适用于非过载减法抖动和独立随机舍入,并实验评估了确定性最近舍入(RTN)。利用乘积保持等价AB=(AT)(T^{-1}B),我们提出了收缩度量预处理:在量化前联合选择因子表示及其共享模式。预处理可以减少乘积误差,但可能需要额外的变换后量化副本:共享变换需要一个副本,块特定变换每块最多一个副本。在正对角度量(折叠)的有界族内,几何规划计算全局最优共享折叠,线性规划判断恒等折叠是否已最优。对于其他族,我们推导了可计算的选取统计量——缩放的尾指数、分区的剖面扩展、旋转的相干性和加权Gram能量、层次深度的切片能量协方差——并给出了排序启发式候选的上界。在来自训练好的三块图像分类器的十二个线性乘积中,抖动模型预测与确定性RTN误差之间的中位产品内秩相关系数在8位时为0.937,在4位时为0.918。GP折叠相对于恒等折叠将留出乘积误差的几何平均值降低了18.0%(8位)和20.5%(4位),在两种精度下均优于SmoothQuant风格网格基准,并在十二个产品中的十个上表现更好,并将组合logit MSE降低了15.4%和26.4%。因此,我们提供了精确的随机乘积误差核算、对角族内的认证选择,以及评估RTN下可重用变换候选的共同目标。
#47Current-Sheet Formation in Electron Magnetohydrodynamics with Split Fractional Dissipation
分裂分数阶耗散电子磁流体动力学中的电流片形成
Ruimeng Hu, Qirui Peng, Xu Yang · 2026-07-21T02:37:09Z
Abstract
Thin current sheets are central small-scale structures in electron magnetohydrodynamics (EMHD), closely associated with energy dissipation and fast magnetic reconnection at electron scales. We study their formation numerically in a $2\frac{1}{2}$-dimensional EMHD system on a periodic domain with split fractional dissipation, where the magnetic potential and the vertical magnetic component are damped separately. The local theory is governed by a symmetric combined damping balance, but the numerical onset of small-scale growth need not follow this symmetry. A scaling analysis identifies the out-of-plane current as the primary concentration observable, since it is regularized only through the magnetic-potential equation. Using a validated Fourier pseudospectral exponential time-differencing solver with resolution-controlled diagnostics, we find a clear decay/concentration dichotomy. The onset boundary is markedly asymmetric: current-sheet formation appears to be controlled mainly by damping of the magnetic potential, rather than by the combined damping strength. The analyticity strip collapses to the grid scale, the concentration sharpens under grid refinement, and the observed growth is consistent with an energy-critical self-similar rate, with exponent near three. These experiments indicate that magnetic-potential damping is the apparent binding constraint for current-sheet concentration, refining the symmetric sum picture.
摘要
薄电流片是电子磁流体动力学(EMHD)中的中心小尺度结构,与电子尺度上的能量耗散和快速磁重联密切相关。我们在具有分裂分数阶耗散的周期域上的二维半EMHD系统中数值研究它们的形成,其中磁势和垂直磁场分量分别被阻尼。局部理论由对称的组合阻尼平衡控制,但小尺度增长的数值开始不一定遵循这种对称性。尺度分析识别出面外电流为主要集中可观测量,因为它仅通过磁势方程被正则化。使用经过验证的傅里叶伪谱指数时间差分求解器和分辨率控制的诊断,我们发现了一个清晰的衰减/集中二分法。开始边界明显不对称:电流片形成似乎主要由磁势阻尼控制,而不是组合阻尼强度。解析性条塌缩到网格尺度,浓缩在网格细化下加剧,观察到的增长与能量临界自相似速率一致,指数接近三。这些实验表明,磁势阻尼是电流片浓缩的明显约束条件,细化了对称和图像。
#48anyakrakusuma: A Python Library for Entropic Schrödinger Bridges on Idealized Geometries
anyakrakusuma:一个用于理想几何上熵Schrödinger桥的Python库
Sandy Hardian Susanto Herho, Dasapta Erwin Irawan, Agus Wahyu Jatmiko, Sito Fossy Biosa, Candrasa Surya Dharma, Edi Riawan, Astyka Pamumpuni, Rendy Dwi Kartiko, Rusmawan Suwarman, Deny Juanda Puradimaja · 2026-07-20T17:23:37Z
Abstract
We present anyakrakusuma, an open-source Python library that solves the discrete static Schrödinger bridge problem, the entropically regularized counterpart of optimal transport, through a log-domain Sinkhorn--Knopp iteration and reconstructs the entropic interpolation between two empirical point clouds. The solver is paired with a diagnostic pipeline that characterizes the optimal coupling and the intermediate distributions through information-theoretic and geometric measures. We exercise the library on four idealized planar cases spanning a circle-to-circle dilation, a spiral-to-mixture fragmentation, a rigid reorientation of two moons, and a Lissajous-to-trefoil deformation. The log-domain formulation is necessary rather than merely convenient at the parameters studied, where the cost-to-regularization ratio reaches four hundred and the Gibbs kernel underflows double precision across most of its range; the iteration nonetheless attains a marginal residual of $10^{-9}$ and unit marginal fidelity in every case. Residual histories decay geometrically over approximately eight decades at per-iteration contraction factors between $0.966$ and $0.976$, which are local rates near the fixed point that lie many orders of magnitude below the worst-case Hilbert-metric bound. The covariance analysis recovers an imposed ninety-degree reorientation to within $0.07^\circ$, roughly forty times smaller than its uncertainty, across a masked interval of near-isotropy on which the principal axis is unobservable. The diagnostics are reported with explicit attention to the regimes in which each is well defined, including the differential entropy, which is meaningful only on the open interpolation interval. The presented cases are constructed rather than measured; quantitative application to empirical point clouds requires further study.
摘要
我们提出了anyakrakusuma,一个开源的Python库,通过对数域Sinkhorn-Knopp迭代求解离散静态Schrödinger桥问题(最优传输的熵正则化对应),并重构两个经验点云之间的熵插值。该求解器配有一个诊断流程,通过信息论和几何度量表征最优耦合和中间分布。我们在四个理想平面案例上测试了该库,包括圆到圆的膨胀、螺旋到混合的碎裂、两个月亮的刚性重定向以及利萨如到三叶结的变形。在所研究的参数下,对数域公式是必要的,而不仅仅是方便的,其中成本与正则化之比达到四百,吉布斯核在大部分范围内下溢双精度;然而,迭代在每个案例中仍达到了$10^{-9}$的边缘残差和单位边缘保真度。残差历史在约八个数量级上几何衰减,每次迭代的收缩因子在0.966到0.976之间,这些是固定点附近的局部速率,比最坏情况下的Hilbert度量界低多个数量级。协方差分析在近各向同性的掩蔽区间内(此时主轴不可观测)恢复了一个施加的九十度重定向,误差在$0.07^\circ$以内,比其不确定性小约四十倍。诊断报告明确关注每个定义良好的区域,包括仅在开放插值区间上有意义的微分熵。呈现的案例是构造的而非测量的;对经验点云的定量应用需要进一步研究。