Search arXivSearch

SEARCH · Search arXiv

Results for “math.IT”

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

5,876 records · Page 8Linked to original sources

Terminating Zero-Balanced Hypergeometric Series Using Divided Differences

There is a close relationship between divided differences and hypergeometric series, and recent studies have shown that divided differences can be used effectively to derive terminating hypergeometric identities. In this paper, we apply this approach to terminating zero-balanced hypergeometric series. Using explicit product evaluations arising from divided differences, we give alternative proofs of zero-balanced ${}_3F_2(1)$ and ${}_4F_3(1)$ summation formulas and then extend the argument to the general terminating ${}_{r+1}F_r(1)$ case. We also combine the Lagrange representation with the Leibniz rule for divided differences to derive a finite convolution transformation for a terminating ${}_{r+2}F_{r+1}(1)$ series. Thus, the divided-difference approach not only reproduces the known zero-balanced summation formulas but also yields a finite convolution transformation for terminating hypergeometric series, together with its zero-balanced specialization.

math.CA

Rigorous Error Certification for Neural PDE Solvers: From Empirical Residuals to Solution Guarantees

Uncertainty quantification for partial differential equations is traditionally grounded in discretization theory, where solution error is controlled via mesh/grid refinement. Physics-informed neural networks fundamentally depart from this paradigm: they approximate solutions by minimizing residual losses at collocation points, introducing new sources of error arising from optimization, sampling, representation, and overfitting. As a result, the generalization error in the solution space remains an open problem. Our main theoretical contribution establishes generalization bounds that connect residual control to solution-space error. We prove that when neural approximations lie in a compact subset of the solution space, vanishing residual error guarantees convergence to the true solution. We derive deterministic and probabilistic convergence results and provide certified generalization bounds translating residual, boundary, and initial errors into explicit solution error guarantees.

cs.LG

Why Multi-Layer Message Passing Works: Completeness Theory for Graph Neural Network Interatomic Potentials

We prove that the Hypergraph Neural Network, an invariant architecture with 3-body message passing, is a universal approximator for potential energy surfaces. Our main contribution is a multi-layer completeness theory. We show that $L$ layers of message passing on sparse, cutoff-based graphs achieve the same representational power as having access to the full $L$-hop neighborhood, provided the configurations are generic, satisfy an overlap condition and a connectivity condition. This provides the first rigorous justification for the common practice of using multi-layer message passing with a per-layer cutoff smaller than the physical interaction range, the setting used by virtually all practical graph neural network based machine-learned interatomic potentials. As immediate consequences, we show that both DPA3 and CHGNet architectures inherit universal approximation.

cs.LG

On the Optimality of Gaussian Code-books for Signaling over a Two-Users Weak Gaussian Interference Channel

This article establishes that the capacity region of the two-user weak Gaussian interference channel can be achieved using single-letter Gaussian codebooks. The converse is established by showing that successive decoding can be employed for at least one of the receivers. It is further shown that the upper concave envelope of the achievable rate region can be attained using at most two time-sharing phases. In one phase, both users transmit simultaneously, while in the second phase, when present, only one user is active. Furthermore, the boundary of the capacity region can be traversed continuously through incremental reallocations of power between the two messages transmitted by each user. Finally, it is proven that the Han-Kobayashi achievable rate region with single-letter Gaussian codebooks attains the optimal boundary of the capacity region.

cs.IT

Numerical Ergodicity and Uniform Estimate of Monotone SPDEs Driven by Multiplicative Noise

We analyze the long-time behavior of numerical schemes for a class of monotone stochastic partial differential equations (SPDEs) driven by multiplicative noise. By deriving several time-independent a priori estimates for the numerical solutions, combined with the ergodic theory of Markov processes, we establish the exponential ergodicity of these schemes with a unique invariant measure, respectively. Applying these results to the stochastic Allen--Cahn equation indicates that these schemes always have at least one invariant measure, respectively, and converge strongly to the exact solution with sharp time-independent rates. We also show that these numerical invariant measures are exponentially ergodic and thus give an affirmative answer to a question proposed in (J. Cui, J. Hong, and L. Sun, Stochastic Process. Appl. (2021): 55--93), provided that the interface thickness is not too small.

math.NA

Bernstein--von Mises theorems for Bayesian probabilistic numerics

We study probabilistic numerical methods for solving nonlinear PDEs from a Bayesian nonparametric perspective. Given noisy evaluations at random collocation points, we place a truncated Gaussian series prior on the unknown solution and establish contraction at the minimax nonparametric rate, up to a logarithmic factor. Our main results give Gaussian approximations of the posterior in positive-order Sobolev spaces and, under suitable conditions, in the uniform topology. This contrasts with classical ill-posed inverse problems, where Bernstein--von Mises theorems typically require substantially weaker topologies. Here, the observation operator is differential rather than smoothing, and inversion of its linearisation gains regularity, making these strong-topology results possible. The posterior may be centred at either the posterior mean or the posterior mode. We further prove that the Gaussian Laplace approximation is asymptotically equivalent to the true posterior at a $\sqrt{N}$-scale.

math.ST

The Neighbor Graph of Linear Complementary Dual (LCD) Codes

Linear complementary dual (LCD) codes form an important class of linear codes with applications in cryptography, classical error correction, and quantum coding theory. In this paper, we study the neighbor relation on LCD codes over finite fields and the graph induced by this relation, where two codes are adjacent whenever they intersect in codimension one. We determine the number of neighbors of an LCD code that are also LCD, and we use this result to analyze the structure of the corresponding neighbor graph. In particular, we prove its regularity over arbitrary finite fields and establish further regularity properties for its main structural subgraphs in the binary and odd-characteristic cases. These results provide a graph-theoretic framework for the study of LCD codes and reveal a strong combinatorial regularity in their neighborhood structure.

math.CO

Hölder-Logarithmic Stability and Convergence Rates for an Inverse Random Source Problem

In this paper, we investigate an inverse random source problem concerned with recovering the strength of a random, uncorrelated acoustic source from correlation measurements of emitted time-harmonic acoustic waves. Such problems arise in applications including aeroacoustics and seismic imaging. Unlike their deterministic counterparts, inverse random source problems are known to be uniquely solvable in the absence of noise. Nevertheless, due to their inherent ill-posedness, regularization is required to stably reconstruct the source strength. We derive conditional Hölder-logarithmic stability estimates under Sobolev smoothness assumptions by employing complex geometrical optics solutions. Moreover, by establishing a variational source condition, we obtain Hölder-logarithmic convergence rates for spectral regularization methods. At fixed frequency, the exponents in the logarithmic stability and convergence estimates grow unboundedly as the Sobolev regularity of the source increases. Finally, we present numerical experiments supporting our theoretical findings.

math.NA

Eleven, twelve, and thirteen lonely runners

Wills conjectured that, for any non-zero integers $u_1,\ldots,u_k$, there is a real number $t$ such that, for all $i=1,\ldots,k$, \[\lVert tu_i\rVert\geq\frac{1}{k+1},\] where $\lVert x\rVert$ is the distance from $x$ to the closest integer. This statement is known as the Lonely Runner Conjecture. A computational method developed by Rosenfeld and the second author verified the conjecture for $k\leq9$. We further refine this method with new sieving techniques and employ a polynomial method argument to show that any $(u_1,\ldots,u_k)\equiv(1,2,\ldots,k)\pmod{p}$ with $\gcd(u_1,\ldots,u_k)=1$ satisfies the conjecture when $k+1$ and $p > k^2+k$ are both odd primes. Ultimately, we provide a computer-assisted proof of the Lonely Runner Conjecture for $k\in\{10,11,12\}$.

math.CO

On the Equality of the ELBO to a Sum of Entropies at Stationary Points of Learning

The variational lower bound (a.k.a. ELBO or free energy) is the central objective for many established as well as for many novel algorithms for unsupervised learning. Such algorithms usually increase the bound until parameters have converged to values close to a stationary point of the learning dynamics. Here we show that (for a very large class of generative models) the variational lower bound is at all stationary points of learning equal to a sum of entropies. Concretely, for standard generative models with one set of latents and one set of observed variables, the sum consists of three entropies: (A) the (average) entropy of the variational distributions, (B) the negative entropy of the model's prior distribution, and (C) the (expected) negative entropy of the observable distribution. The obtained result applies under realistic conditions including: finite numbers of data points, at any stationary point (including saddle points) and for any family of (well behaved) variational distributions. The class of generative models for which we show the equality to entropy sums contains many standard as well as novel generative models including standard (Gaussian) variational autoencoders. The prerequisites we use to show equality to entropy sums are relatively mild. Concretely, the distributions defining a given generative model have to be of the exponential family, and the model has to satisfy a parameterization criterion (which is usually fulfilled). Proving equality of the ELBO to entropy sums at stationary points (under the stated conditions) is the main contribution of this work.

stat.ML

Recovering linear images of sparse signals from indirect observations

In this paper, we develop and analyze techniques for recovering a linear image $Bx$ of an unknown signal $x$ from indirect noisy observation $ω=Ax+ξ$. It is {\em a priori} known that $x\in \cX$, a given convex compact set, and that $x$ is $s$-sparse---has at most $s$ nonvanishing entries. The proposed estimates belong to a large family of recovery routines by $\ell_1$-minimization. However, unlike the classical result describing performance of such estimates, we do not make any special (and hard to check) assumptions about the sensing matrix $A$ such as nullspace or Restricted Isometry condition and the like. As a consequence, parameters of the estimates and the upper bounds on their risks are not available in a closed analytic form, but are delivered instead by efficient computation as solutions to explicit convex optimization problems.

stat.ML

A fast and stable test to check if a weakly diagonally dominant matrix is a nonsingular M-matrix

We present a test for determining if a substochastic matrix is convergent. By establishing a duality between weakly chained diagonally dominant (w.c.d.d.) L-matrices and convergent substochastic matrices, we show that this test can be trivially extended to determine whether a weakly diagonally dominant (w.d.d.) matrix is a nonsingular M-matrix. The test's runtime is linear in the order of the input matrix if it is sparse and quadratic if it is dense. This is a partial strengthening of the cubic test in [J. M. Peña., A stable test to check if a matrix is a nonsingular M-matrix, Math. Comp., 247, 1385-1392, 2004]. As a by-product of our analysis, we prove that a nonsingular w.d.d. M-matrix is a w.c.d.d. L-matrix, a fact whose converse has been known since at least 1964. We point out that this strengthens some recent results on M-matrices in the literature.

math.NA

Asymptotic Bounds on Generalized Covering Radii of Binary Primitive BCH Codes

Fix integers $e\ge2$ and $r\ge1$. In this paper we study the $r$-th generalized covering radius $ρ_r\left(BCH(e,m)\right)$ of the binary primitive $e$-error-correcting BCH code $BCH(e,m)$. By using an algebraic-geometric reformulation of the covering problem together with an explicit Lang-Weil estimate, we prove that \[ρ_r\bigl(\BCH(e,m)\bigr)\le(r+1)e-1\] for all sufficiently large $m$. For $e\ge7$, this improves a recent result of Belinsky--Zabokritskiy. Our proof gives a substantially simpler geometric approach to this upper bound. In particular it implies that \[ρ_2\bigl(BCH(e,m)\bigr)=3e-1\] for all sufficiently large $m$. Previously it was only known that \[ρ_2\bigl(\BCH(e,m)\bigr) \in \left\{3e-1,3e\right\}\] for all sufficiently large $m$.

cs.IT

A multi-class kinetic traffic flow model: discrete-velocity formulation and diffusively-corrected macroscopic limits

This paper introduces a multi-class extension of a discrete-velocity kinetic traffic flow model based on a non-local Prigogine-Herman framework. We derive a hyperbolically scaled system of equations from a continuous kinetic formulation describing interactions between different vehicle classes through braking and relaxation terms. The model is then discretized with respect to the velocity variable for an arbitrary number of vehicle classes, and the structural properties of the resulting formulation are analyzed. In particular, we prove hyperbolicity and total linear degeneracy. Due to the non-conservative structure of the model, we employ a path-conservative finite volume scheme for the numerical approximation of the system. Finally, we derive the corresponding diffusively-corrected macroscopic multi-class model, investigate its stability and present numerical simulations on a single-lane road to illustrate the theoretical findings.

math.NA

Moment-enhanced shallow-water equations with an effective wall closure for no-slip bottoms

Shallow-water equations and low-order shallow-water moment models use vertically coarse representations and therefore cannot, in general, resolve the thin wall-affected region produced by a no-slip bottom. Enforcing the pointwise wall value on a low-order global polynomial reconstruction can introduce stiff relaxation and distort the resolved interior velocity profile. Starting from the incompressible Navier--Stokes equations with Navier bottom friction, we derive a bottom-to-mean relation in a distinguished regular-friction regime and use it to define an endpoint-consistent effective wall-traction closure for the shallow-water equations and the hyperbolic shallow-water moment equations. The closure represents the momentum effect of unresolved near-wall dynamics; it neither resolves the physical boundary layer nor imposes the pointwise no-slip trace on the reconstructed polynomial. It recovers the perfect-slip wall contribution when the friction coefficient vanishes. Because only source terms are changed, the homogeneous principal matrices and their established two-dimensional hyperbolicity classification remain unchanged. We compare the standard and modified reduced models with two-phase incompressible Navier--Stokes computations in OpenFOAM for wet-bed dam-break and three-dimensional collapse tests. In the cases considered, the modified closure reduces the excessive damping of the classical low-order wall source and improves agreement in depth-averaged and resolved-interior velocity diagnostics, but it does not uniformly improve front-propagation speed. The regular-friction asymptotic remainder is not uniform in the large-friction numerical regime; there the effective coefficient is used as a wall-model continuation and assessed empirically.

math.NA

Breakdown of Edgeworth Expansion in Finite-Blocklength Regime and Exact Absorption via $q$-Deformation

This paper addresses the structural breakdown of the Edgeworth expansion in the finite-blocklength (FBL) regime, where conventional asymptotic approximations yield unphysical negative probabilities in the deep-tail region. We propose a $q$-deformed framework that resolves this inconsistency by replacing additive polynomial perturbations with a geometric deformation of the information density space. Motivated by the linearization of nonlinear dynamics, we prove that dynamically scaling the $q$-logarithmic parameter exactly absorbs the third-order skewness while preserving global nonnegativity. We establish a universal asymptotic matching, demonstrating that the framework encapsulates higher-order asymptotic scales. Numerical results confirm that the proposed method matches the state-of-the-art precision of the Cornish-Fisher bound without the risk of negative probabilities. The framework offers a robust and computationally stable foundation for evaluating operational limits in ultra-reliable communications such as 6G and URLLC.

cs.IT

Nonlinear Dynamics In Optimization Landscape of Shallow Neural Networks with Tunable Leaky ReLU

In this work, we study the nonlinear dynamics of a shallow neural network trained with mean-squared loss and leaky ReLU activation. Under Gaussian inputs and equal layer width k, (1) we establish, based on the equivariant gradient degree, a theoretical framework, applicable to any number of neurons k>= 4, to detect bifurcation of critical points with associated symmetries from global minimum as leaky parameter $α$ varies. Typically, our analysis reveals that a multi-mode degeneracy consistently occurs at the critical number 0, independent of k. (2) As a by-product, we further show that such bifurcations are width-independent, arise only for nonnegative $α$ and that the global minimum undergoes no further symmetry-breaking instability throughout the engineering regime $α$ in range (0,1). An explicit example with k=5 is presented to illustrate the framework and exhibit the resulting bifurcation together with their symmetries.

math.OC