Search arXiv⌕ Search

arXiv · 2609.38497

The Mode of Null-A: Compositional Computation of a Generalized Inverse

Abstract

We present a novel algorithm for calculating the preimage of an affine space through a product ${J}={J}_{T-1}\cdots{J}_0$ of matrices ${J}_t$ of special form: finding the largest input space $\mathbf{X}$ such that $\mathbf{x}\in\mathbf{X}$ implies ${J}\mathbf{x}\in \mathbf{Y}$, where $\mathbf{Y}$ is a given output affine space. These special matrices arise in AD, where the Jacobians $J$ describing the linearized computation have precisely this structure: the product of a series of linearized primitive numeric operations. This allows us to use the new algorithm to formulate Null-A mode preimage AD, which finds the affine preimage through the Jacobian or Jacobian transpose of a numeric computation. This is a generalization of the inverse AD problem of solving ${J}\acute{\mathbf{x}}^{\ast}=\acute{\mathbf{y}}^{\ast}$ or ${J}^{T}\grave{\mathbf{y}}^{\ast}=\grave{\mathbf{x}}^{\ast}$. The key is to represent affine spaces in a fashion which lends itself to efficient preimage calculation, in a compositional and quasi-local fashion, through a succession of matrices ${J}_t$. Unlike previous methods, Null-A preimage mode AD allows the ${J}_t$ matrices to be non-square, corresponding to a computer program whose number of active variables swells and shrinks during the computation. When ${J}$ is square and the initial affine space is a single point, this finds the conventional inverse. But in the more general case, having the entire affine space provides freedom which can be leveraged in a problem-specific manner. We apply the method to small problems on-CPU where the $J_t$ are linearized scalar unary or binary numeric functions; and to larger problems on-GPU where the $J_t$ are linearized aggregate array operations like convolution and attention.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Barak A. Pearlmutter, Jeffrey Mark Siskind. 2026-09-29. The Mode of Null-A: Compositional Computation of a Generalized Inverse. https://arxiv.org/abs/2609.38497

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Error Estimates for the Arnoldi Approximation of a Matrix Square Root

The Arnoldi process provides an efficient framework for approximating functions of a matrix applied to a vector, i.e., of the form $f(M)\bm{b}$, by repeated matrix-vector multiplications. In this paper, we derive error estimates for approximating the action of a matrix square root using the Arnoldi process, where the integral representation of the error is reformulated in terms of the error for solving the linear system $M\bm{x}=\bm{b}$. The results extend the error analysis of the Lanczos method for Hermitian matrices in [Chen et al., SIAM J. Matrix Anal. Appl., 2022] to non-Hermitian cases and provide an improved bound for the Hermitian case. Furthermore, in practical settings, the matrix may only be available via approximate or structured representations. Motivated by this, we extend the analysis and establish a generalized error bound for perturbed matrices. The numerical results on matrices with different structures demonstrate that our theoretical analysis yields a reliable upper bound. Finally, simulations on large-scale matrices arising in particulate suspensions, represented in hierarchical matrix form, validate the effectiveness and practicality of the approach.

math.NA↗

Worst-Case Completion of Tensors with Approximately Few ANOVA Terms

In this article, the problem of completing a tensor from some incomplete knowledge of its entries is treated by adopting a worst-case perspective, given the realistic assumption that the tensor's low-order ANOVA terms are dominant. We survey and leverage some recent all-purpose results from the field of Optimal Recovery to provide solutions on a theoretical level. But the accompanying constructions of optimal completion procedures, which often feature semidefinite programs, are not directly applicable in the tensor case due to the huge dimensions involved. To resolve the issue, we put forward a storage-friendly way to produce low-order ANOVA projections based on the fast Fourier transform (FFT), while exploiting the specificities of the completion problem to efficiently compute regularizers and extremal eigenvalues. Numerical experiments on synthetic tensors and real-world datasets demonstrate the accuracy and scalability of our FFT-based method.

math.NA↗

A fully convergent fixed-point fast sweeping method with the WENO-JS local solver for steady state of hyperbolic conservation laws

The fixed-point fast sweeping methods with weighted essentially non-oscillatory (WENO) local solvers are a class of efficient and high-order accuracy numerical methods for solving steady-state solutions of hyperbolic conservation laws. However, with the classical WENO-JS local solver, the iteration residue of high-order fixed-point fast sweeping scheme often has difficulty to settle down to round-off errors. To achieve the full convergence in a fast sweeping method, the fixed-point fast sweeping methods with non-traditional WENO local solvers based on unequal-sized substencils were designed. However, the WENO schemes based on unequal-sized stencils are more complex and in general more expensive in computational costs than the classical WENO-JS schemes. In this paper, we go back to the classical WENO-JS local solver and develop a new fully convergent fifth-order fixed-point fast sweeping method for solving steady-state problems of hyperbolic conservation laws. Based on recent studies on the nonlinear weighting process around discontinuities of solution, we apply the technique of frozen weights for avoiding unnecessary adjustment of nonlinear weights in the WENO-JS local solver, which freezes the nonlinear weights once the residual sequence in the fast sweeping iterations has stabilized. Different from the existing work on frozen weights, we design a simple and robust approach to judge stabilization of iteration residues and determine the iteration step when the nonlinear weights are frozen in the fast sweeping method. Extensive numerical experiments on a wide range of challenging two-dimensional steady-state problems demonstrate that, unlike the previous fast sweeping method with the fifth-order WENO-JS local solver, the proposed new scheme consistently drives the iteration residues to round-off errors and achieves the full convergence.

math.NA↗