Search arXivSearch

arXiv subjects

Tom Needham

Publications and source records attributed to Tom Needham.

At least 19 recordsLinked to original sources

Metric Geometry of Lebesgue, Wasserstein, and Gromov-Wasserstein Spaces: Submetries, Curvature, and Geodesics

A metric space $Z$ gives rise to three natural classes of infinite-dimensional metric spaces associated to $Z$: $p$-Wasserstein spaces of probability measures on $Z$, nonlinear Lebesgue $L^p$-spaces of $Z$-valued maps, and $p$-Gromov-Wasserstein spaces of $Z$-valued kernels. The latter class, referred to as $Z$-Gromov-Wasserstein ($Z$-GW) spaces, extends the classical Gromov-Wasserstein framework from metric measure spaces to more general, possibly attributed, network-like structures, and unifies many GW-type distances that nowadays play a significant role in metric geometry, data science and machine learning. In this article we develop a unified metric-geometric theory of these three classes of spaces, with a particular focus on the $Z$-GW spaces. Our first main result identifies a fundamental submetry structure linking them: the nonlinear Lebesgue space maps via a submetry onto the $Z$-GW space, which in turn maps via a submetry onto the Wasserstein space. This structure provides a mechanism for transferring geometric information among the three spaces. We apply this framework to geodesics and Alexandrov curvature. For $1<p<\infty$, we prove that geodesicity of $Z$ is equivalent to geodesicity of each of the three associated spaces; in the endpoint case $p=1$, all three associated spaces are geodesic, even when $Z$ is not. We also characterize geodesics in the $Z$-GW space as generalized interpolations, extending a known characterization in the classical setting due to Sturm. Finally, we give a complete classification of Alexandrov curvature bounds for these spaces in terms of the curvature of $Z$. Thus, while the main focus of the paper is a new metric-geometric theory of $Z$-GW spaces, the submetry framework also extends classical theorems for Wasserstein and Gromov-Wasserstein spaces and yields new geometric consequences for nonlinear Lebesgue spaces.

math.MG

The Observable Wasserstein Distance

We introduce the observable Wasserstein distance, a framework for deriving lower bounds on the Wasserstein distance between probability measures on Polish metric spaces, designed to bypass the computational intractability of exact optimal transport in large-scale, non-Euclidean datasets. Analogous to the sliced Wasserstein distance in $\mathbb{R}^d$, our approach projects measures onto the real line via 1-Lipschitz observables and computes the Wasserstein distances between the resulting pushforward distributions. We define a hierarchy of pseudo-metrics by restricting observables to a nested chain of subspaces. A central theoretical contribution is an injectivity result linking the metric covering dimension of the support of a measure to the specific order in the hierarchy that guarantees unique recovery. This serves as a metric-space analogue to the Cram\'{e}r-Wold Device for Euclidean distributions. We demonstrate that this hierarchy offers a tunable trade-off between sharpness as a lower bound on the Wasserstein distance and computational efficiency. We also present a discrete computational model for finite grids and numerical experiments validating the efficacy and utility of these approximations.

math.MG

Geometric Perspective on Concentration Phenomena in Frame Theory

Parseval and equal-norm frames play a fundamental role in frame theory and signal processing. It is known that a random frame, with unit vectors drawn independently from the uniform distribution on the sphere, will be nearly Parseval with high probability; asymptotic results go back at least to Goyal, Vetterli, and Thao and a non-asymptotic error bound was proved more recently by Kwok, Lau, and Ramachandran. In this work, we prove a dual result, which shows that random Parseval frames, with respect to the Haar measure, are nearly equal-norm with high probability. Our proofs are geometric in nature, and rely on general measure concentration principles in Riemannian manifolds. Using these techniques, we also give a novel probabilistic upper bound for the Paulsen problem.

math.FA

On the Hausdorff stability of barcodes over posets

The Isometry Theorem of Chazal et al. and Lesnick is a fundamental result in persistence theory, which states that the interleaving distance between two one-parameter persistence modules is equal to the bottleneck distance between their barcodes. Significant effort has been devoted to extending this result to modules defined over more general posets. As these modules do not generally admit nice decompositions, one must restrict attention to the class of interval-decomposable modules in order to define an appropriate notion of bottleneck distance. Even with this assumption, it is known that bottleneck distance may not be equivalent to interleaving distance, but that it is Lipschitz stable under certain, fairly restrictive, assumptions. In this paper, we consider the more basic question of stability of the Hausdorff distance with respect to interleaving distance for interval-decomposable modules. Our main theorem is a Lipschitz stability result, which holds in a fairly general setting of interval-decomposable modules over arbitrary posets, where intervals are assumed to be taken from any family satisfying certain closure conditions. Along the way, we develop some new tools and results for interval-decomposable modules over arbitrary posets, in the form of geometrically-flavored characterizations of the existence of morphisms and interleavings between interval modules.

math.AT

A Benamou-Brenier Proximal Splitting Method for Constrained Unbalanced Optimal Transport

The dynamic formulation of optimal transport, also known as the Benamou-Brenier formulation, has been extended to the unbalanced case by introducing a source term in the continuity equation. When this source term is penalized based on the Fisher-Rao metric, the resulting model is referred to as the Wasserstein-Fisher-Rao (WFR) setting, and allows for the comparison between any two positive measures without the need for equalized total mass. In recent work, we introduced a constrained variant of this model, in which affine integral equality constraints are imposed along the measure path. In the present paper, we propose a further generalization of this framework, which allows for constraints that apply not just to the density path but also to the momentum and source terms, and incorporates affine inequalities in addition to equality constraints. We prove, under suitable assumptions on the constraints, the well-posedness of the resulting class of convex variational problems. The paper is then primarily devoted to developing an effective numerical pipeline that tackles the corresponding constrained optimization problem based on finite difference discretizations and parallel proximal schemes. Our proposed framework encompasses standard balanced and unbalanced optimal transport, as well as a multitude of natural and practically relevant constraints, and we highlight its versatility via several synthetic and real data examples.

math.OC

A Persistent Homology Pipeline for the Analysis of Neural Spike Train Data

In this article, we introduce a Topological Data Analysis (TDA) pipeline for neural spike train data. Understanding how the brain transforms sensory information into perception and behavior requires analyzing coordinated neural population activity. Modern electrophysiology enables simultaneous recording of spike train ensembles, but extracting meaningful information from these datasets remains a central challenge in neuroscience. A fundamental question is how ensembles of neurons discriminate between different stimuli or behavioral states, particularly when individual neurons exhibit weak or no stimulus selectivity, yet their coordinated activity may still contribute to network-level encoding. We describe a TDA framework that identifies stimulus-discriminative structure in spike train ensembles recorded from the mouse insular cortex during presentation of deionized water stimuli at distinct non-nociceptive temperatures. We show that population-level topological signatures effectively differentiate oral thermal stimuli even when individual neurons provide little or no discrimination. These findings demonstrate that ensemble organization can carry perceptually relevant information that standard single-unit analysis may miss. The framework builds on a mathematical representation of spike train ensembles that enables persistent homology to be applied to collections of point processes. At its core is the widely-used Victor-Purpura (VP) distance. Using this metric, we construct persistence-based descriptors that capture multiscale topological features of ensemble geometry. Two key theoretical results support the method: a stability theorem establishing robustness of persistent homology to perturbations in the VP metric parameter, and a probabilistic stability theorem ensuring robustness of topological signatures.

stat.ME

Persistent Homology for Labeled Datasets: Gromov-Hausdorff Stability and Generalized Landscapes

Techniques from metric geometry have become fundamental tools in modern mathematical data science, providing principled methods for comparing datasets modeled as finite metric spaces. Two of the central tools in this area are the Gromov-Hausdorff distance and persistent homology, both of which yield isometry-invariant notions of distance between datasets. However, these frameworks do not account for categorical labels, which are intrinsic to many real-world datasets, such as labeled images, pre-clustered data, and semantically segmented shapes. In this paper, we introduce a general framework for labeled metric spaces and develop new notions of Gromov-Hausdorff distance and persistent homology which are adapted to this setting. Our main result shows that our persistent homology construction is stable with respect to our novel notion of Gromov-Hausdorff distance, extending a classic result in topological data analysis. To facilitate computation, we also introduce a labeled version of persistence landscapes and show that the landscape map is Lipschitz.

math.AT

Metrics for Parametric Families of Networks

We introduce a general framework for analyzing data modeled as parameterized families of networks. Building on a Gromov-Wasserstein variant of optimal transport, we define a family of parameterized Gromov-Wasserstein distances for comparing such parametric data, including time-varying metric spaces induced by collective motion, temporally evolving weighted social networks, and random graph models. We establish foundational properties of these distances, showing that they subsume several existing metrics in the literature, and derive theoretical approximation guarantees. In particular, we develop computationally tractable lower bounds and relate them to graph statistics commonly used in random graph theory. Furthermore, we prove that our distances can be consistently approximated in random graph and random metric space settings via empirical estimates from generative models. Finally, we demonstrate the practical utility of our framework through a series of numerical experiments.

stat.ML

Conic Formulations of Transport Metrics for Unbalanced Measure Networks and Hypernetworks

The Gromov-Wasserstein (GW) variant of optimal transport, designed to compare probability densities defined over distinct metric spaces, has emerged as an important tool for the analysis of data with complex structure, such as ensembles of point clouds or networks. To overcome certain limitations, such as the restriction to comparisons of measures of equal mass and sensitivity to outliers, several unbalanced or partial transport relaxations of the GW distance have been introduced in the recent literature. This paper is concerned with the Conic Gromov-Wasserstein (CGW) distance introduced by S\'{e}journ\'{e}, Vialard, and Peyr\'{e}. We provide a novel formulation in terms of semi-couplings, and extend the framework beyond the metric measure space setting, to compare more general network and hypernetwork structures. With this new formulation, we establish several fundamental properties of the CGW metric, including its scaling behavior under dilation, variational convergence in the limit of volume growth constraints, and comparison bounds with established optimal transport metrics. We further derive quantitative bounds that characterize the robustness of the CGW metric to perturbations in the underlying measures. The hypernetwork formulation of CGW admits a simple and provably convergent block coordinate ascent algorithm for its estimation, and we demonstrate the computational tractability and scalability of our approach through experiments on synthetic and real-world high-dimensional and structured datasets.

stat.ML

Equivalence of Landscape and Erosion Distances for Persistence Diagrams

This paper establishes connections between three of the most prominent metrics used in the analysis of persistence diagrams in topological data analysis: the bottleneck distance, Patel's erosion distance, and Bubenik's landscape distance. Our main result shows that the erosion and landscape distances are equal, thereby bridging the former's natural category-theoretic interpretation with the latter's computationally convenient structure. The proof utilizes the category with a flow framework of de Silva et al., and leads to additional insights into the structure of persistence landscapes. Our equivalence result is applied to prove several results on the geometry of the erosion distance. We show that the erosion distance is not a length metric, and that its intrinsic metric is the bottleneck distance. We also show that the erosion distance does not coarsely embed into any Hilbert space, even when restricted to persistence diagrams arising from degree-0 persistent homology. Moreover, we show that erosion distance agrees with bottleneck distance on this subspace, so that our non-embeddability theorem generalizes several results in the recent literature.

math.MG

Optimization and the Topology of Spaces of Parseval Frames

A Parseval frame is a spanning set for a Hilbert space which satisfies the Parseval identity: a vector can be expressed as a linear combination of the frame whose coefficients are inner products with the frame vectors. There is considerable interest within the signal processing community in the structural properties of the space of finite-dimensional Parseval frames whose vectors all have the same norm, or which satisfy more general prescribed norm constraints. In this paper, we introduce a function on the space of arbitrary spanning sets that jointly measures the failure of a spanning set to satisfy both the Parseval identity and given norm constraints. We show that, despite its nonconvexity, this function has no spurious local minimizers, thereby extending the Benedetto--Fickus theorem to this non-compact setting. In particular, this shows that gradient descent converges to an equal norm Parseval frame when initialized within a dense open set in the associated matrix space. We then apply this result to study the topology of frame spaces. Using our Benedetto--Fickus-type result, we realize spaces of Parseval frames with prescribed norms as deformation retracts of simpler spaces, leading to explicit conditions which guarantee the vanishing of their homotopy groups. These conditions yield new path-connectedness results for spaces of real Parseval frames, generalizing the Frame Homotopy Theorem, which has seen significant interest in recent years.

math.FA

Robust Representation and Estimation of Barycenters and Modes of Probability Measures on Metric Spaces

This paper is concerned with the problem of defining and estimating statistics for distributions on spaces such as Riemannian manifolds and more general metric spaces. The challenge comes, in part, from the fact that statistics such as means and modes may be unstable: for example, a small perturbation to a distribution can lead to a large change in Fr\'echet means on spaces as simple as a circle. We address this issue by introducing a new merge tree representation of barycenters called the barycentric merge tree (BMT), which takes the form of a measured metric graph and summarizes features of the distribution in a multiscale manner. Modes are treated as special cases of barycenters through diffusion distances. In contrast to the properties of classical means and modes, we prove that BMTs are stable -- this is quantified as a Lipschitz estimate involving optimal transport metrics. This stability allows us to derive a consistency result for approximating BMTs from empirical measures, with explicit convergence rates. We also give a provably accurate method for discretely approximating the BMT construction and use this to provide numerical examples for distributions on spheres and shape spaces.

math.ST

Stability of Hypergraph Invariants and Transformations

Graphs are fundamental tools for modeling pairwise interactions in complex systems. However, many real-world systems involve multi-way interactions that cannot be fully captured by standard graphs. Hypergraphs, which generalize graphs by allowing edges to connect any number of vertices, offer a more expressive framework. In this paper, we introduce a new metric on the space of hypergraphs, inspired by the Gromov-Hausdorff distance for metric spaces. We establish Lipschitz properties of common hypergraph transformations, which send hypergraphs to graphs, including a novel graphification method with ties to single linkage hierarchical clustering. Additionally, we derive lower bounds for the hypergraph distance via invariants coming from basic summary statistics and from topological data analysis techniques. Finally, we explore stability properties of cost functions in the context of optimal transport. Our results in this direction consider Lipschitzness of the Hausdorff map and conservation of the non-negative cross curvature property under limits of cost functions.

math.MG

Fused Gromov-Wasserstein Variance Decomposition with Linear Optimal Transport

Wasserstein distances form a family of metrics on spaces of probability measures that have recently seen many applications. However, statistical analysis in these spaces is complex due to the nonlinearity of Wasserstein spaces. One potential solution to this problem is Linear Optimal Transport (LOT). This method allows one to find a Euclidean embedding, called LOT embedding, of measures in some Wasserstein spaces, but some information is lost in this embedding. So, to understand whether statistical analysis relying on LOT embeddings can make valid inferences about original data, it is helpful to quantify how well these embeddings describe that data. To answer this question, we present a decomposition of the Fr\'echet variance of a set of measures in the 2-Wasserstein space, which allows one to compute the percentage of variance explained by LOT embeddings of those measures. We then extend this decomposition to the Fused Gromov-Wasserstein setting. We also present several experiments that explore the relationship between the dimension of the LOT embedding, the percentage of variance explained by the embedding, and the classification accuracy of machine learning classifiers built on the embedded data. We use the MNIST handwritten digits dataset, IMDB-50000 dataset, and Diffusion Tensor MRI images for these experiments. Our results illustrate the effectiveness of low dimensional LOT embeddings in terms of the percentage of variance explained and the classification accuracy of models built on the embedded data.

stat.ME

Metric properties of partial and robust Gromov-Wasserstein distances

The Gromov-Wasserstein (GW) distances define a family of metrics, based on ideas from optimal transport, which enable comparisons between probability measures defined on distinct metric spaces. They are particularly useful in areas such as network analysis and geometry processing, as computation of a GW distance involves solving for registration between the objects which minimizes geometric distortion. Although GW distances have proven useful for various applications in the recent machine learning literature, it has been observed that they are inherently sensitive to outlier noise and cannot accommodate partial matching. This has been addressed by various constructions building on the GW framework; in this article, we focus specifically on a natural relaxation of the GW optimization problem, introduced by Chapel et al., which is aimed at addressing exactly these shortcomings. Our goal is to understand the theoretical properties of this relaxed optimization problem, from the viewpoint of metric geometry. While the relaxed problem fails to induce a metric, we derive precise characterizations of how it fails the axioms of non-degeneracy and triangle inequality. These observations lead us to define a novel family of distances, whose construction is inspired by the Prokhorov and Ky Fan distances, as well as by the recent work of Raghvendra et al.\ on robust versions of classical Wasserstein distance. We show that our new distances define true metrics, that they induce the same topology as the GW distances, and that they enjoy additional robustness to perturbations. These results provide a mathematically rigorous basis for using our robust partial GW distances in applications where outliers and partial matching are concerns.

math.MG

Geometry of the Space of Partitioned Networks: A Unified Theoretical and Computational Framework

Interactions and relations between objects may be pairwise or higher-order in nature, and so network-valued data are ubiquitous in the real world. The "space of networks", however, has a complex structure that cannot be adequately described using conventional statistical tools. We introduce a measure-theoretic formalism for modeling generalized network structures such as graphs, hypergraphs, or graphs whose nodes come with a partition into categorical classes. We then propose a metric that extends the Gromov-Wasserstein distance between graphs and the co-optimal transport distance between hypergraphs. We characterize the geometry of this space, thereby providing a unified theoretical treatment of generalized networks that encompasses the cases of pairwise, as well as higher-order, relations. In particular, we show that our metric is an Alexandrov space of non-negative curvature, and leverage this structure to define gradients for certain functionals commonly arising in geometric data analysis tasks. We extend our analysis to the setting where vertices have additional label information, and derive efficient computational schemes to use in practice. Equipped with these theoretical and computational tools, we demonstrate the utility of our framework in a suite of applications, including hypergraph alignment, clustering and dictionary learning from ensemble data, multi-omics alignment, as well as multiscale network alignment.

math.MG

The Z-Gromov-Wasserstein Distance

The Gromov-Wasserstein (GW) distance is a powerful tool for comparing metric measure spaces which has found broad applications in data science and machine learning. Driven by the need to analyze datasets whose objects have increasingly complex structure (such as node and edge-attributed graphs), several variants of GW distance have been introduced in the recent literature. With a view toward establishing a general framework for the theory of GW-like distances, this paper considers a vast generalization of the notion of a metric measure space: for an arbitrary metric space $Z$, we define a $Z$-network to be a measure space endowed with a kernel valued in $Z$. We introduce a method for comparing $Z$-networks by defining a generalization of GW distance, which we refer to as $Z$-Gromov-Wasserstein ($Z$-GW) distance. This construction subsumes many previously known metrics and offers a unified approach to understanding their shared properties. This paper demonstrates that the $Z$-GW distance defines a metric on the space of $Z$-networks which retains desirable properties of $Z$, such as separability, completeness, and geodesicity. Many of these properties were unknown for existing variants of GW distance that fall under our framework. Our focus is on foundational theory, but our results also include computable lower bounds and approximations of the distance which will be useful for practical applications.

math.MG

Generalized Dimension Reduction Using Semi-Relaxed Gromov-Wasserstein Distance

Dimension reduction techniques typically seek an embedding of a high-dimensional point cloud into a low-dimensional Euclidean space which optimally preserves the geometry of the input data. Based on expert knowledge, one may instead wish to embed the data into some other manifold or metric space in order to better reflect the geometry or topology of the point cloud. We propose a general method for manifold-valued multidimensional scaling based on concepts from optimal transport. In particular, we establish theoretical connections between the recently introduced semi-relaxed Gromov-Wasserstein (srGW) framework and multidimensional scaling by solving the Monge problem in this setting. We also derive novel connections between srGW distance and Gromov-Hausdorff distance. We apply our computational framework to analyze ensembles of political redistricting plans for states with two Congressional districts, achieving an effective visualization of the ensemble as a distribution on a circle which can be used to characterize typical neutral plans, and to flag outliers.

math.OC