Search arXivSearch

arXiv · 1912.10839

Gaussianity and typicality in matrix distributional semantics

Abstract

Constructions in type-driven compositional distributional semantics associate large collections of matrices of size $D$ to linguistic corpora. We develop the proposal of analysing the statistical characteristics of this data in the framework of permutation invariant matrix models. The observables in this framework are permutation invariant polynomial functions of the matrix entries, which correspond to directed graphs. Using the general 13-parameter permutation invariant Gaussian matrix models recently solved, we find, using a dataset of matrices constructed via standard techniques in distributional semantics, that the expectation values of a large class of cubic and quartic observables show high gaussianity at levels between 90 to 99 percent. Beyond expectation values, which are averages over words, the dataset allows the computation of standard deviations for each observable, which can be viewed as a measure of typicality for each observable. There is a wide range of magnitudes in the measures of typicality. The permutation invariant matrix models, considered as functions of random couplings, give a very good prediction of the magnitude of the typicality for different observables. We find evidence that observables with similar matrix model characteristics of Gaussianity and typicality also have high degrees of correlation between the ranked lists of words associated to these observables.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Sanjaye Ramgoolam, Mehrnoosh Sadrzadeh, Lewis Sword. 2019-12-19. Gaussianity and typicality in matrix distributional semantics. https://arxiv.org/abs/1912.10839

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Fortuitous Chaos, BPS Black Holes, and Random Matrices

The ``fortuitous'' Bogomol'nyi-Prasad-Sommerfield (BPS) sector states in gauge theory have been argued to furnish a description, through holography, of generic BPS black hole microstates. They are expected to be strongly chaotic, a necessary feature to capture the black hole dynamics. This dovetails nicely with the existence of various random matrix models of JT supergravity with extended supersymmetry, within which the BPS chaos must be contained as a subsector. This paper identifies and studies a simple random matrix model that underlies all known random matrix models of JT supergravity. It is argued that it captures many essential universal features of fortuitous BPS chaos. The model is topological, naturally interpolating between the Bessel and Airy models, where the gap energy $E_0$ controls the interpolation, and seems to have a simple intersection theory interpretation.

hep-th

Introduction to Generalized Symmetries

These notes were prepared for a series of intensive lectures delivered at Hokkaido University, Nagoya University, Kyoto University, and Kyushu University. We begin with a brief review of higher-form symmetries, anomalies, and discrete gauge theories, before introducing non-invertible symmetries in $(1+1)$-dimensional systems. The basic structure of fusion categories is then discussed, including a discussion of categorical analogs of discrete gauging and representation theory. We subsequently turn to $(3+1)$-dimensional theories, where several physical applications of non-invertible symmetries are discussed. These notes are intended to be largely self-contained, and require no prior familiarity with subjects such as conformal field theory or lattice models.

hep-th

Planar loop integrands from cuts in $D$ dimensions

We present a direct reconstruction formula for planar loop integrands from $D$-dimensional generalized unitarity cuts in any colored theory. The reconstruction combinatorics is separated from the theory-dependent tree amplitudes entering the cuts: for the $L$-loop $n$-point color-ordered amplitude, the integrand is expressed as a sum over admissible non-scaleless scalar graphs dressed by corresponding cuts in $D$ dimensions; the coefficients are given by the universal Möbius-inversion formula of the refinement poset, or equivalently one minus the Euler characteristics of associated complexes. As an application we write down closed-formulas for loop integrands in pure Yang--Mills theory, where the required cuts are generated by gluing $D$-dimensional tree amplitudes and summing over internal gluon states. We also use the two-loop five-point case as a validation, comparing with known integrand data and after integration-by-parts reduction, with known integrated helicity amplitudes. The same framework also produces compact cut-organized data for larger examples, including the two-loop six-point and three-loop four-point cases. We also describe the corresponding simplification in maximally supersymmetric Yang--Mills theory, where the absence of bubble and triangle subgraphs reduces the relevant cut poset substantially.

hep-th