Search arXivSearch

arXiv · 1610.08416

Minimum spanning tree filtering of correlations for varying time scales and size of fluctuations

Abstract

Based on a recently proposed $q$-dependent detrended cross-correlation coefficient $ρ_q$, we generalize the concept of minimum spanning tree (MST) by introducing a family of $q$-dependent minimum spanning trees ($q$MST) that are selective to cross-correlations between different fluctuation amplitudes and different time scales. They inherit this ability directly from the coefficients $ρ_q$ that are processed here to construct a distance matrix. Conventional MST with detrending corresponds in this context to $q=2$. We apply the $q$MSTs to sample empirical data from the stock market and discuss the results. We show that the $q$MST graphs can complement $ρ_q$ in disentangling correlations that cannot be observed by the MST graphs based on $ρ_{\rm DCCA}$ and, therefore, they can be useful in many areas where the multivariate cross-correlations are of interest. We apply our method to data from the stock market and obtain more information about correlation structure of the data than by using $q=2$ only. We show that two sets of signals that differ from each other statistically can give comparable trees for $q=2$, while only by using the trees for $q \ne 2$ we become able to distinguish between these sets. We also show that a family of $q$MSTs for a range of $q$ express the diversity of correlations in a manner resembling the multifractal analysis, where one computes a spectrum of the generalized fractal dimensions, the generalized Hurst exponents, or the multifractal singularity spectra: the more diverse the correlations are, the more variable the tree topology is for different $q$s. Our analysis exhibits that the stocks belonging to the same or similar industrial sectors are correlated via the fluctuations of moderate amplitudes, while the largest fluctuations often happen to synchronize in those stocks that do not necessarily belong to the same industry.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Jaroslaw Kwapien, Pawel Oswiecimka, Marcin Forczek, Stanislaw Drozdz. 2017-05-04. Minimum spanning tree filtering of correlations for varying time scales and size of fluctuations. https://doi.org/10.1103/physreve.95.052313

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Causal Discovery in Financial Markets: A Framework for Nonstationary Time-Series Data

This paper introduces a new causal structure learning method for nonstationary time series data, a common data type found in fields such as finance, economics, healthcare, and environmental science. Our work builds upon the constraint-based causal discovery from nonstationary data algorithm (CD-NOD). We introduce a refined version (CDNOTS) which is designed specifically to account for lagged dependencies in time series data. We compare the performance of different algorithmic choices, such as the type of conditional independence test and the significance level, to help select the best hyperparameters given various scenarios of sample size, problem dimensionality, and availability of computational resources. Using the results from the simulated data, we apply CDNOTS to a broad range of real-world financial applications in order to identify causal connections among nonstationary time series data, thereby illustrating applications in factor-based investing, portfolio diversification, and comprehension of market dynamics.

q-fin.ST

Extreme Value Analysis for Finite, Multivariate and Correlated Systems with Finance as an Example

Extreme values and the tail behavior of probability distributions are essential for quantifying and mitigating risk in complex systems of all kinds. In multivariate settings, accounting for correlations is crucial. Although extreme value analysis for infinite correlated systems remains an open challenge, we propose a practical framework for handling a large but finite number of correlated time series. We develop our approach for finance as a concrete example but emphasize its generality. We study the extremal behavior of high-frequency stock returns after rotating them into the eigenbasis of the correlation matrix. This separates and extracts various collective effects, including information on the correlated market as a whole and on correlated sectoral behavior from idiosyncratic features, while allowing us to use univariate tools of extreme value analysis. This holds even for high-frequency data where discretization effects normally complicate analysis. We employ a peaks-over-threshold approach and thereby fully avoid the analysis of block maxima. We estimate the tail shape of the rotated returns while explicitly accounting for nonstationarity, a key feature in finance and many other complex systems. Our framework facilitates tail risk estimation relative to larger trends and intraday seasonalities at both market and sectoral levels.

q-fin.ST

Does Crypto Sentiment Extremity Widen Estimated Spreads? Evidence Depends on the Specification

We examine whether extreme values of the Crypto Fear & Greed Index are associated with a daily high-low spread estimate for Bitcoin. The sample contains 2,896 BTC/USDT observations from February 2018 to January 2026. We find an unconditional extreme-minus-neutral gap of 61.99 basis points. After close-to-close realised-volatility-quintile demeaning it is 24.79 basis points, although none of the five separate quintile contrasts survives Holm correction. With quadratic realised-volatility and strictly lagged momentum controls, the HAC estimate is 11.81 basis points (95% CI [-2.31,25.93], p=.101). A fixed non-parametric stratification gives 20.44 basis points (p=.0195 under circular shifts), while separate models for a zero-floored estimate's incidence and positive magnitude are imprecise. The results therefore show only a descriptive, specification-dependent association. We conclude that they do not establish a stable or causal liquidity premium.

q-fin.ST