Search arXiv⌕ Search

arXiv · 2009.04729

Non-asymptotic Optimal Prediction Error for Growing-dimensional Partially Functional Linear Models

Abstract

Under the reproducing kernel Hilbert spaces (RKHS), we consider the penalized least-squares of the partially functional linear models (PFLM), whose predictor contains both functional and traditional multivariate parts, and the multivariate part allows a divergent number of parameters. From the non-asymptotic point of view, we focus on the rate-optimal upper and lower bounds of the prediction error. An exact upper bound for the excess prediction risk is shown in a non-asymptotic form under a more general assumption known as the effective dimension to the model, by which we also show the prediction consistency when the number of multivariate covariates $p$ slightly increases with the sample size $n$. Our new finding implies a trade-off between the number of non-functional predictors and the effective dimension of the kernel principal components to ensure prediction consistency in the increasing-dimensional setting. The analysis in our proof hinges on the spectral condition of the sandwich operator of the covariance operator and the reproducing kernel, and on sub-Gaussian and Berstein concentration inequalities for the random elements in Hilbert space. Finally, we derive the non-asymptotic minimax lower bound under the regularity assumption of the Kullback-Leibler divergence of the models.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Huiming Zhang, Xiaoyu Lei. 2022-09-30. Non-asymptotic Optimal Prediction Error for Growing-dimensional Partially Functional Linear Models. https://arxiv.org/abs/2009.04729

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Topological Periodicity Test (TopPT) via Confidence Bound of Time-Delay Embeddings

Time-delay embedding is a fundamental technique in Topological Data Analysis (TDA) for reconstructing phase-space dynamics of time-series data, where persistent homology can reveal loops associated with periodicity. However, rigorous statistical uncertainty quantification for these features remains underdeveloped. First, we analyze the topology of time-delay embeddings, showing that the embedded trajectory is homotopy equivalent to a circle ($S^1$) for periodic signals and contractible for non-periodic ones. We also prove a positive lower bound on the embedding reach, ensuring stable topological features. Second, we develop a subsampling approach to construct confidence bounds for persistence diagrams. Under standard manifold regularity conditions, we derive data-dependent bounds with asymptotic guarantees. Finally, we propose Topological Periodicity Test (TopPT), a hypothesis testing framework for periodicity with asymptotically controlled type I and type II error rates. Experiments on bounded-error synthetic data show that raw TDA detects periodic alternatives while avoiding false rejections on structured non-periodic signals, and that the robust rule is conservative after interpolation-error correction. On PhysioNet Fantasia and BIDMC respiratory waveforms, TDA detects most records but only a subset of local windows, unlike scalar periodogram baselines.

math.ST↗

A note on estimation of quarticity based on spot volatility

We study an estimator of twice the integrated quarticity of a continuous Itô semimartingale, constructed from squared returns and local estimates of spot volatility. For a shrinking estimation window, we establish a functional stable central limit theorem for a bias-corrected version of the estimator. The limiting conditional variance is $56\int_0^t c_s^4\dd s$, where $c$ denotes spot variance. The proof identifies a common martingale approximation for the three components of the estimator and determines their joint covariance structure. The volatility process may have jumps. We also compare the limiting variance with those of the realized-quarticity and quadratic spot-volatility estimators.

math.ST↗

Revisiting the Brunner-Munzel test from the viewpoint of local linear approximation

The Brunner-Munzel (BM) test is a nonparametric test for two independent samples that evaluates whether observations from one group tend to be greater than observations from another group, or vice versa. The BM test has a broader scope of application than the Mann-Whitney $U$ test because it does not assume equal variances between the two groups. However, the meaning of the BM test statistic is difficult to understand intuitively, which may be one of the factors hindering the widespread use of the BM test. To alleviate this problem, in this paper, I introduce an alternative interpretation of the BM test statistic from the viewpoint of local linear approximation. It is shown that the variance estimator for the sample stochastic superiority used in the BM test can be derived using local linear approximation, in which the influence of each observation on the sample stochastic superiority is assumed to be additive. This simple interpretation will help practitioners decide to use the BM test without hesitation.

math.ST↗