Search arXiv⌕ Search

arXiv · 2610.03685

Does a mortality schedule need a Makeham term? Calibrating the likelihood-ratio test

Abstract

Whether a fitted mortality model needs a Makeham term, the non-negative constant representing background mortality, is commonly decided with a likelihood-ratio test. Because the constant cannot be negative, testing its absence places the parameter on the boundary of its range, and the usual chi-squared calibration does not apply. Assuming independent Poisson death counts and standard regularity conditions, we show that when the constant is the only parameter on a boundary the statistic converges to an equal mixture of a point mass at zero and a chi-squared distribution with one degree of freedom, so that at the 5% level the critical value is 2.71 rather than 3.84. We give conditions under which the correction holds for Makeham models, verify them for Gompertz-Makeham and gamma-Gompertz-Makeham, and quantify the information the data carry about the constant, which fixes the local power of the test, the smallest term it can detect, and how fast detectability falls as the age window narrows. When a gamma-frailty variance is estimated and its true value is also zero, the limit is no longer a chi-squared mixture and the usual correction rejects too often. Monte Carlo experiments show that the conventional cutoff rejects at half the nominal level, that a correctly calibrated test can still miss a term at its own detection threshold in most samples, and that excess rejections under an estimated zero frailty variance persist as exposure grows.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Silvio C. Patricio. 2026-10-02. Does a mortality schedule need a Makeham term? Calibrating the likelihood-ratio test. https://arxiv.org/abs/2610.03685

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Direct estimation and inference of differential Granger causality between two high-dimensional time series

Comparing dependence structures between two related multivariate time series is of fundamental interest in many scientific applications, where changes may occur in both directed temporal interactions and contemporaneous connectivity. Modeling each time series by a vector autoregressive (VAR) model, we propose a new framework for estimation and inference of differential Granger causality (DiffGC) and differential network (DiffNet) structures in high dimensions. The proposed method is based on a novel estimating equation derived from the Yule-Walker equations that directly links the difference between VAR transition matrices to the corresponding difference between precision matrices. Unlike separate estimation strategies that estimate the two VAR models individually and then take their difference, the proposed method directly targets the differential structures and requires sparsity only of the differences, thereby accommodating potentially dense individual networks, including hub structures. Building on the direct estimators, we develop selective inference procedures for both DiffNet and DiffGC parameters, providing valid post-selection inference while accounting for the data-driven screening process. Theoretically, we establish convergence rates and support recovery guarantees for the proposed estimators, derive asymptotic distributions for the selective inference targets, and obtain new consistency results for DiffNet estimation under weaker assumptions than existing methods. Simulation studies confirm the theoretical convergence rates and demonstrate accurate support recovery and robust inferential performance. An application to resting-state electroencephalography (EEG) data identifies substantial changes in both contemporaneous and Granger-causal connectivity across recording sessions.

stat.ME↗

Modelling Territorial Dynamics through Statistical Mechanics: An Application to Resident Foreign Population

This study proposes an energy-based framework for exploring local variations around observed territorial configurations in Official Statistics. The observed distribution serves as the initial reference, without assuming equilibrium. A register-derived graph represents structural similarities between municipalities, while socio-economic composite indices are synthesised through Principal Component Analysis into a univariate external field. A Continuous Ising model and a Langevin-based procedure provide complementary stochastic exploration strategies. Simulated Annealing progressively limits exploration, favouring lower-energy configurations under each specification. Recorded energy trajectories and municipal deviations characterise the resulting ensembles. Conformal Prediction provides calibrated intervals whose widths and empirical coverage describe their relationship with fixed observed municipal values under repeated subsampling. The application to resident foreign population shares in North-East Italy yields simulated municipal averages close to the observed values, but substantially different calibrated interval widths between the two procedures. Comparisons across territorial groups reveal heterogeneous associations with socio-economic profiles, without a uniform ordering of interval widths between central and peripheral municipalities. The framework connects territorial interactions, socio-economic information, and stochastic variation within an exploratory analysis. Its contribution is to examine how nearby simulated alternatives relate to the observed configuration and to formulate hypotheses about territorial heterogeneity. Further applications and sensitivity analyses would assess the robustness of these interpretations.

stat.ME↗

Parameter Estimation for Differential Equation Models Using Generalized Profiling: A Computational Tutorial

Parameter estimation connects mathematical models to real-world data and decision making across many scientific and industrial applications. Standard approaches such as maximum likelihood estimation and Markov chain Monte Carlo estimate parameters by repeatedly solving the model, which often requires numerical solutions of differential equation models. In contrast, generalized profiling (also called parameter cascading) focuses directly on the governing differential equation(s), linking the model and data through a penalized likelihood that explicitly measures both the data fit and model fit. Despite several advantages, generalized profiling is relatively rarely used in practice. This tutorial-style article outlines a set of self-directed computational exercises that facilitate skills development in applying generalized profiling to a range of ordinary differential equation models. All calculations can be repeated using reproducible open-source Jupyter notebooks that are available on GitHub.

stat.ME↗