Search arXiv⌕ Search

arXiv · 2609.37873

Target-Aware Sequential Inference: Pooled versus Stratified Anytime-Valid Designs

Abstract

Sequential studies with heterogeneous strata often target a weighted population mean while collecting data under a different, possibly adaptive allocation. This creates two distinct design choices: how observations are allocated and whether inference is performed directly for the target or by aggregating simultaneous stratum-level confidence sequences. We compare these architectures under anytime-valid inference. For a single prespecified target, direct target sampling yields one bounded pooled process. When simultaneous stratum-level reporting, post-hoc reweighting, or robustness over several targets is required, a stratified construction aggregates local confidence sequences. For local half-widths with power-law rate n raised to minus beta, we derive the asymptotically width-optimal allocation; its exponent is 1/(1+beta), and root-n variance-adaptive boundaries yield the 2/3 rule. We establish validity under predictable adaptive sampling, show oracle tracking under a vanishing exploration floor, extend the design to uncertain target distributions and familywise best-system identification, and quantify the first-order cost of unnecessary local multiplicity. Simulations and a public benchmark replay show two robust patterns: variance adaptation can matter more than fine allocation tuning, and pooled target-specific inference can reduce stopping cost dramatically when local simultaneous guarantees are not needed. The framework connects stratified sampling, confidence sequences, adaptive allocation, and ranking and selection through a common target-aware sequential design problem.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Subir Hait. 2026-09-29. Target-Aware Sequential Inference: Pooled versus Stratified Anytime-Valid Designs. https://arxiv.org/abs/2609.37873

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Bayesian Neural-Net-Assisted Multi-Treatment Mixture Cure Survival Model with Application in Pediatric Oncology

Estimating covariate-conditional treatment effects in multi-arm oncology studies is complicated when treatment arms have common distributional features and a non-negligible fraction of patients achieve long-term remission. We propose a joint mixture cure model with covariate-dependent mixtures of log-normal kernels with treatment-specific inclusions. Both linear and neural-network-assisted non-linear covariate links are proposed. Specifically, the susceptible survival distributions use a common finite dictionary of log-normal components, and then a binary inclusion matrix determines which components are active in each treatment arm. All parameters, including the hidden bases of the neural network, are learned jointly, while the output coefficients remain treatment- or component-specific. Posterior inference is performed using gradient-based MCMC, and treatment effects are summarized by covariate-conditional differences in restricted mean survival time (RMST). Variable importance is assessed using thresholded marginal best linear projections with data partitioning. Across two simulation settings, the proposed method demonstrates good finite-sample performance, with lower RMST-based estimation error than flexsurvcure. Compared with pairwise grf fits, the proposed method yields lower RMST-contrast MSE in most comparisons while ensuring mutually coherent multi-treatment contrasts. Finally, the application to the AALL0434 trial reveals covariate-dependent patterns in RMST posterior across methotrexate-based regimens and provides new insights into how these differences vary with patient covariates, highlighting the method's practical utility for studying heterogeneous treatment effects in pediatric oncology trials.

stat.ME↗

A Three-Stage PCA Procedure for Sequentially Arriving High-Dimensional Data

We develop a three-stage adaptive procedure for principal component analysis (PCA) when high-dimensional observations are collected sequentially and additional sampling incurs a cost. The procedure balances PCA compression loss against sampling cost while selecting the retained dimension through a prescribed explained-variance criterion. Starting from a pilot sample, an intermediate stage updates the PCA quantities before determining the final sample size, thereby avoiding reliance on unknown population eigenvalues. Under suitable regularity conditions, we establish both first- and second-order efficiency relative to the population oracle. Comparison with the corresponding two-stage rule shows that the additional recalibration yields sharper second-order control and reduces the influence of the pilot stage on the final sampling decision. The theory allows the ambient dimension to exceed the sample size under appropriate covariance and spectral conditions. Simulation studies demonstrate the strong finite-sample performance of the procedure across increasing dimensions and several dense covariance structures. As a real-data application, we conduct a retrospective study of gene-expression data from 32 cancer-type cohorts in The Cancer Genome Atlas, illustrating both cost-effective early stopping and settings in which additional observations are recommended.

stat.ME↗

Batting Average as the Product of Two Rates: Skill, Luck, and the Disappearance of the .400 Hitter

Batting average factors exactly as BA = c times f, where c = (AB - SO)/AB is the rate of avoiding a strikeout and f = H/(AB - SO) is the rate at which non-strikeout at-bats become hits. Using the 256 major-league hitters with at least 300 at-bats in 2025, we show that the two factors behave very differently. Strikeout avoidance is highly repeatable, with a median year-to-year correlation of 0.86 over 19 consecutive-season pairs from 2004 to 2025. The finishing rate is not: its median correlation is 0.44, and batting average itself (0.44) is no more repeatable than its noisier factor. A bivariate logistic-normal random-effects model fit to the 2025 season estimates the correlation between the two talents at -0.68 (95\% profile interval -0.85 to -0.50), far stronger than the raw correlation of -0.44, and implies single-season reliabilities of 0.91 for c and 0.37 for f. The model also yields a closed-form bivariate shrinkage estimator in which a hitter's strikeout rate informs the estimate of his finishing rate. Applied to decades of American and National League data, the decomposition revisits Gould's explanation for the disappearance of the .400 hitter. Since the dead-ball era, the talent variance of strikeout avoidance has grown roughly 4.6-fold while that of finishing has halved, and the correlation between them has moved from near zero to about -0.53. That emerging trade-off, rather than a general narrowing of talent, accounts for the reduced spread of modern batting averages.

stat.ME↗