Search arXivSearch

arXiv · 2406.08297

Improving subgroup analysis using methods to extend inferences to specific target populations

Abstract

Subgroup analyses are common in epidemiologic and clinical research. Unfortunately, restriction to subgroup members to test for heterogeneity can yield imprecise effect estimates. If the true effect differs between members and non-members due to different distributions of other measured effect measure modifiers (EMMs), leveraging data from non-members can improve the precision of subgroup effect estimates. We obtained data from the PRIME RCT of panitumumab in patients with metastatic colon and rectal cancer from Project Datasphere(TM) to demonstrate this method. We weighted non-Hispanic White patients to resemble Hispanic patients in measured potential EMMs (e.g., age, KRAS distribution, sex), combined Hispanic and weighted non-Hispanic White patients in one data set, and estimated 1-year differences in progression-free survival (PFS). We obtained percentile-based 95% confidence limits for this 1-year difference in PFS from 2,000 bootstraps. To show when the method is less helpful, we also reweighted male patients to resemble female patients and mutant-type KRAS (no treatment benefit) patients to resemble wild-type KRAS (treatment benefit) patients. The PRIME RCT included 795 non-Hispanic White and 42 Hispanic patients with complete data on EMMs. While the Hispanic-only analysis estimated a one-year PFS change of -17% (95% C.I. -45%, 8.8%) with panitumumab, the combined weighted estimate was more precise (-8.7%, 95% CI -22%, 5.3%) while differing from the full population estimate (1.0%, 95% CI: -5.9%, 7.5%). When targeting wild-type KRAS patients the combined weighted estimate incorrectly suggested no benefit (one-year PFS change: 0.9%, 95% CI: -6.0%, 7.2%). Methods to extend inferences from study populations to specific targets can improve the precision of estimates of subgroup effect estimates when their assumptions are met. Violations of those assumptions can lead to bias, however.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Michael Webster-Clark, Anthony A. Matthews, Alan R. Ellis, Alan C. Kinlaw, Robert W. Platt. 2024-07-09. Improving subgroup analysis using methods to extend inferences to specific target populations. https://arxiv.org/abs/2406.08297

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

From Metrics to Decisions in NBA Analytics: A Critical Integrative Review and Decision-Readiness Framework

National Basketball Association (NBA) teams have increasingly detailed metrics, but better predictions do not necessarily improve decisions. This critical integrative review draws on prior reviews, citation tracing, and topic searches across seven research streams: on-court action, player value, role, lineup synergy, availability, draft and development, and contracts and roster construction. An observation-state-action-decision-evaluation chain organizes the synthesis. Six decision-readiness gates guide our assessment: point-in-time validity, uncertainty, context portability, action feasibility, opportunity-set observability, and evaluation, with requirements matched to each claim. The reviewed literature is strongest in measuring and predicting individual components of a decision. Evidence is less developed at interfaces that combine components, transfer them across settings, and compare feasible actions. We outline a proposed deployment workflow, a reporting contract, and a research agenda covering player transport, role substitution, roster fragility, legal action generation, and asset valuation. The 2023 collective bargaining agreement and forthcoming 3-2-1 Draft Lottery illustrate how institutional changes generate research questions. Models should inform evaluable comparisons of feasible choices. While its effect on organizational decision quality remains an empirical question, the framework provides a diagnostic and reporting structure for matching decision claims to evidence requirements.

stat.AP

Bayesian calibration of adaptive-behavior SIR models for multi-wave COVID-19 incidence in New York City

Epidemic incidence reflects both transmission dynamics and adaptive human behavior, yet these mechanisms may be difficult to distinguish from aggregate case data alone. We calibrated four susceptible--infected--recovered (SIR) specifications to weekly confirmed COVID-19 incidence in New York City from June to December 2020, comparing a single continuous SIR trajectory, a wave-initialized SIR model, and two adaptive-behavior models with either shared or wave-specific transmission. Inference was performed using rejection Approximate Bayesian Computation (ABC), and in-sample reconstruction was assessed using root mean squared error (RMSE) and the weighted interval score (WIS). Reinitializing the epidemic state by wave produced the largest structural improvement over the continuous SIR trajectory, reducing mean-based RMSE by 48.6\% and WIS by 17.6\%. Adding delayed prevalence-dependent behavioral adaptation with shared transmission further reduced mean-based RMSE by 23.4\%, but yielded essentially unchanged WIS relative to the wave-initialized SIR model. Allowing transmission to vary by wave did not provide a consistent additional advantage and produced strongly asymmetric posterior-simulation trajectories. Behavioral sensitivity, response midpoint, and delay remained only weakly to partially identified. The clearest posterior structure was a negative association between transmission intensity and the behavioral midpoint, indicating that higher transmission could be compensated by behavioral responses activated at lower prevalence. Sensitivity to ordered behavioral priors further showed that reconstruction and behavioral inference depend materially on structural prior assumptions. These results suggest that adaptive mechanisms can improve multi-wave incidence reconstruction, while aggregate incidence alone is insufficient to sharply separate transmission from behavioral adaptation.

stat.AP

Short-term rental market occupancy - daily time series for 2017-2022 on 500 markets worldwide

Short-term vacation rentals, as promoted by platforms such as Airbnb, Homeaway, Vrbo, etc., are a growing component of the travel industry. This paper provides a unique, large dataset on global market occupancy for the short-term rental market using data from the American company Wheelhouse. The dataset consists of data for $500$ markets around the world. For each market, a daily occupancy time series from January $2017$ to December $2022$ is provided, allowing for studies of local and global patterns in the evolution of the short-term rental market. Additionally, the dataset includes curves representing the booking trajectory of each market and stay date up to one year prior to the stay date. This large dataset comprises a unique combination of time series and survival analysis data, and is suitable as a methodological benchmark for both classical statistical and machine learning models.

stat.AP