Search arXiv⌕ Search

arXiv · 2610.05629

Does the AI Productivity Dividend Diminish? Does the AI Productivity Dividend Diminish? A Validated Framework for Identifying the Shape of Firm-Level Returns to Artificial Intelligence

Abstract

Forecasts of the productivity dividend from artificial intelligence (AI) differ by an order of magnitude, partly because they extrapolate average gains observed among early, highly exposed adopters. Whether those gains scale linearly, flatten, or are competed away is an empirical question about the shape of the return to AI, not its average. This paper develops and validates a design-based framework for measuring that shape with linked firm-worker data. The framework combines a pre-determined, occupation-based measure of firm exposure to generative AI with a continuous-treatment difference-in-differences design around the release of ChatGPT in November 2022, three pre-specified tests of concavity, a decomposition of exposure effects into an adoption margin and a per-adopter return, a doubly robust analysis of use intensity, a staggered-adoption analysis of earlier AI adopters, and a production-function estimator that places AI in the law of motion of productivity. We validate every component on a calibrated synthetic panel of 20,000 enterprises that reproduces the structure of Statistics Canada's linked business microdata and embeds a known concave effect. The estimators recover the true dose-response curve and detect its concavity (slope difference -0.010, standard error 0.003, one-sided p = 0.002); Monte Carlo coverage of the main confidence interval is 0.97, and the concavity test has power of 0.93. The validation also exposes a practical pitfall: pre-trend tests clustered at roughly 40 industries over-reject a true null. The framework, released as open code, is ready for application to Canadian enterprise microdata, where it will inform estimates of potential output and the persistence of Canada's productivity gap.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Louis Agyekum. 2026-10-04. Does the AI Productivity Dividend Diminish? Does the AI Productivity Dividend Diminish? A Validated Framework for Identifying the Shape of Firm-Level Returns to Artificial Intelligence. https://arxiv.org/abs/2610.05629

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

DeepAJM: Deep Association Joint Model for Irregularly Sampled data

Joint Models simultaneously model longitudinal and survival outcomes, leveraging patterns in patients' longitudinal trajectory to improve the prediction of survival outcomes. The classical parametric joint models, however, rely on fixed parametric assumptions, making them susceptible to bias under model misspecification and smaller sample sizes. We propose a deep joint model, DeepAJM, that does not require any parametric assumptions, while retaining a partially interpretable, per-longitudinal-outcome association structure. The joint model uses an encoder-decoder (sequence-to-sequence) architecture to learn the latent structure in patients' time-varying covariate trajectories. The model links the longitudinal processes to the survival processes through a learned interpretable association structure, in which each longitudinal output from the decoder gets remodulated by baseline covariates before it contributes to the risk scores from the survival head of the architecture. The model was evaluated on three datasets ( a cardiovascular-disease EHR cohort, a primary biliary cirrhosis (PBC2) dataset, and a simulated dataset) against a classical parametric joint model, TransformerJM, DA-LSTM and a Cox-based survival-only model. All models were assessed using C-index, integrated brier score (IBS), time-dependent AUROC, and time-dependent AUPRC. Our model achieved the best discrimination in terms of the C-index, time-dependent AUROC, and AUPRC across all datasets.

stat.AP↗

Bayesian Optimization for Dose Finding with Two Agents: Participant Allocation and Final Selection

In two-agent dose-finding trials, the next cohort should help identify a combination for final selection. We studied a constrained knowledge-gradient (cKG) rule with one-cohort lookahead that updates independent Gaussian-process models of efficacy and continuous toxicity, reapplies a probability criterion for mean toxicity, and evaluates the resulting selection. We derived a deterministic calculation over a fixed set of dose combinations, holding fitted model parameters fixed during each hypothetical update. We compared cKG with constrained expected improvement (cEI) and two toxicity-only rules, targeted mean squared error (tMSE) and entropy, in four synthetic scenarios. In the primary obstructive sleep apnea (OSA)-derived scenario, averaged equally over strata and five probability cutoffs, cKG assigned fewer participants to combinations above the true mean-toxicity limit than tMSE (17.92% versus 27.08%), but selected such combinations more often at trial completion (18.80% versus 11.85%). Compared with cEI, cKG had higher mean simulated reduction in the 4%-desaturation apnea-hypopnea index (AHI4) at final selection (7.46 versus 6.72 events/hour), more above-limit final selections (18.80% versus 10.50%), and more above-limit assignments (17.92% versus 15.10%). Across scenarios, its efficacy advantage over cEI was smaller under stricter toxicity criteria. Continuous outcomes, uncalibrated toxicity limits, and a rule that still selects a combination when none meets the criterion limit clinical interpretation. Allocation and final-selection toxicity should be reported separately, alongside efficacy.

stat.AP↗

A Sensitivity-based Framework for Calibrating Coupled Ordinary Differential Equations under Model Discrepancy

Coupled systems of ordinary differential equations (ODEs) are widely used to model complex physical and biological processes. In these applications, ODE model parameters must be calibrated to field data to enable prediction and parameter inference. In practice, the governing ODEs are an imperfect approximation to the true system, leading to non-negligible model discrepancy that must be incorporated to avoid biased parameter estimates. However, it is well known that highly flexible discrepancy models can confound discrepancy and simulator parameters, leading to poor identifiability. As a result, practitioners often must impose problem-specific prior constraints to adequately regularize the discrepancy, which can be challenging and cumbersome. In this work, we propose a novel calibration framework for coupled, multi-output ODE systems that enforces automatic, model-structural constraints on the discrepancy, improving identifiability without requiring application-specific discrepancy priors. Forward-model gradients are computed via sensitivity equations, solved jointly with the ODE system as a coupled initial value problem. Posterior inference is performed using an adaptive Metropolis-within-Gibbs sampler tailored to the resulting constrained posterior. We demonstrate the approach on three model systems: a mass-spring oscillator, a model of infectious disease spread, and a neuron firing model. We show that our method leads to improved parameter inference and predictive accuracy compared to standard calibration approaches.

stat.AP↗