Search arXivSearch

arXiv · 2504.20993

GDP-GFCF Dynamics Across Global Economies: A Comparative Study of Panel Regressions and Random Forest

Abstract

This study examines the relationship between GDP growth and Gross Fixed Capital Formation (GFCF) across developed economies (G7, EU-15, OECD) and emerging markets (BRICS). We integrate Random Forest machine learning (non-linear regression) with traditional econometric models (linear regression) to better capture non-linear interactions in investment analysis. Our findings reveal that while GDP growth positively influences corporate investment, its impact varies significantly by region. Developed economies show stronger GDP-GFCF linkages due to stable financial systems, while emerging markets demonstrate weaker connections due to economic heterogeneity and structural constraints. Random Forest models indicate that GDP growth's importance is lower than suggested by traditional econometrics, with lagged GFCF emerging as the dominant predictor-confirming investment follows path-dependent patterns rather than short-term GDP fluctuations. Regional variations in investment drivers are substantial: taxation significantly influences developed economies but minimally affects BRICS, while unemployment strongly drives investment in BRICS but less so elsewhere. We introduce a parallelized p-value importance algorithm for Random Forest that enhances computational efficiency while maintaining statistical rigor through sequential testing methods (SPRT and SAPT). The research demonstrates that hybrid methodologies combining machine learning with econometric techniques provide more nuanced understanding of investment dynamics, supporting region-specific policy design and improving forecasting accuracy.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Alina Landowska, Robert A. Kłopotek, Dariusz Filip, Konrad Raczkowski. 2025-04-29. GDP-GFCF Dynamics Across Global Economies: A Comparative Study of Panel Regressions and Random Forest. https://arxiv.org/abs/2504.20993

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Decomposing Wage Stagnation: Employment Reallocation, Wage Structure,and Demographics

Japan's average log real hourly wages rose until the mid-1990s, declined through the mid-2010s, and partially recovered thereafter. This paper decomposes these changes over 1980-2024 into four components: demographic change across worker types, changes in relative employment shares across job types, changes in relative log wages across job types, and unweighted mean wage growth. The framework combines a shift-share decomposition across worker types with an extension of the Olley-Pakes decomposition across job types within worker types, separating employment reallocation from changes in relative wage structure. The four components vary across periods. Before 1996, unweighted mean wage growth and changes in relative wage structure contribute positively, while demographic change and employment reallocation contribute negatively. During 1996-2014, all four components are negative. After 2014, the recovery mainly reflects unweighted mean wage growth. Employment reallocation and changes in relative wage structure contribute differently across dimensions of job type.

econ.GN

Complements or Substitutes? Technology Adoption and Clinical Care Utilization: Evidence from Automated Insulin Delivery

Whether medical technology reduces or increases demand for professional care is central to understanding its effects on healthcare utilization and costs. We study automated insulin delivery (AID) adoption among 1,608 adults with type 1 diabetes treated in four specialist clinics of the Italian National Health Service, 283 of whom adopt AID. Because adoption is clinically targeted and staggered, we combine risk-set coarsened exact matching with a staggered event-study design. Matching retains 181 adopters with comparable untreated controls. We further adjust for differential pre-adoption trends by extrapolating the untreated trajectory in event time. After adjustment, AID adoption is associated with greater routine outpatient engagement. The estimated effect on the probability of a diabetologist visit is 16.3 percentage points one semester after adoption and 39.3 percentage points four semesters after adoption, relative to a 55.2% visit probability in the semester before adoption. HbA1c testing and a broader process-of-care index show similar positive patterns, although the evidence is less precise. The visit estimates remain sensitive to the trend-extrapolation assumption and should therefore be interpreted conditionally on that restriction. The findings are consistent with a task-based view in which automation substitutes for routine dosing decisions while complementing clinical labor through interpretive, adjustment, and supervisory tasks. Patient-facing automation may therefore reorganize rather than simply reduce healthcare use.

econ.GN

Mining Meaning: Measurement Error in AI-Assisted Literature Reviews

Researchers increasingly use generative AI, particularly large language models (LLMs), to automate tasks across the research pipeline. We study the reliability of these tools at the reading, classification, and synthesis of large bodies of academic literature. We frame LLM-assisted literature reviews as a measurement problem, treating models as measurement systems and tracing how their errors affect downstream conclusions. As a test case, we use three different implementations of ChatGPT to identify and extract metadata from economics papers that use rainfall as an instrumental variable. We benchmark each implementation against a subset of human-labeled evaluation data, and then deploy those implementations to extract metadata from the full corpus. The LLMs perform well on binary classification, but performance deteriorates as tasks demand greater contextual interpretation. More importantly, how much researchers can rely on model outputs depends not only on the complexity of the reading task but also on the type of claims the data is asked to support. The same amount of measurement error substantially affects paper-level claims while having little effect on broader claims about the literature. Measurement error in LLM-generated data is thus most consequential at precisely the level of detail that constitutes an LLM's principal value added over human reviewers. We conclude that standard model performance metrics are informative about the quality of generated data but do not by themselves establish the credibility of downstream inference. Researchers must also evaluate whether substantive claims are robust to the measurement system used to generate the underlying data.

econ.GN