Search arXivSearch

arXiv · 2510.20100

Factors Associated with Unit-Specific Failure in a University-Level Statistics Course

Abstract

This study investigates the factors associated with failure in each of the four thematic units of a General Statistics course offered at a private university in Colombia. Unlike traditional analyses that treat performance as a single outcome, this research disaggregates results by unit: Exploratory Data Analysis, Probability and Random Variables, Statistical Inference, and Linear Regression -- highlighting distinct challenges across content areas. Based on a sample of 186 undergraduate students from Engineering, Geology, and Interactive Design programs, the study combines exam performance data with self-perceived preparedness surveys to develop unit-specific logistic regression models. The findings reveal consistent structural disadvantages for students from non-engineering programs, especially in concept-heavy units such as Inference and Regression. Academic stage and perception of competence also emerged as important predictors, though their effects varied across units. The results align with prior research on statistical thinking and self-efficacy, and support the need for targeted pedagogical interventions and curricular alignment. This disaggregated approach offers a more nuanced understanding of academic vulnerability in statistics education and contributes to the design of evidence-based, context-sensitive strategies to reduce failure and improve learning outcomes.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Biviana Marcela Suarez Sierra. 2025-10-23. Factors Associated with Unit-Specific Failure in a University-Level Statistics Course. https://arxiv.org/abs/2510.20100

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

See You at the Posterior Line: Learning Bayesian Modeling Through a Car Racing Game

We present an interactive classroom activity designed to address a central challenge in teaching introductory Bayesian statistics: how to formalize subjective knowledge and available information into prior distributions and then update them with empirical data. Role-playing as data analysts for a racing team, students evaluate candidate tires by converting qualitative engineering reports into prior distributions, collecting primary data via a virtual racing game, and using a Beta-Binomial model to inform team strategy. This discovery-based exercise allows small groups to observe directly how different prior choices and sample data jointly shape posterior inference. Student feedback ($n=32$) highlights high enjoyment, engagement and improved conceptual clarity. Open-access materials to implement the activity are provided, alongside recommendations for adapting it to other teaching contexts.

stat.OT

Why is Regularization Underused? An Empirical Study on Trust and Adoption of Statistical Methods

Statistical practice does not automatically follow methodological innovation. Regularization methods, widely advocated to reduce overfitting and stabilize inference, are readily available in modern software, but are not consistently used by data analysts. We investigate this implementation gap in a large-scale empirical study of trust in, and acceptance of, regularization techniques, based on $N = 606$ data analysts. Drawing on measurement frameworks from technology acceptance research, we survey practitioners and embed a randomized experiment to test whether written recommendation of regularization methods increases trust or intended use. We find no evidence of such an effect. Instead, adoption intentions are strongly associated with analysts' perceptions of ease of implementation and practical benefit, such as improved bias control or interpretability. Perceived social norms also emerge as a central driver. These results indicate that uptake of statistical methodology depends less on formal recommendations than on usability, perceived utility, and community practice.

stat.OT

Exact analysis of a split--merge queue with latent Erlang-factor dependent subtask times

This paper studies a two-server split--merge queue with positively dependent subtask service times modeled through a latent-factor bivariate Erlang construction. An exact characterization of the split--merge completion time is obtained, including explicit formulas for its first two moments and the resulting mean waiting time. Under fixed marginal service-time distributions, independence is shown to stochastically increase the completion time and hence overestimate mean waiting time. Numerical illustrations show that this benchmark gap can be substantial.

stat.OT