Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Applied Statistics (1191)
- Social and Behavioral Sciences (1133)
- Statistical Methodology (426)
- Statistical Models (216)
- Survival Analysis (114)
-
- Biostatistics (96)
- Medicine and Health Sciences (80)
- Multivariate Analysis (71)
- Public Health (65)
- Probability (63)
- Life Sciences (57)
- Applied Mathematics (50)
- Genetics and Genomics (50)
- Design of Experiments and Sample Surveys (49)
- Longitudinal Data Analysis and Time Series (45)
- Categorical Data Analysis (44)
- Epidemiology (44)
- Other Statistics and Probability (41)
- Data Science (40)
- Microarrays (34)
- Numerical Analysis and Computation (33)
- Clinical Trials (32)
- Computer Sciences (31)
- Genetics (31)
- Economics (29)
- Econometrics (27)
- Economic Theory (27)
- Institution
-
- Wayne State University (1091)
- COBRA (336)
- Central Bank of Nigeria (26)
- University of Kentucky (12)
- East Tennessee State University (11)
-
- Georgia Southern University (8)
- Stephen F. Austin State University (8)
- Southern Methodist University (7)
- Old Dominion University (6)
- Portland State University (6)
- University of Arkansas, Fayetteville (6)
- University of Denver (6)
- Rochester Institute of Technology (5)
- The British University in Egypt (5)
- University of Nebraska - Lincoln (5)
- California Polytechnic State University, San Luis Obispo (4)
- Claremont Colleges (4)
- Marshall University (4)
- University of Connecticut (4)
- University of Louisville (4)
- University of New Mexico (4)
- University of North Florida (4)
- University at Albany, State University of New York (3)
- University of Central Florida (3)
- University of Malaya (3)
- Utah State University (3)
- Virginia Commonwealth University (3)
- Western Michigan University (3)
- City University of New York (CUNY) (2)
- Kennesaw State University (2)
- Keyword
-
- Bootstrap (40)
- Bias (32)
- Simulation (32)
- Monte Carlo simulation (28)
- Power (27)
-
- Confidence interval (25)
- Statistics (25)
- Mean squared error (20)
- Sample size (20)
- Type I error (20)
- Effect size (19)
- Maximum likelihood estimation (19)
- Monte Carlo (19)
- Robustness (19)
- Missing data (18)
- Permutation test (18)
- Multicollinearity (17)
- Prediction (17)
- Regression (17)
- Confidence intervals (16)
- Estimation (16)
- Model selection (16)
- Logistic regression (15)
- Bayesian (14)
- Efficiency (14)
- Heteroscedasticity (14)
- Longitudinal data (14)
- Reliability (14)
- Causal inference (13)
- SPSS (13)
- Publication Year
- Publication
-
- Journal of Modern Applied Statistical Methods (1091)
- U.C. Berkeley Division of Biostatistics Working Paper Series (116)
- Harvard University Biostatistics Working Paper Series (73)
- UW Biostatistics Working Paper Series (55)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (43)
-
- CBN Journal of Applied Statistics (JAS) (26)
- Electronic Theses and Dissertations (26)
- The University of Michigan Department of Biostatistics Working Paper Series (24)
- COBRA Preprint Series (20)
- Theses and Dissertations--Statistics (11)
- College of Graduate Studies: Theses & Dissertations (8)
- Articles (5)
- Basic Science Engineering (5)
- Graduate Theses and Dissertations (5)
- Theses and Dissertations (5)
- Department of Statistics: Dissertations, Theses, and Student Research (4)
- Dissertations (4)
- Statistical Science Theses and Dissertations (4)
- Theses, Dissertations and Capstones (4)
- Honors Scholar Theses (3)
- Human Rights & Human Welfare (3)
- Memorial Sloan-Kettering Cancer Center, Dept. of Epidemiology & Biostatistics Working Paper Series (3)
- Statistics (3)
- UNF Graduate Theses and Dissertations (3)
- Applications and Applied Mathematics: An International Journal (AAM) (2)
- Branch Mathematics and Statistics Faculty and Staff Publications (2)
- CMC Senior Theses (2)
- Data Science and Data Mining (2)
- Dissertations and Theses (Open Access) (2)
- Electronic Theses & Dissertations (2024 - present) (2)
- Publication Type
Articles 1 - 30 of 1633
Full-Text Articles in Statistical Theory
Machine Learning For Predictive Energy And Emissions Modeling Of Vehicles And Power Grids In The United States, S M Tanvir Faysal Alam Chowdhoury
Machine Learning For Predictive Energy And Emissions Modeling Of Vehicles And Power Grids In The United States, S M Tanvir Faysal Alam Chowdhoury
Dissertations
The environmental benefits of electric vehicle (EV) adoption depend on more than replacing internal combustion engine vehicles with electric powertrains. EV adoption reshapes electricity demand, interacts with regional generation mixes, and influences travel behavior and congestion, creating a coupled transportation-energy system in which vehicle and power-plant emissions must be evaluated together. This dissertation develops machine-learning frameworks for predicting energy consumption and emissions from vehicles and power grids under rising EV adoption. The first component forecasts grid emissions from EV charging. Using simulation data from NREL's Cambium database, a Prophet-based time-series framework predicts carbon dioxide, nitrous oxide, and methane emission rates …
(R2174) A Stationary First-Order Autoregressive Process With New Discrete Lindley Marginal Distribution, Tégawendé Martin Kabore, Jean-Etienne Ouindllassida Ouédraogo
(R2174) A Stationary First-Order Autoregressive Process With New Discrete Lindley Marginal Distribution, Tégawendé Martin Kabore, Jean-Etienne Ouindllassida Ouédraogo
Applications and Applied Mathematics: An International Journal (AAM)
This paper introduces a new stationary first-order autoregressive integer-valued process with a marginal distribution following the New Discrete Lindley distribution, referred to as NDLINAR( 1). The process is developed to model over-dispersed count time series exhibiting a mixed behavior arising from the combination of multiple distributions. Its statistical properties are thoroughly investigated, and the parameters are estimated using conditional maximum likelihood. The asymptotic properties of the estimators are also analyzed. The performance of the proposed model is evaluated by comparing it with other existing INAR(1) processes through applications to real datasets. Results demonstrate that the NDL-INAR(1) process effectively captures the …
A Predictive Coding Account Of Spatial Working Memory Following Prophylactic Levetiracetam Administration Prior To Traumatic Brain Injury, Omeima Mutwali
A Predictive Coding Account Of Spatial Working Memory Following Prophylactic Levetiracetam Administration Prior To Traumatic Brain Injury, Omeima Mutwali
Dissertations, Theses, and Capstone Projects
Traumatic brain injury (TBI) symptom prevention and remediation is an important area of research that would benefit vulnerable groups, including active-duty and veteran soldiers. These patients can sustain penetrative forces in fields of combat or in training, which result in focal lesions that trigger inflammatory and degenerative processes in the brain. Both primary and secondary injuries are associated with changes to cognition, behavior and affective state. This disease poses increased risk of epileptogenesis, as well. Given these outcomes, prior research has evaluated levetiracetam (LEV) as a prophylactic treatment for seizures, cognitive deficits and negative emotionality. LEV acts as a presynaptic …
On M-Estimation: From Theory To Examples, Alexander Yuan
On M-Estimation: From Theory To Examples, Alexander Yuan
Master's Theses
M-estimation provides a unified framework for statistical procedures defined as optimizers of data-dependent criterion functions. This thesis gives an expository account of M-estimation in classical and high-dimensional settings. The classical part develops weak convergence, empirical process tools, and the argmax framework for studying consistency, rates of convergence, and weak limits. Examples including least squares, maximum likelihood, robust location estimation, change-point estimation, and empirical risk minimization illustrate regular and non-regular asymptotic behavior.
The high-dimensional part studies regularized M-estimators, where the focus shifts to finite-sample error bounds and model selection guarantees. Topics include decomposable regularizers, restricted strong convexity, non-convex penalties, and sparsistency. …
An Analytical Framework For Quantifying Urban And Community Resilience To Natural Hazards From Cell-Phone Gps-Location And Traffic-Flow Data, Georgios Chatzikyriakidis
An Analytical Framework For Quantifying Urban And Community Resilience To Natural Hazards From Cell-Phone Gps-Location And Traffic-Flow Data, Georgios Chatzikyriakidis
Civil and Environmental Engineering Theses and Dissertations
Urban areas are increasingly exposed to natural hazards while accommodating a growing share of the global population, yet a consistent science-based framework for quantifying urban and community resilience remains lacking. This dissertation develops a physics-based analytical framework grounded in statistical mechanics and the quantitative theory of Brownian motion. A city is conceptualized as a complex medium in which citizens move analogously to Brownian particles within a viscoelastic environment, influenced by socioeconomic interactions and infrastructure functionality.
A central premise is that urban resilience, interpreted as engineering resilience (an outcome), can be quantified through a single metric: the mean-square displacement MSD=⟨r²(t)⟩, of …
Bayesian Spatiotemporal Model For Counterfactual Estimation In Socioeconomic Studies, Duwani W. Gonzalez
Bayesian Spatiotemporal Model For Counterfactual Estimation In Socioeconomic Studies, Duwani W. Gonzalez
Statistical Science Theses and Dissertations
Impact evaluations of regional development programs often require estimating counterfactual outcomes for a small number of treated regions using survey-based areal data. In practice, evaluators typically rely on two-group quasi-experimental methods such as propensity score matching (PSM) and Difference-in-Differences (DiD). These approaches perform poorly when only a few regions receive treatment, and when the set of observed covariates is limited or only partially relevant. Moreover, they typically do not explicitly exploit the spatial and temporal dependence present in survey-based areal data such as in ACS (American Community Survey). This dissertation develops a family of Bayesian spatial predictive models for directly …
The Impatience Of Winning: An Analysis Of Time Discounting, Predictive Modeling, And The Nba Draft, Alec R. Plante
The Impatience Of Winning: An Analysis Of Time Discounting, Predictive Modeling, And The Nba Draft, Alec R. Plante
Business and Economics Honors Papers
This paper examines whether NBA draft decisions can be better explained by incorporating non-geometric time discounting into a model of general manager decision making. Using a dataset of 285 NBA draft prospects over a 12-year period, the impact of college statistics on Value Over Replacement Player (VORP) is determined, and these impact values are then used to create a “predicted” VORP for the first 4 seasons of each player’s career: a projection of what a general manager might think of a prospect’s future value given their college statistics. Following this, geometric and hyperbolic time discounting models are applied to estimate …
Efficacy Analysis In Clinical Trials: A Comprehensive Review Of Statistical And Machine Learning Approaches, Dhrubajyoti Ghosh, Samhita Pal
Efficacy Analysis In Clinical Trials: A Comprehensive Review Of Statistical And Machine Learning Approaches, Dhrubajyoti Ghosh, Samhita Pal
Faculty Articles
Efficacy testing is a cornerstone of clinical trials, ensuring that medical interventions achieve their intended therapeutic effects. Over the decades, a wide range of statistical methodologies have been developed to address the complexities of clinical trial data, including parametric, nonparametric, Bayesian, and machine learning approaches. Parametric methods, such as t-tests, ANOVA, and LMMs, have traditionally been the foundation of efficacy testing due to their efficiency under well-defined assumptions. Nonparametric techniques, including the Friedman test, Brunner-Munzel test, and modern extensions like nparLD, have emerged as robust alternatives, particularly for skewed, ordinal, or non-normal data. Bayesian methodologies have enabled the incorporation of …
Saturated Hierarchical Atomic Incremental Learning (Shail): A Behavioral Learning Perspective On Staged Mastery And Saturation, Ernest Fokoue
Saturated Hierarchical Atomic Incremental Learning (Shail): A Behavioral Learning Perspective On Staged Mastery And Saturation, Ernest Fokoue
Articles
We introduce Saturated Hierarchical Atomic Incremental Learning (sHAIL), a learning paradigm in which complex tasks are approached through a sequence of simpler atomic subtasks, each mastered to saturation before progression. The central mechanism is a saturation criterion that detects when learning dynamics enter a plateau region, triggering consolidation and subsequent ascent to a higher level of task complexity. We develop a theoretical framework for sHAIL and show that it naturally gives rise to \emph{staircased convergence}: alternating phases of rapid improvement and genuine plateau. Within each level, classical convergence guarantees apply under standard smoothness conditions, while the hierarchical transitions are driven …
No Intelligence Without Statistics: The Invisible Backbone Of Artificial Intelligence, Ernest Fokoue
No Intelligence Without Statistics: The Invisible Backbone Of Artificial Intelligence, Ernest Fokoue
Articles
The rapid ascent of artificial intelligence (AI) is often portrayed as a revolution born from computer science and engineering. This narrative, however, obscures a fundamental truth: the theoretical and methodological core of AI is, and has always been, statistical. This paper systematically argues that the field of statistics provides the indispensable foundation for machine learning and modern AI. We deconstruct AI into nine foundational pillars—Inference, Density Estimation, Sequential Learning, Generalization, Representation Learning, Interpretability, Causality, Optimization, and Unification—demonstrating that each is built upon century-old statistical principles. From the inferential frameworks of hypothesis testing and estimation that underpin model evaluation, to the …
Decorrelation, Diversity, And Emergent Intelligence: The Isomorphism Between Social Insect Colonies And Ensemble Machine Learning, Ernest Fokoue, Gregory Babbitt, Yuval Levental
Decorrelation, Diversity, And Emergent Intelligence: The Isomorphism Between Social Insect Colonies And Ensemble Machine Learning, Ernest Fokoue, Gregory Babbitt, Yuval Levental
Articles
Social insect colonies and ensemble machine learning methods represent two of the most successful examples of decentralized information processing in nature and computation respectively. Here we develop a rigorous mathematical framework demonstrating that ant colony decision-making and random forest learning are isomorphic under a common formalism of stochastic ensemble intelligence. We show that the mechanisms by which genetically identical ants achieve functional differentiation— through stochastic response to local cues and positive feedback—map precisely onto the bootstrap aggregation and random feature subsampling that decorrelate decision trees. Using tools from Bayesian inference, multi-armed bandit theory, and statistical learning theory, we prove that …
A General Weighting Theory For Ensemble Learning: Beyond Variance Reduction Via Spectral And Geometric Structure, Ernest Fokoue
A General Weighting Theory For Ensemble Learning: Beyond Variance Reduction Via Spectral And Geometric Structure, Ernest Fokoue
Articles
Ensemble learning is traditionally justified as a variance-reduction strategy, explaining its strong performance for unstable predictors such as decision trees. This explanation, however, does not account for ensembles constructed from intrinsically stable estimators-including smoothing splines, kernel ridge regression, Gaussian process regression, and other regularized reproducing kernel Hilbert space (RKHS) methods whose variance is already tightly controlled by regularization and spectral shrinkage. This paper develops a general weighting theory for ensemble learning that moves beyond classical variance-reduction arguments. We formalize ensembles as linear operators acting on a hypothesis space and endow the space of weighting sequences with geometric and spectral constraints. …
On Fibonacci Ensembles: An Alternative Approach To Ensemble Learning Inspired By The Timeless Architecture Of The Golden Ratio, Ernest Fokoue
On Fibonacci Ensembles: An Alternative Approach To Ensemble Learning Inspired By The Timeless Architecture Of The Golden Ratio, Ernest Fokoue
Articles
Nature rarely reveals her secrets bluntly, yet in the Fibonacci sequence she grants us a glimpse of her quiet architecture of growth, harmony, and recursive stability \citep{Koshy2001Fibonacci, Livio2002GoldenRatio}. From spiral galaxies to the unfolding of leaves, this humble sequence reflects a universal grammar of balance. In this work, we introduce \emph{Fibonacci Ensembles}, a mathematically principled yet philosophically inspired framework for ensemble learning that complements and extends classical aggregation schemes such as bagging, boosting, and random forests \citep{Breiman1996Bagging, Breiman2001RandomForests, Friedman2001GBM, Zhou2012Ensemble, HastieTibshiraniFriedman2009ESL}. Two intertwined formulations unfold: (1) the use of normalized Fibonacci weights -- tempered through orthogonalization and Rao--Blackwell optimization -- …
Comparative Machine Learning Models For Disease Risk Prediction, Mercy Mawusi Agbley
Comparative Machine Learning Models For Disease Risk Prediction, Mercy Mawusi Agbley
Theses, Dissertations and Capstones
Accurate prediction of disease outcomes is crucial for improving clinical decision-making and enabling early intervention. This study compares the performance of various statistical and machine learning models for clinical risk prediction using two healthcare datasets: diabetic retinopathy and heart disease. The models assessed include Logistic Regression, LASSO, k-Nearest Neighbors (KNN), Support Vector Machines (SVM), Neural Networks, Random Forests, Gradient Boosting Machines (GBM), and a stacked ensemble model. Prior to modeling, datasets were split into train and test sets. Standardization was applied to numeric features whilst categorical features were one-hot encoded. These transformations were later applied to the test set. Principal …
First-Generation Medical School Applicants: A Quantitative Study Designed To Identify Areas Of Educational Support, Bethsabe Romero, Amanda K. Burbage
First-Generation Medical School Applicants: A Quantitative Study Designed To Identify Areas Of Educational Support, Bethsabe Romero, Amanda K. Burbage
EVMS School of Health Professions Faculty Publications
First-generation (First Gen) students are unique medical school applicants. Due to their lived experience, they approach patient care by prioritizing trust, comfort and understanding. They have proven ability to overcome obstacles and were found to be more resilient than their continuing generation (Cont Gen) peers. Despite these notable attributes, they face unique challenges in gaining medical school acceptance. There are very few quantitative studies examining this student subpopulation, and our study identifies characteristics of first-generation medical school applicants while highlighting areas of needed support. This cross-sectional study used deidentified Application and Matriculating Student Questionnaire survey data that was obtained from …
Statistical Quality Control: A Bayesian Framework, Jakia Jaber Tunal
Statistical Quality Control: A Bayesian Framework, Jakia Jaber Tunal
College of Graduate Studies: Theses & Dissertations
In many industries, it is important to assess whether a machine or system is operating within acceptable limits or has gone out of control. This project applies Bayesian statistics to monitor a process over time and detect changes in its behavior. First, initial data are collected to understand the system’s typical performance and to form a starting prior distribution. As new observations arrive over time, the prior is updated through Bayesian inference, combining past information with incoming data. This iterative updating creates a continuous monitoring framework that adapts as more evidence becomes available. When the updated results suggest that the …
Modeling Rank Distribution And The Relative Importance Factor Index In Discrete Power-Law Models: Application To Social Resilience Using The Scopus Database, Brian Llinas, Jose Padilla, Humberto Llinas, Erika Frydenlund, Katherine Palacio
Modeling Rank Distribution And The Relative Importance Factor Index In Discrete Power-Law Models: Application To Social Resilience Using The Scopus Database, Brian Llinas, Jose Padilla, Humberto Llinas, Erika Frydenlund, Katherine Palacio
VMASC Publications
Prior research on power-law distributions has primarily focused on modeling frequency patterns, with less attention given to rank distributions and how ranked positions reflect relative importance among elements. In discrete power-law distributions, frequency-based metrics often provide limited discrimination in the tail, where elements may exhibit similar counts but differ in relative dominance. These patterns are especially evident, for instance, in academic publishing, where keywords, affiliations, and citations commonly exhibit power-law behavior. To address this limitation, we introduce the Relative Importance Factor (RIF) Index, a statistical measure derived from the estimated discrete power-law rank distribution rather than an additional independent parameter. …
A Comparative Study Of Classification Methods For Healthcare Analytics, Xueting Zhao
A Comparative Study Of Classification Methods For Healthcare Analytics, Xueting Zhao
UNF Graduate Theses and Dissertations
This thesis presents a comparative study of logistic regression, Linear Discriminant Analy- sis (LDA), and Quadratic Discriminant Analysis (QDA) for binary classification in healthcare analytics, integrating theoretical derivation, simulation, and real-data application. A facto- rial simulation study crosses the covariance structure (equal vs. unequal), predictor correla- tion (ρ ∈ {0, 0.5, 0.9}), dimensionality (p ∈ {2, 5, 10}) and sample size (n ∈ {50, 100, 200}) across 54 scenarios with 1,000 Monte Carlo replicates each. Three main findings emerge. Logistic regression and LDA are nearly interchangeable when the assumption of equal-covariance holds. QDA achieves substantially better discrimi- nation when class-specific …
(R2130) Cusum-Test For Unconditional Variance Change Detection In Bilinear Garch Models, Edoh Katchekpele, Abdou Kâ Diongue, Ben Célestin Kouassi
(R2130) Cusum-Test For Unconditional Variance Change Detection In Bilinear Garch Models, Edoh Katchekpele, Abdou Kâ Diongue, Ben Célestin Kouassi
Applications and Applied Mathematics: An International Journal (AAM)
We examine CUSUM-type test for detecting changes in unconditional variance within Bilinear GARCH models. We derive the asymptotic distribution of the test statistic under both null and alternative hypotheses and assess test effectiveness in identifying single structural breaks. Simulation studies support our theoretical results and demonstrate the practical utility of the test.
Performance Of The Two Sample Likelihood Ratio Test Under A Nested Dirichlet: A Simulation Study, Edwina Agyeman
Performance Of The Two Sample Likelihood Ratio Test Under A Nested Dirichlet: A Simulation Study, Edwina Agyeman
Electronic Theses and Dissertations
Compositional data analysis (CoDA) addresses multivariate data constrained to a constant sum, such as proportions or percentages. Originating from early warnings regarding misinterpretation by Pearson (1897), the field was formalized by John Aitchison in 1986, whose foundational work remains highly influential. Over time, new modeling techniques and visualization tools have advanced the field, as noted by Greenacre et al. More recently, Turner et al. proposed an approach based on the Nested Dirichlet Distribution (NDD), which accommodates more flexible dependence structures than the standard Dirichlet model. This thesis builds on the methodology of Turner et al. Chapter 1 introduces the nature …
Unified Hybrid Censoring Samples From Power Pratibha Distribution And Its Applications, Mahmoud Mansour, Hebatalla H. Mohammad Dr, Khalaf S. Sultan Prof.
Unified Hybrid Censoring Samples From Power Pratibha Distribution And Its Applications, Mahmoud Mansour, Hebatalla H. Mohammad Dr, Khalaf S. Sultan Prof.
Basic Science Engineering
This paper suggests an extensive inferential method for the Power Pratibha Distribution (PPD) under Unified Hybrid Censoring Schemes (UHCSs), since there is a growing interest in flexible models in both reliability and service operations. This work studies the PPD model using standard Maximum Likelihood Estimation methods and modern Bayesian approaches too. Using a complex architecture, UHCS simulates tests more closely to what is done in practice than by using more basic censoring schemes. Using analysis, the probability and statistical ranges are carefully calculated for the parameters. Tests demonstrate that Bayesian estimation gives better results than many other methods for estimation, …
Welfare Implication Of Alternative Tax Rates Adjustment Policy In Nigeria: A Dsge Analysis, Umar B. Ibrahim, Isah F. Abubakar
Welfare Implication Of Alternative Tax Rates Adjustment Policy In Nigeria: A Dsge Analysis, Umar B. Ibrahim, Isah F. Abubakar
CBN Journal of Applied Statistics (JAS)
This study sets out to determine the desirable policy adjustment in the tax rate for Nigeria that ensures the least welfare cost. A calibrated small open-economy New Keynesian Dynamic Stochastic General Equilibrium (NKDSGE) model of the Nigerian economy is applied to achieve this objective. Within this framework, we examined the impact of an increase in value-added tax (VAT) rate from 7.5 to 15 percent on key macroeconomic variables relative to the impact of an increase in company income tax (CIT) rate from 30 to 35 percent on macroeconomic variables. Furthermore, we examined the welfare costs of the increases in the …
Bayesian Statistics: Origins And Applications, Evelyn Pulla
Bayesian Statistics: Origins And Applications, Evelyn Pulla
Publications and Research
Bayesian Statistics applies Bayes' Theorem to update beliefs through new evidence. In this project, I explored how Bayesian Statistics applies into real supporting decision-making under uncertainty. By solving problems using data, I was able to realize how prior knowledge and evidence collaborate to make better conclusions. The project also demonstrates how Bayesian reasoning corrects our intuition to make decisions based on logical reasoning. Through this project, I was able to learn why using probability to make informed decisions matters both in science and real life.
The Little Diagram That Could: Geometric Properties And Statistical Applications Of Persistence Diagrams In Topological Data Analysis, Eugene Kler
McKelvey School of Engineering Graduate Student Theses & Dissertations
Topological Data Analysis (TDA) is a collection of techniques for data analysis that leverages topological invariants of spaces formed from data points. These methods excel at extracting useful information from noisy or sparse data, making them attractive to many mathematicians, statisticians, and scientists. In this thesis, we explore TDA on three fronts: algebraic foundations, statistical applications, and metric properties. Throughout, the central object of study is the Persistence Diagram (PD), a summary of the changes in homology that occur as one builds simplicial complexes from the data by increasing a parameter.
Evaluating The Performance Of Bayesian Removal Models For Estimating Population Density And Detecting Trends With Variable Detection Probability, David R. Stewart
Evaluating The Performance Of Bayesian Removal Models For Estimating Population Density And Detecting Trends With Variable Detection Probability, David R. Stewart
Mathematics & Statistics ETDs
Removal models have long been used to estimate population abundance by progressively capturing and removing individuals from a closed population. These models provide a valuable tool for ecological monitoring, but their accuracy depends heavily on assumptions about detection probability, which may decline over successive sampling passes. Traditional removal models assume constant detection probabilities, an assumption that is often violated in real-world applications. This thesis aims to advance hierarchical Bayesian models by accounting for variable detection probabilities, improving the reliability of abundance estimates and trend detection. By integrating simulation-based analyses with empirical data from Lahontan Cutthroat Trout (Oncorhynchus clarkia henshawi …
Quaternary Subsurface Characterization Of The Mississippi River Valley Alluvial Aquifer: Insights From Interval Kriging And Airborne Em-Borehole Data Integration, Yuqi Song
LSU Doctoral Dissertations
The Pleistocene period significantly contributed to the formation of alluvial aquifer worldwide. These productive aquifers are crucial for domestic, industrial, and agricultural water supplies. The Mississippi River Valley alluvial aquifer (MRVA), a principal aquifer in the U.S., is crucial for national food security and global agricultural supply. This study aims to characterize the subsurface architecture of the MRVA, thereby enhancing fundamental understanding of sedimentological processes involved in the genesis of glacio-fluvial aquifers worldwide. The research objectives are threefold: (1) to provide a detailed characterization of the MRVA; (2) to develop 3D geostatistical methods for geological modeling; and (3) to develop …
On The Gumbel-Weibull{Cauchy} Distribution, Jennifer D. Pippin
On The Gumbel-Weibull{Cauchy} Distribution, Jennifer D. Pippin
Theses, Dissertations and Capstones
Developing new statistical distributions and seeking higher flexibility in modeling different shapes of data remain a strong emphasis in research. The T-R{Y } framework, introduced in [3], utilizes three statistical distributions in order to generate a new distribution. Many research papers appeared in literature to develop distributions based on the T-R{Y } framework. In this thesis, a member of the T-R{Y } framework, namely the Gumbel-Weibull{Cauchy} (GWC), is introduced. Statistical properties of the GWC are studied, such as the quantile function, the hazard function, transformations, Shannon entropy, the …
Predictive Modeling For Healthcare Data Using Nonlinear Bayesian Methods, Prince Kofi Asare
Predictive Modeling For Healthcare Data Using Nonlinear Bayesian Methods, Prince Kofi Asare
Theses and Dissertations
Unplanned hospital readmissions represent a significant challenge for healthcare systems, contributing to substantial financial burdens and highlighting gaps in patient care coordination. In the U.S., approximately 20% of Medicare beneficiaries are readmitted within 30 days, costing billions annually. Social determinants of health, such as income, housing stability, and social support, account for up to 80% of health outcomes, yet their integration into predictive models remains underexplored. This study introduces a novel Bayesian framework for predicting 30-day readmission risk, combining Gaussian Process models with spike-and-slab priors and Bayesian Lasso regression with Laplace priors. Utilizing Markov Chain Monte Carlo methods, the approach …
Theoretical Analysis Of Cnns For Automatic Seizure Detection In Eeg Signals, Jackson T. Small
Theoretical Analysis Of Cnns For Automatic Seizure Detection In Eeg Signals, Jackson T. Small
Honors Undergraduate Theses
Epilepsy is a common brain disorder where neurons in the brain rapidly fire, causing recurring seizures. The brain activity during a seizure can be detected by electroencephalogram (EEG) signals; however, this process is not only labor-intensive and time-consuming but is also subject to inter-rater variability, with a study showing only moderate agreement when diagnosing patients, even among experts. Convolutional Neural Networks (CNNs) are often proposed to detect seizures automatically, achieving high performance. The focus on performance comes at a cost of losing interpretability, leaving the model as effective but seen as a ’black box’. This thesis confronts the interpretability knowledge …
“Regression To The Mean”: The Confluence Of Eugenics And Statistics In The 19th And 20th Centuries, Emrys G. King
“Regression To The Mean”: The Confluence Of Eugenics And Statistics In The 19th And 20th Centuries, Emrys G. King
Pomona Senior Theses
The work of this thesis is twofold — first, qualitatively characterizing the confluence between the British eugenics and statistics movements in the late 19th and early 20th centuries, and second, quantitatively analyzing the effect of this foundation on pedagogical materials in the growing field of statistics between 1880 and 1970. Towards the first goal, the history of the method of least squares, state statistics, and positive and negative eugenics are outlined, followed by a close reading of the foundational texts authored by Francis Galton and Karl Pearson that introduced linear regression. Towards the latter goal, English-language statistics textbooks published between …