Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- Wayne State University (1091)
- Central Bank of Nigeria (26)
- Stephen F. Austin State University (6)
- University of Kentucky (6)
- Georgia Southern University (5)
-
- Southern Methodist University (4)
- University of Nebraska - Lincoln (4)
- Old Dominion University (3)
- Rochester Institute of Technology (3)
- The British University in Egypt (3)
- University of Louisville (3)
- COBRA (2)
- California Polytechnic State University, San Luis Obispo (2)
- City University of New York (CUNY) (2)
- Claremont Colleges (2)
- Liberty University (2)
- Prairie View A&M University (2)
- University of Central Florida (2)
- University of Connecticut (2)
- University of New Mexico (2)
- University of North Florida (2)
- Utah State University (2)
- Virginia Commonwealth University (2)
- Chapman University (1)
- East Tennessee State University (1)
- Illinois State University (1)
- Kennesaw State University (1)
- Marshall University (1)
- Michigan Technological University (1)
- Minnesota State University, Mankato (1)
- Keyword
-
- Bias (32)
- Simulation (29)
- Bootstrap (28)
- Monte Carlo simulation (27)
- Power (23)
-
- Confidence interval (22)
- Mean squared error (20)
- Effect size (19)
- Sample size (19)
- Type I error (19)
- Monte Carlo (18)
- Robustness (18)
- Multicollinearity (17)
- Permutation test (17)
- Maximum likelihood estimation (16)
- Missing data (16)
- Statistics (15)
- Efficiency (14)
- Reliability (14)
- Confidence intervals (13)
- Heteroscedasticity (13)
- Logistic regression (13)
- SPSS (13)
- Simple random sampling (13)
- P-value (12)
- Regression (12)
- Bayesian (11)
- Estimation (11)
- Mean square error (11)
- Meta-analysis (11)
- Publication Year
- Publication
-
- Journal of Modern Applied Statistical Methods (1091)
- CBN Journal of Applied Statistics (JAS) (26)
- Electronic Theses and Dissertations (9)
- Theses and Dissertations--Statistics (6)
- College of Graduate Studies: Theses & Dissertations (5)
-
- Articles (3)
- Basic Science Engineering (3)
- Department of Statistics: Dissertations, Theses, and Student Research (3)
- Statistical Science Theses and Dissertations (3)
- Theses and Dissertations (3)
- Applications and Applied Mathematics: An International Journal (AAM) (2)
- Branch Mathematics and Statistics Faculty and Staff Publications (2)
- Data Science and Data Mining (2)
- Mathematics & Statistics Theses & Dissertations (2)
- Senior Honors Theses (2)
- Statistics (2)
- UNF Graduate Theses and Dissertations (2)
- All Graduate Plan B and other Reports, Spring 1920 to Spring 2023 (1)
- All Graduate Theses, Dissertations, and Other Capstone Projects (1)
- Business and Economics Honors Papers (1)
- CHIP Documents (1)
- CMC Senior Theses (1)
- COBRA Preprint Series (1)
- Civil and Environmental Engineering Theses and Dissertations (1)
- Community & Environmental Health Faculty Publications (1)
- Department of Management: Faculty Publications (1)
- Dissertations, Master's Theses and Master's Reports (1)
- Dissertations, Theses, and Capstone Projects (1)
- Electronic Theses & Dissertations (2024 - present) (1)
- Faculty Articles (1)
- Publication Type
Articles 1 - 30 of 1191
Full-Text Articles in Statistical Theory
(R2174) A Stationary First-Order Autoregressive Process With New Discrete Lindley Marginal Distribution, Tégawendé Martin Kabore, Jean-Etienne Ouindllassida Ouédraogo
(R2174) A Stationary First-Order Autoregressive Process With New Discrete Lindley Marginal Distribution, Tégawendé Martin Kabore, Jean-Etienne Ouindllassida Ouédraogo
Applications and Applied Mathematics: An International Journal (AAM)
This paper introduces a new stationary first-order autoregressive integer-valued process with a marginal distribution following the New Discrete Lindley distribution, referred to as NDLINAR( 1). The process is developed to model over-dispersed count time series exhibiting a mixed behavior arising from the combination of multiple distributions. Its statistical properties are thoroughly investigated, and the parameters are estimated using conditional maximum likelihood. The asymptotic properties of the estimators are also analyzed. The performance of the proposed model is evaluated by comparing it with other existing INAR(1) processes through applications to real datasets. Results demonstrate that the NDL-INAR(1) process effectively captures the …
A Predictive Coding Account Of Spatial Working Memory Following Prophylactic Levetiracetam Administration Prior To Traumatic Brain Injury, Omeima Mutwali
A Predictive Coding Account Of Spatial Working Memory Following Prophylactic Levetiracetam Administration Prior To Traumatic Brain Injury, Omeima Mutwali
Dissertations, Theses, and Capstone Projects
Traumatic brain injury (TBI) symptom prevention and remediation is an important area of research that would benefit vulnerable groups, including active-duty and veteran soldiers. These patients can sustain penetrative forces in fields of combat or in training, which result in focal lesions that trigger inflammatory and degenerative processes in the brain. Both primary and secondary injuries are associated with changes to cognition, behavior and affective state. This disease poses increased risk of epileptogenesis, as well. Given these outcomes, prior research has evaluated levetiracetam (LEV) as a prophylactic treatment for seizures, cognitive deficits and negative emotionality. LEV acts as a presynaptic …
An Analytical Framework For Quantifying Urban And Community Resilience To Natural Hazards From Cell-Phone Gps-Location And Traffic-Flow Data, Georgios Chatzikyriakidis
An Analytical Framework For Quantifying Urban And Community Resilience To Natural Hazards From Cell-Phone Gps-Location And Traffic-Flow Data, Georgios Chatzikyriakidis
Civil and Environmental Engineering Theses and Dissertations
Urban areas are increasingly exposed to natural hazards while accommodating a growing share of the global population, yet a consistent science-based framework for quantifying urban and community resilience remains lacking. This dissertation develops a physics-based analytical framework grounded in statistical mechanics and the quantitative theory of Brownian motion. A city is conceptualized as a complex medium in which citizens move analogously to Brownian particles within a viscoelastic environment, influenced by socioeconomic interactions and infrastructure functionality.
A central premise is that urban resilience, interpreted as engineering resilience (an outcome), can be quantified through a single metric: the mean-square displacement MSD=⟨r²(t)⟩, of …
Bayesian Spatiotemporal Model For Counterfactual Estimation In Socioeconomic Studies, Duwani W. Gonzalez
Bayesian Spatiotemporal Model For Counterfactual Estimation In Socioeconomic Studies, Duwani W. Gonzalez
Statistical Science Theses and Dissertations
Impact evaluations of regional development programs often require estimating counterfactual outcomes for a small number of treated regions using survey-based areal data. In practice, evaluators typically rely on two-group quasi-experimental methods such as propensity score matching (PSM) and Difference-in-Differences (DiD). These approaches perform poorly when only a few regions receive treatment, and when the set of observed covariates is limited or only partially relevant. Moreover, they typically do not explicitly exploit the spatial and temporal dependence present in survey-based areal data such as in ACS (American Community Survey). This dissertation develops a family of Bayesian spatial predictive models for directly …
The Impatience Of Winning: An Analysis Of Time Discounting, Predictive Modeling, And The Nba Draft, Alec R. Plante
The Impatience Of Winning: An Analysis Of Time Discounting, Predictive Modeling, And The Nba Draft, Alec R. Plante
Business and Economics Honors Papers
This paper examines whether NBA draft decisions can be better explained by incorporating non-geometric time discounting into a model of general manager decision making. Using a dataset of 285 NBA draft prospects over a 12-year period, the impact of college statistics on Value Over Replacement Player (VORP) is determined, and these impact values are then used to create a “predicted” VORP for the first 4 seasons of each player’s career: a projection of what a general manager might think of a prospect’s future value given their college statistics. Following this, geometric and hyperbolic time discounting models are applied to estimate …
Efficacy Analysis In Clinical Trials: A Comprehensive Review Of Statistical And Machine Learning Approaches, Dhrubajyoti Ghosh, Samhita Pal
Efficacy Analysis In Clinical Trials: A Comprehensive Review Of Statistical And Machine Learning Approaches, Dhrubajyoti Ghosh, Samhita Pal
Faculty Articles
Efficacy testing is a cornerstone of clinical trials, ensuring that medical interventions achieve their intended therapeutic effects. Over the decades, a wide range of statistical methodologies have been developed to address the complexities of clinical trial data, including parametric, nonparametric, Bayesian, and machine learning approaches. Parametric methods, such as t-tests, ANOVA, and LMMs, have traditionally been the foundation of efficacy testing due to their efficiency under well-defined assumptions. Nonparametric techniques, including the Friedman test, Brunner-Munzel test, and modern extensions like nparLD, have emerged as robust alternatives, particularly for skewed, ordinal, or non-normal data. Bayesian methodologies have enabled the incorporation of …
Saturated Hierarchical Atomic Incremental Learning (Shail): A Behavioral Learning Perspective On Staged Mastery And Saturation, Ernest Fokoue
Saturated Hierarchical Atomic Incremental Learning (Shail): A Behavioral Learning Perspective On Staged Mastery And Saturation, Ernest Fokoue
Articles
We introduce Saturated Hierarchical Atomic Incremental Learning (sHAIL), a learning paradigm in which complex tasks are approached through a sequence of simpler atomic subtasks, each mastered to saturation before progression. The central mechanism is a saturation criterion that detects when learning dynamics enter a plateau region, triggering consolidation and subsequent ascent to a higher level of task complexity. We develop a theoretical framework for sHAIL and show that it naturally gives rise to \emph{staircased convergence}: alternating phases of rapid improvement and genuine plateau. Within each level, classical convergence guarantees apply under standard smoothness conditions, while the hierarchical transitions are driven …
Decorrelation, Diversity, And Emergent Intelligence: The Isomorphism Between Social Insect Colonies And Ensemble Machine Learning, Ernest Fokoue, Gregory Babbitt, Yuval Levental
Decorrelation, Diversity, And Emergent Intelligence: The Isomorphism Between Social Insect Colonies And Ensemble Machine Learning, Ernest Fokoue, Gregory Babbitt, Yuval Levental
Articles
Social insect colonies and ensemble machine learning methods represent two of the most successful examples of decentralized information processing in nature and computation respectively. Here we develop a rigorous mathematical framework demonstrating that ant colony decision-making and random forest learning are isomorphic under a common formalism of stochastic ensemble intelligence. We show that the mechanisms by which genetically identical ants achieve functional differentiation— through stochastic response to local cues and positive feedback—map precisely onto the bootstrap aggregation and random feature subsampling that decorrelate decision trees. Using tools from Bayesian inference, multi-armed bandit theory, and statistical learning theory, we prove that …
A General Weighting Theory For Ensemble Learning: Beyond Variance Reduction Via Spectral And Geometric Structure, Ernest Fokoue
A General Weighting Theory For Ensemble Learning: Beyond Variance Reduction Via Spectral And Geometric Structure, Ernest Fokoue
Articles
Ensemble learning is traditionally justified as a variance-reduction strategy, explaining its strong performance for unstable predictors such as decision trees. This explanation, however, does not account for ensembles constructed from intrinsically stable estimators-including smoothing splines, kernel ridge regression, Gaussian process regression, and other regularized reproducing kernel Hilbert space (RKHS) methods whose variance is already tightly controlled by regularization and spectral shrinkage. This paper develops a general weighting theory for ensemble learning that moves beyond classical variance-reduction arguments. We formalize ensembles as linear operators acting on a hypothesis space and endow the space of weighting sequences with geometric and spectral constraints. …
Comparative Machine Learning Models For Disease Risk Prediction, Mercy Mawusi Agbley
Comparative Machine Learning Models For Disease Risk Prediction, Mercy Mawusi Agbley
Theses, Dissertations and Capstones
Accurate prediction of disease outcomes is crucial for improving clinical decision-making and enabling early intervention. This study compares the performance of various statistical and machine learning models for clinical risk prediction using two healthcare datasets: diabetic retinopathy and heart disease. The models assessed include Logistic Regression, LASSO, k-Nearest Neighbors (KNN), Support Vector Machines (SVM), Neural Networks, Random Forests, Gradient Boosting Machines (GBM), and a stacked ensemble model. Prior to modeling, datasets were split into train and test sets. Standardization was applied to numeric features whilst categorical features were one-hot encoded. These transformations were later applied to the test set. Principal …
A Comparative Study Of Classification Methods For Healthcare Analytics, Xueting Zhao
A Comparative Study Of Classification Methods For Healthcare Analytics, Xueting Zhao
UNF Graduate Theses and Dissertations
This thesis presents a comparative study of logistic regression, Linear Discriminant Analy- sis (LDA), and Quadratic Discriminant Analysis (QDA) for binary classification in healthcare analytics, integrating theoretical derivation, simulation, and real-data application. A facto- rial simulation study crosses the covariance structure (equal vs. unequal), predictor correla- tion (ρ ∈ {0, 0.5, 0.9}), dimensionality (p ∈ {2, 5, 10}) and sample size (n ∈ {50, 100, 200}) across 54 scenarios with 1,000 Monte Carlo replicates each. Three main findings emerge. Logistic regression and LDA are nearly interchangeable when the assumption of equal-covariance holds. QDA achieves substantially better discrimi- nation when class-specific …
(R2130) Cusum-Test For Unconditional Variance Change Detection In Bilinear Garch Models, Edoh Katchekpele, Abdou Kâ Diongue, Ben Célestin Kouassi
(R2130) Cusum-Test For Unconditional Variance Change Detection In Bilinear Garch Models, Edoh Katchekpele, Abdou Kâ Diongue, Ben Célestin Kouassi
Applications and Applied Mathematics: An International Journal (AAM)
We examine CUSUM-type test for detecting changes in unconditional variance within Bilinear GARCH models. We derive the asymptotic distribution of the test statistic under both null and alternative hypotheses and assess test effectiveness in identifying single structural breaks. Simulation studies support our theoretical results and demonstrate the practical utility of the test.
Performance Of The Two Sample Likelihood Ratio Test Under A Nested Dirichlet: A Simulation Study, Edwina Agyeman
Performance Of The Two Sample Likelihood Ratio Test Under A Nested Dirichlet: A Simulation Study, Edwina Agyeman
Electronic Theses and Dissertations
Compositional data analysis (CoDA) addresses multivariate data constrained to a constant sum, such as proportions or percentages. Originating from early warnings regarding misinterpretation by Pearson (1897), the field was formalized by John Aitchison in 1986, whose foundational work remains highly influential. Over time, new modeling techniques and visualization tools have advanced the field, as noted by Greenacre et al. More recently, Turner et al. proposed an approach based on the Nested Dirichlet Distribution (NDD), which accommodates more flexible dependence structures than the standard Dirichlet model. This thesis builds on the methodology of Turner et al. Chapter 1 introduces the nature …
Unified Hybrid Censoring Samples From Power Pratibha Distribution And Its Applications, Mahmoud Mansour, Hebatalla H. Mohammad Dr, Khalaf S. Sultan Prof.
Unified Hybrid Censoring Samples From Power Pratibha Distribution And Its Applications, Mahmoud Mansour, Hebatalla H. Mohammad Dr, Khalaf S. Sultan Prof.
Basic Science Engineering
This paper suggests an extensive inferential method for the Power Pratibha Distribution (PPD) under Unified Hybrid Censoring Schemes (UHCSs), since there is a growing interest in flexible models in both reliability and service operations. This work studies the PPD model using standard Maximum Likelihood Estimation methods and modern Bayesian approaches too. Using a complex architecture, UHCS simulates tests more closely to what is done in practice than by using more basic censoring schemes. Using analysis, the probability and statistical ranges are carefully calculated for the parameters. Tests demonstrate that Bayesian estimation gives better results than many other methods for estimation, …
Welfare Implication Of Alternative Tax Rates Adjustment Policy In Nigeria: A Dsge Analysis, Umar B. Ibrahim, Isah F. Abubakar
Welfare Implication Of Alternative Tax Rates Adjustment Policy In Nigeria: A Dsge Analysis, Umar B. Ibrahim, Isah F. Abubakar
CBN Journal of Applied Statistics (JAS)
This study sets out to determine the desirable policy adjustment in the tax rate for Nigeria that ensures the least welfare cost. A calibrated small open-economy New Keynesian Dynamic Stochastic General Equilibrium (NKDSGE) model of the Nigerian economy is applied to achieve this objective. Within this framework, we examined the impact of an increase in value-added tax (VAT) rate from 7.5 to 15 percent on key macroeconomic variables relative to the impact of an increase in company income tax (CIT) rate from 30 to 35 percent on macroeconomic variables. Furthermore, we examined the welfare costs of the increases in the …
Bayesian Statistics: Origins And Applications, Evelyn Pulla
Bayesian Statistics: Origins And Applications, Evelyn Pulla
Publications and Research
Bayesian Statistics applies Bayes' Theorem to update beliefs through new evidence. In this project, I explored how Bayesian Statistics applies into real supporting decision-making under uncertainty. By solving problems using data, I was able to realize how prior knowledge and evidence collaborate to make better conclusions. The project also demonstrates how Bayesian reasoning corrects our intuition to make decisions based on logical reasoning. Through this project, I was able to learn why using probability to make informed decisions matters both in science and real life.
“Regression To The Mean”: The Confluence Of Eugenics And Statistics In The 19th And 20th Centuries, Emrys G. King
“Regression To The Mean”: The Confluence Of Eugenics And Statistics In The 19th And 20th Centuries, Emrys G. King
Pomona Senior Theses
The work of this thesis is twofold — first, qualitatively characterizing the confluence between the British eugenics and statistics movements in the late 19th and early 20th centuries, and second, quantitatively analyzing the effect of this foundation on pedagogical materials in the growing field of statistics between 1880 and 1970. Towards the first goal, the history of the method of least squares, state statistics, and positive and negative eugenics are outlined, followed by a close reading of the foundational texts authored by Francis Galton and Karl Pearson that introduced linear regression. Towards the latter goal, English-language statistics textbooks published between …
Theoretical Foundations And Applied Performance Of Periodicity-Aware Imputation: Variable Bandpass Block Bootstrap Methods For Incomplete Time Series, Asmaa Ahmad
Electronic Theses & Dissertations (2024 - present)
Time series data are prevalent across a wide range of disciplines, including health surveillance, public policy, and environmental monitoring. In the presence of underlying cyclical patterns, the integrity of time series analysis depends critically on the ability to detect, model, and impute structured missing data without compromising the temporal structure. This dissertation introduces and validates a novel imputation framework that integrates the Variable Bandpass Periodic Block Bootstrap (VBPBB) into multiple imputation procedures, improving the accuracy, robustness, and interpretability of time series models under high rates of missingness and noise. The overarching goal of this dissertation was to develop and evaluate …
Predictive Modeling For Healthcare Data Using Nonlinear Bayesian Methods, Prince Kofi Asare
Predictive Modeling For Healthcare Data Using Nonlinear Bayesian Methods, Prince Kofi Asare
Theses and Dissertations
Unplanned hospital readmissions represent a significant challenge for healthcare systems, contributing to substantial financial burdens and highlighting gaps in patient care coordination. In the U.S., approximately 20% of Medicare beneficiaries are readmitted within 30 days, costing billions annually. Social determinants of health, such as income, housing stability, and social support, account for up to 80% of health outcomes, yet their integration into predictive models remains underexplored. This study introduces a novel Bayesian framework for predicting 30-day readmission risk, combining Gaussian Process models with spike-and-slab priors and Bayesian Lasso regression with Laplace priors. Utilizing Markov Chain Monte Carlo methods, the approach …
A Uniformly Most Powerful Test For The Mean Of A Beta Distribution, Richard Ntiamoah Kyei
A Uniformly Most Powerful Test For The Mean Of A Beta Distribution, Richard Ntiamoah Kyei
Electronic Theses and Dissertations
The beta distribution is used in numerous real-world applications, including areas such as manufacturing (quality control) and analyzing patient outcomes in health care. It also plays a key role in statistical theory, including multivariate analysis of variance (MANOVA) and Bayesian statistics. It is a flexible distribution that can account for many different characteristics of real data. To our surprise, there has been very little work or discussion on performing statistical hypothesis testing for the mean when it is reasonable to assume that the population is beta distributed. Many analysts conduct traditional analyses using a t-test or nonparametric approach, try transformations, …
Simulation Study On Confidence Interval Estimation For Standard Deviation With Non-Normal Distributions, Theophilus Oppong Kyeremeh
Simulation Study On Confidence Interval Estimation For Standard Deviation With Non-Normal Distributions, Theophilus Oppong Kyeremeh
Electronic Theses and Dissertations
This study explores innovative approaches to constructing confidence intervals for the population standard deviation, σ, in non-normal data scenarios. While the sample standard deviation, s, is widely used, its reliability is compromised when dealing with skewed or heavy-tailed distributions and exhibits sensitivity to outliers. Our research addresses these limitations by investigating alternative estimation methods that offer greater robustness and accuracy.
Contributions To Nonparametric Testing In Clustered Data, Hasika Kalani Wickrama Senevirathne
Contributions To Nonparametric Testing In Clustered Data, Hasika Kalani Wickrama Senevirathne
Mathematics & Statistics Theses & Dissertations
Clustered data refers to a specific kind of correlated data where units within the same cluster are correlated while units from different clusters are independent. The number of units in each cluster, known as the cluster size, can be associated with the cluster’s outcome. This is known as the informative cluster size (ICS) and affects the inference drawn from clustered data. Recently, a hypothesis testing method has been developed to detect the presence of ICS. However, considering ICS alone may not be sufficient when comparing outcomes across multiple groups of units within clustered data. The size of a group within …
Comparison Of Value At Risk Using Historical And Monte Carlo Methods On Pt Xyz Stock Portofolio, Eka Fitriani, Yulial Hikmah, Ira Rosianal Hikmah
Comparison Of Value At Risk Using Historical And Monte Carlo Methods On Pt Xyz Stock Portofolio, Eka Fitriani, Yulial Hikmah, Ira Rosianal Hikmah
Jurnal Administrasi Bisnis Terapan
One way to achieve profits in a company is through investment activities. However, everything has risks. Investing can also be risky. Therefore, the relationship between risk and investment is important because it will influence the determination of investment selection. The problem faced by investors is choosing an efficient portfolio, or a portfolio that provides the smallest risk. This risk can be done by measuring risk, one of which is using the Value at Risk (VaR) measure. Measurement using Value at Risk has several methods that are quite popular, namely the Historical Method, Variance-Covariance, and Monte Carlo. In this research, the …
Trade Liberalization, Non-Oil Export And Economic Growth In Nigeria, Jerome T. Andohol, Terhemen Tarzoor, Dennis T. Nomor
Trade Liberalization, Non-Oil Export And Economic Growth In Nigeria, Jerome T. Andohol, Terhemen Tarzoor, Dennis T. Nomor
CBN Journal of Applied Statistics (JAS)
The study examines the impact of trade liberalization and non-oil exports on economic growth in Nigeria from 1986 to 2021. The study utilizes an autoregressive distributed lag model and found the combined effect of trade liberalization and non-oil exports to be positive and statistical significant. While trade liberalization alone may have negative consequences, its synergy with a robust non-oil export can drive sustainable economic growth. The study recommends that strategies to enhance non-oil exports should be encouraged to support the effectiveness of trade liberalization in promoting growth.
Stock Market Volatility In The United Kingdom: Simulating Post-Covid-19 Recovery, Bala A. Dahiru, Mohammed Shuaibu, Najibullah Hassanov
Stock Market Volatility In The United Kingdom: Simulating Post-Covid-19 Recovery, Bala A. Dahiru, Mohammed Shuaibu, Najibullah Hassanov
CBN Journal of Applied Statistics (JAS)
This paper investigates the time it would take for the FTSE-100 index to reach its post-COVID-19 peak. The paper utilises an exponential generalised autoregressive conditional heteroscedasticity (EGARCH) model that accounts for leverage effect and asymmetries. The preferred models amongst competing variants was the Autoregressive Moving Average (ARMA)-EGARCH(2,1) specification and was used to predict daily FTSE-100 data from 5th January 2000 to 21st June 2024. The empirical exercise showed that the COVID-19-induced financial crisis negatively affected the United Kingdom’s stock market performance. The results show that the FTSE100 index could reach its post-pandemic peak around 27th August, 2024 (two months after …
"Who Wrote The Epistle, God Only Knows": A Statistical Authorial Analysis Of Hebrews In Comparison With Pauline And Lukan Literature, Benjamin J. Erickson
"Who Wrote The Epistle, God Only Knows": A Statistical Authorial Analysis Of Hebrews In Comparison With Pauline And Lukan Literature, Benjamin J. Erickson
Senior Honors Theses
The authorship of Hebrews has been a point of contention for scholars for the past two millennia. While the epistle is traditionally attributed to Paul, many scholars assert that it carries thematic, structural, and stylistic differences from the remainder of his extant epistles; therefore, many other possible authors have been proposed. Of these, only Luke has other New Testament writings. Therefore, this project conducts a statistical comparison of Hebrews to the Pauline and Lukan corpora using stylometric authorial analysis methods. This analysis demonstrates that Hebrews is stylistically closer to Lukan literature than Pauline (but not to a significant degree), and …
Predicting Superconducting Critical Temperature Using Regression Analysis, Roland Fiagbe
Predicting Superconducting Critical Temperature Using Regression Analysis, Roland Fiagbe
Data Science and Data Mining
This project estimates a regression model to predict the superconducting critical temperature based on variables extracted from the superconductor’s chemical formula. The regression model along with the stepwise variable selection gives a reasonable and good predictive model with a lower prediction error (MSE). Variables extracted based on atomic radius, valence, atomic mass and thermal conductivity appeared to have the most contribution to the predictive model.
Machine Learning Approaches For Cyberbullying Detection, Roland Fiagbe
Machine Learning Approaches For Cyberbullying Detection, Roland Fiagbe
Data Science and Data Mining
Cyberbullying refers to the act of bullying using electronic means and the internet. In recent years, this act has been identifed to be a major problem among young people and even adults. It can negatively impact one’s emotions and lead to adverse outcomes like depression, anxiety, harassment, and suicide, among others. This has led to the need to employ machine learning techniques to automatically detect cyberbullying and prevent them on various social media platforms. In this study, we want to analyze the combination of some Natural Language Processing (NLP) algorithms (such as Bag-of-Words and TFIDF) with some popular machine learning …
Accounting For Variability Due To Resampling Using Bootstrapping, Dipendra Phuyal
Accounting For Variability Due To Resampling Using Bootstrapping, Dipendra Phuyal
College of Graduate Studies: Theses & Dissertations
Bradley Efron (1979) introduced bootrapping. Typically a researcher is interested in studying a process which generates individuals. The collection of individuals the process has(actual) or could have (conceptual) generated is the population. The collection of conceptual members of the population is an uncountable collection. Hence, the population is anuncountable collection of individuals. The collection of individuals the process has generated (actual individuals) is representative of what the process can generate and will bereferred to as the representative sample. The size of this sample is a nonnegative integervalued random variable N which may be a constant random variable such as in …
The Distribution Of The Significance Level, Paul O. Monnu
The Distribution Of The Significance Level, Paul O. Monnu
College of Graduate Studies: Theses & Dissertations
Reporting the p-value is customary when conducting a test of hypothesis or significance. The likelihood of getting a fictitious second sample and presuming the null hypothesis is correct is the p-value. The significance level is a statistic that interests us to investigate. Being a statistic, it has a distribution. For the F-test in a one-way ANOVA and the t-tests for population means, we define the significance level, its observed value, and the observed significance level. It is possible to derive the significance level distribution. The t-test and the F-test are not without controversy. Specifically, we demonstrate that as sample size …