Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- COBRA (331)
- Central Bank of Nigeria (25)
- University of Kentucky (9)
- Georgia Southern University (5)
- Southern Methodist University (4)
-
- Stephen F. Austin State University (4)
- University of Arkansas, Fayetteville (3)
- University of Louisville (3)
- University of North Florida (3)
- Virginia Commonwealth University (3)
- California Polytechnic State University, San Luis Obispo (2)
- Claremont Colleges (2)
- Marshall University (2)
- University of Central Florida (2)
- University of Nebraska - Lincoln (2)
- University of New Mexico (2)
- Utah State University (2)
- Chapman University (1)
- City University of New York (CUNY) (1)
- Clemson University (1)
- Department of Primary Industries and Regional Development, Western Australia (1)
- East Tennessee State University (1)
- Florida Institute of Technology (1)
- Kennesaw State University (1)
- Louisiana State University (1)
- Michigan Technological University (1)
- Old Dominion University (1)
- Portland State University (1)
- Rochester Institute of Technology (1)
- South Dakota State University (1)
- Keyword
-
- Prediction (13)
- Bootstrap (12)
- Causal inference (11)
- Genetics (11)
- Model selection (11)
-
- Statistics (9)
- Type I error rate (9)
- Adjusted p-value (8)
- Counterfactual (8)
- Cross-validation (8)
- Multiple testing (8)
- False discovery rate (7)
- Gene expression (7)
- Censored data (6)
- Classification (6)
- Counting process (6)
- Estimating equation (6)
- Loss function (6)
- Null distribution (6)
- Survival analysis (6)
- Asymptotic control (5)
- Current status data (5)
- Density estimation (5)
- Diagnostic tests (5)
- Generalized family-wise error rate (5)
- Longitudinal data (5)
- Microarray (5)
- Multiple hypothesis testing (5)
- Proportion of false positives (5)
- Regression (5)
- Publication Year
- Publication
-
- U.C. Berkeley Division of Biostatistics Working Paper Series (113)
- Harvard University Biostatistics Working Paper Series (73)
- UW Biostatistics Working Paper Series (55)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (42)
- CBN Journal of Applied Statistics (JAS) (25)
-
- The University of Michigan Department of Biostatistics Working Paper Series (23)
- COBRA Preprint Series (20)
- Theses and Dissertations--Statistics (9)
- Electronic Theses and Dissertations (7)
- College of Graduate Studies: Theses & Dissertations (5)
- Theses and Dissertations (4)
- Graduate Theses and Dissertations (3)
- Memorial Sloan-Kettering Cancer Center, Dept. of Epidemiology & Biostatistics Working Paper Series (3)
- Statistical Science Theses and Dissertations (3)
- UNF Graduate Theses and Dissertations (3)
- Data Science and Data Mining (2)
- Statistics (2)
- Theses, Dissertations and Capstones (2)
- All Dissertations (1)
- All Graduate Theses and Dissertations, Spring 1920 to Summer 2023 (1)
- All other publications (1)
- Articles (1)
- Bioconductor Project Working Papers (1)
- Branch Mathematics and Statistics Faculty and Staff Publications (1)
- Business and Economics Honors Papers (1)
- CHIP Documents (1)
- CMC Senior Theses (1)
- Community & Environmental Health Faculty Publications (1)
- Department of Management: Faculty Publications (1)
- Department of Statistics: Dissertations, Theses, and Student Research (1)
- Publication Type
Articles 31 - 60 of 426
Full-Text Articles in Statistical Theory
Bayesian Methods For Graphical Models With Neighborhood Selection., Sagnik Bhadury
Bayesian Methods For Graphical Models With Neighborhood Selection., Sagnik Bhadury
Electronic Theses and Dissertations
Graphical models determine associations between variables through the notion of conditional independence. Gaussian graphical models are a widely used class of such models, where the relationships are formalized by non-null entries of the precision matrix. However, in high-dimensional cases, covariance estimates are typically unstable. Moreover, it is natural to expect only a few significant associations to be present in many realistic applications. This necessitates the injection of sparsity techniques into the estimation method. Classical frequentist methods, like GLASSO, use penalization techniques for this purpose. Fully Bayesian methods, on the contrary, are slow because they require iteratively sampling over a quadratic …
Fiscal And Monetary Policy Interactions In A Developing Economy: A Dsge-Based Evidence From Nigeria, Queen E. Oye, Philip O. Alege
Fiscal And Monetary Policy Interactions In A Developing Economy: A Dsge-Based Evidence From Nigeria, Queen E. Oye, Philip O. Alege
CBN Journal of Applied Statistics (JAS)
This study characterizes the nature of fiscal-monetary interaction in Nigeria and gauges its macroeconomic effects by estimating a New Keynesian Dynamic Stochastic General Equilibrium (NK DSGE) model. Two policy simulations were also conducted. The first experiment considers the desirable active-passive policy mix while the second experiment ranks alternative monetary policy rules among the differing objectives of price, output and exchange rate stabilization. The study finds that fiscal and monetary policies interact as complements in an active monetary and passive fiscal policy mix over the sample period. The result from the first policy simulation reveals that the active monetary and passive …
Functional Data Analysis Of Covid-19, Nichole L. Fluke
Functional Data Analysis Of Covid-19, Nichole L. Fluke
Mathematics & Statistics ETDs
This thesis deals with Functional Data Analysis (FDA) on COVID data. The Data involves counts for new COVID cases, hospitalized COVID patients, and new COVID deaths. The data used is for all the states and regions in the United States. The data starts in March 1st, 2020 and goes through March 31st, 2021. The FDA smooths the data and looks to see if there are similarities or differences between the states and regions in the data. The data also shows which states and regions stand out from the others and which ones are similar. Also shown …
Effect Of Monetary Policy Rate On Market Interest Rates In Nigeria: A Threshold And Nardl Approach, Oluwafemi E. Awopegba, Joseph O. Afolabi, Lydia T. Adeoye, Godwin O. Akpokodje
Effect Of Monetary Policy Rate On Market Interest Rates In Nigeria: A Threshold And Nardl Approach, Oluwafemi E. Awopegba, Joseph O. Afolabi, Lydia T. Adeoye, Godwin O. Akpokodje
CBN Journal of Applied Statistics (JAS)
This study examines the effect of monetary policy rate (MPR) on market interest rates in Nigeria. For parsimony, we develop two indexes called the short-term interest rate (SINT) and Lending interest rate (LINT) to represent deposit and lending rates respectively. The nonlinear autoregressive distributed lag (NARDL) and threshold regression models are adopted. The study uses monthly data from 2002:M1 to 2019:M12. The results of the threshold regression model indicate that the degree of the effect of MPR on SINT and LINT above the estimated threshold of 11 and 13 percent respectively is greater and significant than if MPR were to …
Social Dimension Of Inclusive Growth In Ecowas: Implication For Poverty Reduction, Toriola K. Anu, Goerge O. Emmanuel, Ajayi O. Felix
Social Dimension Of Inclusive Growth In Ecowas: Implication For Poverty Reduction, Toriola K. Anu, Goerge O. Emmanuel, Ajayi O. Felix
CBN Journal of Applied Statistics (JAS)
This study investigates the implication of the social dimension of inclusive growth on poverty reduction in Economic Community of West African States (ECOWAS) countries. It specifically examines how social indices of inclusive growth comprising of income inequality, education, and health outcomes affect poverty reduction. The study uses a panel dataset of the six (6) lower-middle income countries in ECOWAS which was analysed via panel Difference Generalised Method of Moment (D-GMM). The results show that GDP per capita exerts significant negative effect on poverty while inequality, education and health outcomes do not show significant effect on poverty. Although, the estimates of …
Effect Of Fdi Inflows On Employment Generation In Selected Ecowas Countries: Heterogeneous Panel Analysis, Timothy A. Aderemi, Olawunmi Omitogun, Bukonla G. Osisanwo
Effect Of Fdi Inflows On Employment Generation In Selected Ecowas Countries: Heterogeneous Panel Analysis, Timothy A. Aderemi, Olawunmi Omitogun, Bukonla G. Osisanwo
CBN Journal of Applied Statistics (JAS)
The aim of this study is to examine the effect of FDI on employment in ECOWAS sub region between 1990 and 2019. The study utilizes a panel autoregressive distributed lag model to analyse the short run and long run relationship between FDI and employment across ECOWAS sub region. In the short run, the impact of FDI on employment is negative and statistically not significant. Meanwhile, in the long run FDI has a positive and statistically significant impact on employment rate. This implies that FDI has the capacity to generate employment in countries in ECOWAS sub region. Therefore, this study recommends …
Impact Of Covid-19 Pandemic On The Nigeria Stock Market: A Sectoral Stock Prices Analysis, Peter A. Adekunle, Yakubu A. Bello, Udochukwu G. Nwachukwu
Impact Of Covid-19 Pandemic On The Nigeria Stock Market: A Sectoral Stock Prices Analysis, Peter A. Adekunle, Yakubu A. Bello, Udochukwu G. Nwachukwu
CBN Journal of Applied Statistics (JAS)
This study examines the impact of the COVID-19 pandemic on sectoral stock prices in Nigeria stock market using daily data covering from February 28, 2020 to June 26, 2020. Applying the autoregressive distributed lag (ARDL) bounds test, the study finds that COVID-19 pandemic had adverse impact on the stock market indices in the short run. Furthermore, the study documents negative response of sectoral stock prices to the pandemic while the stock prices of the banking sub-sector are the worst hit. Compared to the consumer goods, and industrial subsector indices, the speed of adjustment to long run equilibrium is faster for …
Advancements In Gaussian Process Learning For Uncertainty Quantification, John C. Nicholson
Advancements In Gaussian Process Learning For Uncertainty Quantification, John C. Nicholson
All Dissertations
Gaussian processes are among the most useful tools in modeling continuous processes in machine learning and statistics. The research presented provides advancements in uncertainty quantification using Gaussian processes from two distinct perspectives. The first provides a more fundamental means of constructing Gaussian processes which take on arbitrary linear operator constraints in much more general framework than its predecessors, and the other from the perspective of calibration of state-aware parameters in computer models. If the value of a process is known at a finite collection of points, one may use Gaussian processes to construct a surface which interpolates these values to …
Aberrant Responding With Underlying Dominance And Unfolding Response Processes: Examining Model Fit And Performance Of Person-Fit Statistics, Jennifer A. Reimers
Aberrant Responding With Underlying Dominance And Unfolding Response Processes: Examining Model Fit And Performance Of Person-Fit Statistics, Jennifer A. Reimers
Graduate Theses and Dissertations
Researchers have recognized that respondents may not answer items in a way that accurately reflects their attitude or trait level being measured. The resulting response data that deviates from what would be expected has been shown to have significant effects on the psychometric properties of a scale and analytical results. However, many studies that have investigated the detection of aberrant data and its effects have done so using dominance item response theory (IRT) models. It is unknown whether the impacts of aberrant data and the methodology used to identify aberrant responding when using dominance IRT models apply similarly when scales …
Upper-Sided Ewma-Based Distribution-Specific Tolerance Limits, Owen Visser
Upper-Sided Ewma-Based Distribution-Specific Tolerance Limits, Owen Visser
UNF Graduate Theses and Dissertations
Tolerance limits are constructed from sample data to ascertain if a proportion of a process is within specification limits. There exists multiple methods of calculating the sample size requirements for tolerance limits under various assumptions. In this research, a distribution-specific algorithm that utilizes the exponentially weighted moving average technique (EWMA), first introduced by Sa and Razaila (2004), is reconstructed. The algorithm is used to calculate the required sample sizes for continuous construction of upper-sided tolerance limits. The sample sizes and intervals constructed from them are compared to three existing methods for various distributions. The distribution-specific algorithm was observed to reduce …
Multi-Level Small Area Estimation Based On Calibrated Hierarchical Likelihood Approach Through Bias Correction With Applications To Covid-19 Data, Nirosha Rathnayake
Multi-Level Small Area Estimation Based On Calibrated Hierarchical Likelihood Approach Through Bias Correction With Applications To Covid-19 Data, Nirosha Rathnayake
Theses & Dissertations
Small area estimation (SAE) has been widely used in a variety of applications to draw estimates in geographic domains represented as a metropolitan area, district, county, or state. The direct estimation methods provide accurate estimates when the sample size of study participants within each area unit is sufficiently large, but it might not always be realistic to have large sample sizes of study participants when considering small geographical regions. Meanwhile, high dimensional socio-ecological data exist at the community level, providing an opportunity for model-based estimation by incorporating rich auxiliary information at the individual and area levels. Thus, it is critical …
Statistical Approaches Of Gene Set Analysis With Quantitative Trait Loci For High-Throughput Genomic Studies., Samarendra Das
Statistical Approaches Of Gene Set Analysis With Quantitative Trait Loci For High-Throughput Genomic Studies., Samarendra Das
Electronic Theses and Dissertations
Recently, gene set analysis has become the first choice for gaining insights into the underlying complex biology of diseases through high-throughput genomic studies, such as Microarrays, bulk RNA-Sequencing, single cell RNA-Sequencing, etc. It also reduces the complexity of statistical analysis and enhances the explanatory power of the obtained results. Further, the statistical structure and steps common to these approaches have not yet been comprehensively discussed, which limits their utility. Hence, a comprehensive overview of the available gene set analysis approaches used for different high-throughput genomic studies is provided. The analysis of gene sets is usually carried out based on …
Sensitivity Analysis For Incomplete Data And Causal Inference, Heng Chen
Sensitivity Analysis For Incomplete Data And Causal Inference, Heng Chen
Statistical Science Theses and Dissertations
In this dissertation, we explore sensitivity analyses under three different types of incomplete data problems, including missing outcomes, missing outcomes and missing predictors, potential outcomes in \emph{Rubin causal model (RCM)}. The first sensitivity analysis is conducted for the \emph{missing completely at random (MCAR)} assumption in frequentist inference; the second one is conducted for the \emph{missing at random (MAR)} assumption in likelihood inference; the third one is conducted for one novel assumption, the ``sixth assumption'' proposed for the robustness of instrumental variable estimand in causal inference.
A Monte Carlo Analysis Of Standard Error-Based Methods For Computing Confidence Intervals, Elayna Wichert
A Monte Carlo Analysis Of Standard Error-Based Methods For Computing Confidence Intervals, Elayna Wichert
Masters Theses & Specialist Projects
The objective of this study is to empirically test existing techniques to calculate the likely range of values for a Classical Test Theory true score given an observed score. The traditional method for forming these confidence intervals has used the standard error of measurement (SEM) as the basis for this confidence interval. An alternate equation, the standard error of estimate (SEE), has been recommended in place of the SEM for this purpose, yet it remains overlooked in the field of psychometrics. It is important that the correct equation be used in various applications in personnel psychology. Monte Carlo analyses were …
Accounting For The Uncertainty Due To Chemicals Below The Detection Limit In Mixture Analysis, Paul M. Hargarten
Accounting For The Uncertainty Due To Chemicals Below The Detection Limit In Mixture Analysis, Paul M. Hargarten
Theses and Dissertations
Humans are exposed to multiple chemicals every day. Epidemiological studies have shown that chemical mixtures are associated with cancers, allergies, neurodevelopmental disorders, and other adverse health effects. To assess these associations, investigators are increasingly using chemical mixture approaches like weighted quantile sum (WQS) regression. In these studies, the research objectives are to determine whether a mixture of correlated chemicals is associated with an adverse health outcome and to identify the important chemicals. However, as experimental equipment measures each exposure to a chemical-specific detection limit, the exposures are unknown between zero and the detection limit. Indeed, the number of exposures below …
How Machine Learning And Probability Concepts Can Improve Nba Player Evaluation, Harrison Miller
How Machine Learning And Probability Concepts Can Improve Nba Player Evaluation, Harrison Miller
CMC Senior Theses
In this paper I will be breaking down a scholarly article, written by Sameer K. Deshpande and Shane T. Jensen, that proposed a new method to evaluate NBA players. The NBA is the highest level professional basketball league in America and stands for the National Basketball Association. They proposed to build a model that would result in how NBA players impact their teams chances of winning a game, using machine learning and probability concepts. I preface that by diving into these concepts and their mathematical backgrounds. These concepts include building a linear model using ordinary least squares method, the bias …
Generalized Matrix Decomposition Regression: Estimation And Inference For Two-Way Structured Data, Yue Wang, Ali Shojaie, Tim Randolph, Jing Ma
Generalized Matrix Decomposition Regression: Estimation And Inference For Two-Way Structured Data, Yue Wang, Ali Shojaie, Tim Randolph, Jing Ma
UW Biostatistics Working Paper Series
Analysis of two-way structured data, i.e., data with structures among both variables and samples, is becoming increasingly common in ecology, biology and neuro-science. Classical dimension-reduction tools, such as the singular value decomposition (SVD), may perform poorly for two-way structured data. The generalized matrix decomposition (GMD, Allen et al., 2014) extends the SVD to two-way structured data and thus constructs singular vectors that account for both structures. While the GMD is a useful dimension-reduction tool for exploratory analysis of two-way structured data, it is unsupervised and cannot be used to assess the association between such data and an outcome of interest. …
Statistical Inference For Networks Of High-Dimensional Point Processes, Xu Wang, Mladen Kolar, Ali Shojaie
Statistical Inference For Networks Of High-Dimensional Point Processes, Xu Wang, Mladen Kolar, Ali Shojaie
UW Biostatistics Working Paper Series
Fueled in part by recent applications in neuroscience, high-dimensional Hawkes process have become a popular tool for modeling the network of interactions among multivariate point process data. While evaluating the uncertainty of the network estimates is critical in scientific applications, existing methodological and theoretical work have only focused on estimation. To bridge this gap, this paper proposes a high-dimensional statistical inference procedure with theoretical guarantees for multivariate Hawkes process. Key to this inference procedure is a new concentration inequality on the first- and second-order statistics for integrated stochastic processes, which summarizes the entire history of the process. We apply this …
Optimal Design For A Causal Structure, Zaher Kmail
Optimal Design For A Causal Structure, Zaher Kmail
Department of Statistics: Dissertations, Theses, and Student Research
Linear models and mixed models are important statistical tools. But in many natural phenomena, there is more than one endogenous variable involved and these variables are related in a sophisticated way. Structural Equation Modeling (SEM) is often used to model the complex relationships between the endogenous and exogenous variables. It was first implemented in research to estimate the strength and direction of direct and indirect effects among variables and to measure the relative magnitude of each causal factor.
Historically, traditional optimal design theory focuses on univariate linear, nonlinear, and mixed models. There is no current literature on the subject of …
Interpreting Patient Reported Outcomes In Orthopaedic Surgery: A Systematic Review, Shgufta Docter, Zina Fathalla, Michael Lukacs, Michaela Khan, Morgan Jennings, Shu-Hsuan Liu, Dong Zi, Dianne Bryant
Interpreting Patient Reported Outcomes In Orthopaedic Surgery: A Systematic Review, Shgufta Docter, Zina Fathalla, Michael Lukacs, Michaela Khan, Morgan Jennings, Shu-Hsuan Liu, Dong Zi, Dianne Bryant
Western Research Forum
Background: Reporting methods of patient reported outcome measures (PROMs) vary in orthopaedic surgery literature. While most studies report statistical significance, the interpretation of results would be improved if authors reported confidence intervals (CIs), the minimally clinically important difference (MCID), and number needed to treat (NNT).
Objective: To assess the quality and interpretability of reporting the results of PROMs. To evaluate reporting, we will assess the proportion of studies that reported (1) 95% CIs, (2) MCID, and (3) NNT. To evaluate interpretation, we will assess the proportion of studies that discussed results using the MCID or the effect sizes and how …
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes, Lauren Spirko-Burns, Karthik Devarajan
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes, Lauren Spirko-Burns, Karthik Devarajan
COBRA Preprint Series
One of the major goals in large-scale genomic studies is to identify genes with a prognostic impact on time-to-event outcomes which provide insight into the disease's process. With rapid developments in high-throughput genomic technologies in the past two decades, the scientific community is able to monitor the expression levels of tens of thousands of genes and proteins resulting in enormous data sets where the number of genomic features is far greater than the number of subjects. Methods based on univariate Cox regression are often used to select genomic features related to survival outcome; however, the Cox model assumes proportional hazards …
Neural Shrubs: Using Neural Networks To Improve Decision Trees, Kyle Caudle, Randy Hoover, Aaron Alphonsus
Neural Shrubs: Using Neural Networks To Improve Decision Trees, Kyle Caudle, Randy Hoover, Aaron Alphonsus
SDSU Data Science Symposium
Decision trees are a method commonly used in machine learning to either predict a categorical response or a continuous response variable. Once the tree partitions the space, the response is either determined by the majority vote – classification trees, or by averaging the response values – regression trees. This research builds a standard regression tree and then instead of averaging the responses, we train a neural network to determine the response value. We have found that our approach typically increases the predicative capability of the decision tree. We have 2 demonstrations of this approach that we wish to present as …
Controlling For Confounding Via Propensity Score Methods Can Result In Biased Estimation Of The Conditional Auc: A Simulation Study, Hadiza I. Galadima, Donna K. Mcclish
Controlling For Confounding Via Propensity Score Methods Can Result In Biased Estimation Of The Conditional Auc: A Simulation Study, Hadiza I. Galadima, Donna K. Mcclish
Community & Environmental Health Faculty Publications
In the medical literature, there has been an increased interest in evaluating association between exposure and outcomes using nonrandomized observational studies. However, because assignments to exposure are not random in observational studies, comparisons of outcomes between exposed and nonexposed subjects must account for the effect of confounders. Propensity score methods have been widely used to control for confounding, when estimating exposure effect. Previous studies have shown that conditioning on the propensity score results in biased estimation of conditional odds ratio and hazard ratio. However, research is lacking on the performance of propensity score methods for covariate adjustment when estimating the …
Some New Generalized Distribution Via Lindley-Weibuli And Lindley-Log-Logistic Distributions With Applications, Soliu A. Raheem
Some New Generalized Distribution Via Lindley-Weibuli And Lindley-Log-Logistic Distributions With Applications, Soliu A. Raheem
College of Graduate Studies: Theses & Dissertations
In this thesis, new generalized distributions, namely Beta Lindley-Log-Logistic (BLLLoG) distribution, Marshall-Olkin Lindley-Weibull (MOLW) distribution, and Gamma LindleyWeibull (GLW) distribution as well as related sub-distributions are proposed. Series expansion of the densities are obtained. Statistical properties of these distributions, including hazard function, reverse hazard function, moments, reliability, quantile function, mean deviations, Bonferroni and Lorenz curves, entropy and Fisher information are derived. Method of maximum likelihood is used to estimate the parameters of the new distributions. Monte Carlo simulation is employed to examine the performance of the proposed distributions. Applications of the generalized distributions to real lifetime data are presented to …
Modeling Stochastically Intransitive Relationships In Paired Comparison Data, Ryan Patrick Alexander Mcshane
Modeling Stochastically Intransitive Relationships In Paired Comparison Data, Ryan Patrick Alexander Mcshane
Statistical Science Theses and Dissertations
If the Warriors beat the Rockets and the Rockets beat the Spurs, does that mean that the Warriors are better than the Spurs? Sophisticated fans would argue that the Warriors are better by the transitive property, but could Spurs fans make a legitimate argument that their team is better despite this chain of evidence?
We first explore the nature of intransitive (rock-scissors-paper) relationships with a graph theoretic approach to the method of paired comparisons framework popularized by Kendall and Smith (1940). Then, we focus on the setting where all pairs of items, teams, players, or objects have been compared to …
Statistical Designs For Network A/B Testing, Victoria V. Pokhilko
Statistical Designs For Network A/B Testing, Victoria V. Pokhilko
Theses and Dissertations
A/B testing refers to the statistical procedure of experimental design and analysis to compare two treatments, A and B, applied to different testing subjects. It is widely used by technology companies such as Facebook, LinkedIn, and Netflix, to compare different algorithms, web-designs, and other online products and services. The subjects participating in these online A/B testing experiments are users who are connected in different scales of social networks. Two connected subjects are similar in terms of their social behaviors, education and financial background, and other demographic aspects. Hence, it is only natural to assume that their reactions to online products …
Statistical Modeling Of Influenza-Like-Illness In Montana Using Spatial And Temporal Methods, Benjamin A. Stark
Statistical Modeling Of Influenza-Like-Illness In Montana Using Spatial And Temporal Methods, Benjamin A. Stark
Graduate Student Theses, Dissertations, & Professional Papers
Studying air pollution and public health has been a historically important question in science. It has long been hypothesized that severe air pollution conditions lead to negative implications in basic human health. Primarily, areas thats are prone to severe degrees of human pollution are the focus of such studies. Such research relating to less populated areas are scarce, and this scarcity raises the question of how such pollution dynamics (human-made and natural) influence human health in more rural areas.
The aim of this study is to explore this hole in research; in particular we explore possible links between air pollution …
Minimizing The Perceived Financial Burden Due To Cancer, Hassan Azhar, Zoheb Allam, Gino Varghese, Daniel W. Engels, Sajiny John
Minimizing The Perceived Financial Burden Due To Cancer, Hassan Azhar, Zoheb Allam, Gino Varghese, Daniel W. Engels, Sajiny John
SMU Data Science Review
In this paper, we present a regression model that predicts perceived financial burden that a cancer patient experiences in the treatment and management of the disease. Cancer patients do not fully understand the burden associated with the cost of cancer, and their lack of understanding can increase the difficulties associated with living with the disease, in particular coping with the cost. The relationship between demographic characteristics and financial burden were examined in order to better understand the characteristics of a cancer patient and their burden, while all subsets regression was used to determine the best predictors of financial burden. Age, …
Evaluation Of Using The Bootstrap Procedure To Estimate The Population Variance, Nghia Trong Nguyen
Evaluation Of Using The Bootstrap Procedure To Estimate The Population Variance, Nghia Trong Nguyen
Electronic Theses and Dissertations
The bootstrap procedure is widely used in nonparametric statistics to generate an empirical sampling distribution from a given sample data set for a statistic of interest. Generally, the results are good for location parameters such as population mean, median, and even for estimating a population correlation. However, the results for a population variance, which is a spread parameter, are not as good due to the resampling nature of the bootstrap method. Bootstrap samples are constructed using sampling with replacement; consequently, groups of observations with zero variance manifest in these samples. As a result, a bootstrap variance estimator will carry a …
The Family Of Conditional Penalized Methods With Their Application In Sufficient Variable Selection, Jin Xie
The Family Of Conditional Penalized Methods With Their Application In Sufficient Variable Selection, Jin Xie
Theses and Dissertations--Statistics
When scientists know in advance that some features (variables) are important in modeling a data, then these important features should be kept in the model. How can we utilize this prior information to effectively find other important features? This dissertation is to provide a solution, using such prior information. We propose the Conditional Adaptive Lasso (CAL) estimates to exploit this knowledge. By choosing a meaningful conditioning set, namely the prior information, CAL shows better performance in both variable selection and model estimation. We also propose Sufficient Conditional Adaptive Lasso Variable Screening (SCAL-VS) and Conditioning Set Sufficient Conditional Adaptive Lasso Variable …