Generalizing Multistage Partition Procedures For Two-Parameter Exponential Populations,
2018
University of New Orleans
Generalizing Multistage Partition Procedures For Two-Parameter Exponential Populations, Rui Wang
LSU New Orleans Theses and Dissertations
ANOVA analysis is a classic tool for multiple comparisons and has been widely used in numerous disciplines due to its simplicity and convenience. The ANOVA procedure is designed to test if a number of different populations are all different. This is followed by usual multiple comparison tests to rank the populations. However, the probability of selecting the best population via ANOVA procedure does not guarantee the probability to be larger than some desired prespecified level. This lack of desirability of the ANOVA procedure was overcome by researchers in early 1950's by designing experiments with the goal of selecting the best …
Development Of A Statistical Model For Discrimination Of Rupture Status In Posterior Communicating Artery Aneurysms,
2018
George Mason University
Development Of A Statistical Model For Discrimination Of Rupture Status In Posterior Communicating Artery Aneurysms, Felicitas J. Detmer, Bong Jae Chung, Fernando Mut, Michael Pritz, Martin Slawski, Farid Hamzei-Sichani, David Kallmes, Christopher Putman, Carlos Jimenez, Juan R. Cebral
Department of Applied Mathematics and Statistics Faculty Scholarship and Creative Works
Background: Intracranial aneurysms at the posterior communicating artery (PCOM) are known to have high rupture rates compared to other locations. We developed and internally validated a statistical model discriminating between ruptured and unruptured PCOM aneurysms based on hemodynamic and geometric parameters, angio-architectures, and patient age with the objective of its future use for aneurysm risk assessment. Methods: A total of 289 PCOM aneurysms in 272 patients modeled with image-based computational fluid dynamics (CFD) were used to construct statistical models using logistic group lasso regression. These models were evaluated with respect to discrimination power and goodness of fit using tenfold nested …
Bayesian Analytical Approaches For Metabolomics : A Novel Method For Molecular Structure-Informed Metabolite Interaction Modeling, A Novel Diagnostic Model For Differentiating Myocardial Infarction Type, And Approaches For Compound Identification Given Mass Spectrometry Data.,
2018
University of Louisville
Bayesian Analytical Approaches For Metabolomics : A Novel Method For Molecular Structure-Informed Metabolite Interaction Modeling, A Novel Diagnostic Model For Differentiating Myocardial Infarction Type, And Approaches For Compound Identification Given Mass Spectrometry Data., Patrick J. Trainor
Electronic Theses and Dissertations
Metabolomics, the study of small molecules in biological systems, has enjoyed great success in enabling researchers to examine disease-associated metabolic dysregulation and has been utilized for the discovery biomarkers of disease and phenotypic states. In spite of recent technological advances in the analytical platforms utilized in metabolomics and the proliferation of tools for the analysis of metabolomics data, significant challenges in metabolomics data analyses remain. In this dissertation, we present three of these challenges and Bayesian methodological solutions for each. In the first part we develop a new methodology to serve a basis for making higher order inferences in metabolomics, …
Comparison Of Correlation, Partial Correlation, And Conditional Mutual Information For Interaction Effects Screening In Generalized Linear Models,
2018
University of Arkansas, Fayetteville
Comparison Of Correlation, Partial Correlation, And Conditional Mutual Information For Interaction Effects Screening In Generalized Linear Models, Ji Li
Graduate Theses and Dissertations
Numerous screening techniques have been developed in recent years for genome-wide association studies (GWASs) (Moore et al., 2010). In this thesis, a novel model-free screening method was developed and validated by an extensive simulation study. Many screening methods were mainly focused on main effects, while very few studies considered the models containing both main effects and interaction effects. In this work, the interaction effects were fully considered and three different methods (Pearson’s Correlation Coefficient, Partial Correlation, and Conditional Mutual Information) were tested and their prediction accuracies were compared.
Pearson’s Correlation Coefficient method, which is a direct interaction screening (DIS) procedure, …
Clustering Mixed Data: An Extension Of The Gower Coefficient With Weighted L2 Distance,
2018
East Tennessee State University
Clustering Mixed Data: An Extension Of The Gower Coefficient With Weighted L2 Distance, Augustine Oppong
Electronic Theses and Dissertations
Sorting out data into partitions is increasing becoming complex as the constituents of data is growing outward everyday. Mixed data comprises continuous, categorical, directional functional and other types of variables. Clustering mixed data is based on special dissimilarities of the variables. Some data types may influence the clustering solution. Assigning appropriate weight to the functional data may improve the performance of the clustering algorithm. In this paper we use the extension of the Gower coefficient with judciously chosen weight for the L2 to cluster mixed data.The benefits of weighting are demonstrated both in in applications to the Buoy data set …
Pretrial Release And Failure-To-Appear In Mclean County, Il,
2018
Illinois State University
Pretrial Release And Failure-To-Appear In Mclean County, Il, Jonathan Monsma
Student Research – Stevenson Center
Actuarial risk assessment tools increasingly have been employed in jurisdictions across the U.S. to assist courts in the decision of whether someone charged with a crime should be detained or released prior to their trial. These tools should be continually monitored and researched by independent 3rd parties to ensure that these powerful tools are being administered properly and used in the most proficient way as to provide socially optimal results. McLean County, Illinois began using the Public Safety Assessment-CourtTM (PSA-Court or simply PSA) risk assessment tool beginning in 2016. This study culls data from the McLean County Jail …
A Distance Based Method For Solving Multi-Objective Optimization Problems,
2018
Aligarh Muslim University
A Distance Based Method For Solving Multi-Objective Optimization Problems, Murshid Kamal, Syed Aqib Jalil, Syed Mohd Muneeb, Irfan Ali
Journal of Modern Applied Statistical Methods
A new model for the weighted method of goal programming is proposed based on minimizing the distances between ideal objectives to feasible objective space. It provides the best compromised solution for Multi Objective Linear Programming Problems (MOLPP). The proposed model tackles MOLPP by solving a series of single objective sub-problems, where the objectives are transformed into constraints. The compromise solution so obtained may be improved by defining priorities in terms of the weight. A criterion is also proposed for deciding the best compromise solution. Applications of the algorithm are discussed for transportation and assignment problems involving multiple and conflicting objectives. …
Goalie Analytics: Statistical Evaluation Of Context-Specific Goalie Performance Measures In The National Hockey League,
2018
Southern Methodist University
Goalie Analytics: Statistical Evaluation Of Context-Specific Goalie Performance Measures In The National Hockey League, Marc Naples, Logan Gage, Amy Nussbaum
SMU Data Science Review
In this paper, we attempt to improve upon the classic formulation of save percentage in the NHL by controlling the context of the shots and use alternative measures than save percentage. In particular, we find save percentage to be both a weakly repeatable skill and predictor of future performance, and we seek other goalie performance calculations that are more robust. To do so, we use three primary tests to test intra-season consistency, intra-season predictability, and inter-season consistency, and extend the analysis to disentangle team effects on goalie statistics. We find that there are multiple ways to improve upon classic save …
Data Scientist’S Analysis Toolbox: Comparison Of Python, R, And Sas Performance,
2018
Southern Methodist University
Data Scientist’S Analysis Toolbox: Comparison Of Python, R, And Sas Performance, Jim Brittain, Mariana Cendon, Jennifer Nizzi, John Pleis
SMU Data Science Review
A quantitative analysis will be performed on experiments utilizing three different tools used for Data Science. The analysis will include replication of analysis along with comparisons of code length, output, and results. Qualitative data will supplement the quantitative findings. The conclusion will provide data support guidance on the correct tool to use for common situations in the field of Data Science.
Estimation Of Finite Population Mean By Using Minimum And Maximum Values In Stratified Random Sampling,
2018
Quaid-i-Azam University
Estimation Of Finite Population Mean By Using Minimum And Maximum Values In Stratified Random Sampling, Umer Daraz, Javid Shabbir, Hina Khan
Journal of Modern Applied Statistical Methods
In this paper we have suggested an improved class of ratio type estimators in estimating the finite population mean when information on minimum and maximum values of the auxiliary variable is known. The properties of the suggested class of estimators in terms of bias and mean square error are obtained up to first order of approximation. Two data sets are used for efficiency comparisons.
A Bayesian Beta-Mixture Model For Nonparametric Irt (Bbm-Irt),
2018
University of Illinois at Chicago
A Bayesian Beta-Mixture Model For Nonparametric Irt (Bbm-Irt), Ethan A. Arenson, George Karabatsos
Journal of Modern Applied Statistical Methods
Item response models typically assume that the item characteristic (step) curves follow a logistic or normal cumulative distribution function, which are strictly monotone functions of person test ability. Such assumptions can be overly-restrictive for real item response data. A simple and more flexible Bayesian nonparametric IRT model for dichotomous items is introduced, which constructs monotone item characteristic (step) curves by a finite mixture of beta distributions, which can support the entire space of monotone curves to any desired degree of accuracy. An adaptive random-walk Metropolis-Hastings algorithm is proposed to estimate the posterior distribution of the model parameters. The Bayesian IRT …
Robust Estimation And Inference On Current Status Data With Applications To Phase Iv Cancer Trial,
2018
St. Jude Children's Research Hospital
Robust Estimation And Inference On Current Status Data With Applications To Phase Iv Cancer Trial, Deo Kumar Srivastava, Liang Zhu, Melissa M. Hudson, Jianmin Pan, Shesh N. Rai
Journal of Modern Applied Statistical Methods
The use of piecewise exponential distributions was proposed by Rai et al. (2013) for analyzing cardiotoxicity data. Some parametric models are proposed, but the focus is on the Weibull distribution, which overcomes the limitation of piecewise exponential.
Robust Heteroscedasticity Consistent Covariance Matrix Estimator Based On Robust Mahalanobis Distance And Diagnostic Robust Generalized Potential Weighting Methods In Linear Regression,
2018
Universiti Putra Malaysia
Robust Heteroscedasticity Consistent Covariance Matrix Estimator Based On Robust Mahalanobis Distance And Diagnostic Robust Generalized Potential Weighting Methods In Linear Regression, M. Habshah, Muhammad Sani, Jayanthi Arasan
Journal of Modern Applied Statistical Methods
The violation of the assumption of homoscedasticity and the presence of high leverage points (HLPs) are common in the use of regression models. The weighted least squares can provide the solution to heteroscedastic regression model if the heteroscedastic error structures are known. Based on Furno (1996), two robust weighting methods are proposed based on HLP detection measures (robust Mahalanobis distance based on minimum volume ellipsoid and diagnostic robust generalized potential based on index set equality (DRGP(ISE)) on robust heteroscedasticity consistent covariance matrix estimators. Results obtained from a simulation study and real data sets indicated the DRGP(ISE) method is superior.
Masked Instability: Within-Sector Financial Risk In The Presence Of Wealth Inequality,
2018
Montclair State University
Masked Instability: Within-Sector Financial Risk In The Presence Of Wealth Inequality, Youngna Choi
Department of Applied Mathematics and Statistics Faculty Scholarship and Creative Works
We investigate masked financial instability caused by wealth inequality. When an economic sector is decomposed into two subsectors that possess a severe wealth inequality, the sector in entirety can look financially stable while the two subsectors possess extreme financially instabilities of opposite nature, one from excessive equity, the other from lack thereof. The unstable subsector can result in further financial distress and even trigger a financial crisis. The market instability indicator, an early warning system derived from dynamical systems applied to agent-based models, is used to analyze the subsectoral financial instabilities. Detailed mathematical analysis is provided to explain what financial …
Internal Consistency Reliability In Measurement: Aggregate And Multilevel Approaches,
2018
Harvard Medical School
Internal Consistency Reliability In Measurement: Aggregate And Multilevel Approaches, Georgios Sideridis, Abdullah Saddaawi, Khaleel Al-Harbi
Journal of Modern Applied Statistical Methods
The purpose of the present paper was to evaluate the internal consistency reliability of the General Teacher Test assuming clustered and non-clustered data using commercial software (Mplus). Participants were 2,000 testees who were selected using random sampling from a larger pool of examinees (more than 65k). The measure involved four factors, namely: (a) planning for learning, (b) promoting learning, (c) supporting learning, and (d) professional responsibilities and was hypothesized to comprise a unidimensional instrument assessing generalized skills and competencies. Intra-class correlation coefficients and variance ratio statistics suggested the need to incorporate a clustering variable (i.e., university) when evaluating the factor …
Fitting The Rasch Model Under The Logistic Regression Framework To Reduce Estimation Bias,
2018
Pearson
Fitting The Rasch Model Under The Logistic Regression Framework To Reduce Estimation Bias, Tianshu Pan
Journal of Modern Applied Statistical Methods
This article showed how and why the Rasch model can be fitted under the logistic regression framework. Then a penalized maximum likelihood (Firth 1993) for logistic regression models can also be used to reduce ML biases when fitting the Rasch model. These conclusions are supported by a simulation study.
Regressions Regularized By Correlations,
2018
GfK North America
Regressions Regularized By Correlations, Stan Lipovetsky
Journal of Modern Applied Statistical Methods
The regularization of multiple regression by proportionality to correlations of predictors with dependent variable is applied to the least squares objective and normal equations to relax the exact equalities and to get a robust solution. This technique produces models not prone to multicollinearity and is very useful in practical applications.
An Explanatory Study On The Non-Parametric Multivariate T2 Control Chart,
2018
Department of Statistics, University of Isfahan, Isfahan, Iran
An Explanatory Study On The Non-Parametric Multivariate T2 Control Chart, Abdolrasoul Mostajeran, Nasrolah Iranpanah, Rassoul Noorossana
Journal of Modern Applied Statistical Methods
Most control charts require the assumption of normal distribution for observations. When distribution is not normal, one can use non-parametric control charts such as sign control chart. A deficiency of such control charts could be the loss of information due to replacing an observation with its sign or rank. Furthermore, because the chart statistics of T2 are correlated, the T2 chart is not a desire performance. Non-parametric bootstrap algorithm could help to calculate control chart parameters using the original observations while no assumption regarding the distribution is needed. In this paper, first, a bootstrap multivariate control chart is …
Optimum Stratification In Bivariate Auxiliary Variables Under Neyman Allocation,
2018
Sher-e-Kashmir University of Agricultural Sciences and Technology of Jammu, J&K India
Optimum Stratification In Bivariate Auxiliary Variables Under Neyman Allocation, Faizan Danish, S.E.H. Rizvi
Journal of Modern Applied Statistical Methods
In several situations complete data set of the study variable is unknown that becomes a stumbling block in various stratification techniques in order to obtain stratification points on two way stratification method. In this paper a technique has been proposed under Neyman allocation when the stratification is done oj the two auxiliary variable having one estimation variable under consideration. Due to complexities created by minimal equations approximate optimum strata boundaries has been obtained. Empirical study has been done to illustrate the proposed method when the auxiliary variables have standard Cauchy and power distributions.
Estimation Of Zero-Inflated Population Mean: A Bootstrapping Approach,
2018
University of Wisconsin-Whitewater
Estimation Of Zero-Inflated Population Mean: A Bootstrapping Approach, Khyam Paneru, R. Noah Padgett, Hanfeng Chen
Journal of Modern Applied Statistical Methods
A mixture model was adopted from the maximum pseudo-likelihood approach under complex sampling designs to estimate the mean of zero-inflated population. To overcome the complexity and assumptions of asymptotic distribution, the maximum pseudo-likelihood function was used, but a bootstrapping procedure was proposed as an alternative. Bootstrap confidence intervals consistently capture the true means of zero-inflated populations of the simulation studies.
