Open Access. Powered by Scholars. Published by Universities.®

Applied Statistics Commons™

Open Access. Powered by Scholars. Published by Universities.®

Social and Behavioral Sciences

Institution
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 541 - 570 of 1395

Full-Text Articles in Applied Statistics

Test For The Equality Of Partial Correlation Coefficients For Two Populations, Madhusudan Bhandary, Arjun K. Gupta May 2015

Test For The Equality Of Partial Correlation Coefficients For Two Populations, Madhusudan Bhandary, Arjun K. Gupta

Journal of Modern Applied Statistical Methods

A likelihood ratio test for the equality of two partial correlation coefficients based on two independent multinormal samples has been derived. The large sample Z-test for the same problem has also been discussed. The power analysis of the two tests is obtained. It has been found that the approximate likelihood ratio (ALR) test showed consistently better results than Z -test in terms of power. The size of the ALR test is slightly more than the alpha level. The ALR test is recommended strongly for use in practice.


Comparison Of Model Fit Indices Used In Structural Equation Modeling Under Multivariate Normality, Sengul Cangur, Ilker Ercan May 2015

Comparison Of Model Fit Indices Used In Structural Equation Modeling Under Multivariate Normality, Sengul Cangur, Ilker Ercan

Journal of Modern Applied Statistical Methods

The purpose of this study is to investigate the impact of estimation techniques and sample sizes on model fit indices in structural equation models constructed according to the number of exogenous latent variables under multivariate normality. The performances of fit indices are compared by considering effects of related factors. The Ratio Chi-square Test Statistic to Degree of Freedom, Root Mean Square Error of Approximation, and Comparative Fit Index are the least affected indices by estimation technique and sample size under multivariate normality, especially with large sample size.


Method Of Estimation In The Presence Of Non-Response And Measurement Errors Simultaneously, Rajesh Singh Singh, Prayas Sharma May 2015

Method Of Estimation In The Presence Of Non-Response And Measurement Errors Simultaneously, Rajesh Singh Singh, Prayas Sharma

Journal of Modern Applied Statistical Methods

The problem of estimating the finite population mean of in simple random sampling in the presence of non-response and response error was considered. The estimators use auxiliary information to improve efficiency, assuming non–response and measurement error are present in both the study and auxiliary variables. A class of estimators was proposed and its properties studied in the simultaneous presence of non-response and response errors. It was shown that the proposed class of estimators is more efficient than the usual unbiased estimator, ratio and product estimators under non-response and response error together. A numerical study was carried out to compare its …


Pseudo-Random Number Generators For Vector Processors And Multicore Processors, Agner Fog May 2015

Pseudo-Random Number Generators For Vector Processors And Multicore Processors, Agner Fog

Journal of Modern Applied Statistical Methods

Large scale Monte Carlo applications need a good pseudo-random number generator capable of utilizing both the vector processing capabilities and multiprocessing capabilities of modern computers in order to get the maximum performance. The requirements for such a generator are discussed. New ways of avoiding overlapping subsequences by combining two generators are proposed. Some fundamental philosophical problems in proving independence of random streams are discussed. Remedies for hitherto ignored quantization errors are offered. An open source C++ implementation is provided for a generator that meets these needs.


Estimating The Accuracy Of Automated Classification Systems Using Only Expert Ratings That Are Less Accurate Than The System, Paul E. Lehner May 2015

Estimating The Accuracy Of Automated Classification Systems Using Only Expert Ratings That Are Less Accurate Than The System, Paul E. Lehner

Journal of Modern Applied Statistical Methods

A method is presented to estimate the accuracy of an automated classification system based only on expert ratings on test cases, where the system may be substantially more accurate than the raters. In this method an estimate of overall rater accuracy is derived from the level of inter-rater agreement, Bayesian updating based on estimated rater accuracy is applied to estimate a ground truth probability for each classification on each test case, and then overall system accuracy is estimated by comparing the relative frequency that the system agrees with the most probable classification at different probability levels. A simulation analysis provides …


Modeling Probability Of Causal And Random Impacts, Stan Lipovetsky, Igor Mandel May 2015

Modeling Probability Of Causal And Random Impacts, Stan Lipovetsky, Igor Mandel

Journal of Modern Applied Statistical Methods

The method of the estimation of the probability of an event occurring under the influence of the causal and random effects is considered. Epistemological differences from the traditional approaches to causality are discussed, and a new model of the statistical estimation of the parameters of each effect is proposed. The simple and effective algorithms of the model parameters estimation are presented, and numerical simulations are performed. A practical marketing example is analyzed. The results support the validity of the estimation procedure and open the perspective for the application of the method for various decision making problems, where different causes can …


Estimation For The Parameters Of The Exponentiated Exponential Distribution Using A Median Ranked Set Sampling, Monjed H. Samuh, Areen Qtait May 2015

Estimation For The Parameters Of The Exponentiated Exponential Distribution Using A Median Ranked Set Sampling, Monjed H. Samuh, Areen Qtait

Journal of Modern Applied Statistical Methods

The method of maximum likelihood estimation based on Median Ranked Set Sampling (MRSS) was used to estimate the shape and scale parameters of the Exponentiated Exponential Distribution (EED). They were compared with the conventional estimators. The relative efficiency was used for comparison. The amount of information (in Fisher's sense) available from the MRSS about the parameters of the EED were be evaluated. Confidence intervals for the parameters were constructed using MRSS.


Estimating The Strength Of An Association Based On A Robust Smoother, Rand Wilcox May 2015

Estimating The Strength Of An Association Based On A Robust Smoother, Rand Wilcox

Journal of Modern Applied Statistical Methods

It is known that the more obvious parametric approaches to fitting a regression line to data are often not flexible enough to provide an adequate approximation of the true regression line. Many nonparametric regression estimators, often called smoothers, have been derived that are aimed at dealing with this problem. The paper deals with the issue of estimating the strength of an association based on the fit obtained by a robust smoother. A simple approach, already known, is to estimate explanatory power in a fairly obvious manner. This approach has been found to perform reasonably well when using the smoother LOESS. …


Per Family Or Familywise Type I Error Control: "Eether, Eyether, Neether, Nyther, Let's Call The Whole Thing Off!", H. J. Keselman May 2015

Per Family Or Familywise Type I Error Control: "Eether, Eyether, Neether, Nyther, Let's Call The Whole Thing Off!", H. J. Keselman

Journal of Modern Applied Statistical Methods

Frane (2015) pointed out the difference between per-family and familywise Type I error control and how different multiple comparison procedures control one method but not necessarily the other. He then went on to demonstrate in the context of a two group multivariate design containing different numbers of dependent variables and correlations between variables how the per-family rate inflates beyond the level of significance. In this article I reintroduce other newer better methods of Type I error control. These newer methods provide more power to detect effects than the per-family and familywise techniques of control yet maintain the overall rate of …


Comparison Of Bayesian Credible Intervals To Frequentist Confidence Intervals, Kathy Gray, Brittany Hampton, Tony Silveti-Falls, Allison Mcconnell, Casey Bausell May 2015

Comparison Of Bayesian Credible Intervals To Frequentist Confidence Intervals, Kathy Gray, Brittany Hampton, Tony Silveti-Falls, Allison Mcconnell, Casey Bausell

Journal of Modern Applied Statistical Methods

Frequentist confidence intervals were compared with Bayesian credible intervals under a variety of scenarios to determine when Bayesian credible intervals outperform frequentist confidence intervals. Results indicated that Bayesian interval estimation frequently produces results with precision greater than or equal to the frequentist method.


Special Education Distributions And Analysis, Valerie Felder, Shlomo S. Sawilowsky May 2015

Special Education Distributions And Analysis, Valerie Felder, Shlomo S. Sawilowsky

Journal of Modern Applied Statistical Methods

Micceri (1989) examined the distributional characteristics of 440 large sample general education achievement and psychometric measures. All the distributions were found to be statistically significantly different from the normal distribution. In this study, 395 special education datasets were examined. Although there were some normally distributed datasets, most were not, and some were markedly different in shape from those found by Micceri (1989). Implications for statistical testing and making special education policy decisions were given.


Vol. 14, No. 1 (Full Issue), Jmasm Editors May 2015

Vol. 14, No. 1 (Full Issue), Jmasm Editors

Journal of Modern Applied Statistical Methods

.


A Comparison Of Semi-Parametric And Nonparametric Methods For Estimating Mean Time To Event For Randomly Left Censored Data, Farzana Chowdhury, Jahida Gulshan, Syed Shahadat Hossain May 2015

A Comparison Of Semi-Parametric And Nonparametric Methods For Estimating Mean Time To Event For Randomly Left Censored Data, Farzana Chowdhury, Jahida Gulshan, Syed Shahadat Hossain

Journal of Modern Applied Statistical Methods

The aim of this study was to make a comparison among existing estimation methods (Kaplan-Meier, Nelson-Aalen and Regression on Ordered Statistics (ROS)) for randomly left censored time to event data under selected distributions and for different level of censoring and sample sizes in order to determine the strength of these methods based on simulated data. Comparisons among the methods are made on the basis of unbiasedness and Monte Carlo Standard Error of the summary statistics (mean time to event) obtained by those methods under different conditions.


A Comparison Of Population-Averaged And Cluster-Specific Approaches In The Context Of Unequal Probabilities Of Selection, Natalie A. Koziol May 2015

A Comparison Of Population-Averaged And Cluster-Specific Approaches In The Context Of Unequal Probabilities Of Selection, Natalie A. Koziol

College of Education and Human Sciences: Dissertations, Theses, and Student Research

Sampling designs of large-scale, federally funded studies are typically complex, involving multiple design features (e.g., clustering, unequal probabilities of selection). Researchers must account for these features in order to obtain unbiased point estimators and make valid inferences about population parameters. Single-level (i.e., population-averaged) and multilevel (i.e., cluster-specific) methods provide two alternatives for modeling clustered data. Single-level methods rely on the use of adjusted variance estimators to account for dependency due to clustering, whereas multilevel methods incorporate the dependency into the specification of the model.

Although the literature comparing single-level and multilevel approaches is vast, comparisons have been limited to the …


Characteristics Of Stem Success: A Survival Analysis Model Of Factors Influencing Time To Graduation Among Undergraduate Stem Majors, Riley K. Acton Apr 2015

Characteristics Of Stem Success: A Survival Analysis Model Of Factors Influencing Time To Graduation Among Undergraduate Stem Majors, Riley K. Acton

Business and Economics Honors Papers

Producing more graduates in Science, Technology, Engineering, and Mathematics (STEM), as well as ensuring students complete college in a timely manner are both areas of national public policy interest. In order to improve these two outcomes, it is imperative to understand what factors lead undergraduate students to persist in, and ultimately graduate with STEM degrees. This paper uses data from the Beginning Postsecondary Students Longitudinal Study, provided by The National Center of Education Statistics, to model the time to baccalaureate degree among STEM majors using a Cox proportional hazard model.


Mathematical Modeling Of Trending Topics On Twitter, Jonathan S. Skaza Apr 2015

Mathematical Modeling Of Trending Topics On Twitter, Jonathan S. Skaza

Honors Projects in Mathematics

Created in 2006, Twitter is an online social networking service in which users share and read 140-character messages called Tweets. The site has approximately 288 million monthly active users who produce about 500 million Tweets per day. This study applies dynamical and statistical modeling strategies to quantify the spread of information on Twitter. Parameter estimates for the rates of infection and recovery are obtained using Bayesian Markov Chain Monte Carlo (MCMC) methods. The methodological strategy employed is an extension of techniques traditionally used in an epidemiological and biomedical context (particularly in the spread of infectious disease). This study, which addresses …


Best Practice Recommendations For Data Screening, Justin A. Desimone, Peter D. Harms, Alice J. Desimone Feb 2015

Best Practice Recommendations For Data Screening, Justin A. Desimone, Peter D. Harms, Alice J. Desimone

Department of Management: Faculty Publications

Survey respondents differ in their levels of attention and effort when responding to items. There are a number of methods researchers may use to identify respondents who fail to exert sufficient effort in order to increase the rigor of analysis and enhance the trustworthiness of study results. Screening techniques are organized into three general categories, which differ in impact on survey design and potential respondent awareness. Assumptions and considerations regarding appropriate use of screening techniques are discussed along with descriptions of each technique. The utility of each screening technique is a function of survey design and administration. Each technique has …


Equate: Observed-Score Linking And Equating In R, Anthony D. Albano Jan 2015

Equate: Observed-Score Linking And Equating In R, Anthony D. Albano

Department of Educational Psychology: Faculty Publications

Linking and equating are statistical procedures used to convert scores from one measurement scale to another. These procedures are most often used in testing programs that involve multiple test forms, where adjustments are made for form difficulty differences when creating a measurement scale that is common across forms. Linking and equating methods are traditionally distinguished by the type of scores they are applied to, whether observed scores or scores from an item response theory model. Methods are also distinguished by the study design under which measurements are taken. The R package equate (Albano, 2014) is free, open-source software for conducting …


Applications Of Monte Carlo Methods In Statistical Inference Using Regression Analysis, Ji Young Huh Jan 2015

Applications Of Monte Carlo Methods In Statistical Inference Using Regression Analysis, Ji Young Huh

CMC Senior Theses

This paper studies the use of Monte Carlo simulation techniques in the field of econometrics, specifically statistical inference. First, I examine several estimators by deriving properties explicitly and generate their distributions through simulations. Here, simulations are used to illustrate and support the analytical results. Then, I look at test statistics where derivations are costly because of the sensitivity of their critical values to the data generating processes. Simulations here establish significance and necessity for drawing statistical inference. Overall, the paper examines when and how simulations are needed in studying econometric theories.


Analyzing Alcohol Behavior In San Luis Obispo, Ariana Montes Dec 2014

Analyzing Alcohol Behavior In San Luis Obispo, Ariana Montes

Statistics

No abstract provided.


A Vegetation Analysis On Horn Island, Mississippi, Ca. 1940 Using Characteristic Dimensions Derived From Historical Aerial Photography, Guy Wilburn Jeter Jr. Dec 2014

A Vegetation Analysis On Horn Island, Mississippi, Ca. 1940 Using Characteristic Dimensions Derived From Historical Aerial Photography, Guy Wilburn Jeter Jr.

Master's Theses

Horn Island is part of the MS/AL barrier island chain in the northern Gulf of Mexico located approximately 18kn off the coast of Mississippi. This island’s habitats have undergone many transitions over the last several decades. The goal of this study was to quantify habitat change over a seventy year period using historical black and white photography from 1940. Using present NAIP imagery from the USDA, habitat structure was estimated by using geo-statistics, and second order statistics, from a co-occurrence matrix, to characterize texture for habitat classification. Percent land cover was then calculated to determine overall land cover change over …


Top Of The Order: Modeling The Optimal Locations Of Minor League Baseball Teams, W. Coleman Conley Nov 2014

Top Of The Order: Modeling The Optimal Locations Of Minor League Baseball Teams, W. Coleman Conley

Undergraduate Economic Review

Over the last twenty-five years, minor league baseball franchises have defined firm mobility. Revisiting the work of Michael C. Davis (2006), I construct a logistic regression model to predict which cities house minor league baseball teams. Six variables are tested for inclusion in the model, including population, income level, the number of major-league professional sports teams in a city, five-year population change, and distance from the closest professional team. Based on the model's predicted probabilities, cities are ranked in order of highest probability of having a team at each of the different levels from Class A to Class AAA.


Retained-Components Factor Transformation: Factor Loadings And Factor Score Predictors In The Column Space Of Retained Components, André Beauducel, Frank Spohn Nov 2014

Retained-Components Factor Transformation: Factor Loadings And Factor Score Predictors In The Column Space Of Retained Components, André Beauducel, Frank Spohn

Journal of Modern Applied Statistical Methods

Factor loadings optimally account for the non-diagonal elements of the covariance matrix of observed variables. Principal component analysis leads to components accounting for a maximum of the variance of the observed variables. Retained-components factor transformation is proposed in order to combine the advantages of factor analysis and principal component analysis.


Pairwise Comparison In Repeated Measures, I.C.A. Oyeka, C. C. Nnanatu Nov 2014

Pairwise Comparison In Repeated Measures, I.C.A. Oyeka, C. C. Nnanatu

Journal of Modern Applied Statistical Methods

Sometimes a random sample of subjects or patients may be exposed to a battery of diagnostic tests or medication over time and interest is on determining whether there is progressive remission of condition, disease or symptom. Also perhaps early in a program or experiment, subjects or candidates may be required to significantly improve in their performance rates at the current trial relative to an immediately preceding trial, otherwise they may have to withdraw from or drop out. The research interest would then be to determine some critical minimum marginal success rate to guide the management in decision making as well …


A Comparison Of Methods For Group Prediction With High Dimensional Data, Holmes Finch Nov 2014

A Comparison Of Methods For Group Prediction With High Dimensional Data, Holmes Finch

Journal of Modern Applied Statistical Methods

High dimensional data is the situation in which the number of variables included in an analysis approaches or exceeds the sample size. In the context of group classification, researchers are typically interested in finding a model that can be used to correctly place an individual into their appropriate group; e.g. correctly diagnose individuals with depression. However, when the size of the training sample is small and the number of predictors used to differentiate the groups is larger, standard approaches such as discriminant analysis may not work well. In order to address this issue, statisticians have developed a number of tools …


Objective Priors For Estimation Of Extended Exponential Geometric Distribution, Pedro L. Ramos, Fernando A. Moala, Jorge A. Achcar Nov 2014

Objective Priors For Estimation Of Extended Exponential Geometric Distribution, Pedro L. Ramos, Fernando A. Moala, Jorge A. Achcar

Journal of Modern Applied Statistical Methods

A Bayesian analysis was developed with different noninformative prior distributions such as Jeffreys, Maximal Data Information, and Reference. The aim was to investigate the effects of each prior distribution on the posterior estimates of the parameters of the extended exponential geometric distribution, based on simulated data and a real application.


Some General Guidelines For Choosing Missing Data Handling Methods In Educational Research, Jehanzeb R. Cheema Nov 2014

Some General Guidelines For Choosing Missing Data Handling Methods In Educational Research, Jehanzeb R. Cheema

Journal of Modern Applied Statistical Methods

The effect of a number of factors, such as the choice of analytical method, the handling method for missing data, sample size, and proportion of missing data, were examined to evaluate the effect of missing data treatment on accuracy of estimation. A methodological approach involving simulated data was adopted. One outcome of the statistical analyses undertaken in this study is the formulation of easy-to-implement guidelines for educational researchers that allows one to choose one of the following factors when all others are given: sample size, proportion of missing data in the sample, method of analysis, and missing data handling method.


Bayesian Inference For Volatility Of Stock Prices, Juliet G. D'Cunha, K. A. Rao Nov 2014

Bayesian Inference For Volatility Of Stock Prices, Juliet G. D'Cunha, K. A. Rao

Journal of Modern Applied Statistical Methods

Lognormal distribution is widely used in the analysis of failure time data and stock prices. Maximum likelihood and Bayes estimator of the coefficient of variation of lognormal distribution along with confidence/credible intervals are developed. The utility of Bayes procedure is illustrated by analyzing prices of selected stocks.


Local Bandwidths For Improving Performance Statistics Of Model-Robust Regression 2, Efosa Edionwe, Julian L. Mbegbu Nov 2014

Local Bandwidths For Improving Performance Statistics Of Model-Robust Regression 2, Efosa Edionwe, Julian L. Mbegbu

Journal of Modern Applied Statistical Methods

Model-Robust Regression 2 (MRR2) method is a semi-parametric regression approach that combines parametric and nonparametric fits. The bandwidth controls the smoothness of the nonparametric portion. We present a methodology for deriving data-driven local bandwidth that enhances the performance of MRR2 method for fitting curves to data generated from designed experiments.


Contrast Of Bayesian And Classical Sample Size Determination, Farhana Sadia, Syed S. Hossain Nov 2014

Contrast Of Bayesian And Classical Sample Size Determination, Farhana Sadia, Syed S. Hossain

Journal of Modern Applied Statistical Methods

Sample size determination is a prerequisite for statistical surveys. A comprehensive overview of the Bayesian approach for computation of the sample size, and a comparison with classical approaches, is presented. Two surveys are taken as example to illustrate the accuracy and efficiency of each approach, and to make recommendations about which method is preferred. The Bayesian approach of sample size determination may require fewer subjects if proper prior information is available.