Open Access. Powered by Scholars. Published by Universities.®

Statistics and Probability Commons™

Open Access. Powered by Scholars. Published by Universities.®

2013

Discipline
Institution
Keyword
Publication
Publication Type
File Type

Articles 301 - 330 of 555

Full-Text Articles in Statistics and Probability

Randomization Test P-Values Versus Significance Levels, Bryan Manly May 2013

Randomization Test P-Values Versus Significance Levels, Bryan Manly

Journal of Modern Applied Statistical Methods

Bryan Manly responds to Richard Anderson's article Conceptual Distinction between the Critical p Value and the Type I Error Rate in Permutation Testing.


Conceptual Distinction Between The Critical P Value And The Type I Error Rate In Permutation Testing: Author Response To Peer Comments, Richard B. Anderson May 2013

Conceptual Distinction Between The Critical P Value And The Type I Error Rate In Permutation Testing: Author Response To Peer Comments, Richard B. Anderson

Journal of Modern Applied Statistical Methods

Richard Anderson responds to comments regarding his target article Conceptual Distinction between the Critical p Value and the Type I Error Rate in Permutation Testing.


Estimation Of Variance Using Known Coefficient Of Variation And Median Of An Auxiliary Variable, J. Subramani, G. Kumarapandiyan May 2013

Estimation Of Variance Using Known Coefficient Of Variation And Median Of An Auxiliary Variable, J. Subramani, G. Kumarapandiyan

Journal of Modern Applied Statistical Methods

A modified ratio type variance estimator for estimating population variance of a study variable when the population median and coefficient of variation of an auxiliary variable are known is proposed. The bias and mean squared error of the proposed estimator are derived and conditions under which the proposed estimator performs better than the traditional ratio type variance estimators and modified ratio type variance estimators are obtained. Using a numerical study results show that the proposed estimator performs better than the traditional ratio type variance estimator and existing modified ratio type variance estimators.


Priorities In Thurstone Scaling And Steady-State Probabilities In Markov Stochastic Modeling, Stan Lipovetsky May 2013

Priorities In Thurstone Scaling And Steady-State Probabilities In Markov Stochastic Modeling, Stan Lipovetsky

Journal of Modern Applied Statistical Methods

Thurstone scaling is widely used in marketing and advertising research where various methods of applied psychology are utilized. This article considers several analytical tools useful for positioning a set of items on a Thurstone scale via regression modeling and Markov stochastic processing in the form of Chapman-Kolmogorov equations. These approaches produce interval and ratio scales of preferences and enrich the possibilities of paired comparison estimation applied for solving practical problems of prioritization and probability of choice modeling.


On The Gamma-Half Normal Distribution And Its Applications, Ayman Alzaatreh, Kristen Knight May 2013

On The Gamma-Half Normal Distribution And Its Applications, Ayman Alzaatreh, Kristen Knight

Journal of Modern Applied Statistical Methods

A new distribution, the gamma-half normal distribution, is proposed and studied. Various structural properties of the gamma-half normal distribution are derived. The shape of the distribution may be unimodal or bimodal. Results for moments, limit behavior, mean deviations and Shannon entropy are provided. To estimate the model parameters, the method of maximum likelihood estimation is proposed. Three real-life data sets are used to illustrate the applicability of the gamma-half normal distribution.


The Length-Biased Versus Random Sampling For The Binomial And Poisson Events, Makarand V. Ratnaparkhi, Uttara V. Naik-Nimbalkar May 2013

The Length-Biased Versus Random Sampling For The Binomial And Poisson Events, Makarand V. Ratnaparkhi, Uttara V. Naik-Nimbalkar

Journal of Modern Applied Statistical Methods

The equivalence between the length-biased and the random sampling on a non-negative, discrete random variable is established. The length-biased versions of the binomial and Poisson distributions are discussed.


Improved Estimators In Finite Population Surveys: Theory And Applications, Sunil Kumar May 2013

Improved Estimators In Finite Population Surveys: Theory And Applications, Sunil Kumar

Journal of Modern Applied Statistical Methods

Improved estimators are proposed for estimating the population mean Y̅ of the study variable y using auxiliary variable x in simple random sampling. Explicit expression for the bias and MSE of the proposed family are derived to the first order of approximation. The proposed estimators are compared with other estimators and theoretical findings are illustrated by two numerical examples.


An Approach For Dealing With Statuses Of Non-Statistically Significant Interactions Between Treatments, Zakaria M. Sawan May 2013

An Approach For Dealing With Statuses Of Non-Statistically Significant Interactions Between Treatments, Zakaria M. Sawan

Journal of Modern Applied Statistical Methods

A field experiment on cotton yield resulted in a non-statistically significant interaction. An approach for follow-up examination between treatments based on least significant difference values was suggested to identify the effect regardless of insignificance. It was found that the classical formula used in calculating the significance of interactions suffers a possible shortage that can be eliminated by applying a suggested revision.


An Application Of Machine Learning Methods To The Derivation Of Exposure-Response Curves For Respiratory Outcomes, Ekaterina Eliseeva, Alan E. Hubbard, Ira B. Tager May 2013

An Application Of Machine Learning Methods To The Derivation Of Exposure-Response Curves For Respiratory Outcomes, Ekaterina Eliseeva, Alan E. Hubbard, Ira B. Tager

U.C. Berkeley Division of Biostatistics Working Paper Series

Analyses of epidemiological studies of the association between short-term changes in air pollution and health outcomes have not sufficiently discussed the degree to which the statistical models chosen for these analyses reflect what is actually known about the true data-generating distribution. We present a method to estimate population-level ambient air pollution (NO2) exposure-health (wheeze in children with asthma) response functions that is not dependent on assumptions about the data-generating function that underlies the observed data and which focuses on a specific scientific parameter of interest (the marginal adjusted association of exposure on probability of wheeze, over a grid of possible …


Section Abstracts: Statistics May 2013

Section Abstracts: Statistics

Virginia Journal of Science

Abstracts of the Statistics Section for the 91st Annual Virginia Journal of Science Meeting, May 2013


Analyzing And Solving Non-Linear Stochastic Dynamic Models On Non-Periodic Discrete Time Domains, Gang Cheng May 2013

Analyzing And Solving Non-Linear Stochastic Dynamic Models On Non-Periodic Discrete Time Domains, Gang Cheng

Masters Theses & Specialist Projects

Stochastic dynamic programming is a recursive method for solving sequential or multistage decision problems. It helps economists and mathematicians construct and solve a huge variety of sequential decision making problems in stochastic cases. Research on stochastic dynamic programming is important and meaningful because stochastic dynamic programming reflects the behavior of the decision maker without risk aversion; i.e., decision making under uncertainty. In the solution process, it is extremely difficult to represent the existing or future state precisely since uncertainty is a state of having limited knowledge. Indeed, compared to the deterministic case, which is decision making under certainty, the stochastic …


The Torsion Angle Of Random Walks, Mu He May 2013

The Torsion Angle Of Random Walks, Mu He

Masters Theses & Specialist Projects

In this thesis, we study the expected mean of the torsion angle of an n-step
equilateral random walk in 3D. We consider the random walk is generated within a confining sphere or without a confining sphere: given three consecutive vectors →e1 , →e2 , and →e3 of the random walk then the vectors →e1 and →e2 define a plane and the vectors →e2 and →e3 define a second plane. The angle between the two planes is called the torsion angle of the three vectors. Algorithms are …


Penalized Smoothed Partial Rank Estimator For The Nonparametric Transformation Survival Model With High-Dimensional Covariates, Wei Dai, Yi Li May 2013

Penalized Smoothed Partial Rank Estimator For The Nonparametric Transformation Survival Model With High-Dimensional Covariates, Wei Dai, Yi Li

The University of Michigan Department of Biostatistics Working Paper Series

Microarray technology has the potential to lead to a better understanding of biological processes and diseases such as cancer. When failure time outcomes are also available, one might be interested in relating gene expression profiles to the survival outcome such as time to cancer recurrence or time to death. This is statistically challenging because the number of covariates greatly exceeds the number of observations. While the majority of work has focused on regularized Cox regression model and accelerated failure time model, they may be restrictive in practice. We relax the model assumption and and consider a nonparametric transformation model that …


Bayesian Hypothesis Testing And Variable Selection In High Dimensional Regression, Min Wang May 2013

Bayesian Hypothesis Testing And Variable Selection In High Dimensional Regression, Min Wang

All Dissertations

This dissertation consists of three distinct but related research projects. First of all, we study the Bayesian approach to model selection in the class of normal regression models. We propose an explicit closed-form expression of the Bayes factor with the use of Zellner's g-prior and the beta-prime prior for g. Noting that linear models with a growing number of unknown parameters have recently gained increasing popularity in practice, such as the spline problem, we shall thus be particularly interested in studying the model selection consistency of the Bayes factor under the scenario in which the dimension of the parameter space …


Conceptual Distinction Between The Critical P Value And The Type I Error Rate In Permutation Testing, Richard B. Anderson May 2013

Conceptual Distinction Between The Critical P Value And The Type I Error Rate In Permutation Testing, Richard B. Anderson

Journal of Modern Applied Statistical Methods

To counter past assertions that permutation testing is not distribution-free, this article clarifies that the critical p value (alpha) in permutation testing is not a Type I error rate and that a test's validity is independent of the concept of Type I error.


A Response To Anderson's (2013) Conceptual Distinction Between The Critical P Value And Type I Error Rate In Permutation Testing, Fortunato Pesarin, Stefano Bonnini May 2013

A Response To Anderson's (2013) Conceptual Distinction Between The Critical P Value And Type I Error Rate In Permutation Testing, Fortunato Pesarin, Stefano Bonnini

Journal of Modern Applied Statistical Methods

Pesarin and Bonnini respond to Anderson's (2013) Conceptual Distinction between the Critical p value and Type I Error Rate in Permutation Testing


A Monte Carlo Simulation Of The Robust Rank-Order Test Under Various Population Symmetry Conditions, William T. Mickelson May 2013

A Monte Carlo Simulation Of The Robust Rank-Order Test Under Various Population Symmetry Conditions, William T. Mickelson

Journal of Modern Applied Statistical Methods

The Type I Error Rate of the Robust Rank Order test under various population symmetry conditions is explored through Monte Carlo simulation. Findings indicate the test has difficulty controlling Type I error under generalized Behrens-Fisher conditions for moderately sized samples.


Constructing A More Powerful Test In Two-Level Block Randomized Designs, Spyros Konstantopoulos May 2013

Constructing A More Powerful Test In Two-Level Block Randomized Designs, Spyros Konstantopoulos

Journal of Modern Applied Statistical Methods

A more powerful test is proposed for the treatment effect in two-level block randomized designs where random assignment takes place at the first level. When clustering at the second level is assumed to be known, the proposed test produces higher estimates of power than the typical test.


Bayesian Inference Of Pair-Copula Constriction For Multivariate Dependency Modeling Of Iran’S Macroeconomic Variables, M. R. Zadkarami, O. Chatrabgoun May 2013

Bayesian Inference Of Pair-Copula Constriction For Multivariate Dependency Modeling Of Iran’S Macroeconomic Variables, M. R. Zadkarami, O. Chatrabgoun

Journal of Modern Applied Statistical Methods

Bayesian inference of pair-copula constriction (PCC) is used for multivariate dependency modeling of Iran’s macroeconomics variables: oil revenue, economic growth, total consumption and investment. These constructions are based on bivariate t-copulas as building blocks and can model the nature of extreme events in bivariate margins individually. The model parameter was estimated based on Markov chain Monte Carlo (MCMC) methods. A MCMC algorithm reveals unconditional as well as conditional independence in Iran’s macroeconomic variables, which can simplify resulting PCC’s for these data.


Error Covariance Matrix Estimation In High Dimensional Approximate Factor Models Using Adaptive Thresholding: A Simulation Study, Paul Chimenti May 2013

Error Covariance Matrix Estimation In High Dimensional Approximate Factor Models Using Adaptive Thresholding: A Simulation Study, Paul Chimenti

All Theses

Approximate factor models are popular in finance and economics. A key to effectively utilizing such a model is to accurately estimate the error covariance matrix. Errors related to certain predictors are expected to be correlated and this must be modeled effectively. Adaptive thresholding is a method for estimating the error covariance matrix of such a model. This method is described in detail and a simulation study sheds light on the behavior of this method under different sample sizes and parameterizations.


Using The Bootstrap For Estimating The Sample Size In Statistical Experiments, Maher Qumsiyeh May 2013

Using The Bootstrap For Estimating The Sample Size In Statistical Experiments, Maher Qumsiyeh

Journal of Modern Applied Statistical Methods

Efron’s (1979) Bootstrap has been shown to be an effective method for statistical estimation and testing. It provides better estimates than normal approximations for studentized means, least square estimates and many other statistics of interest. It can be used to select the active factors - factors that have an effect on the response - in experimental designs. This article shows that the bootstrap can be used to determine sample size or the number of runs required to achieve a certain confidence level in statistical experiments.


The X-Alter Algorithm: A Parameter-Free Method Of Unsupervised Clustering, Thomas Laloë, Rémi Servien May 2013

The X-Alter Algorithm: A Parameter-Free Method Of Unsupervised Clustering, Thomas Laloë, Rémi Servien

Journal of Modern Applied Statistical Methods

Using quantization techniques, Laloë (2010) defined a new clustering algorithm called Alter. This L1-based algorithm is shown to be convergent but suffers two major flaws. The number of clusters, K, must be supplied by the user and the computational cost is high. This article adapts the X-means algorithm (Pelleg & Moore, 2000) to solve both problems.


Estimating Heterogeneous Intra-Class Correlation Coefficients In Dyadic Ecological Momentary Assessment, Emily A. Blood, Leslie A. Kalish, Lydia A. Shrier May 2013

Estimating Heterogeneous Intra-Class Correlation Coefficients In Dyadic Ecological Momentary Assessment, Emily A. Blood, Leslie A. Kalish, Lydia A. Shrier

Journal of Modern Applied Statistical Methods

A method is described for estimating and testing predictors for influence on the variance of momentary behaviors in dyadic ecological momentary assessment data. Results show that the method allows intraclass correlations of momentary observations from two members of the same couple to vary by observation-level, individual-level and couple-level predictors.


Jmasm 32: Multiple Imputation Of Missing Multilevel, Longitudinal Data: A Case When Practical Considerations Trump Best Practices?, Jennifer E. V. Lloyd, Jelena Obradović, Richard M. Carpiano, Frosso Motti-Stefanidi May 2013

Jmasm 32: Multiple Imputation Of Missing Multilevel, Longitudinal Data: A Case When Practical Considerations Trump Best Practices?, Jennifer E. V. Lloyd, Jelena Obradović, Richard M. Carpiano, Frosso Motti-Stefanidi

Journal of Modern Applied Statistical Methods

A pedagogical tool is presented for applied researchers dealing with incomplete multilevel, longitudinal data. It explains why such data pose special challenges regarding missingness. Syntax created to perform a multiply-imputed growth modeling procedure in Stata Version 11 (StataCorp, 2009) is also described.


Bootstrap Interval Estimation Of Reliability Via Coefficient Omega, Miguel A. Padilla, Jasmin Divers May 2013

Bootstrap Interval Estimation Of Reliability Via Coefficient Omega, Miguel A. Padilla, Jasmin Divers

Journal of Modern Applied Statistical Methods

Three different bootstrap confidence intervals (CIs) for coefficient omega were investigated. The CIs were assessed through a simulation study with conditions not previously investigated. All methods performed well; however, the normal theory bootstrap (NTB) CI had the best performance because it had more consistent acceptable coverage under the simulation conditions investigated.


Fitting Proportional Odds Models To Educational Data With Complex Sampling Designs In Ordinal Logistic Regression, Xing Liu, Hari Koirala May 2013

Fitting Proportional Odds Models To Educational Data With Complex Sampling Designs In Ordinal Logistic Regression, Xing Liu, Hari Koirala

Journal of Modern Applied Statistical Methods

The conventional proportional odds (PO) model assumes that data are collected using simple random sampling by which each sampling unit has the equal probability of being selected from a population. However, when complex survey sampling designs are used, such as stratified sampling, clustered sampling or unequal selection probabilities, it is inappropriate to conduct ordinal logistic regression analyses without taking sampling design into account. Failing to do so may lead to biased estimates of parameters and incorrect corresponding variances. This study illustrates the use of PO models with complex survey data to predict mathematics proficiency levels using Stata and compare the …


Compound Identification Using Penalized Linear Regression., Ruiqi Liu May 2013

Compound Identification Using Penalized Linear Regression., Ruiqi Liu

Electronic Theses and Dissertations

In this study, we propose a new method for compound identification using penalized linear regression. Compound identification is often achieved by matching the experimental mass spectra to the mass spectra stored in a reference library based on mass spectral similarity. In the context of the linear regression, the response variable is an experimental mass spectrum (i.e., query) and all the compounds in the reference library are the independent variables. However, the number of compounds in the reference library is much larger than the range of m/z values so that the data become high dimensional data with suffering from singularity. For …


Geometric Framework For Evaluating Rare Variant Tests Of Association, Keli Liu, Shannon Fast, Matthew Zawistowski, Nathan L. Tintle May 2013

Geometric Framework For Evaluating Rare Variant Tests Of Association, Keli Liu, Shannon Fast, Matthew Zawistowski, Nathan L. Tintle

Faculty Work Comprehensive List

The wave of next-generation sequencing data has arrived. However, many questions still remain about how to best analyze sequence data, particularly the contribution of rare genetic variants to human disease. Numerous statistical methods have been proposed to aggregate association signals across multiple rare variant sites in an effort to increase statistical power; however, the precise relation between the tests is often not well understood. We present a geometric representation for rare variant data in which rare allele counts in case and control samples are treated as vectors in Euclidean space. The geometric framework facilitates a rigorous classification of existing rare …


Optimal Methods For Using Posterior Probabilities In Association Testing, Keli Liu, Alexander Luedtke, Nathan L. Tintle May 2013

Optimal Methods For Using Posterior Probabilities In Association Testing, Keli Liu, Alexander Luedtke, Nathan L. Tintle

Faculty Work Comprehensive List

Objective: The use of haplotypes to impute the genotypes of unmeasured single nucleotide variants continues to rise in popularity. Simulation results suggest that the use of the dosage as a one-dimensional summary statistic of imputation posterior probabilities may be optimal both in terms of statistical power and computational efficiency; however, little theoretical understanding is available to explain and unify these simulation results. In our analysis, we provide a theoretical foundation for the use of the dosage as a one-dimensional summary statistic of genotype posterior probabilities from any technology. Methods: We analytically evaluate the dosage, mode and the more general set …


The Association Of Diet And Physical Activity With Stroke Mortality Results From The Adventist Health Study-1, Tahereh Zamansani May 2013

The Association Of Diet And Physical Activity With Stroke Mortality Results From The Adventist Health Study-1, Tahereh Zamansani

Loma Linda University Electronic Theses, Dissertations & Projects

Stroke is one of the leading causes of death and disability, with major global public health implications. Stroke ranks No. 4 among all causes of death, behind heart disease, cancer, and chronic lower respiratory disease (CLRD). Stroke accounts for almost 1 of every 18 deaths in the United States. Women accounted for 60.6% of stroke deaths. Death certificate data show that the mean age at stroke death was 79.6 years; males had a younger mean age (76.3) than females.

There is still a great scientific uncertainty among researchers and epidemiologists about the magnitude of any preventive effect, mechanisms of action …