Open Access. Powered by Scholars. Published by Universities.®

Statistical Theory Commons™

Open Access. Powered by Scholars. Published by Universities.®

Discipline
Institution
Keyword
Publication Year
Publication
Publication Type

Articles 1321 - 1350 of 1633

Full-Text Articles in Statistical Theory

Bias Affiliated With Two Variants Of Cohen’S D When Determining U1 As A Measure Of The Percent Of Non-Overlap, David A. Walker May 2005

Bias Affiliated With Two Variants Of Cohen’S D When Determining U1 As A Measure Of The Percent Of Non-Overlap, David A. Walker

Journal of Modern Applied Statistical Methods

Variants of Cohen’s d, in this instance dt and dadj, has the largest influence on U1 measures used with smaller sample sizes, specifically when n1 and n2 = 10. This study indicated that bias for variants of d, which influence U1 measures, tends to subside and become more manageable, in terms of precision of estimation, around 1% to 2% when n1 and n2 = 20. Thus, depending on the direction of the influence, both dt and dadj are likely to manage bias in the U1 measure quite well for smaller to …


Some Guidelines For Using Nonparametric Methods For Modeling Data From Response Surface Designs, Christine M. Anderson-Cook, Kathryn Prewitt May 2005

Some Guidelines For Using Nonparametric Methods For Modeling Data From Response Surface Designs, Christine M. Anderson-Cook, Kathryn Prewitt

Journal of Modern Applied Statistical Methods

Traditional response surface methodology focuses on modeling responses using parametric models with designs chosen to balance cost with adequate estimation of parameters and prediction in the design space. Using nonparametric smoothing to approximate the response surface offers both opportunities as well as problems. This article explores some conditions under which these methods can be appropriately used to increase the flexibility of surfaces modeled. The Box and Draper (1987) printing ink study is considered to illustrate the methods.


Determining The Correct Number Of Components To Extract From A Principal Components Analysis: A Monte Carlo Study Of The Accuracy Of The Scree Plot, Gibbs Y. Kanyongo May 2005

Determining The Correct Number Of Components To Extract From A Principal Components Analysis: A Monte Carlo Study Of The Accuracy Of The Scree Plot, Gibbs Y. Kanyongo

Journal of Modern Applied Statistical Methods

This article pertains to the accuracy of the of the scree plot in determining the correct number of components to retain under different conditions of sample size, component loading and variable-tocomponent ratio. The study employs use of Monte Carlo simulations in which the population parameters were manipulated, and data were generated, and then the scree plot applied to the generated scores.


Bayesian Wavelet Estimation Of Long Memory Parameter, Leming Qu May 2005

Bayesian Wavelet Estimation Of Long Memory Parameter, Leming Qu

Journal of Modern Applied Statistical Methods

A Bayesian wavelet estimation method for estimating parameters of a stationary I(d) process is represented as an useful alternative to the existing frequentist wavelet estimation methods. The effectiveness of the proposed method is demonstrated through Monte Carlo simulations. The sampling from the posterior distribution is through the Markov Chain Monte Carlo (MCMC) easily implemented in the WinBUGS software package.


Enhancing The Performance Of A Short Run Multivariate Control Chart For The Process Mean, Michael B. C. Khoo, T. F. Ng May 2005

Enhancing The Performance Of A Short Run Multivariate Control Chart For The Process Mean, Michael B. C. Khoo, T. F. Ng

Journal of Modern Applied Statistical Methods

No abstract provided.


Jmasm16: Pseudo-Random Number Generation In R For Some Univariate Distributions, Hakan Demirtas May 2005

Jmasm16: Pseudo-Random Number Generation In R For Some Univariate Distributions, Hakan Demirtas

Journal of Modern Applied Statistical Methods

An increasing number of practitioners and applied researchers started using the R programming system in recent years for their computing and data analysis needs. As far as pseudo-random number generation is concerned, the built-in generator in R does not contain some important univariate distributions. In this article, complementary R routines that could potentially be useful for simulation and computation purposes are provided.


Jmasm17: An Algorithm And Code For Computing Exact Critical Values For Friedman’S Nonparametric Anova, Sikha Bagui, Sbuhash Bagui May 2005

Jmasm17: An Algorithm And Code For Computing Exact Critical Values For Friedman’S Nonparametric Anova, Sikha Bagui, Sbuhash Bagui

Journal of Modern Applied Statistical Methods

Provided in this article is an algorithm and code for computing exact critical values (or percentiles) for Friedman’s nonparametric rank test for k related treatment populations using Visual Basic (VB.NET). This program has the ability to calculate critical values for any number of treatment populations ( k ) and block sizes (b) at any significance level (α ) . We developed an exact critical value table for k = 2(1)5 and b = 2(1)15. This table will be useful to practitioners since it is not available in standard nonparametric statistics texts. The program can also be used to compute any …


Jmasm19: A Spss Matrix For Determining Effect Sizes From Three Categories: R And Functions Of R, Differences Between Proportions, And Standardized Differences Between Means, David A. Walker May 2005

Jmasm19: A Spss Matrix For Determining Effect Sizes From Three Categories: R And Functions Of R, Differences Between Proportions, And Standardized Differences Between Means, David A. Walker

Journal of Modern Applied Statistical Methods

The program is intended to provide editors, manuscript reviewers, students, and researchers with an SPSS matrix to determine an array of effect sizes not reported or the correctness of those reported, such as rrelated indices, r-related squared indices, and measures of association, when the only data provided in the manuscript or article are the n, M, and SD (and sometimes proportions and t and F (1) values) for twogroup designs. This program can create an internal matrix table to assist researchers in determining the size of an effect for commonly utilized r-related, mean difference, and difference in proportions indices when …


Letter To The Editor, Jmasm Editors May 2005

Letter To The Editor, Jmasm Editors

Journal of Modern Applied Statistical Methods

No abstract provided.


An Exploration Of Using Data Mining In Educational Research, Yonghong Jade Xu May 2005

An Exploration Of Using Data Mining In Educational Research, Yonghong Jade Xu

Journal of Modern Applied Statistical Methods

Technology advances popularized large databases in education. Traditional statistics have limitations for analyzing large quantities of data. This article discusses data mining by analyzing a data set with three models: multiple regression, data mining, and a combination of the two. It is concluded that data mining is applicable in educational research.


Effect Of Position Of An Outlier On The Influence Curve Of The Measures Of Preferred Direction For Circular Data, B. Sango Otieno, Christine M. Anderson-Cook May 2005

Effect Of Position Of An Outlier On The Influence Curve Of The Measures Of Preferred Direction For Circular Data, B. Sango Otieno, Christine M. Anderson-Cook

Journal of Modern Applied Statistical Methods

Circular or angular data occur in many fields of applied statistics. A common problem of interest in circular data is estimating a preferred direction and its corresponding distribution. It is complicated by the wrap-around effect on the circle, which exists because there is no natural minimum or maximum. The usual statistics employed for linear data are inappropriate for directional data, as they do not account for its circular nature. The robustness of the three common choices for summarizing the preferred direction (the sample circular mean, sample circular median and a circular analog of the Hodges-Lehmann estimator) are evaluated via their …


An Empirical Evaluation Of The Retrospective Pretest: Are There Advantages To Looking Back?, Paul A. Nakonezny, Joseph Lee Rodgers May 2005

An Empirical Evaluation Of The Retrospective Pretest: Are There Advantages To Looking Back?, Paul A. Nakonezny, Joseph Lee Rodgers

Journal of Modern Applied Statistical Methods

This article builds on research regarding response shift effects and retrospective self-report ratings. Results suggest moderate evidence of a response shift bias in the conventional pretest-posttest treatment design in the treatment group. The use of explicitly worded anchors on response scales, as well as the measurement of knowledge ratings (a cognitive construct) in an evaluation methodology setting, helped to mitigate the magnitude of a response shift bias. The retrospective pretest-posttest design provides a measure of change that is more in accord with the objective measure of change than is the conventional pretest-posttest treatment design with the objective measure of change, …


Regression By Data Segments Via Discriminant Analysis, Stan Lipovetsky, Michael Conklin May 2005

Regression By Data Segments Via Discriminant Analysis, Stan Lipovetsky, Michael Conklin

Journal of Modern Applied Statistical Methods

It is known that two-group linear discriminant function can be constructed via binary regression. In this article, it is shown that the opposite relation is also relevant – it is possible to present multiple regression as a linear combination of a main part, based on the pooled variance, and Fisher discriminators by data segments. Presenting regression as an aggregate of the discriminators allows one to decompose coefficients of the model into sum of several vectors related to segments. Using this technique provides an understanding of how the total regression model is composed of the regressions by the segments with possible …


Inferences About Regression Interactions Via A Robust Smoother With An Application To Cannabis Problems, Rand R. Wilcox, Mitchell Earleywine May 2005

Inferences About Regression Interactions Via A Robust Smoother With An Application To Cannabis Problems, Rand R. Wilcox, Mitchell Earleywine

Journal of Modern Applied Statistical Methods

A flexible approach to testing the hypothesis of no regression interaction is to test the hypothesis that a generalized additive model provides a good fit to the data, where the components are some type of robust smoother. A practical concern, however, is that there are no published results on how well this approach controls the probability of a Type I error. Simulation results, reported here, indicate that an appropriate choice for the span of the smoother is required so that the actual probability of a Type I error is reasonably close to the nominal level. The technique is illustrated with …


Local Power For Combining Independent Tests In The Presence Of Nuisance Parameters For The Logistic Distribution, Walid A. Abu-Dayyeh, Z. R. Al-Rawi, M. M. A. Al-Momani May 2005

Local Power For Combining Independent Tests In The Presence Of Nuisance Parameters For The Logistic Distribution, Walid A. Abu-Dayyeh, Z. R. Al-Rawi, M. M. A. Al-Momani

Journal of Modern Applied Statistical Methods

Four combination methods of independent tests for testing a simple hypothesis versus one-sided alternative are considered viz. Fisher, the logistic, the sum of P-values and the inverse normal method in case of logistic distribution. These methods are compared via local power in the presence of nuisance parameters for some values of α using simple random sample.


A Comparison Of Parametric And Coarsened Bayesian Interval Estimation In The Presence Of A Known Mean-Variance Relationship, Kent Koprowicz, Scott S. Emerson, Peter Hoff Apr 2005

A Comparison Of Parametric And Coarsened Bayesian Interval Estimation In The Presence Of A Known Mean-Variance Relationship, Kent Koprowicz, Scott S. Emerson, Peter Hoff

UW Biostatistics Working Paper Series

While the use of Bayesian methods of analysis have become increasingly common, classical frequentist hypothesis testing still holds sway in medical research - especially clinical trials. One major difference between a standard frequentist approach and the most common Bayesian approaches is that even when a frequentist hypothesis test is derived from parametric models, the interpretation and operating characteristics of the test may be considered in a distribution-free manner. Bayesian inference, on the other hand, is often conducted in a parametric setting where the interpretation of the results is dependent on the parametric model. Here we consider a Bayesian counterpart to …


Causal Inference In Longitudinal Studies With History-Restricted Marginal Structural Models, Romain Neugebauer, Mark J. Van Der Laan, Ira B. Tager Apr 2005

Causal Inference In Longitudinal Studies With History-Restricted Marginal Structural Models, Romain Neugebauer, Mark J. Van Der Laan, Ira B. Tager

U.C. Berkeley Division of Biostatistics Working Paper Series

Causal Inference based on Marginal Structural Models (MSMs) is particularly attractive to subject-matter investigators because MSM parameters provide explicit representations of causal effects. We introduce History-Restricted Marginal Structural Models (HRMSMs) for longitudinal data for the purpose of defining causal parameters which may often be better suited for Public Health research. This new class of MSMs allows investigators to analyze the causal effect of a treatment on an outcome based on a fixed, shorter and user-specified history of exposure compared to MSMs. By default, the latter represents the treatment causal effect of interest based on a treatment history defined by the …


The Bayesian Two-Sample T-Test, Mithat Gonen, Wesley O. Johnson, Yonggang Lu, Peter H. Westfall Apr 2005

The Bayesian Two-Sample T-Test, Mithat Gonen, Wesley O. Johnson, Yonggang Lu, Peter H. Westfall

Memorial Sloan-Kettering Cancer Center, Dept. of Epidemiology & Biostatistics Working Paper Series

In this article we show how the pooled-variance two-sample t-statistic arises from a Bayesian formulation of the two-sided point null testing problem, with emphasis on teaching. We identify a reasonable and useful prior giving a closed-form Bayes factor that can be written in terms of the distribution of the two-sample t-statistic under the null and alternative hypotheses respectively. This provides a Bayesian motivation for the two-sample t-statistic, which has heretofore been buried as a special case of more complex linear models, or given only roughly via analytic or Monte Carlo approximations. The resulting formulation of the Bayesian test is easy …


Resampling Based Multiple Testing Procedure Controlling Tail Probability Of The Proportion Of False Positives, Mark J. Van Der Laan, Merrill D. Birkner, Alan E. Hubbard Mar 2005

Resampling Based Multiple Testing Procedure Controlling Tail Probability Of The Proportion Of False Positives, Mark J. Van Der Laan, Merrill D. Birkner, Alan E. Hubbard

U.C. Berkeley Division of Biostatistics Working Paper Series

Simultaneously testing a collection of null hypotheses about a data generating distribution based on a sample of independent and identically distributed observations is a fundamental and important statistical problem involving many applications. In this article we propose a new resampling based multiple testing procedure asymptotically controlling the probability that the proportion of false positives among the set of rejections exceeds q at level alpha, where q and alpha are user supplied numbers. The procedure involves 1) specifying a conditional distribution for a guessed set of true null hypotheses, given the data, which asymptotically is degenerate at the true set of …


Bayesian Evaluation Of Group Sequential Clinical Trial Designs, Scott S. Emerson, John M. Kittelson, Daniel L. Gillen Mar 2005

Bayesian Evaluation Of Group Sequential Clinical Trial Designs, Scott S. Emerson, John M. Kittelson, Daniel L. Gillen

UW Biostatistics Working Paper Series

Clincal trial designs often incorporate a sequential stopping rule to serve as a guide in the early termination of a study. When choosing a particular stopping rule, it is most common to examine frequentist operating characteristics such as type I error, statistical power, and precision of confi- dence intervals (Emerson, et al. [1]). Increasingly, however, clinical trials are designed and analyzed in the Bayesian paradigm. In this paper we describe how the Bayesian operating characteristics of a particular stopping rule might be evaluated and communicated to the scientific community. In particular, we consider a choice of probability models and a …


Implementation Of Estimating-Function Based Inference Procedures With Mcmc Sampler, Lu Tian, Jun S. Liu, L. J. Wei Feb 2005

Implementation Of Estimating-Function Based Inference Procedures With Mcmc Sampler, Lu Tian, Jun S. Liu, L. J. Wei

Harvard University Biostatistics Working Paper Series

No abstract provided.


Fixed-Width Output Analysis For Markov Chain Monte Carlo, Galin L. Jones, Murali Haran, Brian S. Caffo, Ronald Neath Feb 2005

Fixed-Width Output Analysis For Markov Chain Monte Carlo, Galin L. Jones, Murali Haran, Brian S. Caffo, Ronald Neath

Johns Hopkins University, Dept. of Biostatistics Working Papers

Markov chain Monte Carlo is a method of producing a correlated sample in order to estimate features of a complicated target distribution via simple ergodic averages. A fundamental question in MCMC applications is when should the sampling stop? That is, when are the ergodic averages good estimates of the desired quantities? We consider a method that stops the MCMC sampling the first time the width of a confidence interval based on the ergodic averages is less than a user-specified value. Hence calculating Monte Carlo standard errors is a critical step in assessing the output of the simulation. In particular, we …


Multiple Testing Procedures And Applications To Genomics, Merrill D. Birkner, Katherine S. Pollard, Mark J. Van Der Laan, Sandrine Dudoit Jan 2005

Multiple Testing Procedures And Applications To Genomics, Merrill D. Birkner, Katherine S. Pollard, Mark J. Van Der Laan, Sandrine Dudoit

U.C. Berkeley Division of Biostatistics Working Paper Series

This chapter proposes widely applicable resampling-based single-step and stepwise multiple testing procedures (MTP) for controlling a broad class of Type I error rates, in testing problems involving general data generating distributions (with arbitrary dependence structures among variables), null hypotheses, and test statistics (Dudoit and van der Laan, 2005; Dudoit et al., 2004a,b; van der Laan et al., 2004a,b; Pollard and van der Laan, 2004; Pollard et al., 2005). Procedures are provided to control Type I error rates defined as tail probabilities for arbitrary functions of the numbers of Type I errors, V_n, and rejected hypotheses, R_n. These error rates include: …


Robust Inferences For Covariate Effects On Survival Time With Censored Linear Regression Models, Larry Leon, Tianxi Cai, L. J. Wei Jan 2005

Robust Inferences For Covariate Effects On Survival Time With Censored Linear Regression Models, Larry Leon, Tianxi Cai, L. J. Wei

Harvard University Biostatistics Working Paper Series

Various inference procedures for linear regression models with censored failure times have been studied extensively. Recent developments on efficient algorithms to implement these procedures enhance the practical usage of such models in survival analysis. In this article, we present robust inferences for certain covariate effects on the failure time in the presence of "nuisance" confounders under a semiparametric, partial linear regression setting. Specifically, the estimation procedures for the regression coefficients of interest are derived from a working linear model and are valid even when the function of the confounders in the model is not correctly specified. The new proposals are …


Multiple Testing Procedures For Controlling Tail Probability Error Rates, Sandrine Dudoit, Mark J. Van Der Laan, Merrill D. Birkner Dec 2004

Multiple Testing Procedures For Controlling Tail Probability Error Rates, Sandrine Dudoit, Mark J. Van Der Laan, Merrill D. Birkner

U.C. Berkeley Division of Biostatistics Working Paper Series

The present article discusses and compares multiple testing procedures (MTP) for controlling Type I error rates defined as tail probabilities for the number (gFWER) and proportion (TPPFP) of false positives among the rejected hypotheses. Specifically, we consider the gFWER- and TPPFP-controlling MTPs proposed recently by Lehmann & Romano (2004) and in a series of four articles by Dudoit et al. (2004), van der Laan et al. (2004b,a), and Pollard & van der Laan (2004). The former Lehmann & Romano (2004) procedures are marginal, in the sense that they are based solely on the marginal distributions of the test statistics, i.e., …


Multiple Testing Procedures: R Multtest Package And Applications To Genomics, Katherine S. Pollard, Sandrine Dudoit, Mark J. Van Der Laan Dec 2004

Multiple Testing Procedures: R Multtest Package And Applications To Genomics, Katherine S. Pollard, Sandrine Dudoit, Mark J. Van Der Laan

U.C. Berkeley Division of Biostatistics Working Paper Series

The Bioconductor R package multtest implements widely applicable resampling-based single-step and stepwise multiple testing procedures (MTP) for controlling a broad class of Type I error rates, in testing problems involving general data generating distributions (with arbitrary dependence structures among variables), null hypotheses, and test statistics. The current version of multtest provides MTPs for tests concerning means, differences in means, and regression parameters in linear and Cox proportional hazards models. Procedures are provided to control Type I error rates defined as tail probabilities for arbitrary functions of the numbers of false positives and rejected hypotheses. These error rates include tail probabilities …


Semiparametric Regression In Capture-Recapture Modelling, O. Gimenez, C. Barbraud, Ciprian M. Crainiceanu, S. Jenouvrier, B.T. Morgan Dec 2004

Semiparametric Regression In Capture-Recapture Modelling, O. Gimenez, C. Barbraud, Ciprian M. Crainiceanu, S. Jenouvrier, B.T. Morgan

Johns Hopkins University, Dept. of Biostatistics Working Papers

Capture-recapture models were developed to estimate survival using data arising from marking and monitoring wild animals over time. Variation in the survival process may be explained by incorporating relevant covariates. We develop nonparametric and semiparametric regression models for estimating survival in capture-recapture models. A fully Bayesian approach using MCMC simulations was employed to estimate the model parameters. The work is illustrated by a study of Snow petrels, in which survival probabilities are expressed as nonlinear functions of a climate covariate, using data from a 40-year study on marked individuals, nesting at Petrels Island, Terre Adelie.


Semi-Parametric Single-Index Two-Part Regression Models, Xiao-Hua Zhou, Hua Liang Dec 2004

Semi-Parametric Single-Index Two-Part Regression Models, Xiao-Hua Zhou, Hua Liang

UW Biostatistics Working Paper Series

In this paper, we proposed a semi-parametric single-index two-part regression model to weaken assumptions in parametric regression methods that were frequently used in the analysis of skewed data with additional zero values. The estimation procedure for the parameters of interest in the model was easily implemented. The proposed estimators were shown to be consistent and asymptotically normal. Through a simulation study, we showed that the proposed estimators have reasonable finite-sample performance. We illustrated the application of the proposed method in one real study on the analysis of health care costs.


A Bayesian Mixture Model Relating Dose To Critical Organs And Functional Complication In 3d Conformal Radiation Therapy, Tim Johnson, Jeremy Taylor, Randall K. Ten Haken, Avraham Eisbruch Nov 2004

A Bayesian Mixture Model Relating Dose To Critical Organs And Functional Complication In 3d Conformal Radiation Therapy, Tim Johnson, Jeremy Taylor, Randall K. Ten Haken, Avraham Eisbruch

The University of Michigan Department of Biostatistics Working Paper Series

A goal of radiation therapy is to deliver maximum dose to the target tumor while minimizing complications due to irradiation of critical organs. Technological advances in 3D conformal radiation therapy has allowed great strides in realizing this goal, however complications may still arise. Critical organs may be adjacent to tumors or in the path of the radiation beam. Several mathematical models have been proposed that describe a relationship between dose and observed functional complication, however only a few published studies have successfully fit these models to data using modern statistical methods which make efficient use of the data. One complication …


Choice Of Monitoring Mechanism For Optimal Nonparametric Functional Estimation For Binary Data, Nicholas P. Jewell, Mark J. Van Der Laan, Stephen Shiboski Nov 2004

Choice Of Monitoring Mechanism For Optimal Nonparametric Functional Estimation For Binary Data, Nicholas P. Jewell, Mark J. Van Der Laan, Stephen Shiboski

U.C. Berkeley Division of Biostatistics Working Paper Series

Optimal designs of dose levels in order to estimate parameters from a model for binary response data have a long and rich history. These designs are based on parametric models. Here we consider fully nonparametric models with interest focused on estimation of smooth functionals using plug-in estimators based on the nonparametric maximum likelihood estimator. An important application of the results is the derivation of the optimal choice of the monitoring time distribution function for current status observation of a survival distribution. The optimal choice depends in a simple way on the dose response function and the form of the functional. …