Open Access. Powered by Scholars. Published by Universities.®

Statistics and Probability Commons™

Open Access. Powered by Scholars. Published by Universities.®

2014

Discipline
Institution
Keyword
Publication
Publication Type
File Type

Articles 271 - 300 of 546

Full-Text Articles in Statistics and Probability

Inference For The Rayleigh Distribution Based On Progressive Type-Ii Fuzzy Censored Data, Abbas Pak, Gholam Ali Parham, Mansour Saraj May 2014

Inference For The Rayleigh Distribution Based On Progressive Type-Ii Fuzzy Censored Data, Abbas Pak, Gholam Ali Parham, Mansour Saraj

Journal of Modern Applied Statistical Methods

Classical statistical analysis of the Rayleigh distribution deals with precise information. However, in real world situations, experimental performance results cannot always be recorded or measured precisely, but each observable event may only be identified with a fuzzy subset of the sample space. Therefore, the conventional procedures used for estimating the Rayleigh distribution parameter will need to be adapted to the new situation. This article discusses different estimation methods for the parameters of the Rayleigh distribution on the basis of a progressively type-II censoring scheme when the available observations are described by means of fuzzy information. They include the maximum likelihood …


Distance Correlation Coefficient: An Application With Bayesian Approach In Clinical Data Analysis, Atanu Bhattacharjee May 2014

Distance Correlation Coefficient: An Application With Bayesian Approach In Clinical Data Analysis, Atanu Bhattacharjee

Journal of Modern Applied Statistical Methods

The distance correlation coefficient – based on the product-moment approach – is one method by which to explore the relationship between variables. The Bayesian approach is a powerful tool to determine statistical inferences with credible intervals. Prior information about the relationship between BP and Serum cholesterol was applied to formulate the distance correlation between the two variables. The conjugate prior is considered to formulate the posterior estimates of the distance correlations. The illustrated method is simple and is suitable for other experimental studies.


Oscillation Theorems For Fourth-Order Half-Linear Delay Dynamic Equations With Damping, Ravi P. Agarwal, Martin Bohner, Tongxing Li, Chenghui Zhang May 2014

Oscillation Theorems For Fourth-Order Half-Linear Delay Dynamic Equations With Damping, Ravi P. Agarwal, Martin Bohner, Tongxing Li, Chenghui Zhang

Mathematics and Statistics Faculty Research & Creative Works

This article is concerned with oscillatory behavior of a class of fourth-order half-linear delay dynamic equations with damping on a time scale. Some new oscillation criteria are established. © 2013 Springer Basel.


A Comparison Of Prenatal Alcohol, Tobacco, And Other Drug Use Between San Luis Obispo County And Ventura County, Dana M. Williamson May 2014

A Comparison Of Prenatal Alcohol, Tobacco, And Other Drug Use Between San Luis Obispo County And Ventura County, Dana M. Williamson

Statistics

Prenatal substance abuse is a growing issue in America. It can lead to fetal alcohol spectrum disorder, long term growth, behavior, and executive functioning problems, and creates a predisposition for drug use for the child.

This project summarizes the statistical analyses comparing alcohol, tobacco, and other drug use by pregnant women between San Luis Obispo County and Ventura County. The main goal of these analyses is to determine if there is a difference between San Luis Obispo County and Ventura County. This is an interesting comparison because these counties are neighboring counties, and past data have shown that the rate …


Some New Probability Distributions Based On Random Extrema And Permutation Patterns, Jie Hao May 2014

Some New Probability Distributions Based On Random Extrema And Permutation Patterns, Jie Hao

Electronic Theses and Dissertations

In this paper, we study a new family of random variables, that arise as the distribution of extrema of a random number N of independent and identically distributed random variables X1,X2, ..., XN, where each Xi has a common continuous distribution with support on [0,1]. The general scheme is first outlined, and SUG and CSUG models are introduced in detail where Xi is distributed as U[0,1]. Some features of the proposed distributions can be studied via its mean, variance, moments and moment-generating function. Moreover, we make some other choices for …


Family-Wise Error Rate Control In Quantitative Trait Loci (Qtl) Mapping And Gene Ontology Graphs With Remarks On Family Selection, Garrett Saunders May 2014

Family-Wise Error Rate Control In Quantitative Trait Loci (Qtl) Mapping And Gene Ontology Graphs With Remarks On Family Selection, Garrett Saunders

All Graduate Theses and Dissertations, Spring 1920 to Summer 2023

One of the great aims of statistics, the science of collecting, analyzing, and interpreting data, is to protect against the probability of falsely rejecting an accepted claim, or hypothesis, given observed data stemming from some experiment. This is generally known as protecting against a Type I Error, or controlling the Type I Error rate. The extension of this protection against Type I Errors to the situation where thousands upon thousands of hypotheses are examined simultaneously is known as multiple hypothesis testing. This dissertation presents an improvement to an existing multiple hypothesis testing approach, the Focus Level method, specific to gene …


The Financial Crisis Was Good For Something: Improved Nonprofit Efficiency, Caitlin Paige Britt May 2014

The Financial Crisis Was Good For Something: Improved Nonprofit Efficiency, Caitlin Paige Britt

Finance Undergraduate Honors Theses

This study explores the need for financial performance measures in the nonprofit sector and the impact the 2008-2009 Financial Crisis had upon nonprofits’ efficiency. This analysis tests the hypothesis that the financial crisis actually improved nonprofit efficiency by forcing nonprofits to eliminate unnecessary costs, continue to produce their services, thus improving operational efficiency, despite decreased donor contributions and increased user need. Entries reported on nonprofits’ IRS 990 forms from 2003-2010 determined whether nonprofit efficiency was significantly different after the crisis. The efficiencies used to measure the impact of the Financial Crisis include: Program Expense Efficiency, Administrative Expense Efficiency, Fundraising Expense …


Mvgst: Tools For Multivariate And Directional Gene Set Testing, Dennis S. Mecham May 2014

Mvgst: Tools For Multivariate And Directional Gene Set Testing, Dennis S. Mecham

All Graduate Plan B and other Reports, Spring 1920 to Spring 2023

There are many platforms available for simultaneously measuring the relative activity, or expression, levels of all genes in an organism. Genes that have systematically dier- ent expression levels between experimental factor levels are called \dierentially expressed". Because genes are annotated based on their known roles in biological processes (BP), molec- ular functions (MF) and cellular components (CC), gene expression levels can be used to determine relative activity levels of individual BP, MF, or CC between experimental fac- tor levels (this is called gene set testing). Often multiple experimental dierences are of interest simultaneously, which necessitates multivariate gene set testing. Only …


Median Based Modified Ratio Estimators With Known Quartiles Of An Auxiliary Variable, Jambulingam Subramani, G Prabavathy May 2014

Median Based Modified Ratio Estimators With Known Quartiles Of An Auxiliary Variable, Jambulingam Subramani, G Prabavathy

Journal of Modern Applied Statistical Methods

New median based modified ratio estimators for estimating a finite population mean using quartiles and functions of an auxiliary variable are proposed. The bias and mean squared error of the proposed estimators are obtained and the mean squared error of the proposed estimators are compared with the usual simple random sampling without replacement (SRSWOR) sample mean, ratio estimator, a few existing modified ratio estimators, the linear regression estimator and median based ratio estimator for certain natural populations. A numerical study shows that the proposed estimators perform better than existing estimators; in addition, it is shown that the proposed median based …


On The Exponentiated Weibull Distribution For Modeling Wind Speed In South Western Nigeria, Olanrewaju I. Shittu, K A. Adepoju May 2014

On The Exponentiated Weibull Distribution For Modeling Wind Speed In South Western Nigeria, Olanrewaju I. Shittu, K A. Adepoju

Journal of Modern Applied Statistical Methods

One of the bases for assessment of wind energy potential for a specified region is the probability distribution of wind speed. Thus, appropriate and adequate specification of the probability distribution of wind speed becomes increasingly important. Several distributions have been proposed for describing wind distribution. Among the most popular distributions is the Weibull whose choice is due to its flexibility. An exponentiated Weibull distribution is proposed as an alternative to model wind speed data with a view to comparing it with the existing Weibull distribution. Results indicate that the proposed distribution outperforms the existing Weibull distribution for modeling wind speed …


Robust Regression Analysis For Non-Normal Situations Under Symmetric Distributions Arising In Medical Research, S S. Ganguly May 2014

Robust Regression Analysis For Non-Normal Situations Under Symmetric Distributions Arising In Medical Research, S S. Ganguly

Journal of Modern Applied Statistical Methods

In medical research, while carrying out regression analysis, it is usually assumed that the independent (covariates) and dependent (response) variables follow a multivariate normal distribution. In some situations, the covariates may not have normal distribution and instead may have some symmetric distribution. In such a situation, the estimation of the regression parameters using Tiku’s Modified Maximum Likelihood (MML) method may be more appropriate. The method of estimating the parameters is discussed and the applications of the method are illustrated using real sets of data from the field of public health.


Jmasm 33: A Two Dependent Samples Maximum Test Calculator: Excel, Saverpierre Maggio, Shlomo Sawilowsky May 2014

Jmasm 33: A Two Dependent Samples Maximum Test Calculator: Excel, Saverpierre Maggio, Shlomo Sawilowsky

Journal of Modern Applied Statistical Methods

An Excel Macro was created to provide researchers with an easy to use resource in order to calculate the two dependent samples maximum test as provided in Maggio and Sawilowsky (2014), which permits conducting both the two dependent samples t-test and Wilcoxon signed-ranks test on the same data while eliminating concerns related to Type I error inflation and choice of statistical tests.


Vol. 13, No. 1 (Full Issue), Jmasm Editors May 2014

Vol. 13, No. 1 (Full Issue), Jmasm Editors

Journal of Modern Applied Statistical Methods

No abstract provided.


Relative Importance Of Predictors In Multilevel Modeling, Yan Liu, Bruno D. Zumbo, Amery D. Wu May 2014

Relative Importance Of Predictors In Multilevel Modeling, Yan Liu, Bruno D. Zumbo, Amery D. Wu

Journal of Modern Applied Statistical Methods

The Pratt index is a useful and practical strategy for day-to-day researchers when ordering predictors in a multiple regression analysis. The purposes of this study are to introduce and demonstrate the use of the Pratt index to assess the relative importance of predictors for a random intercept multilevel model.


Predicting Survival Time Of Localized Melanoma Patients Using Discrete Survival Time Method, Taysseer Sharaf, Chris P. Tsokos May 2014

Predicting Survival Time Of Localized Melanoma Patients Using Discrete Survival Time Method, Taysseer Sharaf, Chris P. Tsokos

Journal of Modern Applied Statistical Methods

Melanoma is the most fatal type of skin cancer. It is ranked first in death of skin cancer diseases. This study establishes a statistical model that can predict the survival time of localized melanoma patients, as a function of age at diagnosis, tumor thickness, and extension of the tumor (tumor invasion). The discrete time survival method was used to build the statistical model. The patients involved in the current study were observed from the SEER database. Patients were divided into nine groups according to age at diagnosis. Variation in survival time was found to be significant among some of the …


Population Mean Estimation With Sub Sampling The Non-Respondents Using Two Phase Sampling, Sunil Kumar, M Viswanathaiah May 2014

Population Mean Estimation With Sub Sampling The Non-Respondents Using Two Phase Sampling, Sunil Kumar, M Viswanathaiah

Journal of Modern Applied Statistical Methods

The problem of non-response in double (or two phase) sampling is dealt with combined ratio, product and regression estimators. Expressions of bias and MSE for these estimators are obtained. Comparisons of a proposed strategy with a usual unbiased estimator and other estimators are carried out and results obtained are illustrated numerically using an empirical sample.


Estimation And Testing In Type-Ii Generalized Half Logistic Distribution, R R. L. Kantam, V Ramakrishna, M S. Ravikumar May 2014

Estimation And Testing In Type-Ii Generalized Half Logistic Distribution, R R. L. Kantam, V Ramakrishna, M S. Ravikumar

Journal of Modern Applied Statistical Methods

A generalization of the Half Logistic Distribution is developed through exponentiation of its survival function and named the Type II Generalized Half Logistic Distribution (GHLD). The distributional characteristics are presented and estimation of its parameters using maximum likelihood and modified maximum likelihood methods is studied with comparisons. Discrimination between Type II GHLD and exponential distribution in pairs is conducted via likelihood ratio criterion.


A Compound Of Geeta Distribution With Generalized Beta Distribution, Adil Rashid, T R. Jan May 2014

A Compound Of Geeta Distribution With Generalized Beta Distribution, Adil Rashid, T R. Jan

Journal of Modern Applied Statistical Methods

A compound of Geeta distribution with Generalized Beta distribution (GBD) is obtained and the compound is specialized for different values of β. The first order factorial moments of some special compound distributions are also obtained. A chronological overview of recent developments in the compounding of distributions is provided in the introduction.


Hierarchical Clustering With Simple Matching And Joint Entropy Dissimilarity Measure, A Mete ÇilingtüRk, ÖZlem ErgüT May 2014

Hierarchical Clustering With Simple Matching And Joint Entropy Dissimilarity Measure, A Mete ÇilingtüRk, ÖZlem ErgüT

Journal of Modern Applied Statistical Methods

Conventional clustering algorithms are restricted for use with data containing ratio or interval scale variables; hence, distances are used. As social studies require merely categorical data, the literature is enriched with more complicated clustering techniques and algorithms of categorical data. These techniques are based on similarity or dissimilarity matrices. The algorithms are using density based or pattern based approaches. A probabilistic nature to similarity structure is proposed. The entropy dissimilarity measure has comparable results with simple matching dissimilarity at hierarchical clustering. It overcomes dimension increase through binarization of the categorical data. This approach is also functional with the clustering methods, …


An Exploratory Graphical Method For Identifying Associations In R X C Contingency Tables, Martin L. Lesser, Meredith B. Akerman May 2014

An Exploratory Graphical Method For Identifying Associations In R X C Contingency Tables, Martin L. Lesser, Meredith B. Akerman

Journal of Modern Applied Statistical Methods

On finding a significant association between rows and columns of an r x c contingency table, the next step is to study the nature of the association in more detail. The use of a scree plot to visualize the largest contributions to Χ2 among all cells in the table in order to determine the nature of the association in more detail is proposed.


Separate Ratio-Type Estimators Of Population Mean In Stratified Random Sampling, Rajesh Tailor, Hilal A. Lone May 2014

Separate Ratio-Type Estimators Of Population Mean In Stratified Random Sampling, Rajesh Tailor, Hilal A. Lone

Journal of Modern Applied Statistical Methods

Separate ratio-type estimators for population mean with their properties are considered. Some separate ratio-type estimators for population mean using known parameters of auxiliary variate are proposed. The bias and mean squared error of the proposed estimators are obtained up to the first degree of approximation. It is shown that the proposed estimators are more efficient than unbiased estimators in stratified random sampling and usual separate ratio estimators under certain obtained conditions. To judge the merits of the proposed estimators, an empirical study was conducted.


Evaluation Of Area Under The Constant Shape Bi-Weibull Roc Curve, Sudesh Pundir, R Amala May 2014

Evaluation Of Area Under The Constant Shape Bi-Weibull Roc Curve, Sudesh Pundir, R Amala

Journal of Modern Applied Statistical Methods

The Receiver Operating Characteristic (ROC) curve generated based on assuming a constant shape Bi-Weibull distribution is studied. In the context of ROC curve analysis, it is assumed that biomarker values from controls and cases follow some specific distribution and the accuracy is evaluated by using the ROC model developed from that specified distribution. This article assumes that the biomarker values from the two groups follow Weibull distributions with equal shape parameter and different scale parameters. The ROC model, area under the ROC curve (AUC), asymptotic and bootstrap confidence intervals for the AUC are derived. Theoretical results are validated by simulation …


Investigating The Feasibility Of Using Mplus In The Estimation Of Growth Mixture Models, Ming Li, Jeffrey R. Harring, George B. Macready May 2014

Investigating The Feasibility Of Using Mplus In The Estimation Of Growth Mixture Models, Ming Li, Jeffrey R. Harring, George B. Macready

Journal of Modern Applied Statistical Methods

Hipp and Bauer (2006) investigated the issues of singularities and local maximum solutions within growth mixture models (GMMs) and made recommendations regarding the use of multiple starting values. Building on their work, this simulation study investigates the feasibility of estimating GMMs within Mplus as measured by convergence to proper, but local solutions.


Poisson Distributed Individuals Control Charts With Optimal Limits, Negin Enayaty Ahangar May 2014

Poisson Distributed Individuals Control Charts With Optimal Limits, Negin Enayaty Ahangar

Graduate Theses and Dissertations

The conventional method used in attribute control charts is the Shewhart three sigma limits. The implicit assumption of the Normal distribution in this approach is not appropriate for skewed distributions such as Poisson, Geometric and Negative Binomial. Normal approximations perform poorly in the tail area of the these distributions. In this research, a type of attribute control chart is introduced to monitor the processes that provide count data. The economic objective of this chart is to minimize the cost of its errors which is determined by the designer. This objective is a linear function of type I and II errors. …


A Reduced Bias Method Of Estimating Variance Components In Generalized Linear Mixed Models, Elizabeth A. Claassen May 2014

A Reduced Bias Method Of Estimating Variance Components In Generalized Linear Mixed Models, Elizabeth A. Claassen

Department of Statistics: Dissertations, Theses, and Student Research

In small samples it is well known that the standard methods for estimating variance components in a generalized linear mixed model (GLMM), pseudo-likelihood and maximum likelihood, yield estimates that are biased downward. An important consequence of this is that inferences on fixed effects will have inflated Type I error rates because their precision is overstated. We introduce a new method for estimating parameters in GLMMs that applies a Firth bias adjustment to the maximum likelihood-based GLMM estimating algorithm. We apply this technique to one- and two-treatment logistic regression models with a single random effect. We show simulation results that demonstrate …


High Frequency Data: Modeling Durations Via The Acd And Log Acd Models, Lilian Cheung May 2014

High Frequency Data: Modeling Durations Via The Acd And Log Acd Models, Lilian Cheung

Honors Scholar Theses

This thesis proposes a method of finding initial parameter estimates in the Log ACD1 model for use in recursive estimation. The recursive estimating equations method is applied to the Log ACD1 model to find recursive estimates for the unknown parameters in the model. A literature review is provided on the ACD and Log ACD models, and on the theory of estimating equations. Monte Carlo simulations indicate that the proposed method of finding initial parameter estimates is viable. The parameter estimation process is demonstrated by fitting an ACD model and a Log ACD model to a set of IBM …


Comparison Of Different Methods For Estimating Log-Normal Means, Qi Tang May 2014

Comparison Of Different Methods For Estimating Log-Normal Means, Qi Tang

Electronic Theses and Dissertations

The log-normal distribution is a popular model in many areas, especially in biostatistics and survival analysis where the data tend to be right skewed. In our research, a total of ten different estimators of log-normal means are compared theoretically. Simulations are done using different values of parameters and sample size. As a result of comparison, ``A degree of freedom adjusted" maximum likelihood estimator and Bayesian estimator under quadratic loss are the best when using the mean square error (MSE) as a criterion. The ten estimators are applied to a real dataset, an environmental study from Naval Construction Battalion Center (NCBC), …


Are Highly Dispersed Variables More Extreme? The Case Of Distributions With Compact Support, Benedict E. Adjogah May 2014

Are Highly Dispersed Variables More Extreme? The Case Of Distributions With Compact Support, Benedict E. Adjogah

Electronic Theses and Dissertations

We consider discrete and continuous symmetric random variables X taking values in [0; 1], and thus having expected value 1/2. The main thrust of this investigation is to study the correlation between the variance, Var(X) of X and the value of the expected maximum E(Mn) = E(X1,...,Xn) of n independent and identically distributed random variables X1,X2,...,Xn, each distributed as X. Many special cases are studied, some leading to very interesting alternating sums, and some progress is made towards a general theory.


Discovering Predictors Of Readmission For Acute Myocardial Infarction In A Medicare Population: A Data Mining Approach, Daniel Macdonald Knowles Ms May 2014

Discovering Predictors Of Readmission For Acute Myocardial Infarction In A Medicare Population: A Data Mining Approach, Daniel Macdonald Knowles Ms

All Student Scholarship

Health care costs in the United States have risen at rates far exceeding the cost of living for many years. Previous attempts to control these costs have proven futile. Studies have shown that high per-capita spending in the U.S. does not equate to consistent quality of care or better outcomes.


Modeling Asset Volatility Using Various Resources, Isaac G. Blackhurst May 2014

Modeling Asset Volatility Using Various Resources, Isaac G. Blackhurst

All Graduate Plan B and other Reports, Spring 1920 to Spring 2023

Volatility is of central interest in modern financial econometrics. This thesis evaluates three different methods of measuring volatility.