Open Access. Powered by Scholars. Published by Universities.®

Statistics and Probability Commons™

Open Access. Powered by Scholars. Published by Universities.®

12,804 Full-Text Articles 23,873 Authors 9,922,835 Downloads 282 Institutions

All Articles in Statistics and Probability

Faceted Search

12,804 full-text articles. Page 437 of 486.

The Interacting Multiple Models Algorithm With State-Dependent Value Assignment, Rastin Rastgoufard 2012 University of New Orleans

The Interacting Multiple Models Algorithm With State-Dependent Value Assignment, Rastin Rastgoufard

LSU New Orleans Theses and Dissertations

The value of a state is a measure of its worth, so that, for example, waypoints have high value and regions inside of obstacles have very small value. We propose two methods of incorporating world information as state-dependent modifications to the interacting multiple models (IMM) algorithm, and then we use a game's player-controlled trajectories as ground truths to compare the normal IMM algorithm to versions with our proposed modifications. The two methods involve modifying the model probabilities in the update step and modifying the transition probability matrix in the mixing step based on the assigned values of different target states. …


Support Vector Machines For Classification And Imputation, Spencer David Rogers 2012 Brigham Young University - Provo

Support Vector Machines For Classification And Imputation, Spencer David Rogers

Theses and Dissertations

Support vector machines (SVMs) are a powerful tool for classification problems. SVMs have only been developed in the last 20 years with the availability of cheap and abundant computing power. SVMs are a non-statistical approach and make no assumptions about the distribution of the data. Here support vector machines are applied to a classic data set from the machine learning literature and the out-of-sample misclassification rates are compared to other classification methods. Finally, an algorithm for using support vector machines to address the difficulty in imputing missing categorical data is proposed and its performance is demonstrated under three different scenarios …


Families Or Unrelated: The Evolving Debate In Genetic Association Studies, David W. Fardo, Richard Charnigo, Michael P. Epstein 2012 University of Kentucky

Families Or Unrelated: The Evolving Debate In Genetic Association Studies, David W. Fardo, Richard Charnigo, Michael P. Epstein

Biostatistics Faculty Publications

To help uncover the genetic determinants of complex disease, a scientist often designs an association study using either unrelated subjects or family members within pedigrees. But which of these two subject recruitment paradigms is preferable? This editorial addresses the debate over the relative merits of family- and population-based genetic association studies. We begin by briefly recounting the evolution of genetic epidemiology and the rich crossroads of statistics and genetics. We then detail the arguments for the two aforementioned paradigms in recent and current applications. Finally, we speculate on how the debate may progress with the emergence of next-generation sequencing technologies.


Bayesian And Related Methods: Techniques Based On Bayes' Theorem, Mehmet Vurkaç 2012 Portland State University

Bayesian And Related Methods: Techniques Based On Bayes' Theorem, Mehmet Vurkaç

Systems Science Friday Noon Seminar Series

Bayes' theorem is a simple algebraic consequence of conditional probability. Yet, its consequences are critical to philosophy, society, and technology. Starting from its simple derivation, we will show how its interpretation in terms of base rates (priors) and class-conditional likelihoods illuminates everyday problems in medicine and law, and provides signal processing, communications, machine learning, model selection, and other applications of statistics with powerful classification and estimation tools. Next, we will briefly examine some of the ways in which this theorem can be adopted to include multiple attributes, contexts, hypotheses, and levels of risk. Methods derived from or related to Bayes’ …


Dna Methylation Arrays As Surrogate Measures Of Cell Mixture Distribution, Eugene Houseman, William P. Accomando, Devin C. Koestler, Brock C. Christensen, Carmen J. Marsit 2012 Oregon State University

Dna Methylation Arrays As Surrogate Measures Of Cell Mixture Distribution, Eugene Houseman, William P. Accomando, Devin C. Koestler, Brock C. Christensen, Carmen J. Marsit

Dartmouth Scholarship

There has been a long-standing need in biomedical research for a method that quantifies the normally mixed composition of leukocytes beyond what is possible by simple histological or flow cytometric assessments. The latter is restricted by the labile nature of protein epitopes, requirements for cell processing, and timely cell analysis. In a diverse array of diseases and following numerous immune-toxic exposures, leukocyte composition will critically inform the underlying immuno-biology to most chronic medical conditions. Emerging research demonstrates that DNA methylation is responsible for cellular differentiation, and when measured in whole peripheral blood, serves to distinguish cancer cases from controls.


A Methodology For The Analysis Of Fly Activity Data., Ruoying Wang 2012 East Tennessee State University

A Methodology For The Analysis Of Fly Activity Data., Ruoying Wang

Undergraduate Honors Theses

Experiments to learn about the effect of light, sex, and diet on the activity of flies generate great quantities of data that is necessary to analyze. Since different researches and students participate in the analysis of those experiments, it is convenient to have a methodology to analyze the experimental data using software so that the data can be analyzed in a uniform way. Being a double major in mathematics and biology, I am interested in:

  • Deciding which statistical procedure to use to analyze the data so that the research questions of the researchers in biology are answered.
  • To recommend how …


Modeling Antibiotic Resistance When Adding A New Antibiotic To A Hospital Setting., Brandi N. Canter 2012 East Tennessee State University

Modeling Antibiotic Resistance When Adding A New Antibiotic To A Hospital Setting., Brandi N. Canter

Undergraduate Honors Theses

As of now, not many pharmaceutical companies are producing new categories of antibiotics to fight bacterial infections. Therefore, bacteria are building up a resistance to the medications commonly used. Often, antibiotic resistance begins within a hospital. To combat resistance, researchers completed several studies using cycling of the medications that are already in place, but they found either no improvement or the resistance increased with this type of setting. In addition, although preventative infection control measures have been shown to decrease antibiotic resistance for some antibiotics, the level of antibiotic resistance found in hospitals is still extremely high. This motivates the …


Confidence Intervals For The Selected Population In Randomized Trials That Adapt The Population Enrolled, Michael Rosenblum 2012 Johns Hopkins Bloomberg School of Public Health, Department of Biostatistics

Confidence Intervals For The Selected Population In Randomized Trials That Adapt The Population Enrolled, Michael Rosenblum

Johns Hopkins University, Dept. of Biostatistics Working Papers

It is a challenge to design randomized trials when it is suspected that a treatment may benefit only certain subsets of the target population. In such situations, trial designs have been proposed that modify the population enrolled based on an interim analysis, in a preplanned manner. For example, if there is early evidence that the treatment only benefits a certain subset of the population, enrollment may then be restricted to this subset. At the end of such a trial, it is desirable to draw inferences about the selected population. We focus on constructing confidence intervals for the average treatment effect …


Why Odds Ratio Estimates Of Gwas Are Almost Always Close To 1.0, Yutaka Yasui 2012 University of Alberta

Why Odds Ratio Estimates Of Gwas Are Almost Always Close To 1.0, Yutaka Yasui

COBRA Preprint Series

“Missing heritability” in genome-wide association studies (GWAS) refers to the seeming inability for GWAS data to capture the great majority of genetic causes of a disease in comparison to the known degree of heritability for the disease, in spite of GWAS’ genome-wide measures of genetic variations. This paper presents a simple mathematical explanation for this phenomenon, assuming that the heritability information exists in GWAS data. Specifically, it focuses on the fact that the great majority of association measures (in the form of odds ratios) from GWAS are consistently close to the value that indicates no association, explains why this occurs, …


Analyzing Multiple Independent Spatial Point Processes, Neal Grantham 2012 California Polytechnic State University, San Luis Obispo

Analyzing Multiple Independent Spatial Point Processes, Neal Grantham

Statistics

No abstract provided.


A Study Of Women Working In The Actuarial Field, Jillian Emberg 2012 Actuarial Mathematics Concentration

A Study Of Women Working In The Actuarial Field, Jillian Emberg

Honors Projects in Mathematics

The goal of this project is to examine how women fit into the actuarial career path and how cultural expectations, biological factors, and personal aspirations affect their experiences in the field. Dramatic changes in the profession have occurred since its emergence in the nineteenth century to become more welcoming to women who choose to enter the profession. However, despite the equalizing demographic shifts of the field, it is still a male-dominated profession. This paper attempts to analyze why some of the changes in the demographics of the field have occurred as well as explain what factors contribute to women’s underrepresentation …


The Impact Of Violating Factor Scaling Method Assumptions On Latent Mean Difference Testing In Structured Means Models, Dandan Wang, Tiffany A. Whittaker, S. Natasha Beretvas 2012 The University of Texas at Austin

The Impact Of Violating Factor Scaling Method Assumptions On Latent Mean Difference Testing In Structured Means Models, Dandan Wang, Tiffany A. Whittaker, S. Natasha Beretvas

Journal of Modern Applied Statistical Methods

Type I error rates and power of the likelihood ratio test and bias of the standardized effect size measure associated with the latent mean difference in structured means modeling are examined when violating the assumptions underlying the two available factor scaling methods under various conditions. Implications and recommendations are discussed.


New Approximate Bayesian Confidence Intervals For The Coefficient Of Variation Of A Gaussian Distribution, Vincent A. R. Camara 2012 Research Center for Bayesian Applications, Inc., Largo, FL

New Approximate Bayesian Confidence Intervals For The Coefficient Of Variation Of A Gaussian Distribution, Vincent A. R. Camara

Journal of Modern Applied Statistical Methods

Confidence intervals are constructed for the coefficient of variation of a Gaussian distribution. Considering the square error and the Higgins-Tsokos loss functions, approximate Bayesian models are derived and compared to a published classical model. The models are shown to have great coverage accuracy. The classical model does not always yield the best confidence intervals; the proposed models often perform better.


A Poisson Regression Model For Female Radium Dial Workers, Tze-San Lee 2012 Western Illinois University

A Poisson Regression Model For Female Radium Dial Workers, Tze-San Lee

Journal of Modern Applied Statistical Methods

A Poisson regression model with interaction terms was applied to study the dose response relationship for radium-induced skeletal cancers. The model showed that the expected frequency count of bone tumors depended not only on the logarithmic dose and the time since first exposure, but also on the interaction between the logarithmic dose and the time since first exposure, whereas the dose-response model for head tumors depended only on the logarithmic dose.


Jmasm 32: Sas Template For Single-Subject Experimental Designs, Hyewon Chung, Jiseon Kim, Ryoungsun Park 2012 Chungnam National University, Deajeon, Korea

Jmasm 32: Sas Template For Single-Subject Experimental Designs, Hyewon Chung, Jiseon Kim, Ryoungsun Park

Journal of Modern Applied Statistical Methods

Meta-analysis has been used to synthesize research findings and to evaluate the effectiveness of treatments or the accuracy of diagnostic tools. Although meta-analytic techniques were developed to synthesize the results of several studies, controversy exists as to how to quantify the results from singlesubject experimental designs (SSEDs). The most commonly used metrics are reviewed, including nonregression and regression based methods. The application of the SAS template is demonstrated through simulated data sets. The SAS templates can be modified to accommodate a more complex data structure.


The Length-Biased Lognormal Distribution And Its Application In The Analysis Of Data From Oil Field Exploration Studies, Makarand V. Ratnaparkhi, Uttara V. Naik-Nimbalkar 2012 Wright State University

The Length-Biased Lognormal Distribution And Its Application In The Analysis Of Data From Oil Field Exploration Studies, Makarand V. Ratnaparkhi, Uttara V. Naik-Nimbalkar

Journal of Modern Applied Statistical Methods

The length-biased version of the lognormal distribution and related estimation problems are considered and sized-biased data arising in the exploration of oil fields is analyzed. The properties of the estimators are studied using simulations and the use of sample mode as an estimate of the lognormal parameter is discussed.


Four Period Crossover Designs, James F. Reed III 2012 Christiana Care Hospital System, Newark, Delaware

Four Period Crossover Designs, James F. Reed Iii

Journal of Modern Applied Statistical Methods

In higher-order four period crossover designs with two treatments, sixteen possible treatment sequences can result: AAAA, AAAB, AABA, AABB, ABAA, ABAB, ABBA, ABBB and their duals. Higher-order crossover designs are useful for several reasons: they allow estimation of a treatment effect even in the presence of a carry-over effect, they provide estimates of intra-subject variability and they draw inference on the carry-over effect. The real question related to a two-treatment four-period crossover design is the real world application of these designs. This article considers four designs: Design I: ABBA and its dual; Design II: ABBA, AABB and their duals, Design …


Gamma-Pareto Distribution And Its Applications, Ayman Alzaatreh, Felix Famoye, Carl Lee 2012 Austin Peay State University, Clarksville, TN

Gamma-Pareto Distribution And Its Applications, Ayman Alzaatreh, Felix Famoye, Carl Lee

Journal of Modern Applied Statistical Methods

A new distribution, the gamma-Pareto, is defined and studied and various properties of the distribution are obtained. Results for moments, limiting behavior and entropies are provided. The method of maximum likelihood is proposed for estimating the parameters and the distribution is applied to fit three real data sets.


A Weighted Exponential Detection Function Model For Line Transect Data, Faisal Ababneh, Omar M. Eidous 2012 Al-Hussian Bin Talal University, Ma’an, Jordan

A Weighted Exponential Detection Function Model For Line Transect Data, Faisal Ababneh, Omar M. Eidous

Journal of Modern Applied Statistical Methods

A new parametric model is proposed for modeling the density function of perpendicular distances in line transects sampling. The model can be considered a weighted exponential model in the sense that it combines two exponential models with different weights. The proposed model is appealing because it is monotone decreasing with distance from transect line; in contrast to the classical exponential model, it satisfies the shoulder condition at the origin. Simulation results for a wide range of target densities show reasonable and good performances of the weighted exponential model in most considered cases compared to the classical exponential and the half-normal …


Ordinal Regression Analysis: Using Generalized Ordinal Logistic Regression Models To Estimate Educational Data, Xing Liu, Hari Koirala 2012 Eastern Connecticut State University

Ordinal Regression Analysis: Using Generalized Ordinal Logistic Regression Models To Estimate Educational Data, Xing Liu, Hari Koirala

Journal of Modern Applied Statistical Methods

The proportional odds (PO) assumption for ordinal regression analysis is often violated because it is strongly affected by sample size and the number of covariate patterns. To address this issue, the partial proportional odds (PPO) model and the generalized ordinal logit model were developed. However, these models are not typically used in research. One likely reason for this is the restriction of current statistical software packages: SPSS cannot perform the generalized ordinal logit model analysis and SAS requires data restructuring. This article illustrates the use of generalized ordinal logistic regression models to predict mathematics proficiency levels using Stata and compares …


Digital Commons powered by bepress