Open Access. Powered by Scholars. Published by Universities.®

Statistics and Probability Commons™

Open Access. Powered by Scholars. Published by Universities.®

Statistical Methodology

Institution
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 1081 - 1110 of 1562

Full-Text Articles in Statistics and Probability

Binomial Regression With A Misclassified Covariate And Outcome, Sheng Luo, Wenyaw Chan, Michelle A Detry, Paul J Massman, R S. Doody Feb 2016

Binomial Regression With A Misclassified Covariate And Outcome, Sheng Luo, Wenyaw Chan, Michelle A Detry, Paul J Massman, R S. Doody

Faculty, Staff and Students Publications

Misclassification occurring in either outcome variables or categorical covariates or both is a common issue in medical science. It leads to biased results and distorted disease-exposure relationships. Moreover, it is often of clinical interest to obtain the estimates of sensitivity and specificity of some diagnostic methods even when neither gold standard nor prior knowledge about the parameters exists. We present a novel Bayesian approach in binomial regression when both the outcome variable and one binary covariate are subject to misclassification. Extensive simulation results under various scenarios and a real clinical example are given to illustrate the proposed approach. This approach …


Models For Hsv Shedding Must Account For Two Levels Of Overdispersion, Amalia Magaret Jan 2016

Models For Hsv Shedding Must Account For Two Levels Of Overdispersion, Amalia Magaret

UW Biostatistics Working Paper Series

We have frequently implemented crossover studies to evaluate new therapeutic interventions for genital herpes simplex virus infection. The outcome measured to assess the efficacy of interventions on herpes disease severity is the viral shedding rate, defined as the frequency of detection of HSV on the genital skin and mucosa. We performed a simulation study to ascertain whether our standard model, which we have used previously, was appropriately considering all the necessary features of the shedding data to provide correct inference. We simulated shedding data under our standard, validated assumptions and assessed the ability of 5 different models to reproduce the …


Open Access!: Review Of Online Statistics: An Interactive Multimedia Course Of Study By David Lane, Samuel L. Tunstall Jan 2016

Open Access!: Review Of Online Statistics: An Interactive Multimedia Course Of Study By David Lane, Samuel L. Tunstall

Numeracy

David M. Lane (project leader). Online Statistics Education: An Interactive Multimedia Course of Study (http://onlinestatbook.com/)
Also: David M. Lane (primary author and editor), with David Scott, Mikki Hebl, Rudy Guerra, Dan Osherson, and Heidi Zimmer. Introduction to Statistics. Online edition (http://onlinestatbook.com/Online_Statistics_Education.pdf), 694 pp.

It is rare that students receive high-quality textbooks for free, but David Lane's Online Statistics: An Interactive Multimedia Course of Study permits precisely that. This review gives an overview of the many features in Lane's online textbook, including the Java Applets, the textbook itself, and the resources available for instructors. A discussion …


The Myth Of Making Inferences For An Overall Treatment Efficacy With Data From Multiple Comparative Studies Via Meta-Analysis, Takahiro Hasegawa, Brian Claggett, Lu Tian, Scott D. Solomon, Marc A. Pfeffer, Lee-Jen Wei Jan 2016

The Myth Of Making Inferences For An Overall Treatment Efficacy With Data From Multiple Comparative Studies Via Meta-Analysis, Takahiro Hasegawa, Brian Claggett, Lu Tian, Scott D. Solomon, Marc A. Pfeffer, Lee-Jen Wei

Harvard University Biostatistics Working Paper Series

Meta analysis techniques, if applied appropriately, can provide a summary of the totality of evidence regarding an overall difference between a new treatment and a control group using data from multiple comparative clinical studies. The standard meta analysis procedures, however, may not give a meaningful between-group difference summary measure or identify a meaningful patient population of interest, especially when the fixed effect model assumption is not met. Moreover, a single between-group comparison measure without a reference value obtained from patients in the control arm would likely not be informative enough for clinical decision making. In this paper, we propose a …


What's In A Name? A Critical Review Of Definitions Of Quantitative Literacy, Numeracy, And Quantitative Reasoning, Gizem Karaali, Edwin H Villafane Hernandez '18, Jeremy Alexander Taylor '18 Jan 2016

What's In A Name? A Critical Review Of Definitions Of Quantitative Literacy, Numeracy, And Quantitative Reasoning, Gizem Karaali, Edwin H Villafane Hernandez '18, Jeremy Alexander Taylor '18

Pomona Faculty Publications and Research

This article aims to bring together various threads in the eclectic literature that make up the scholarship around the theme of Quantitative Literacy. In investigating the meanings of terms like "quantitative literacy," "quantitative reasoning," and "numeracy," we seek common ground, common themes, common goals and aspirations of a community of practitioners. A decade ago, these terms were relatively new in the public sphere; today policy makers and accrediting agencies are routinely inserting them into general education conversations. Having good, representative, and perhaps even compact and easily digestible definitions of these terms might come in handy in public relations contexts as …


Design & Analysis Of A Computer Experiment For An Aerospace Conformance Simulation Study, Ryan W. Gryder Jan 2016

Design & Analysis Of A Computer Experiment For An Aerospace Conformance Simulation Study, Ryan W. Gryder

Theses and Dissertations

Within NASA's Air Traffic Management Technology Demonstration # 1 (ATD-1), Interval Management (IM) is a flight deck tool that enables pilots to achieve or maintain a precise in-trail spacing behind a target aircraft. Previous research has shown that violations of aircraft spacing requirements can occur between an IM aircraft and its surrounding non-IM aircraft when it is following a target on a separate route. This research focused on the experimental design and analysis of a deterministic computer simulation which models our airspace configuration of interest. Using an original space-filling design and Gaussian process modeling, we found that aircraft delay assignments …


Development In Normal Mixture And Mixture Of Experts Modeling, Meng Qi Jan 2016

Development In Normal Mixture And Mixture Of Experts Modeling, Meng Qi

Theses and Dissertations--Statistics

In this dissertation, first we consider the problem of testing homogeneity and order in a contaminated normal model, when the data is correlated under some known covariance structure. To address this problem, we developed a moment based homogeneity and order test, and design weights for test statistics to increase power for homogeneity test. We applied our test to microarray about Down’s syndrome. This dissertation also studies a singular Bayesian information criterion (sBIC) for a bivariate hierarchical mixture model with varying weights, and develops a new data dependent information criterion (sFLIC).We apply our model and criteria to birth- weight and gestational …


Statistical Methods For Handling Intentional Inaccurate Responders, Kristen J. Mcquerry Jan 2016

Statistical Methods For Handling Intentional Inaccurate Responders, Kristen J. Mcquerry

Theses and Dissertations--Statistics

In self-report data, participants who provide incorrect responses are known as intentional inaccurate responders. This dissertation provides statistical analyses for address intentional inaccurate responses in the data.

Previous work with adolescent self-report, labeled survey participants who intentionally provide inaccurate answers as mischievous responders. This phenomenon also occurs in clinical research. For example, pregnant women who smoke may report that they are nonsmokers. Our advantage is that we do not solely have self-report answers and can verify responses with lab values. Currently, there is no clear method for handling these intentional inaccurate respondents when it comes to making statistical inferences.

We …


Improved Models For Differential Analysis For Genomic Data, Hong Wang Jan 2016

Improved Models For Differential Analysis For Genomic Data, Hong Wang

Theses and Dissertations--Statistics

This paper intend to develop novel statistical methods to improve genomic data analysis, especially for differential analysis. We considered two different data type: NanoString nCounter data and somatic mutation data. For NanoString nCounter data, we develop a novel differential expression detection method. The method considers a generalized linear model of the negative binomial family to characterize count data and allows for multi-factor design. Data normalization is incorporated in the model framework through data normalization parameters, which are estimated from control genes embedded in the nCounter system. For somatic mutation data, we develop beta-binomial model-based approaches to identify highly or lowly …


Empirical Likelihood And Differentiable Functionals, Zhiyuan Shen Jan 2016

Empirical Likelihood And Differentiable Functionals, Zhiyuan Shen

Theses and Dissertations--Statistics

Empirical likelihood (EL) is a recently developed nonparametric method of statistical inference. It has been shown by Owen (1988,1990) and many others that empirical likelihood ratio (ELR) method can be used to produce nice confidence intervals or regions. Owen (1988) shows that -2logELR converges to a chi-square distribution with one degree of freedom subject to a linear statistical functional in terms of distribution functions. However, a generalization of Owen's result to the right censored data setting is difficult since no explicit maximization can be obtained under constraint in terms of distribution functions. Pan and Zhou (2002), instead, study the …


Improved Parameter Estimation Of The Log-Logistic Distribution With Applications, Joseph Reath Jan 2016

Improved Parameter Estimation Of The Log-Logistic Distribution With Applications, Joseph Reath

Dissertations, Master's Theses and Master's Reports

In this report, we work with parameter estimation of the log-logistic distribution. We first consider one of the most common methods encountered in the literature, the maximum likelihood (ML) method. However, it is widely known that the maximum likelihood estimators (MLEs) are usually biased with a finite sample size. This motivates a study of obtaining unbiased or nearly unbiased estimators for this distribution. Specifically, we consider a certain `corrective' approach and Efron's bootstrap resampling method, which both can reduce the biases of the MLEs to the second order of magnitude. As a comparison, we also consider the generalized moments (GM) …


Comparison Of Option Price From Black-Scholes Model To Actual Values, Matthew J. Krznaric Jan 2016

Comparison Of Option Price From Black-Scholes Model To Actual Values, Matthew J. Krznaric

Williams Honors College, Honors Research Projects

The Black-Scholes model is a widely used method for pricing European-style options in a straightforward way, through the use of calculations and ideal market assumptions. Due to certain unrealistic ideal conditions exercised by the model, The Black-Scholes technique of pricing options may not be entirely accurate in implementation. This paper addresses these problems due to the model limitations, determining how The Black-Scholes method compares to the results when using the actual data. Using a mix of historical S&P500 data and generated normal distributions, we first calculated and graphed option prices through the Black-Scholes formulas. With the help of R, we …


Black Cloud Randomization Test, Nicholas S. Vanni Jan 2016

Black Cloud Randomization Test, Nicholas S. Vanni

Williams Honors College, Honors Research Projects

The Black Cloud Randomization Test looks at a nontraditional question and attempts to answer the question using unique statistics. The purpose of this paper is to apply what has been learned throughout the years and apply this knowledge to a final project. Data for this project follows an emergency room’s on call schedule, as well as the number of traumas that came in during each day shift. The project builds on what has been already learned and helps to open a different way of working with statistics. The project was coded in the R software. With different restrictions, there are …


Missing Data In Clinical Trial: A Critical Look At The Proportionality Of Mnar And Mar Assumptions For Multiple Imputation, Theophile B. Dipita Jan 2016

Missing Data In Clinical Trial: A Critical Look At The Proportionality Of Mnar And Mar Assumptions For Multiple Imputation, Theophile B. Dipita

College of Graduate Studies: Theses & Dissertations

Randomized control trial is a gold standard of research studies. Randomization helps reduce bias and infer causality. One constraint of these studies is that it depends on participants to obtain the desired data. Whatever the researcher can do, there is a possibility to end up with incomplete data. The problem is more relevant in clinical trials when missing data can be related to the condition under study. The benefits of randomization is compromised by missing data. Multiple imputation is a valid method of treating missing data under the assumption of MAR. Unfortunately this is an unverified assumptions. Current practice advise …


In-Shoe Plantar Pressure System To Investigate Ground Reaction Force Using Android Platform, Ahmed A. Mostfa Jan 2016

In-Shoe Plantar Pressure System To Investigate Ground Reaction Force Using Android Platform, Ahmed A. Mostfa

Theses and Dissertations

Human footwear is not yet designed to optimally relieve pressure on the heel of the foot. Proper foot pressure assessment requires personal training and measurements by specialized machinery. This research aims to investigate and hypothesize about Preferred Transition Speed (PTS) and to classify the gait phase of explicit variances in walking patterns between different subjects. An in-shoe wearable pressure system using Android application was developed to investigate walking patterns and collect data on Activities of Daily Living (ADL). In-shoe circuitry used Flexi-Force A201 sensors placed at three major areas: heel contact, 1st metatarsal, and 5th metatarsal with a PIC16F688 microcontroller …


Dimension Reduction And Variable Selection, Hossein Moradi Rekabdarkolaee Jan 2016

Dimension Reduction And Variable Selection, Hossein Moradi Rekabdarkolaee

Theses and Dissertations

High-dimensional data are becoming increasingly available as data collection technology advances. Over the last decade, significant developments have been taking place in high-dimensional data analysis, driven primarily by a wide range of applications in many fields such as genomics, signal processing, and environmental studies. Statistical techniques such as dimension reduction and variable selection play important roles in high dimensional data analysis. Sufficient dimension reduction provides a way to find the reduced space of the original space without a parametric model. This method has been widely applied in many scientific fields such as genetics, brain imaging analysis, econometrics, environmental sciences, etc. …


A New Right Tailed Test Of The Ratio Of Variances, Elizabeth Rochelle Lesser Jan 2016

A New Right Tailed Test Of The Ratio Of Variances, Elizabeth Rochelle Lesser

UNF Graduate Theses and Dissertations

It is important to be able to compare variances efficiently and accurately regardless of the parent populations. This study proposes a new right tailed test for the ratio of two variances using the Edgeworth’s expansion. To study the Type I error rate and Power performance, simulation was performed on the new test with various combinations of symmetric and skewed distributions. It is found to have more controlled Type I error rates than the existing tests. Additionally, it also has sufficient power. Therefore, the newly derived test provides a good robust alternative to the already existing methods.


Inequality In Treatment Benefits: Can We Determine If A New Treatment Benefits The Many Or The Few?, Emily Huang, Ethan Fang, Daniel Hanley, Michael Rosenblum Dec 2015

Inequality In Treatment Benefits: Can We Determine If A New Treatment Benefits The Many Or The Few?, Emily Huang, Ethan Fang, Daniel Hanley, Michael Rosenblum

Johns Hopkins University, Dept. of Biostatistics Working Papers

The primary analysis in many randomized controlled trials focuses on the average treatment effect and does not address whether treatment benefits are widespread or limited to a select few. This problem affects many disease areas, since it stems from how randomized trials, often the gold standard for evaluating treatments, are designed and analyzed. Our goal is to learn about the fraction who benefit from a treatment, based on randomized trial data. We consider the case where the outcome is ordinal, with binary outcomes as a special case. In general, the fraction who benefit is a non-identifiable parameter, and the best …


Preparedness Of Hospitals In The Republic Of Ireland For An Influenza Pandemic, An Infection Control Perspective, Mary Reidy, Fiona Ryan, Dervla Hogan, Seán Lacey, Claire Buckley Sep 2015

Preparedness Of Hospitals In The Republic Of Ireland For An Influenza Pandemic, An Infection Control Perspective, Mary Reidy, Fiona Ryan, Dervla Hogan, Seán Lacey, Claire Buckley

Department of Mathematics Publications

When an influenza pandemic occurs most of the population is susceptible and attack rates can range as high as 40–50 %. The most important failure in pandemic planning is the lack of standards or guidelines regarding what it means to be ‘prepared’. The aim of this study was to assess the preparedness of acute hospitals in the Republic of Ireland for an influenza pandemic from an infection control perspective.


C-Learning: A New Classification Framework To Estimate Optimal Dynamic Treatment Regimes, Baqun Zhang, Min Zhang Aug 2015

C-Learning: A New Classification Framework To Estimate Optimal Dynamic Treatment Regimes, Baqun Zhang, Min Zhang

The University of Michigan Department of Biostatistics Working Paper Series

Personalizing treatment to accommodate patient heterogeneity and the evolving nature of a disease over time has received considerable attention lately. A dynamic treatment regime is a set of decision rules, each corresponding to a decision point, that determine that next treatment based on each individual’s own available characteristics and treatment history up to that point. We show that identifying the optimal dynamic treatment regime can be recast as a sequential classification problem and is equivalent to sequentially minimizing a weighted expected misclassification error. This general classification perspective targets the exact goal of optimally individualizing treatments and is new and fundamentally …


Supervised Classification Using Copula And Mixture Copula, Sumen Sen Jul 2015

Supervised Classification Using Copula And Mixture Copula, Sumen Sen

Mathematics & Statistics Theses & Dissertations

Statistical classification is a field of study that has developed significantly after 1960's. This research has a vast area of applications. For example, pattern recognition has been proposed for automatic character recognition, medical diagnostic and most recently in data mining. Classical discrimination rule assumes normality. However in many situations, this assumption is often questionable. In fact for some data, the pattern vector is a mixture of discrete and continuous random variables. In this dissertation, we use copula densities to model class conditional distributions. Such types of densities are useful when the marginal densities of a pattern vector are not normally …


A Study Of The Parametric And Nonparametric Linear-Circular Correlation Coefficient, Robin Tu Jun 2015

A Study Of The Parametric And Nonparametric Linear-Circular Correlation Coefficient, Robin Tu

Statistics

Circular statistics are specialized statistical methods that deal specifically with directional data. Data that is angular require specialized techniques due to the modulo 2π (in radians) or modulo 360◦ (in degrees) nature of angles.

Correlation, typically in terms of Pearson’s correlation coefficient, is a measure of association between two linear random variables x and y. In this paper, the specific circular technique of the parametric and nonparametric linear-circular correlation coefficient will be explored where correlation is no longer between two linear variables x and y, but between a linear random variable x and circular random variable θ.

A simulation …


The Effects Of Quantitative Easing In The United States: Implications For Future Central Bank Policy Makers, Matthew Q. Rubino May 2015

The Effects Of Quantitative Easing In The United States: Implications For Future Central Bank Policy Makers, Matthew Q. Rubino

Senior Honors Projects, 2010-2019

The purpose of this thesis is to examine the effects of the Federal Reserve’s recent bond buying programs, specifically Quantitative Easing 1, Quantitative Easing 2, Operation Twist (or the Fed’s Maturity Extension Program), and Quantitative Easing 3. In this study, I provide a picture of the economic landscape leading up to the deployment of the programs, an overview of quantitative easing including each program’s respective objectives, and how and why the Fed decided to implement the programs. Using empirical analysis, I measure each program’s effectiveness by applying four models including a yield curve model, an inflation model, a money supply …


The Effects Of A Planned Missingness Design On Examinee Motivation And Psychometric Quality, Matthew S. Swain May 2015

The Effects Of A Planned Missingness Design On Examinee Motivation And Psychometric Quality, Matthew S. Swain

Dissertations, 2014-2019

Assessment practitioners in higher education face increasing demands to collect assessment and accountability data to make important inferences about student learning and institutional quality. The validity of these high-stakes decisions is jeopardized, particularly in low-stakes testing contexts, when examinees do not expend sufficient motivation to perform well on the test. This study introduced planned missingness as a potential solution. In planned missingness designs, data on all items are collected but each examinee only completes a subset of items, thus increasing data collection efficiency, reducing examinee burden, and potentially increasing data quality. The current scientific reasoning test served as the Long …


Examining The Performance Of The Metropolis-Hastings Robbins-Monro Algorithm In The Estimation Of Multilevel Multidimensional Irt Models, Bozhidar M. Bashkov May 2015

Examining The Performance Of The Metropolis-Hastings Robbins-Monro Algorithm In The Estimation Of Multilevel Multidimensional Irt Models, Bozhidar M. Bashkov

Dissertations, 2014-2019

The purpose of this study was to review the challenges that exist in the estimation of complex (multidimensional) models applied to complex (multilevel) data and to examine the performance of the recently developed Metropolis-Hastings Robbins-Monro (MH-RM) algorithm (Cai, 2010a, 2010b), designed to overcome these challenges and implemented in both commercial and open-source software programs. Unlike other methods, which either rely on high-dimensional numerical integration or approximation of the entire multidimensional response surface, MH-RM makes use of Fisher’s Identity to employ stochastic imputation (i.e., data augmentation) via the Metropolis-Hastings sampler and then apply the stochastic approximation method of Robbins and Monro …


Do Footprint-Based Cafe Standards Make Car Models Bigger?, Brianna Marie Jean May 2015

Do Footprint-Based Cafe Standards Make Car Models Bigger?, Brianna Marie Jean

Economics

Corporate Average Fuel Economy (CAFE) standards have historically been set equal across all manufacturer fleets of the same type. Concerns about varying costs across firms and safety implications of standards that are set homogeneously across firms and models resulted in a policy shift towards footprint-based standards. Under this type of standard, individual car models face targets based on the size of the area between the wheelbase and wheel track, so that larger models face less stringent standards, and manufacturers who make, on average, larger cars will face a lighter fleet standard. Theoretical models have shown that this type of policy …


Scientific Awareness At Ursinus College, Frank G. Devone Apr 2015

Scientific Awareness At Ursinus College, Frank G. Devone

Mathematics Honors Papers

Ursinus College prides itself on creating well-rounded students, and recent initiatives, such as the Fellowships in the Ursinus Transition to the Undergraduate Research Experience Program and the Center for Science and the Common Good suggest that science is a vital part of the Ursinus liberal arts mission. A scientific awareness pilot survey was administered to a sample of Ursinus students drawn from the Class of 2014 and students residing at Ursinus during summer 2014. Experience and data collected from this pilot were used to create a final survey which was made available to all students at Ursinus College. The survey …


Adaptive Enrichment Designs For Randomized Trials With Delayed Endpoints, Using Locally Efficient Estimators To Improve Precision, Michael Rosenblum, Tianchen Qian, Yu Du, Huitong Qiu Apr 2015

Adaptive Enrichment Designs For Randomized Trials With Delayed Endpoints, Using Locally Efficient Estimators To Improve Precision, Michael Rosenblum, Tianchen Qian, Yu Du, Huitong Qiu

Johns Hopkins University, Dept. of Biostatistics Working Papers

Adaptive enrichment designs involve preplanned rules for modifying enrollment criteria based on accrued data in an ongoing trial. For example, enrollment of a subpopulation where there is sufficient evidence of treatment efficacy, futility, or harm could be stopped, while enrollment for the remaining subpopulations is continued. Most existing methods for constructing adaptive enrichment designs are limited to situations where patient outcomes are observed soon after enrollment. This is a major barrier to the use of such designs in practice, since for many diseases the outcome of most clinical importance does not occur shortly after enrollment. We propose a new class …


Global Network Inference From Ego Network Samples: Testing A Simulation Approach, Jeffrey A. Smith Apr 2015

Global Network Inference From Ego Network Samples: Testing A Simulation Approach, Jeffrey A. Smith

Department of Sociology: Faculty Publications

Network sampling poses a radical idea: that it is possible to measure global network structure without the full population coverage assumed in most network studies. Network sampling is only useful, however, if a researcher can produce accurate global network estimates. This article explores the practicality of making network inference, focusing on the approach introduced in Smith (2012). The method uses sampled ego network data and simulation techniques to make inference about the global features of the true, unknown network. The validity check here includes more difficult scenarios than previous tests, including those that go beyond the initial scope conditions of …


Relationship Between High School Math Course Selection And Retention Rates At Otterbein University, Lauren A. Fisher Apr 2015

Relationship Between High School Math Course Selection And Retention Rates At Otterbein University, Lauren A. Fisher

Undergraduate Honors Thesis Projects

Binary logistic regression was used to study the relationship between high school math course selection and retention rates at Otterbein University. Graduation rates from postsecondary institutions are low in the United States and, more specifically, at Otterbein. This study is important in helping to determine what can raise retention rates, and ultimately, graduation rates. It directs focus toward high school math course selection and what should be changed before entering a post-secondary institution. Otterbein will have a better idea of what type of students to recruit and which students may be good candidates with some extra help. Recruiting is expensive, …