Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Social and Behavioral Sciences (39)
- Statistical Models (35)
- Statistical Theory (30)
- Biostatistics (17)
- Statistical Methodology (17)
-
- Computer Sciences (16)
- Multivariate Analysis (16)
- Medicine and Health Sciences (15)
- Mathematics (14)
- Applied Mathematics (13)
- Engineering (13)
- Life Sciences (12)
- Artificial Intelligence and Robotics (10)
- Categorical Data Analysis (10)
- Longitudinal Data Analysis and Time Series (10)
- Probability (9)
- Numerical Analysis and Scientific Computing (8)
- Environmental Sciences (7)
- Oceanography and Atmospheric Sciences and Meteorology (7)
- Survival Analysis (7)
- Design of Experiments and Sample Surveys (6)
- Earth Sciences (6)
- Theory and Algorithms (6)
- Bioinformatics (5)
- Civil and Environmental Engineering (5)
- Economics (5)
- Education (5)
- Institution
-
- Wayne State University (23)
- Southern Methodist University (11)
- Montclair State University (7)
- Claremont Colleges (6)
- University of Arkansas, Fayetteville (6)
-
- Georgia Southern University (4)
- Indian Statistical Institute (4)
- Old Dominion University (4)
- University of Montana (4)
- Virginia Commonwealth University (4)
- California Polytechnic State University, San Luis Obispo (3)
- City University of New York (CUNY) (3)
- East Tennessee State University (3)
- Illinois State University (3)
- Louisiana State University (3)
- Technological University Dublin (3)
- The University of San Francisco (3)
- University of Kentucky (3)
- Bowling Green State University (2)
- COBRA (2)
- Embry-Riddle Aeronautical University (2)
- Michigan Technological University (2)
- South Dakota State University (2)
- University of Nevada, Las Vegas (2)
- Air Force Institute of Technology (1)
- Boise State University (1)
- Bucknell University (1)
- Central Washington University (1)
- Chapman University (1)
- Colby College (1)
- Keyword
-
- Statistics (9)
- Machine Learning (5)
- Deep Learning (4)
- Bayesian (3)
- Bias (3)
-
- Machine learning (3)
- Power (3)
- Regression (3)
- Analytics (2)
- Bayesian analysis (2)
- Big data (2)
- Biogeochemistry (2)
- Classification (2)
- Cloud Computing (2)
- Data Science (2)
- Depression (2)
- Feature selection (2)
- Hemodynamics (2)
- Heteroscedasticity (2)
- Logistic regression (2)
- Mean square error (2)
- Misclassification (2)
- Model (2)
- Morphology (2)
- PCA (2)
- Random Forest (2)
- Risk (2)
- Simulation (2)
- Survey (2)
- A/B testing (1)
- Publication
-
- Journal of Modern Applied Statistical Methods (23)
- SMU Data Science Review (9)
- Graduate Theses and Dissertations (6)
- Department of Applied Mathematics and Statistics Faculty Scholarship and Creative Works (5)
- College of Graduate Studies: Theses & Dissertations (4)
-
- Electronic Theses and Dissertations (4)
- Graduate Student Theses, Dissertations, & Professional Papers (4)
- Journal Articles (4)
- Theses and Dissertations (4)
- Annual Symposium on Biomathematics and Ecology Education and Research (3)
- Articles (2)
- CMC Senior Theses (2)
- Creative Activity and Research Day - CARD (2)
- Dissertations, Master's Theses and Master's Reports (2)
- Honors Projects (2)
- International Journal of Aviation, Aeronautics, and Aerospace (2)
- LSU Master's Theses (2)
- Mathematics & Statistics Theses & Dissertations (2)
- Pomona Economics (2)
- SDSU Data Science Symposium (2)
- STAR Program Research Presentations (2)
- Statistical Science Theses and Dissertations (2)
- Theses, Dissertations and Culminating Projects (2)
- All Undergraduate Projects (1)
- Biology and Medicine Through Mathematics Conference (1)
- Boise State University Theses and Dissertations (1)
- Book/Book Chapter (1)
- CGU Theses & Dissertations (1)
- COBRA Preprint Series (1)
- Capstone Experience: Master of Public Health (1)
- Publication Type
- File Type
Articles 91 - 120 of 136
Full-Text Articles in Applied Statistics
Striving For Simple But Effective Advice For Comparing The Central Tendency Of Two Populations, Graeme Ruxton, Markus Neuhäuser
Striving For Simple But Effective Advice For Comparing The Central Tendency Of Two Populations, Graeme Ruxton, Markus Neuhäuser
Journal of Modern Applied Statistical Methods
Nguyen et al. (2016) offered advice to researchers in the commonly-encountered situation where they are interested in testing for a difference in central tendency between two populations. Their data and the available literature support very simple advice that strikes the best balance between ease of implementation, power and reliability. Specifically, apply Satterthwaite’s test, with preliminary ranking of the data if a strong deviation from normality is expected, or is suggested by visual inspection of the data. This simple guideline will serve well except when dealing with small samples of discrete data, when more sophisticated treatment may be required.
Logistic Regression: An Inferential Method For Identifying The Best Predictors, Rand Wilcox
Logistic Regression: An Inferential Method For Identifying The Best Predictors, Rand Wilcox
Journal of Modern Applied Statistical Methods
When dealing with a logistic regression model, there is a simple method for estimating the strength of the association between the jth covariate and the dependent variable when all covariates are entered into the model. There is the issue of determining whether the jth independent variable has a stronger or weaker association than the kth independent variable. This note describes a method for dealing with this issue that was found to perform reasonably well in simulations.
Data Analytics Pipeline For Rna Structure Analysis Via Shape, Quinn Nelson
Data Analytics Pipeline For Rna Structure Analysis Via Shape, Quinn Nelson
UNO Student Research and Creative Activity Fair
Coxsackievirus B3 (CVB3) is a cardiovirulent enterovirus from the family Picornaviridae. The RNA genome houses an internal ribosome entry site (IRES) in the 5’ untranslated region (5’UTR) that enables cap-independent translation. Ample evidence suggests that the structure of the 5’UTR is a critical element for virulence. We probe RNA structure in solution using base-specific modifying agents such as dimethyl sulfate as well as backbone targeting agents such as N-methylisatoic anhydride used in Selective 2’-Hydroxyl Acylation Analyzed by Primer Extension (SHAPE). We have developed a pipeline that merges and evaluates base-specific and SHAPE data together with statistical analyses that provides confidence …
Sustainable Energy Governance In South Tyrol (Italy): A Probabilistic Bipartite Network Model, Jessica Belest, Laura Secco, Elena Pisani, Alberto Caimo
Sustainable Energy Governance In South Tyrol (Italy): A Probabilistic Bipartite Network Model, Jessica Belest, Laura Secco, Elena Pisani, Alberto Caimo
Articles
At the national scale, almost all of the European countries have already achieved energy transition targets, while at the regional and local scales, there is still some potential to further push sustainable energy transitions. Regions and localities have the support of political, social, and economic actors who make decisions for meeting existing social, environmental and economic needs recognising local specificities.
These actors compose the sustainable energy governance that is fundamental to effectively plan and manage energy resources. In collaborative relationships, these actors share, save, and protect several kinds of resources, thereby making energy transitions deeper and more effective.
This research …
Session: 4 Multilinear Subspace Learning And Its Applications To Machine Learning, Randy Hoover, Kyle Caudle Dr., Karen Braman Dr.
Session: 4 Multilinear Subspace Learning And Its Applications To Machine Learning, Randy Hoover, Kyle Caudle Dr., Karen Braman Dr.
SDSU Data Science Symposium
Multi-dimensional data analysis has seen increased interest in recent years. With more and more data arriving as 2-dimensional arrays (images) as opposed to 1-dimensioanl arrays (signals), new methods for dimensionality reduction, data analysis, and machine learning have been pursued. Most notably have been the Canonical Decompositions/Parallel Factors (commonly referred to as CP) and Tucker decompositions (commonly regarded as a high order SVD: HOSVD). In the current research we present an alternate method for computing singular value and eigenvalue decompositions on multi-way data through an algebra of circulants and illustrate their application to two well-known machine learning methods: Multi-Linear Principal Component …
Predicting Unplanned Medical Visits Among Patients With Diabetes Using Machine Learning, Arielle Selya, Eric L. Johnson
Predicting Unplanned Medical Visits Among Patients With Diabetes Using Machine Learning, Arielle Selya, Eric L. Johnson
SDSU Data Science Symposium
Diabetes poses a variety of medical complications to patients, resulting in a high rate of unplanned medical visits, which are costly to patients and healthcare providers alike. However, unplanned medical visits by their nature are very difficult to predict. The current project draws upon electronic health records (EMR’s) of adult patients with diabetes who received care at Sanford Health between 2014 and 2017. Various machine learning methods were used to predict which patients have had an unplanned medical visit based on a variety of EMR variables (age, BMI, blood pressure, # of prescriptions, # of diagnoses on problem list, A1C, …
Nonparametric Depth And Quantile Regression For Functional Data, Joydeep Chowdhury, Probal Chaudhuri
Nonparametric Depth And Quantile Regression For Functional Data, Joydeep Chowdhury, Probal Chaudhuri
Journal Articles
We investigate nonparametric regression methods based on spatial depth and quantiles when the response and the covariate are both functions. As in classical quantile regression for finite dimensional data, regression techniques developed here provide insight into the influence of the functional covariate on different parts, like the center as well as the tails, of the conditional distribution of the functional response. Depth and quantile based nonparametric regression methods are useful to detect heteroscedasticity in functional regression. We derive the asymptotic behavior of the nonparametric depth and quantile regression estimates, which depend on the small ball probabilities in the covariate space. …
Pedestrian Safety -- Fundamental To A Walkable City, Joshua Herrera, Patrick Mcdevitt, Preeti Swaminathan, Raghuram Srinivas
Pedestrian Safety -- Fundamental To A Walkable City, Joshua Herrera, Patrick Mcdevitt, Preeti Swaminathan, Raghuram Srinivas
SMU Data Science Review
In this paper, we present a method to identify urban areas with a higher likelihood of pedestrian safety related events. Pedestrian safety related events are pedestrian-vehicle interactions that result in fatalities, injuries, accidents without injury, or near--misses between pedestrians and vehicles. To develop a solution to this problem of identifying likely event locations, we assemble data, primarily from the City of Cincinnati and Hamilton County, that include safety reports from a five year period, geographic information for these events, citizen survey of pedestrian reported concerns, non-emergency requests for service for any cause in the city, property values and public transportation …
Improving Vix Futures Forecasts Using Machine Learning Methods, James Hosker, Slobodan Djurdjevic, Hieu Nguyen, Robert Slater
Improving Vix Futures Forecasts Using Machine Learning Methods, James Hosker, Slobodan Djurdjevic, Hieu Nguyen, Robert Slater
SMU Data Science Review
The problem of forecasting market volatility is a difficult task for most fund managers. Volatility forecasts are used for risk management, alpha (risk) trading, and the reduction of trading friction. Improving the forecasts of future market volatility assists fund managers in adding or reducing risk in their portfolios as well as in increasing hedges to protect their portfolios in anticipation of a market sell-off event. Our analysis compares three existing financial models that forecast future market volatility using the Chicago Board Options Exchange Volatility Index (VIX) to six machine/deep learning supervised regression methods. This analysis determines which models provide best …
Ample Provision: A Preliminary Study Relating Budget Composition And High School Graduation Rates In Select Washington State Public School Districts, Gregory P. Gadow
Ample Provision: A Preliminary Study Relating Budget Composition And High School Graduation Rates In Select Washington State Public School Districts, Gregory P. Gadow
All Undergraduate Projects
How to allocate scarce resources for an optimal outcome is of keen interest to those who set the budgets in public education. Simply throwing money at schools is not enough; it is important that money is spent where it will do the most good. This study considers Washington State public school districts and examines how the share of per-student expenditures in seven budget categories relates to on-time high school graduation rates. It is an investigative study, exploring whether there is enough evidence to merit further, more in-depth research. Using budget and graduation information from academic years 1997-98 through 2016-17 for …
Comparative Analysis Of Students’ Performance Between Online And On Campus In An Introductory Statistics Course, Kendal Mcdonald
Comparative Analysis Of Students’ Performance Between Online And On Campus In An Introductory Statistics Course, Kendal Mcdonald
The Corinthian
In this research, we compare students’ performance in an online and on-campus introductory statistics and probability course at Georgia College. MyStatLab is the learning management system used in both the online and on-campus courses for homework and quizzes. The online data is produced by five summer courses between Summer 2014 to Summer 2017 and the on-campus data is produced from nine on-campus courses from Spring 2014, Spring 2016, and Spring 2017. For homework, the research compares the scores made between online and on-campus. For quizzes, we test if there is a difference between the scores and the number of attempts …
Step Away From Stepwise, Gary N. Smith
Step Away From Stepwise, Gary N. Smith
Pomona Economics
Stepwise regression is a popular data-mining tool that uses statistical significance to select the explanatory variables to be used in a multiple-regression model. A fundamental problem with stepwise regression is that some real explanatory variables that have causal effects on the dependent variable may happen to not be statistically significant, while nuisance variables may be coincidentally significant. As a result, the model may fit the data well in-sample, but do poorly out-of-sample. Many Big-Data researchers believe that, the larger the number of possible explanatory variables, the more useful is stepwise regression for selecting explanatory variables. The reality is that stepwise …
Controlling For Confounding Via Propensity Score Methods Can Result In Biased Estimation Of The Conditional Auc: A Simulation Study, Hadiza I. Galadima, Donna K. Mcclish
Controlling For Confounding Via Propensity Score Methods Can Result In Biased Estimation Of The Conditional Auc: A Simulation Study, Hadiza I. Galadima, Donna K. Mcclish
Community & Environmental Health Faculty Publications
In the medical literature, there has been an increased interest in evaluating association between exposure and outcomes using nonrandomized observational studies. However, because assignments to exposure are not random in observational studies, comparisons of outcomes between exposed and nonexposed subjects must account for the effect of confounders. Propensity score methods have been widely used to control for confounding, when estimating exposure effect. Previous studies have shown that conditioning on the propensity score results in biased estimation of conditional odds ratio and hazard ratio. However, research is lacking on the performance of propensity score methods for covariate adjustment when estimating the …
The Dark Sky Character Of Archaeological Landscapes: Cultural Meaning And Conservation Strategies, Frank Prendergast
The Dark Sky Character Of Archaeological Landscapes: Cultural Meaning And Conservation Strategies, Frank Prendergast
Book/Book Chapter
This paper presents the first ever study of light pollution at selected Irish prehistoric archaeological landscapes. The concepts of cosmology and landscape are first briefly described and followed by a summary of early human settlement of the island. Building on this, the extant corpus of early prehistoric megalithic burial tombs is illustrated to show their contrasting distribution patterns and typology. Analysis of tomb locations using nearest-neighbour statistical methods reveals evidence of intentional clustering. Further geo-statistical analysis identifies the geographical locations and the density ranking of these nucleated clusters - a feature especially evident in the passage tomb tradition on this …
Be Wary Of Black-Box Trading Algorithms, Gary N. Smith
Be Wary Of Black-Box Trading Algorithms, Gary N. Smith
Pomona Economics
Black-box algorithms now account for nearly a third of all U. S. stock trades. It is a mistake to think that these algorithms possess superhuman intelligence. In reality, computers do not have the common sense and wisdom that humans have accumulated by living. Trading algorithms are particularly dangerous because they are so efficient at discovering statistical patterns—but so utterly useless in judging whether the discovered patterns are meaningful.
The Scaling Limit Of The Membrane Model, Alessandra Cipriani, Biltu Dan, Rajat Subhra Hazra
The Scaling Limit Of The Membrane Model, Alessandra Cipriani, Biltu Dan, Rajat Subhra Hazra
Journal Articles
On the integer lattice, we consider the discrete membrane model, a random interface in which the field has Laplacian interaction. We prove that, under appropriate rescaling, the discrete membrane model converges to the continuum membrane model in d ≥ 2. Namely, it is shown that the scaling limit in d = 2, 3 is a Holder continuous random field, while in d ≥ 4 the membrane model converges to a random distribution. As a by-product of the proof in d = 2, 3, we obtain the scaling limit of the maximum. This work complements the analogous results of Caravenna and …
The Correlation Between Sleep And Lifespan In Drosophila Melanogaster, Joshua Randall Lisse
The Correlation Between Sleep And Lifespan In Drosophila Melanogaster, Joshua Randall Lisse
Masters Theses
”Adequate sleep is associated with an individual’s health. Too little sleep is associated with many health problems, including cardiovascular disease, obesity, and a general increase in all-cause mortality. Yet the molecular changes that link poor sleep and changes in health are still not well understood. Individuals have a unique daily need for sleep, and deviations from the animal’s regular sleeping patterns can be indicative of, or result in, underlying changes in its health. Therefore, we hypothesize that changes in the sleep architecture in Drosophila melanogaster reflect changes in the fly’s health.
We determined sleep architecture in wild-type male flies over …
Local Hemodynamic Conditions Associated With Focal Changes In The Intracranial Aneurysm Wall, Juan R. Cebral, F. Detmer, Bong Jae Chung, J. Choque-Velasquez, B. Rezai, H. Lehto, R. Tulamo, J. Hernesniemi, M. Niemela, A. Yu, R. Williamson, Khaled Aziz, S. Sakur, S. Amin-Hanjani, F. Charbel, Y. Tobe, A. Robertson, J. Frösen
Local Hemodynamic Conditions Associated With Focal Changes In The Intracranial Aneurysm Wall, Juan R. Cebral, F. Detmer, Bong Jae Chung, J. Choque-Velasquez, B. Rezai, H. Lehto, R. Tulamo, J. Hernesniemi, M. Niemela, A. Yu, R. Williamson, Khaled Aziz, S. Sakur, S. Amin-Hanjani, F. Charbel, Y. Tobe, A. Robertson, J. Frösen
Department of Applied Mathematics and Statistics Faculty Scholarship and Creative Works
BACKGROUND AND PURPOSE: Aneurysm hemodynamics has been associated with wall histology and inflammation. We investigated associations between local hemodynamics and focal wall changes visible intraoperatively. MATERIALS AND METHODS: Computational fluid dynamics models were constructed from 3D images of 65 aneurysms treated surgically. Aneurysm regions with different visual appearances were identified in intraoperative videos: 1) “atherosclerotic” (yellow), 2) “hyperplastic” (white), 3) “thin” (red), 4) rupture site, and 5) “normal” (similar to parent artery), They were marked on 3D reconstructions. Regional hemodynamics was characterized by the following: wall shear stress, oscillatory shear index, relative residence time, wall shear stress gradient and divergence, …
Essays On Mixture Models, Trevor R. Camper
Essays On Mixture Models, Trevor R. Camper
College of Graduate Studies: Theses & Dissertations
When considering statistical scenarios where one can sample from populations that are not of interest for the purposes of a study, bivariate mixture models can be used to study the effect that this missampling can have on parameter estimation. In this thesis, we will examine the behavior that bivariate mixture models have on two statistical constructs: Cronbach's alpha \cite{C51}, and Spearman's rho \cite{S04}. Chapter 1 will introduce notions of mixture models and the definition of bias under mixture models which will serve as the central concept of this thesis. Chapter 2 will investigate a particular psychometric issue known as insufficient …
Some New Generalized Distribution Via Lindley-Weibuli And Lindley-Log-Logistic Distributions With Applications, Soliu A. Raheem
Some New Generalized Distribution Via Lindley-Weibuli And Lindley-Log-Logistic Distributions With Applications, Soliu A. Raheem
College of Graduate Studies: Theses & Dissertations
In this thesis, new generalized distributions, namely Beta Lindley-Log-Logistic (BLLLoG) distribution, Marshall-Olkin Lindley-Weibull (MOLW) distribution, and Gamma LindleyWeibull (GLW) distribution as well as related sub-distributions are proposed. Series expansion of the densities are obtained. Statistical properties of these distributions, including hazard function, reverse hazard function, moments, reliability, quantile function, mean deviations, Bonferroni and Lorenz curves, entropy and Fisher information are derived. Method of maximum likelihood is used to estimate the parameters of the new distributions. Monte Carlo simulation is employed to examine the performance of the proposed distributions. Applications of the generalized distributions to real lifetime data are presented to …
Variable Selection In Accelerated Failure Time (Aft) Frailty Models: An Application Of Penalized Quasi-Likelihood, Sarbesh R. Pandeya
Variable Selection In Accelerated Failure Time (Aft) Frailty Models: An Application Of Penalized Quasi-Likelihood, Sarbesh R. Pandeya
College of Graduate Studies: Theses & Dissertations
Variable selection is one of the standard ways of selecting models in large scale datasets. It has applications in many fields of research study, especially in large multi-center clinical trials. One of the prominent methods in variable selection is the penalized likelihood, which is both consistent and efficient. However, the penalized selection is significantly challenging under the influence of random (frailty) covariates. It is even more complicated when there is involvement of censoring as it may not have a closed-form solution for the marginal log-likelihood. Therefore, we applied the penalized quasi-likelihood (PQL) approach that approximates the solution for such a …
Spatiotemporal Dynamics Of Nitrogen And Carbon Biogeochemistry In A Wetland-Stream Sequence, Patrick E. Hurley
Spatiotemporal Dynamics Of Nitrogen And Carbon Biogeochemistry In A Wetland-Stream Sequence, Patrick E. Hurley
Graduate Student Theses, Dissertations, & Professional Papers
Studies of aquatic ecosystems often segregate streams from the influential ponds, lakes, and wetland zones that act as important transitions between terrestrial and fluvial systems. Across the aquatic landscape, these zones interact to form linked ecosystems that function as discrete nutrient processing domains, shifting biogeochemical signals due to spatial and temporal variability in hydrologic and biologic controls. Using a mass-balance approach, we profiled nutrient dynamics along a 23-km wetland-stream sequence over three seasons. Hydrologic, morphologic, and biologic conditions, as well as landscape attributes, were quantified to determine potential controls on biogeochemical cycling in a tributary of the Upper Clark Fork …
High Dimensional Outlier Detection, Omid Khormali
High Dimensional Outlier Detection, Omid Khormali
Graduate Student Theses, Dissertations, & Professional Papers
In statistics and data science, outliers are data points that differ greatly from other observations in a data set. They are important attributes of the data because they can dramatically influence patterns and relationships manifested by non-outliers. It is therefore very important to detect and adequately deal with outliers. Recently, a novel algorithm, the ROMA algorithm, has been proposed [11]. In this paper, we propose a modification of the ROMA algorithm that reduces its computational complexity from $O(n^2 m)$ to $O((n/(2^m-o(1)))^2 m)$ where $n$ is the number of data points and $m$ is the dimension of the space. And as …
Biodiversity And Distribution Of Benthic Foraminifera In Harrington Sound, Bermuda: The Effects Of Physical And Geochemical Factors On Dominant Taxa, Nam Le
Honors Theses
Harrington Sound, Bermuda, is a nearly enclosed lagoon acting as a subtropical/tropical, carbonate-rich basin in which carbonate sediments, reef patches, and carbonate-producing organisms accumulate. Here, one of the most important calcareous groups is the Foraminifera. Analyses of common benthic orders, including miliolids (Quinqueloculina and Triloculina spp.) and rotaliids (Homotrema rubrum, Elphidium spp., and Ammonia beccarii), are essential in understanding past and present environmental conditions affecting the island's coastal environment. These taxa have been studied previously; however, factors explaining their individual patterns of abundance in the Sound are not well detailed. The goal of this study is …
Modeling Stochastically Intransitive Relationships In Paired Comparison Data, Ryan Patrick Alexander Mcshane
Modeling Stochastically Intransitive Relationships In Paired Comparison Data, Ryan Patrick Alexander Mcshane
Statistical Science Theses and Dissertations
If the Warriors beat the Rockets and the Rockets beat the Spurs, does that mean that the Warriors are better than the Spurs? Sophisticated fans would argue that the Warriors are better by the transitive property, but could Spurs fans make a legitimate argument that their team is better despite this chain of evidence?
We first explore the nature of intransitive (rock-scissors-paper) relationships with a graph theoretic approach to the method of paired comparisons framework popularized by Kendall and Smith (1940). Then, we focus on the setting where all pairs of items, teams, players, or objects have been compared to …
On Cluster Robust Models, José Bayoán Santiago Calderón
On Cluster Robust Models, José Bayoán Santiago Calderón
CGU Theses & Dissertations
Cluster robust models are a kind of statistical models that attempt to estimate parameters considering potential heterogeneity in treatment effects. Absent heterogeneity in treatment effects, the partial and average treatment effect are the same. When heterogeneity in treatment effects occurs, the average treatment effect is a function of the various partial treatment effects and the composition of the population of interest. The first chapter explores the performance of common estimators as a function of the presence of heterogeneity in treatment effects and other characteristics that may influence their performance for estimating average treatment effects. The second chapter examines various approaches …
Serial Testing For Detection Of Multilocus Genetic Interactions, Zaid T. Al-Khaledi
Serial Testing For Detection Of Multilocus Genetic Interactions, Zaid T. Al-Khaledi
Theses and Dissertations--Statistics
A method to detect relationships between disease susceptibility and multilocus genetic interactions is the Multifactor-Dimensionality Reduction (MDR) technique pioneered by Ritchie et al. (2001). Since its introduction, many extensions have been pursued to deal with non-binary outcomes and/or account for multiple interactions simultaneously. Studying the effects of multilocus genetic interactions on continuous traits (blood pressure, weight, etc.) is one case that MDR does not handle. Culverhouse et al. (2004) and Gui et al. (2013) proposed two different methods to analyze such a case. In their research, Gui et al. (2013) introduced the Quantitative Multifactor-Dimensionality Reduction (QMDR) that uses the overall …
Data Patterns Discovery Using Unsupervised Learning, Rachel A. Lewis
Data Patterns Discovery Using Unsupervised Learning, Rachel A. Lewis
College of Graduate Studies: Theses & Dissertations
Self-care activities classification poses significant challenges in identifying children’s unique functional abilities and needs within the exceptional children healthcare system. The accuracy of diagnosing a child's self-care problem, such as toileting or dressing, is highly influenced by an occupational therapists’ experience and time constraints. Thus, there is a need for objective means to detect and predict in advance the self-care problems of children with physical and motor disabilities. We use clustering to discover interesting information from self-care problems, perform automatic classification of binary data, and discover outliers. The advantages are twofold: the advancement of knowledge on identifying self-care problems in …
Non-Marginal Decisions: A Novel Bayesian Multiple Testing Procedure, Noirrit Kiran Chandra, Sourabh Bhattacharya
Non-Marginal Decisions: A Novel Bayesian Multiple Testing Procedure, Noirrit Kiran Chandra, Sourabh Bhattacharya
Journal Articles
In this paper, we consider the problem of multiple testing where the hypotheses are dependent. In most of the existing literature, either Bayesian or non-Bayesian, the decision rules mainly focus on the validity of the test procedure rather than actually utilizing the dependency to increase efficiency. Moreover, the decisions regarding different hypotheses are marginal in the sense that they do not depend upon each other directly. However, in realistic situations, the hypotheses are usually dependent, and hence it is desirable that the decisions regarding the dependent hypotheses are taken jointly. In this article, we develop a novel Bayesian multiple testing …
Bayesian Hierarchical Meta-Analysis Of Asymptomatic Ebola Seroprevalence, Peter Brody-Moore
Bayesian Hierarchical Meta-Analysis Of Asymptomatic Ebola Seroprevalence, Peter Brody-Moore
CMC Senior Theses
The continued study of asymptomatic Ebolavirus infection is necessary to develop a more complete understanding of Ebola transmission dynamics. This paper conducts a meta-analysis of eight studies that measure seroprevalence (the number of subjects that test positive for anti-Ebolavirus antibodies in their blood) in subjects with household exposure or known case-contact with Ebola, but that have shown no symptoms. In our two random effects Bayesian hierarchical models, we find estimated seroprevalences of 8.76% and 9.72%, significantly higher than the 3.3% found by a previous meta-analysis of these eight studies. We also produce a variation of this meta-analysis where we exclude …