Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- University of Kentucky (16)
- Virginia Commonwealth University (13)
- COBRA (10)
- University of Louisville (10)
- Southern Methodist University (9)
-
- Old Dominion University (7)
- Kennesaw State University (6)
- Michigan Technological University (6)
- Illinois State University (5)
- University of Texas Rio Grande Valley (5)
- University of Arkansas, Fayetteville (4)
- California Polytechnic State University, San Luis Obispo (3)
- Chapman University (3)
- Claremont Colleges (3)
- Stephen F. Austin State University (3)
- University at Albany, State University of New York (3)
- University of Montana (3)
- University of Nebraska - Lincoln (3)
- University of New Mexico (3)
- Clemson University (2)
- East Tennessee State University (2)
- Georgia Southern University (2)
- HCA Healthcare (2)
- The Texas Medical Center Library (2)
- University of Connecticut (2)
- University of Missouri, St. Louis (2)
- University of Nebraska Medical Center (2)
- Utah State University (2)
- Brigham Young University (1)
- Bucknell University (1)
- Keyword
-
- Statistics (19)
- Machine learning (7)
- Biostatistics (5)
- Regression (5)
- Meta-analysis (4)
-
- Bayesian (3)
- Bayesian statistics (3)
- Bioinformatics (3)
- Diabetes (3)
- Epidemiology (3)
- Logistic regression (3)
- Maximum likelihood (3)
- Simulation (3)
- Survival (3)
- Survival Analysis (3)
- Alzheimer's Disease (2)
- Biomarker (2)
- Breast cancer (2)
- Causal inference (2)
- Clinical trials (2)
- Copula (2)
- Cross Validation (2)
- EM algorithm (2)
- GWAS (2)
- Gene expression (2)
- Generalized linear models (2)
- Genetics (2)
- Influenza (2)
- Longitudinal (2)
- Machine Learning (2)
- Publication Year
- Publication
-
- Electronic Theses and Dissertations (16)
- Theses and Dissertations (15)
- Theses and Dissertations--Statistics (10)
- Statistical Science Theses and Dissertations (8)
- Dissertations, Master's Theses and Master's Reports (6)
-
- Mathematics & Statistics Theses & Dissertations (6)
- Theses and Dissertations--Epidemiology and Biostatistics (5)
- COBRA Preprint Series (4)
- Graduate Theses and Dissertations (4)
- Symposium of Student Scholars (4)
- Annual Symposium on Biomathematics and Ecology Education and Research (3)
- Electronic Theses & Dissertations (2024 - present) (3)
- Graduate Student Theses, Dissertations, & Professional Papers (3)
- All Dissertations (2)
- All Graduate Plan B and other Reports, Spring 1920 to Spring 2023 (2)
- CHIP Documents (2)
- CMC Senior Theses (2)
- College of Graduate Studies: Theses & Dissertations (2)
- Computational and Data Sciences (PhD) Dissertations (2)
- Dissertations and Theses (Open Access) (2)
- HCA Healthcare Journal of Medicine (2)
- Mathematics & Statistics ETDs (2)
- Research Symposium (2)
- School of Mathematical & Statistical Sciences Faculty Publications (2)
- Statistics (2)
- The University of Michigan Department of Biostatistics Working Paper Series (2)
- Theses & Dissertations (2)
- U.C. Berkeley Division of Biostatistics Working Paper Series (2)
- Articles (1)
- Biology and Medicine Through Mathematics Conference (1)
- Publication Type
- File Type
Articles 91 - 120 of 155
Full-Text Articles in Applied Statistics
Identifying Risk Factors Related To Premature Birth Through Binary Logistic And Proportional Odds Ordinal Logistic Regression, Clayton Elwood
Identifying Risk Factors Related To Premature Birth Through Binary Logistic And Proportional Odds Ordinal Logistic Regression, Clayton Elwood
Electronic Theses and Dissertations
Premature birth has been identified as the single greatest cause of death worldwide in children under the age of five. This thesis will implement binary logistic regression and proportional odds ordinal logistic regression to predict different levels of premature birth and identify associated risk factors. The models will be built from the Center for Disease Control and Prevention's 2014 Vital Statistics Natality Birth Data containing nearly 4 million live births within the United States. Odds ratios and confidence intervals on risk factors were produced utilizing binary logistic regression.
Spatio-Temporal Analysis Of Tree Ring Chronology And Precipitation, Ruizhe Yin
Spatio-Temporal Analysis Of Tree Ring Chronology And Precipitation, Ruizhe Yin
Graduate Theses and Dissertations
Tree ring chronology data is known to reflect regional climate due to the strong impact of rainfall and temperature. Therefore, tree ring data can be used to reconstruct historical climate in order to understand how climate changed in the past and make prediction about the future behavior of the climate. For simplicity, this research only considers the influence of precipitation on tree ring growth within the New England area. A total of 94 measurement sites are used to record tree ring width over 881 years and corresponding precipitation data are given at some locations for 121 years. We developed a …
Statin Prescription For Patients With Atherosclerotic Cardiovascular Disease From National Survey Data, Kristina Vatcheva, Vicente Aparicio, Ayesha Araya, Eduardo Gonzalez, Susan T. Laing
Statin Prescription For Patients With Atherosclerotic Cardiovascular Disease From National Survey Data, Kristina Vatcheva, Vicente Aparicio, Ayesha Araya, Eduardo Gonzalez, Susan T. Laing
School of Mathematical & Statistical Sciences Faculty Publications
Despite strong evidence for the use of statins for patients with atherosclerotic cardiovascular disease (ASCVD), statin prescription is still suboptimal. We aimed to determine the rates and factors that influence statin prescription using national survey data. This is a cross-sectional retrospective study on 8,468 patients with clinical ASCVD who were drawn from the National Ambulatory Medical Care Survey and the National Hospital Ambulatory Medical Care Survey from years 2011 to 2015. Survey-weighted analysis was conducted to estimate weighted prevalence and odds ratios for statin prescription. There was a significant increase in statin prescription from the years 2011 to 2015. Nevertheless, …
Copula-Based Zero-Inflated Count Time Series Models, Mohammed Sulaiman Alqawba
Copula-Based Zero-Inflated Count Time Series Models, Mohammed Sulaiman Alqawba
Mathematics & Statistics Theses & Dissertations
Count time series data are observed in several applied disciplines such as in environmental science, biostatistics, economics, public health, and finance. In some cases, a specific count, say zero, may occur more often than usual. Additionally, serial dependence might be found among these counts if they are recorded over time. Overlooking the frequent occurrence of zeros and the serial dependence could lead to false inference. In this dissertation, we propose two classes of copula-based time series models for zero-inflated counts with the presence of covariates. Zero-inflated Poisson (ZIP), zero-inflated negative binomial (ZINB), and zero-inflated Conway-Maxwell-Poisson (ZICMP) distributed marginals of the …
Spatio-Temporal Cluster Detection And Local Moran Statistics Of Point Processes, Jennifer L. Matthews
Spatio-Temporal Cluster Detection And Local Moran Statistics Of Point Processes, Jennifer L. Matthews
Mathematics & Statistics Theses & Dissertations
Moran's index is a statistic that measures spatial dependence, quantifying the degree of dispersion or clustering of point processes and events in some location/area. Recognizing that a single Moran's index may not give a sufficient summary of the spatial autocorrelation measure, a local indicator of spatial association (LISA) has gained popularity. Accordingly, we propose extending LISAs to time after partitioning the area and computing a Moran-type statistic for each subarea. Patterns between the local neighbors are unveiled that would not otherwise be apparent. We consider the measures of Moran statistics while incorporating a time factor under simulated multilevel Palm distribution, …
Tobacco Smoking And Dementia In A Kentucky Cohort: A Competing Risk Analysis, Erin L. Abner, Peter T. Nelson, Gregory A. Jicha, Gregory E. Cooper, David W. Fardo, Frederick A. Schmitt, Richard J. Kryscio
Tobacco Smoking And Dementia In A Kentucky Cohort: A Competing Risk Analysis, Erin L. Abner, Peter T. Nelson, Gregory A. Jicha, Gregory E. Cooper, David W. Fardo, Frederick A. Schmitt, Richard J. Kryscio
Epidemiology and Environmental Health Faculty Publications
Tobacco smoking was examined as a risk for dementia and neuropathological burden in 531 initially cognitively normal older adults followed longitudinally at the University of Kentucky’s Alzheimer’s Disease Center. The cohort was followed for an average of 11.5 years; 111 (20.9%) participants were diagnosed with dementia, while 242 (45.6%) died without dementia. At baseline, 49 (9.2%) participants reported current smoking (median pack-years = 47.3) and 231 (43.5%) former smoking (median pack-years = 24.5). The hazard ratio (HR) for dementia for former smokers versus never smokers based on the Cox model was 1.64 (95% CI: 1.09, 2.46), while the HR for …
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes, Lauren Spirko-Burns, Karthik Devarajan
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes, Lauren Spirko-Burns, Karthik Devarajan
COBRA Preprint Series
One of the major goals in large-scale genomic studies is to identify genes with a prognostic impact on time-to-event outcomes which provide insight into the disease's process. With rapid developments in high-throughput genomic technologies in the past two decades, the scientific community is able to monitor the expression levels of tens of thousands of genes and proteins resulting in enormous data sets where the number of genomic features is far greater than the number of subjects. Methods based on univariate Cox regression are often used to select genomic features related to survival outcome; however, the Cox model assumes proportional hazards …
Data Analytics Pipeline For Rna Structure Analysis Via Shape, Quinn Nelson
Data Analytics Pipeline For Rna Structure Analysis Via Shape, Quinn Nelson
UNO Student Research and Creative Activity Fair
Coxsackievirus B3 (CVB3) is a cardiovirulent enterovirus from the family Picornaviridae. The RNA genome houses an internal ribosome entry site (IRES) in the 5’ untranslated region (5’UTR) that enables cap-independent translation. Ample evidence suggests that the structure of the 5’UTR is a critical element for virulence. We probe RNA structure in solution using base-specific modifying agents such as dimethyl sulfate as well as backbone targeting agents such as N-methylisatoic anhydride used in Selective 2’-Hydroxyl Acylation Analyzed by Primer Extension (SHAPE). We have developed a pipeline that merges and evaluates base-specific and SHAPE data together with statistical analyses that provides confidence …
Controlling For Confounding Via Propensity Score Methods Can Result In Biased Estimation Of The Conditional Auc: A Simulation Study, Hadiza I. Galadima, Donna K. Mcclish
Controlling For Confounding Via Propensity Score Methods Can Result In Biased Estimation Of The Conditional Auc: A Simulation Study, Hadiza I. Galadima, Donna K. Mcclish
Community & Environmental Health Faculty Publications
In the medical literature, there has been an increased interest in evaluating association between exposure and outcomes using nonrandomized observational studies. However, because assignments to exposure are not random in observational studies, comparisons of outcomes between exposed and nonexposed subjects must account for the effect of confounders. Propensity score methods have been widely used to control for confounding, when estimating exposure effect. Previous studies have shown that conditioning on the propensity score results in biased estimation of conditional odds ratio and hazard ratio. However, research is lacking on the performance of propensity score methods for covariate adjustment when estimating the …
Variable Selection In Accelerated Failure Time (Aft) Frailty Models: An Application Of Penalized Quasi-Likelihood, Sarbesh R. Pandeya
Variable Selection In Accelerated Failure Time (Aft) Frailty Models: An Application Of Penalized Quasi-Likelihood, Sarbesh R. Pandeya
College of Graduate Studies: Theses & Dissertations
Variable selection is one of the standard ways of selecting models in large scale datasets. It has applications in many fields of research study, especially in large multi-center clinical trials. One of the prominent methods in variable selection is the penalized likelihood, which is both consistent and efficient. However, the penalized selection is significantly challenging under the influence of random (frailty) covariates. It is even more complicated when there is involvement of censoring as it may not have a closed-form solution for the marginal log-likelihood. Therefore, we applied the penalized quasi-likelihood (PQL) approach that approximates the solution for such a …
Serial Testing For Detection Of Multilocus Genetic Interactions, Zaid T. Al-Khaledi
Serial Testing For Detection Of Multilocus Genetic Interactions, Zaid T. Al-Khaledi
Theses and Dissertations--Statistics
A method to detect relationships between disease susceptibility and multilocus genetic interactions is the Multifactor-Dimensionality Reduction (MDR) technique pioneered by Ritchie et al. (2001). Since its introduction, many extensions have been pursued to deal with non-binary outcomes and/or account for multiple interactions simultaneously. Studying the effects of multilocus genetic interactions on continuous traits (blood pressure, weight, etc.) is one case that MDR does not handle. Culverhouse et al. (2004) and Gui et al. (2013) proposed two different methods to analyze such a case. In their research, Gui et al. (2013) introduced the Quantitative Multifactor-Dimensionality Reduction (QMDR) that uses the overall …
Bayesian Hierarchical Meta-Analysis Of Asymptomatic Ebola Seroprevalence, Peter Brody-Moore
Bayesian Hierarchical Meta-Analysis Of Asymptomatic Ebola Seroprevalence, Peter Brody-Moore
CMC Senior Theses
The continued study of asymptomatic Ebolavirus infection is necessary to develop a more complete understanding of Ebola transmission dynamics. This paper conducts a meta-analysis of eight studies that measure seroprevalence (the number of subjects that test positive for anti-Ebolavirus antibodies in their blood) in subjects with household exposure or known case-contact with Ebola, but that have shown no symptoms. In our two random effects Bayesian hierarchical models, we find estimated seroprevalences of 8.76% and 9.72%, significantly higher than the 3.3% found by a previous meta-analysis of these eight studies. We also produce a variation of this meta-analysis where we exclude …
Methods For Evaluating Dropout Attrition In Survey Data, Camille J. Hochheimer
Methods For Evaluating Dropout Attrition In Survey Data, Camille J. Hochheimer
Theses and Dissertations
As researchers increasingly use web-based surveys, the ease of dropping out in the online setting is a growing issue in ensuring data quality. One theory is that dropout or attrition occurs in phases that can be generalized to phases of high dropout and phases of stable use. In order to detect these phases, several methods are explored. First, existing methods and user-specified thresholds are applied to survey data where significant changes in the dropout rate between two questions is interpreted as the start or end of a high dropout phase. Next, survey dropout is considered as a time-to-event outcome and …
Genome-Wide Systems Genetics Of Alcohol Consumption And Dependence, Kristin Mignogna
Genome-Wide Systems Genetics Of Alcohol Consumption And Dependence, Kristin Mignogna
Theses and Dissertations
Widely effective treatment for alcohol use disorder is not yet available, because the exact biological mechanisms that underlie this disorder are not completely understood. One way to gain a better understanding of these mechanisms is to examine the genetic frameworks that contribute to the risk for developing this disorder. This dissertation examines genetic association data in combination with gene expression networks in the brain to identify functional groups of genes associated with alcohol consumption and dependence.
The first study took advantage of the behavioral complexity of human samples, and experimental capabilities provided by mouse models, by co-analyzing gene expression networks …
Statistical Methods For Joint Analysis Of Multiple Phenotypes And Their Applications For Phewas, Xueling Li
Statistical Methods For Joint Analysis Of Multiple Phenotypes And Their Applications For Phewas, Xueling Li
Dissertations, Master's Theses and Master's Reports
Genome-wide association studies (GWAS) have successfully detected tens of thousands of robust SNP-trait associations. Earlier researches have primarily focused on association studies of genetic variants and some well-defined functions or phenotypic traits. Emerging evidence suggests that pleiotropy, the phenomenon of one genetic variant affects multiple phenotypes, is widespread, especially in complex human diseases. Therefore, individual phenotype analyses may lose statistical power to identify the underlying genetic mechanism. Contrasting with single phenotype analyses, joint analysis of multiple phenotypes exploits the correlations between phenotypes and aggregates multiple weak marginal effects and is therefore likely to provide new insights into the functional consequences …
Statistical Modeling Of Influenza-Like-Illness In Montana Using Spatial And Temporal Methods, Benjamin A. Stark
Statistical Modeling Of Influenza-Like-Illness In Montana Using Spatial And Temporal Methods, Benjamin A. Stark
Graduate Student Theses, Dissertations, & Professional Papers
Studying air pollution and public health has been a historically important question in science. It has long been hypothesized that severe air pollution conditions lead to negative implications in basic human health. Primarily, areas thats are prone to severe degrees of human pollution are the focus of such studies. Such research relating to less populated areas are scarce, and this scarcity raises the question of how such pollution dynamics (human-made and natural) influence human health in more rural areas.
The aim of this study is to explore this hole in research; in particular we explore possible links between air pollution …
A Generative Statistical Approach For Data Classification In A Biologically Inspired Design Tool, Marvin Manuel Arroyo Rujano
A Generative Statistical Approach For Data Classification In A Biologically Inspired Design Tool, Marvin Manuel Arroyo Rujano
Graduate Theses and Dissertations
The objective of the research this thesis describes is to find a way to classify text-based descriptions of biological adaption to support Biologically Inspired design. Biologically inspired design is a fairly new field with ongoing research. There are different tools to assist designers and biologists in bio-inspired design. Some of the most common are BioTRIZ and AskNature. In recent years, more tools have been proposed to aid and make research in the field easier, for example, the Biologically Inspired Adaptive System Design (BIASD) tool. This tool was designed with the goal of helping designers in early design stages generate more …
Spatio-Temporal Reconstruction Of Remote Sensing Observations, Kamrul Khan
Spatio-Temporal Reconstruction Of Remote Sensing Observations, Kamrul Khan
Graduate Theses and Dissertations
The USDA Forest Service aims to use satellite imagery for monitoring and predicting changes in forest conditions over time within the country. We specifically focus on a 230, 400 hectares region in north-central Wisconsin between 2003 - 2012. The auxiliary data collected from the satellite imagery of this region are relatively dense in space and time and can be used to efficiently predict how the forest condition changed over that decade. However, these records have a significant proportion of missing values due to weather conditions and system failures. To fill in these missing values, we build spaciotemporal models based on …
Analysis Of Ranked Gene Tree Probability Distributions Under The Coalescent Process For Detecting Anomaly Zones, Anastasiia Kim
Analysis Of Ranked Gene Tree Probability Distributions Under The Coalescent Process For Detecting Anomaly Zones, Anastasiia Kim
Shared Knowledge Conference
In phylogenetic studies, gene trees are used to reconstruct species tree. Under the multispecies coalescent model, gene trees topologies may differ from that of species trees. The incorrect gene tree topology (one that does not match the species tree) that is more probable than the correct one is termed anomalous gene tree (AGT). Species trees that can generate such AGTs are said to be in the anomaly zone (AZ). In this region, the method of choosing the most common gene tree as the estimate of the species tree will be inconsistent and will converge to an incorrect species tree when …
Bayesian Analytical Approaches For Metabolomics : A Novel Method For Molecular Structure-Informed Metabolite Interaction Modeling, A Novel Diagnostic Model For Differentiating Myocardial Infarction Type, And Approaches For Compound Identification Given Mass Spectrometry Data., Patrick J. Trainor
Electronic Theses and Dissertations
Metabolomics, the study of small molecules in biological systems, has enjoyed great success in enabling researchers to examine disease-associated metabolic dysregulation and has been utilized for the discovery biomarkers of disease and phenotypic states. In spite of recent technological advances in the analytical platforms utilized in metabolomics and the proliferation of tools for the analysis of metabolomics data, significant challenges in metabolomics data analyses remain. In this dissertation, we present three of these challenges and Bayesian methodological solutions for each. In the first part we develop a new methodology to serve a basis for making higher order inferences in metabolomics, …
Developing Statistical Methods For Data From Platforms Measuring Gene Expression, Gaoxiang Jia
Developing Statistical Methods For Data From Platforms Measuring Gene Expression, Gaoxiang Jia
Statistical Science Theses and Dissertations
This research contains two topics: (1) PBNPA: a permutation-based non-parametric analysis of CRISPR screen data; (2) RCRnorm: an integrated system of random-coefficient hierarchical regression models for normalizing NanoString nCounter data from FFPE samples.
Clustered regularly-interspaced short palindromic repeats (CRISPR) screens are usually implemented in cultured cells to identify genes with critical functions. Although several methods have been developed or adapted to analyze CRISPR screening data, no single spe- cific algorithm has gained popularity. Thus, rigorous procedures are needed to overcome the shortcomings of existing algorithms. We developed a Permutation-Based Non-Parametric Analysis (PBNPA) algorithm, which computes p-values at the gene level …
The Impact Of Truncating Data On The Predictive Ability For Single-Step Genomic Best Linear Unbiased Prediction, Jeremy T. Howard, Thomas A. Rathje, Caitlyn E. Bruns, Danielle F. Wilson-Wells, Stephen D. Kachman, Matthew L. Spangler
The Impact Of Truncating Data On The Predictive Ability For Single-Step Genomic Best Linear Unbiased Prediction, Jeremy T. Howard, Thomas A. Rathje, Caitlyn E. Bruns, Danielle F. Wilson-Wells, Stephen D. Kachman, Matthew L. Spangler
Department of Animal Science: Faculty Publications
Simulated and swine industry data sets were utilized to assess the impact of removing older data on the predictive ability of selection candidate estimated breeding values (EBV) when using single-step genomic best linear unbiased prediction (ssGBLUP). Simulated data included thirty replicates designed to mimic the structure of swine data sets. For the simulated data, varying amounts of data were truncated based on the number of ancestral generations back from the selection candidates. The swine data sets consisted of phenotypic and genotypic records for three traits across two breeds on animals born from 2003 to 2017. Phenotypes and genotypes were iteratively …
Improved Methods And Selecting Classification Types For Time-Dependent Covariates In The Marginal Analysis Of Longitudinal Data, I-Chen Chen
Theses and Dissertations--Epidemiology and Biostatistics
Generalized estimating equations (GEE) are popularly utilized for the marginal analysis of longitudinal data. In order to obtain consistent regression parameter estimates, these estimating equations must be unbiased. However, when certain types of time-dependent covariates are presented, these equations can be biased unless an independence working correlation structure is employed. Moreover, in this case regression parameter estimation can be very inefficient because not all valid moment conditions are incorporated within the corresponding estimating equations. Therefore, approaches using the generalized method of moments or quadratic inference functions have been proposed for utilizing all valid moment conditions. However, we have found that …
Improved Standard Error Estimation For Maintaining The Validities Of Inference In Small-Sample Cluster Randomized Trials And Longitudinal Studies, Whitney Ford Tanner
Improved Standard Error Estimation For Maintaining The Validities Of Inference In Small-Sample Cluster Randomized Trials And Longitudinal Studies, Whitney Ford Tanner
Theses and Dissertations--Epidemiology and Biostatistics
Data arising from Cluster Randomized Trials (CRTs) and longitudinal studies are correlated and generalized estimating equations (GEE) are a popular analysis method for correlated data. Previous research has shown that analyses using GEE could result in liberal inference due to the use of the empirical sandwich covariance matrix estimator, which can yield negatively biased standard error estimates when the number of clusters or subjects is not large. Many techniques have been presented to correct this negative bias; However, use of these corrections can still result in biased standard error estimates and thus test sizes that are not consistently at their …
Penalized Mixed-Effects Ordinal Response Models For High-Dimensional Genomic Data In Twins And Families, Amanda E. Gentry
Penalized Mixed-Effects Ordinal Response Models For High-Dimensional Genomic Data In Twins And Families, Amanda E. Gentry
Theses and Dissertations
The Brisbane Longitudinal Twin Study (BLTS) was being conducted in Australia and was funded by the US National Institute on Drug Abuse (NIDA). Adolescent twins were sampled as a part of this study and surveyed about their substance use as part of the Pathways to Cannabis Use, Abuse and Dependence project. The methods developed in this dissertation were designed for the purpose of analyzing a subset of the Pathways data that includes demographics, cannabis use metrics, personality measures, and imputed genotypes (SNPs) for 493 complete twin pairs (986 subjects.) The primary goal was to determine what combination of SNPs and …
Joint Analysis Of Multiple Phenotypes In Association Studies, Xiaoyu Liang
Joint Analysis Of Multiple Phenotypes In Association Studies, Xiaoyu Liang
Dissertations, Master's Theses and Master's Reports
Genome-wide association studies (GWAS) have become a very effective research tool to identify genetic variants of underlying various complex diseases. In spite of the success of GWAS in identifying thousands of reproducible associations between genetic variants and complex disease, in general, the association between genetic variants and a single phenotype is usually weak. It is increasingly recognized that joint analysis of multiple phenotypes can be potentially more powerful than the univariate analysis, and can shed new light on underlying biological mechanisms of complex diseases. Therefore, developing statistical methods to test for genetic association with multiple phenotypes has become increasingly important. …
Step-Selection Functions For Modeling Animal Movement -- Case Study: African Buffalo, Maia Adar
Step-Selection Functions For Modeling Animal Movement -- Case Study: African Buffalo, Maia Adar
CMC Senior Theses
Understanding what factors influence wildlife movement allows landscape planners to make informed decisions that benefit both animals and humans. New quantitative methods, such as step-selection functions, provide valuable objective analyses of wildlife connectivity. This paper provides a framework for creating a step-selection function and demonstrates its use in a case study. The first section provides a general introduction about wildlife connectivity research. The second section explains the math behind the step-selection function using a simple example. The last section gives the results of a step-selection model for African buffalo in the Kavango Zambezi Transfrontier Conservation Area. Buffalo were found to …
Application Of Support Vector Machine Modeling And Graph Theory Metrics For Disease Classification, Jessica M. Rudd
Application Of Support Vector Machine Modeling And Graph Theory Metrics For Disease Classification, Jessica M. Rudd
Published and Grey Literature from PhD Candidates
Disease classification is a crucial element of biomedical research. Recent studies have demonstrated that machine learning techniques, such as Support Vector Machine (SVM) modeling, produce similar or improved predictive capabilities in comparison to the traditional method of Logistic Regression. In addition, it has been found that social network metrics can provide useful predictive information for disease modeling. In this study, we combine simulated social network metrics with SVM to predict diabetes in a sample of data from the Behavioral Risk Factor Surveillance System. In this dataset, Logistic Regression outperformed SVM with ROC index of 81.8 and 81.7 for models with …
Burden Of Atopic Dermatitis In The United States: Analysis Of Healthcare Claims Data In The Commercial, Medicare, And Medi-Cal Databases, Sulena Shrestha, Raymond Miao, Li Wang, Jingdong Chao, Huseyin Yuce, Wenhui Wei
Burden Of Atopic Dermatitis In The United States: Analysis Of Healthcare Claims Data In The Commercial, Medicare, And Medi-Cal Databases, Sulena Shrestha, Raymond Miao, Li Wang, Jingdong Chao, Huseyin Yuce, Wenhui Wei
Publications and Research
Comparative data on the burden of atopic dermatitis (AD) in adults relative to the general population are limited. We performed a large-scale evaluation of the burden of disease among US adults with AD relative to matched non-AD controls, encompassing comorbidities, healthcare resource utilization (HCRU), and costs, using healthcare claims data. The impact of AD disease severity on these outcomes was also evaluated.
Comparison Of Survival Curves Between Cox Proportional Hazards, Random Forests, And Conditional Inference Forests In Survival Analysis, Brandon Weathers
Comparison Of Survival Curves Between Cox Proportional Hazards, Random Forests, And Conditional Inference Forests In Survival Analysis, Brandon Weathers
All Graduate Plan B and other Reports, Spring 1920 to Spring 2023
Survival analysis methods are a mainstay of the biomedical fields but are finding increasing use in other disciplines including finance and engineering. A widely used tool in survival analysis is the Cox proportional hazards regression model. For this model, all the predicted survivor curves have the same basic shape, which may not be a good approximation to reality. In contrast the Random Survival Forests does not make the proportional hazards assumption and has the flexibility to model survivor curves that are of quite different shapes for different groups of subjects. We applied both techniques to a number of publicly available …