Statistical Downscaling With Spatial
Misalignment: Application To Wildland Fire
Pm2.5 Concentration Forecasting,
2020
North Carolina State University
Statistical Downscaling With Spatial Misalignment: Application To Wildland Fire Pm2.5 Concentration Forecasting, Suman Majumder, Yawen Guan, Brian J. Reich, Susan O’Neill, Ana G. Rappold
Department of Statistics: Faculty Publications
Fine particulate matter, PM2.5, has been documented to have adverse health effects, and wildland fires are a major contributor to PM2.5 air pollution in the USA. Forecasters use numerical models to predict PM2.5 concentrations to warn the public of impending health risk. Statistical methods are needed to calibrate the numerical model forecast using monitor data to reduce bias and quantify uncertainty. Typical model calibration techniques do not allow for errors due to misalignment of geographic locations. We propose a spatiotemporal downscaling methodology that uses image registration techniques to identify the spatial misalignment and accounts for and …
The Effects Of Adverse Childhood Experiences On Behavioral Outcomes,
2020
University of Denver
The Effects Of Adverse Childhood Experiences On Behavioral Outcomes, Jennifer Thomas
Electronic Theses and Dissertations
This study intends to explore the intersection of two vulnerable populations, early childhood development and risks associated with exposure to adverse childhood experiences (ACEs). This study examines how age plays a role in the long-term relationship between ACEs and internal and external behaviors. This study seeks to answer the question of: How does age influence the relationship between number of ACEs and internal and external behaviors? The participants in this study include those aged 0 – 16 from the National Survey of Child and adolescent Well-Being (NSCAW) dataset. The NSCAW study consists of five waves of data where Wave I …
Comprehensive Research Synthesis: An Approach To Mixed Methods Research Syntheses,
2020
University of Denver
Comprehensive Research Synthesis: An Approach To Mixed Methods Research Syntheses, Lilian Linialy Chimuma
Electronic Theses and Dissertations
Mixed methods research synthesis (MMRS) is an emerging application of both mixed-methods research (MMR) and review research. MMRS promises to comprehensively address intricate contemporary research and evaluation questions given diverse evidence sources (across quantitative, qualitative, and MMR primary studies). The significance of concurrently addressing methodological issues for new research developments is widely noted in the literature. Current efforts attempt to streamline methodological practices along with application of the MMRS approach. Researchers have proposed conceptual frameworks to guide the application and practice of MMRS studies. Despite these efforts, complications and disagreements persist. In response to these concerns, this study developed a …
Evaluating An Ordinal Output Using Data Modeling, Algorithmic Modeling, And Numerical Analysis,
2020
Murray State University
Evaluating An Ordinal Output Using Data Modeling, Algorithmic Modeling, And Numerical Analysis, Martin Keagan Wynne Brown
Murray State Theses and Dissertations
Data and algorithmic modeling are two different approaches used in predictive analytics. The models discussed from these two approaches include the proportional odds logit model (POLR), the vector generalized linear model (VGLM), the classification and regression tree model (CART), and the random forests model (RF). Patterns in the data were analyzed using trigonometric polynomial approximations and Fast Fourier Transforms. Predictive modeling is used frequently in statistics and data science to find the relationship between the explanatory (input) variables and a response (output) variable. Both approaches prove advantageous in different cases depending on the data set. In our case, the data …
Aggregate Loss Model With Poisson-Tweedie Loss Frequency,
2020
Wilfrid Laurier University
Aggregate Loss Model With Poisson-Tweedie Loss Frequency, Si Chen
Theses and Dissertations (Comprehensive)
The aggregate loss model has applications in various areas such as financial risk management and actuarial science. The aggregate loss is the summation of all random losses occurred in a period, and it is governed by both the loss severity and the loss frequency. While the impact of the loss severity on aggregate loss is well studied, less focus is paid on the influence of loss frequency on aggregate loss, which motivates our study. In this thesis, we enrich the aggregate loss framework by introducing the Poisson-Tweedie distribution as a candidate for modelling loss frequency, prove the closedness of Poisson-Tweedie …
Identifying Customer Churn In After-Market Operations Using Machine Learning Algorithms,
2019
Southern Methodist University
Identifying Customer Churn In After-Market Operations Using Machine Learning Algorithms, Vitaly Briker, Richard Farrow, William Trevino, Brent Allen
SMU Data Science Review
This paper presents a comparative study on machine learning methods as they are applied to product associations, future purchase predictions, and predictions of customer churn in aftermarket operations. Association rules are used help to identify patterns across products and find correlations in customer purchase behaviour. Studying customer behaviour as it pertains to Recency, Frequency, and Monetary Value (RFM) helps inform customer segmentation and identifies customers with propensity to churn. Lastly, Flowserve’s customer purchase history enables the establishment of churn thresholds for each customer group and assists in constructing a model to predict future churners. The aim of this model is …
Ordinal Hyperplane Loss,
2019
Kennesaw State University
Ordinal Hyperplane Loss, Bob Vanderheyden
Doctor of Data Science and Analytics Dissertations
This research presents the development of a new framework for analyzing ordered class data, commonly called “ordinal class” data. The focus of the work is the development of classifiers (predictive models) that predict classes from available data. Ratings scales, medical classification scales, socio-economic scales, meaningful groupings of continuous data, facial emotional intensity and facial age estimation are examples of ordinal data for which data scientists may be asked to develop predictive classifiers. It is possible to treat ordinal classification like any other classification problem that has more than two classes. Specifying a model with this strategy does not fully utilize …
Implications Of The Modifiable Areal Unit Problem For Wildfire Analyses,
2019
University of New Mexico
Implications Of The Modifiable Areal Unit Problem For Wildfire Analyses, Timothy P. Nagle-Mcnaughton, Xi Gong, Jose A. Constantine
Geography and Environmental Studies Faculty Publications
Wildfires pose a danger to both ecologies and communities. To this end, many large-scale analyses of wildfire patterns and behavior rely on the aggregation of point data to polygons, typically those based on distinct disparate ecological areas. However, the sizes, shapes, andorientations of the polygons to which data are aggregated are not neutral factors in the resulting analysis. The influence of the aggregation polygons on calculated results is known as the modifiable areal unit problem (MAUP), which is well-documented in the spatial statistics literature. Despite the documentation of the MAUP, relatively few wildfire studies consider the effects of the MAUP …
Machine Learning In Support Of Electric Distribution Asset Failure Prediction,
2019
Southern Methodist University
Machine Learning In Support Of Electric Distribution Asset Failure Prediction, Robert D. Flamenbaum, Thomas Pompo, Christopher Havenstein, Jade Thiemsuwan
SMU Data Science Review
In this paper, we present novel approaches to predicting as- set failure in the electric distribution system. Failures in overhead power lines and their associated equipment in particular, pose significant finan- cial and environmental threats to electric utilities. Electric device failure furthermore poses a burden on customers and can pose serious risk to life and livelihood. Working with asset data acquired from an electric utility in Southern California, and incorporating environmental and geospatial data from around the region, we applied a Random Forest methodology to predict which overhead distribution lines are most vulnerable to fail- ure. Our results provide evidence …
Is Corequisite Developmental Math Effective At East Tennessee State University?,
2019
East Tennessee State University
Is Corequisite Developmental Math Effective At East Tennessee State University?, Christine Padden
Electronic Theses and Dissertations
This thesis looks at the corequisite developmental math program at East Tennessee State University (ETSU) and compares the effectiveness to the previous developmental math program by comparing the student outcomes in MATH 1530. MATH 1530 is a non-calculus based statistic and probability course that satisfies most majors’ general education math requirements. ETSU sees approximately 1,000 students a year pass through MATH 1530 which is around 6.7% of the total enrollment at ETSU[9]. We are interested in the last five years of the developmental math program before it was changed to corequisite developmental math and the first five years of corequisite …
Mathematics Versus Statistics,
2019
Valparaiso University
Mathematics Versus Statistics, Mindy B. Capaldi
Journal of Humanistic Mathematics
Mathematics and statistics are both important and useful subjects, but the former has maintained prominence in the American education system. On the other hand, statistics is more prevalent in daily life and is an increasingly marketable subject to know. This article gives a personal history of one mathematician’s bumpy road to learning and teaching statistics. Additionally, arguments for how and why to include statistics in the K-12 and college curricula are provided.
Probabilistic Modeling Of Personalized Drug
Combinations From Integrated Chemical
Screen And Molecular Data In Sarcoma,
2019
Children's Cancer Therapy Development Institut
Probabilistic Modeling Of Personalized Drug Combinations From Integrated Chemical Screen And Molecular Data In Sarcoma, Noah E. Berlow, Rishi Rikhi, Mathew Geltzeiler, Jinu Abraham, Matthew N. Svalina, Lara E. Davis, Erin Wise, Maria Mancini, Jonathan Noujaim, Atiya Mansoor, Michael J. Quist, Kevin L. Matlock, Martin W. Goros, Brian S. Hernandez, Yee C. Doung, Khin Thway, Tomohide Tsukahara, Jun Nishio, Elaine T. Huang, Susan Airhart, Carol J. Bult, Regina Gandour-Edwards, Robert G. Maki, Robin L. Jones, Joel E. Michalek, Milan Milovancev, Souparno Ghosh, Ranadip Pal, Charles Keller
Department of Statistics: Faculty Publications
Background: Cancer patients with advanced disease routinely exhaust available clinical regimens and lack actionable genomic medicine results, leaving a large patient population without effective treatments options when their disease inevitably progresses. To address the unmet clinical need for evidence-based therapy assignment when standard clinical approaches have failed, we have developed a probabilistic computational modeling approach which integrates molecular sequencing data with functional assay data to develop patient-specific combination cancer treatments. Methods: Tissue taken from a murine model of alveolar rhabdomyosarcoma was used to perform single agent drug screening and DNA/RNA sequencing experiments; results integrated via our computational modeling approach identified …
Cs + Sociology: Using Big Data To Identify And Understand Educational Inequality In America (1),
2019
CUNY Lehman College
Cs + Sociology: Using Big Data To Identify And Understand Educational Inequality In America (1), Joseph Cleary, Elin Waring
Open Educational Resources
This is the first of two lessons/labs for teaching and learning of computer science and sociology. Either and be used on their own or they can be used in sequence, in which case this should be used first.
Students will develop CS skills and behaviors including but not limited to: learning what an API is, learning how to access and utilize data on an API, and developing their R coding skills and knowledge. Students will also learn basic, but important, sociological principles such as how poverty is related to educational opportunities in America. Although prior knowledge of CS and sociology …
Development Of A School Boredom Proneness Scale For Children,
2019
James Madison University
Development Of A School Boredom Proneness Scale For Children, Taylor Carrington
Educational Specialist, 2009-2019
One common phrase heard from students is, “I’m bored.” However, there is no real understanding of what this actually means. In this study, elementary-age students were asked to respond to a newly developed School Boredom Proneness Scale (SBPS) including questions relating to a five-factor model of boredom. Students were also asked to rate how often they become bored at school and how bored they seem compared to classmates. In addition to student responses, parents and teachers were asked to rate how bored they thought the student was, and teachers were additionally asked to rate students’ level of work completion. The …
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes,
2019
Temple University
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes, Lauren Spirko-Burns, Karthik Devarajan
COBRA Preprint Series
One of the major goals in large-scale genomic studies is to identify genes with a prognostic impact on time-to-event outcomes which provide insight into the disease's process. With rapid developments in high-throughput genomic technologies in the past two decades, the scientific community is able to monitor the expression levels of tens of thousands of genes and proteins resulting in enormous data sets where the number of genomic features is far greater than the number of subjects. Methods based on univariate Cox regression are often used to select genomic features related to survival outcome; however, the Cox model assumes proportional hazards …
Informative Group Testing For Multiplex Assays,
2019
University of Nebraska - Lincoln
Informative Group Testing For Multiplex Assays, Christopher R. Bilder, Joshua M. Tebbs, Christopher S. Mcmahan
Department of Statistics: Faculty Publications
Infectious disease testing frequently takes advantage of two tools–group testing and multiplex assays–to make testing timely and cost effective. Until the work of Tebbs et al. (2013) and Hou et al. (2017), there was no research available to understand how best to apply these tools simultaneously. This recent work focused on applications where each individual is considered to be identical in terms of the probability of disease. However, risk-factor information, such as past behavior and presence of symptoms, is very often available on each individual to allow one to estimate individual-specific probabilities. The purpose of our paper is to propose …
Functional Random Forest With
Applications In Dose-Response
Predictions,
2019
Texas Tech University
Functional Random Forest With Applications In Dose-Response Predictions, Raziur Rahman, Saugato Rahman Dhruba, Souparno Ghosh, Ranadip Pal
Department of Statistics: Faculty Publications
Drug sensitivity prediction for individual tumors is a significant challenge in personalized medicine. Current modeling approaches consider prediction of a single metric of the drug response curve such as AUC or IC50. However, the single summary metric of a dose-response curve fails to provide the entire drug sensitivity profile which can be used to design the optimal dose for a patient. In this article, we assess the problem of predicting the complete dose-response curve based on genetic characterizations. We propose an enhancement to the popular ensemble-based Random Forests approach that can directly predict the entire functional profile of …
Cost-Effective Surveillance For Infectious Diseases Through Specimen Pooling And Multiplex Assays,
2019
University of Nebraska - Lincoln
Cost-Effective Surveillance For Infectious Diseases Through Specimen Pooling And Multiplex Assays, Christopher Bilder, Joshua Tebbs, Christopher Mcmahan
Department of Statistics: Faculty Publications
To develop specimen pooling algorithms that reduce the number of tests needed to test individuals for infectious diseases with multiplex assays.
Genomic Prediction Using Canopy Coverage Image
And Genotypic Information In Soybean Via A Hybrid
Model,
2019
University of Nebraska-Lincoln
Genomic Prediction Using Canopy Coverage Image And Genotypic Information In Soybean Via A Hybrid Model, Reka Howard, Diego Jarquin
Department of Statistics: Faculty Publications
Prediction techniques are important in plant breeding as they provide a tool for selection that is more efficient and economical than traditional phenotypic and pedigree based selection. The conventional genomic prediction models include molecular marker information to predict the phenotype. With the development of new phenomics techniques we have the opportunity to collect image data on the plants, and extend the traditional genomic prediction models where we incorporate diverse set of information collected on the plants. In our research, we developed a hybrid matrix model that incorporates molecular marker and canopy coverage information as a weighted linear combination to predict …
Post-Er Stress Biogenesis Of Golgi Is Governed By Giantin,
2019
University of Nebraska Medical Center
Post-Er Stress Biogenesis Of Golgi Is Governed By Giantin, Cole P. Frisbie, Alexander Y. Lushnikov, Alexey V. Krasnoslobodtsev, Jean-Jack Riethoven, Jennifer L. Clarke, Elena I. Stepchenkova, Armen Petrosyan
Department of Statistics: Faculty Publications
Background: The Golgi apparatus undergoes disorganization in response to stress, but it is able to restore compact and perinuclear structure under recovery. This self-organization mechanism is significant for cellular homeostasis, but remains mostly elusive, as does the role of giantin, the largest Golgi matrix dimeric protein. Methods: In HeLa and different prostate cancer cells, we used the model of cellular stress induced by Brefeldin A (BFA). The conformational structure of giantin was assessed by proximity ligation assay and atomic force microscopy. The post-BFA distribution of Golgi resident enzymes was examined by 3D SIM high-resolution microscopy. Results: We detected that giantin …
