Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Statistical Models (162)
- Medicine and Health Sciences (154)
- Applied Statistics (142)
- Statistical Methodology (141)
- Social and Behavioral Sciences (134)
-
- Life Sciences (98)
- Data Science (73)
- Longitudinal Data Analysis and Time Series (71)
- Statistical Theory (71)
- Biostatistics (67)
- Categorical Data Analysis (66)
- Computer Sciences (61)
- Business (60)
- Public Health (58)
- Medical Specialties (53)
- Dentistry (51)
- Design of Experiments and Sample Surveys (51)
- Economics (51)
- Other Statistics and Probability (48)
- Probability (48)
- Econometrics (45)
- Engineering (45)
- Genetics and Genomics (33)
- Artificial Intelligence and Robotics (32)
- Macroeconomics (32)
- Mathematics (32)
- Bioinformatics (31)
- Institution
-
- Loma Linda University (90)
- COBRA (78)
- Central Bank of Nigeria (29)
- University of Kentucky (21)
- Hunan Provincial Institute of Scientific and Technology Information (13)
-
- City University of New York (CUNY) (12)
- California Polytechnic State University, San Luis Obispo (10)
- Southern Methodist University (10)
- University of Nevada, Las Vegas (10)
- Michigan Technological University (9)
- Stephen F. Austin State University (9)
- Virginia Commonwealth University (9)
- West Virginia University (9)
- East Tennessee State University (7)
- University of Arkansas, Fayetteville (7)
- University of Louisville (7)
- Air Force Institute of Technology (6)
- Claremont Colleges (6)
- Kennesaw State University (6)
- Louisiana State University (6)
- Portland State University (6)
- Dartmouth College (5)
- Georgia Southern University (5)
- Murray State University (5)
- University of Nebraska - Lincoln (5)
- Duquesne University (4)
- Illinois State University (4)
- LSU New Orleans (4)
- Missouri State University (4)
- South Dakota State University (4)
- Keyword
-
- Classification (13)
- Statistics (12)
- Machine learning (11)
- Data mining (9)
- Prediction (9)
-
- Machine Learning (8)
- Regression (8)
- American Southeast (7)
- Caddo (7)
- Factor analysis (7)
- Archaeology (6)
- Clustering (6)
- Cross-validation (6)
- Forecasting (6)
- Gene expression (6)
- Logistic regression (6)
- Ceramics (5)
- Cluster analysis (5)
- GIS (5)
- Genetics (5)
- Geochemistry (5)
- MANOVA (5)
- Model selection (5)
- Modeling (5)
- Survival analysis (5)
- Time series (5)
- Bootstrap (4)
- COVID-19 (4)
- Chemometrics (4)
- INAA (4)
- Publication Year
- Publication
-
- Loma Linda University Electronic Theses, Dissertations & Projects (90)
- CBN Journal of Applied Statistics (JAS) (29)
- Electronic Theses and Dissertations (23)
- U.C. Berkeley Division of Biostatistics Working Paper Series (20)
- UW Biostatistics Working Paper Series (18)
-
- Harvard University Biostatistics Working Paper Series (17)
- Theses and Dissertations (17)
- Theses and Dissertations--Statistics (15)
- Journal of Scientific Information Research (13)
- COBRA Preprint Series (10)
- Dissertations, Master's Theses and Master's Reports (9)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (9)
- SMU Data Science Review (8)
- CRHR: Archaeology (7)
- Graduate Theses, Dissertations, and Problem Reports (ETD) (7)
- Statistics (6)
- Complex Systems Faculty Publications and Presentations (5)
- Dissertations, Theses, and Capstone Projects (5)
- Graduate Theses and Dissertations (5)
- Master's Theses (5)
- College of Graduate Studies: Theses & Dissertations (4)
- Dartmouth Scholarship (4)
- Dissertations (4)
- Dissertations and Theses (Open Access) (4)
- Graduate Theses/Dissertations (4)
- LSU Master's Theses (4)
- LSU New Orleans Theses and Dissertations (4)
- SDSU Data Science Symposium (4)
- UNLV Theses, Dissertations, Professional Papers, and Capstones (4)
- Williams Honors College, Honors Research Projects (4)
- Publication Type
- File Type
Articles 241 - 270 of 522
Full-Text Articles in Multivariate Analysis
Essentials Of Structural Equation Modeling, Mustafa Emre Civelek
Essentials Of Structural Equation Modeling, Mustafa Emre Civelek
Zea E-Books Collection
Structural Equation Modeling is a statistical method increasingly used in scientific studies in the fields of Social Sciences. It is currently a preferred analysis method, especially in doctoral dissertations and academic researches. However, since many universities do not include this method in the curriculum of undergraduate and graduate courses, students and scholars try to solve the problems they encounter by using various books and internet resources.
This book aims to guide the researcher who wants to use this method in a way that is free from math expressions. It teaches the steps of a research program using structured equality modeling …
Building A Better Risk Prevention Model, Steven Hornyak
Building A Better Risk Prevention Model, Steven Hornyak
National Youth Advocacy & Resilience Conference
This presentation chronicles the work of Houston County Schools in developing a risk prevention model built on more than ten years of longitudinal student data. In its second year of implementation, Houston At-Risk Profiles (HARP), has proven effective in identifying those students most in need of support and linking them to interventions and supports that lead to improved outcomes and significantly reduces the risk of failure.
A Preliminary Study Of Smithport Plain Bottle Morphology In The Southern Caddo Area, Robert Z. Selden Jr.
A Preliminary Study Of Smithport Plain Bottle Morphology In The Southern Caddo Area, Robert Z. Selden Jr.
CRHR: Archaeology
This study expands upon a previous analysis of the Clarence H. Webb collection, which resulted in the identification of two discrete shapes used in the manufacture of the base and body of Smithport Plain bottles. The sample includes the Smithport Plain bottles from the Webb collection, and four new bottles: two previously repatriated specimens in the Pohler Collection, and two from the Mitchell site (41BW4) to test whether those specimens align morphologically with the Belcher Mound or Smithport Landing specimens. Results indicate significant allometry and a significant difference in Smithport Plain body and base shapes for bottles produced at the …
Effect Of Socioeconomic And Demographic Factors On Kentucky Crashes, Aaron Berry Cambron
Effect Of Socioeconomic And Demographic Factors On Kentucky Crashes, Aaron Berry Cambron
Theses and Dissertations--Civil Engineering
The goal of this research was to examine the potential predictive ability of socioeconomic and demographic data for drivers on Kentucky crash occurrence. Identifying unique background characteristics of at-fault drivers that contribute to crash rates and crash severity may lead to improved and more specific interventions to reduce the negative impacts of motor vehicle crashes. The driver-residence zip code was used as a spatial unit to connect five years of Kentucky crash data with socioeconomic factors from the U.S. Census, such as income, employment, education, age, and others, along with terrain and vehicle age. At-fault driver crash counts, normalized over …
Psychometric Properties Of A Working Memory Span Task, Juan M. Alzate Vanegas
Psychometric Properties Of A Working Memory Span Task, Juan M. Alzate Vanegas
Honors Undergraduate Theses
The intent of this thesis is to examine the psychometric properties of a complex span task (CST) developed to measure working memory capacity (WMC) using measurements obtained from a sample of 68 undergraduate students at the University of Central Florida. The Grocery List Task (GLT) promises several design improvements over traditional CSTs in a prior study about individual differences in WMC and distraction effects on driving performance, and it offers potential benefits for studying WMC as well as the serial-position effect. Currently, the working memory system is composed of domain-general memorial storage processes and information-processing, which involves the use of …
A Quantitative Analysis Of Intermediate Forms Within Astarte From The Atlantic Coastal Plain, Philip Roberson
A Quantitative Analysis Of Intermediate Forms Within Astarte From The Atlantic Coastal Plain, Philip Roberson
Murray State Theses and Dissertations
The Atlantic Coastal Plain has long been recognized as a natural laboratory useful for testing hypotheses about various environmental and ecological effects on marine fauna. For studies such as these to continue being conducted in a rigorous and easily repeatable manner, a reliable taxonomy must be established for genera within this physiographic province. The bivalve genus, Astarte, is a cosmopolitan genus that is commonly found within the Atlantic Coastal Plain. This genus has many formally recognized species, even though it lacks many features that would encourage diversification, marking it as a taxonomic group in need of potential revision. The …
Prediction Intervals For Functional Data, Nicholas Rios
Prediction Intervals For Functional Data, Nicholas Rios
Theses, Dissertations and Culminating Projects
The prediction of functional data samples has been the focus of several functional data analysis endeavors. This work describes the use of dynamic function-on-function regression for dynamic prediction of the future trajectory as well as the construction of dynamic prediction intervals for functional data. The overall goals of this thesis are to assess the efficacy of Dynamic Penalized Function-on-Function Regression (DPFFR) and to compare DPFFR prediction intervals with those of other dynamic prediction methods. To make these comparisons, metrics are used that measure prediction error, prediction interval width, and prediction interval coverage. Simulations and applications to financial stock data from …
High Dimensional Multivariate Inference Under General Conditions, Xiaoli Kong
High Dimensional Multivariate Inference Under General Conditions, Xiaoli Kong
Theses and Dissertations--Statistics
In this dissertation, we investigate four distinct and interrelated problems for high-dimensional inference of mean vectors in multi-groups.
The first problem concerned is the profile analysis of high dimensional repeated measures. We introduce new test statistics and derive its asymptotic distribution under normality for equal as well as unequal covariance cases. Our derivations of the asymptotic distributions mimic that of Central Limit Theorem with some important peculiarities addressed with sufficient rigor. We also derive consistent and unbiased estimators of the asymptotic variances for equal and unequal covariance cases respectively.
The second problem considered is the accurate inference for high-dimensional repeated …
Using The Qbest Equation To Evaluate Ellagic Acid Safety Data: Generating A Qnoael With Confidence Levels From Disparate Literature, Cynthia Rose Dickerson
Using The Qbest Equation To Evaluate Ellagic Acid Safety Data: Generating A Qnoael With Confidence Levels From Disparate Literature, Cynthia Rose Dickerson
Theses and Dissertations--Pharmacy
QBEST, a novel statistical method, can be applied to the problem of estimating the No Observed Adverse Effect Level (NOAEL or QNOAEL) of a New Molecular Entity (NME) in order to anticipate a safe starting dose for beginning clinical trials. The NOAEL from QBEST (called the QNOAEL) can be calculated using multiple disparate studies in the literature and/or from the lab. The QNOAEL is similar in some ways to the Benchmark Dose Method (BMD) used widely in toxicological research, but is superior to the BMD in some ways. The QNOAEL simulation generates an intuitive curve that is comparable to the …
Understanding The Novice Decision-Making Process In Forensic Footwear Examinations: Accuracy And Decision Rules, Madonna A. Nobel
Understanding The Novice Decision-Making Process In Forensic Footwear Examinations: Accuracy And Decision Rules, Madonna A. Nobel
Graduate Theses, Dissertations, and Problem Reports (ETD)
The reproducibility of experienced-based forensic pattern interpretation is founded on the notion that domain-specific knowledge can be successfully distributed and applied among experts within a group. This assumption persists, even when the examination is complicated by variations in case circumstances, such as impression clarity and totality, as well as media, substrate, collection mechanism and enhancement. While it is further theorized that many of these factors (as well as additional confounding factors) are at play during an examination, the manner and extent to which these sources of variability affect the examination of footwear evidence remain unclear. In order to explore this …
Non-Linear Machine Learning With Active Sampling For Mox Drift Compensation, Tamara Matthews, Muhammad Iqbal, Horacio Gonzalez-Velez
Non-Linear Machine Learning With Active Sampling For Mox Drift Compensation, Tamara Matthews, Muhammad Iqbal, Horacio Gonzalez-Velez
Conference papers
Abstract—Metal oxide (MOX) gas detectors based on SnO2 provide low-cost solutions for real-time sensing of complex gas mixtures for indoor ambient monitoring. With high sensitivity under ideal conditions, MOX detectors may have poor longterm response accuracy due to environmental factors (humidity and temperature) along with sensor aging, leading to calibration drifts. Finding a simple and efficient solution to correct such calibration drifts has been the subject of numerous studies but remains an open problem. In this work, we present an efficient approach to MOX calibration using active and transfer sampling techniques coupled with non-linear machine learning algorithms, namely neural networks, …
Algorithms For Reconstruction Of Gene Regulatory Networks From High -Throughput Gene Expression Data, Wenping Deng
Algorithms For Reconstruction Of Gene Regulatory Networks From High -Throughput Gene Expression Data, Wenping Deng
Dissertations, Master's Theses and Master's Reports
Understanding gene interactions in complex living systems is one of the central tasks in system biology. With the availability of microarray and RNA-Seq technologies, a multitude of gene expression datasets has been generated towards novel biological knowledge discovery through statistical analysis and reconstruction of gene regulatory networks (GRN). Reconstruction of GRNs can reveal the interrelationships among genes and identify the hierarchies of genes and hubs in networks. The new algorithms I developed in this dissertation are specifically focused on the reconstruction of GRNs with increased accuracy from microarray and RNA-Seq high-throughput gene expression data sets.
The first algorithm (Chapter 2) …
Offline And Online Density Estimation For Large High-Dimensional Data, Aref Majdara
Offline And Online Density Estimation For Large High-Dimensional Data, Aref Majdara
Dissertations, Master's Theses and Master's Reports
Density estimation has wide applications in machine learning and data analysis techniques including clustering, classification, multimodality analysis, bump hunting and anomaly detection. In high-dimensional space, sparsity of data in local neighborhood makes many of parametric and nonparametric density estimation methods mostly inefficient.
This work presents development of computationally efficient algorithms for high-dimensional density estimation, based on Bayesian sequential partitioning (BSP). Copula transform is used to separate the estimation of marginal and joint densities, with the purpose of reducing the computational complexity and estimation error. Using this separation, a parallel implementation of the density estimation algorithm on a 4-core CPU is …
Making Models With Bayes, Pilar Olid
Making Models With Bayes, Pilar Olid
Electronic Theses, Projects, and Dissertations
Bayesian statistics is an important approach to modern statistical analyses. It allows us to use our prior knowledge of the unknown parameters to construct a model for our data set. The foundation of Bayesian analysis is Bayes' Rule, which in its proportional form indicates that the posterior is proportional to the prior times the likelihood. We will demonstrate how we can apply Bayesian statistical techniques to fit a linear regression model and a hierarchical linear regression model to a data set. We will show how to apply different distributions to Bayesian analyses and how the use of a prior affects …
A Cross-Sectional Exploration Of Household Financial Reactions And Homebuyer Awareness Of Registered Sex Offenders In A Rural, Suburban, And Urban County., John Charles Navarro
A Cross-Sectional Exploration Of Household Financial Reactions And Homebuyer Awareness Of Registered Sex Offenders In A Rural, Suburban, And Urban County., John Charles Navarro
Electronic Theses and Dissertations
As stigmatized persons, registered sex offenders betoken instability in communities. Depressed home sale values are associated with the presence of registered sex offenders even though the public is largely unaware of the presence of registered sex offenders. Using a spatial multilevel approach, the current study examines the role registered sex offenders influence sale values of homes sold in 2015 for three U.S. counties (rural, suburban, and urban) located in Illinois and Kentucky within the social disorganization framework. Homebuyers were surveyed to examine whether awareness of local registered sex offenders and the homebuyer’s community type operate as moderators between home selling …
Burden Of Atopic Dermatitis In The United States: Analysis Of Healthcare Claims Data In The Commercial, Medicare, And Medi-Cal Databases, Sulena Shrestha, Raymond Miao, Li Wang, Jingdong Chao, Huseyin Yuce, Wenhui Wei
Burden Of Atopic Dermatitis In The United States: Analysis Of Healthcare Claims Data In The Commercial, Medicare, And Medi-Cal Databases, Sulena Shrestha, Raymond Miao, Li Wang, Jingdong Chao, Huseyin Yuce, Wenhui Wei
Publications and Research
Comparative data on the burden of atopic dermatitis (AD) in adults relative to the general population are limited. We performed a large-scale evaluation of the burden of disease among US adults with AD relative to matched non-AD controls, encompassing comorbidities, healthcare resource utilization (HCRU), and costs, using healthcare claims data. The impact of AD disease severity on these outcomes was also evaluated.
Failure Of Care Acquisition: Identifying Risk Factors In American Health Disparities, Nicholas Downing, Mamunur Rashid
Failure Of Care Acquisition: Identifying Risk Factors In American Health Disparities, Nicholas Downing, Mamunur Rashid
Student Research
We examined the effects of various demographic and socioeconomic risk factors that influence an adult's decision not to obtain medical care in the United States utilizing data from the 2015 National Health Interview Survey (NHIS). Bivariate analysis and multivariate logistic regression revealed that family income, insurance status and whether one worries about paying medical bills make individuals nearly 80% less likely to obtain care than their counterparts. This study provides evidence that certain risk factors, especially those directly related to one's socioeconomic status, may put individuals at greater risk for failure to obtain care. Interventions in policy may be needed …
An Investigation Of The Accuracy Of Parallel Analysis For Determining The Number Of Factors In A Factor Analysis, Mandy Matsumoto
An Investigation Of The Accuracy Of Parallel Analysis For Determining The Number Of Factors In A Factor Analysis, Mandy Matsumoto
Mahurin Honors College Capstone Experience/Thesis Projects
Exploratory factor analysis is an analytic technique used to determine the number of factors in a set of data (usually items on a questionnaire) for which the factor structure has not been previously analyzed. Parallel analysis (PA) is a technique used to determine the number of factors in a factor analysis. There are a number of factors that affect the results of a PA: the choice of the eigenvalue percentile, the strength of the factor loadings, the number of variables, and the sample size of the study. Although PA is the most accurate method to date to determine which factors …
Marketing The Mountain State: A Large N Study Of User Engagement On Twitter, Kirk Richardson
Marketing The Mountain State: A Large N Study Of User Engagement On Twitter, Kirk Richardson
Capstone Projects – Politics and Government
Much of the evolving research on the use of social media in destination marketing emphasizes how information diffusion influences the reputational image of place. The present study uses Twitter data to focus on the relative differences in user engagement across discrete account types. Specifically, this is done to examine how the official destination marketing organization of Montana—the Montana Office of Tourism (MTOT)—performs relative to other account types. Several regression analyses conducted on Twitter data associated with an ongoing MTOT place branding campaign reveal that tweets sent from ‘official’ accounts are more likely to be retweeted, and are estimated to receive …
Performance Of Imputation Algorithms On Artificially Produced Missing At Random Data, Tobias O. Oketch
Performance Of Imputation Algorithms On Artificially Produced Missing At Random Data, Tobias O. Oketch
Electronic Theses and Dissertations
Missing data is one of the challenges we are facing today in modeling valid statistical models. It reduces the representativeness of the data samples. Hence, population estimates, and model parameters estimated from such data are likely to be biased.
However, the missing data problem is an area under study, and alternative better statistical procedures have been presented to mitigate its shortcomings. In this paper, we review causes of missing data, and various methods of handling missing data. Our main focus is evaluating various multiple imputation (MI) methods from the multiple imputation of chained equation (MICE) package in the statistical software …
Detecting And Evaluating Therapy Induced Changes In Radiomics Features Measured From Non-Small Cell Lung Cancer To Predict Patient Outcomes, Xenia J. Fave
Dissertations and Theses (Open Access)
The purpose of this study was to investigate whether radiomics features measured from weekly 4-dimensional computed tomography (4DCT) images of non-small cell lung cancers (NSCLC) change during treatment and if those changes are prognostic for patient outcomes or dependent on treatment modality. Radiomics features are quantitative metrics designed to evaluate tumor heterogeneity from routine medical imaging. Features that are prognostic for patient outcome could be used to monitor tumor response and identify high-risk patients for adaptive treatment. This would be especially valuable for NSCLC due to the high prevalence and mortality of this disease.
A novel process was designed to …
Network Exploration Of Correlated Multivariate Protein Data For Alzheimer's Disease Association, Matthew J. Lane
Network Exploration Of Correlated Multivariate Protein Data For Alzheimer's Disease Association, Matthew J. Lane
Theses
Alzheimer Disease (AD) is difficult to diagnose by using genetic testing or other traditional methods. Unlike diseases with simple genetic risk components, there exists no single marker determining as to whether someone will develop AD. Furthermore, AD is highly heterogeneous and different subgroups of individuals develop the disease due to differing factors. Traditional diagnostic methods using perceivable cognitive deficiencies are often too little too late due to the brain having suffered damage from decades of disease progression. In order to observe AD at early stages prior to the observation of cognitive deficiencies, biomarkers with greater accuracy are required. By using …
Statistically Analyzing Assembly Line Processing Times Through Incorporation Of Product Variation, Kyle Rehr, Matthew Farr
Statistically Analyzing Assembly Line Processing Times Through Incorporation Of Product Variation, Kyle Rehr, Matthew Farr
Scholars Week
Timing methods and performance metrics are important in the heavily industrialized world we live in. Industrial plants use metrics to measure quality of production, help make decisions, and drive the strategy of the organization. However, there are many factors to be considered when measuring performance based on a metric; of which we will be analyzing the importance of product variation. We will be analyzing assembly line timings, whilst controlling for product variance, to show the importance differences between products makes in one’s ability to predict performance. In addition, we will be analyzing the current “statistical” methods used by an industrial …
Disability In Long-Term Care Residents Explained By Prevalent Geriatric Syndromes, Not Long-Term Care Home Characteristics: A Cross-Sectional Study, Natasha E. Lane, Walter P. Wodchis, Cynthia M. Boyd, Thérèse A. Stukel
Disability In Long-Term Care Residents Explained By Prevalent Geriatric Syndromes, Not Long-Term Care Home Characteristics: A Cross-Sectional Study, Natasha E. Lane, Walter P. Wodchis, Cynthia M. Boyd, Thérèse A. Stukel
Dartmouth Scholarship
Self-care disability is dependence on others to conduct activities of daily living, such as bathing, eating and dressing. Among long-term care residents, self-care disability lowers quality of life and increases health care costs. Understanding the correlates of self-care disability in this population is critical to guide clinical care and ongoing research in Geriatrics. This study examines which resident geriatric syndromes and chronic conditions are associated with residents’ self-care disability and whether these relationships vary across strata of age, sex and cognitive status. It also describes the proportion of variance in residents’ self-care disability that is explained by residents’ geriatric syndromes …
Studying The Optimal Scheduling For Controlling Prostate Cancer Under Intermittent Androgen Suppression, Sunil K. Dhar, Hans R. Chaudhry, Bruce G. Bukiet, Zhiming Ji, Nan Gao, Thomas W. Findley
Studying The Optimal Scheduling For Controlling Prostate Cancer Under Intermittent Androgen Suppression, Sunil K. Dhar, Hans R. Chaudhry, Bruce G. Bukiet, Zhiming Ji, Nan Gao, Thomas W. Findley
Harvard University Biostatistics Working Paper Series
This retrospective study shows that the majority of patients’ correlations between PSA and Testosterone during the on-treatment period is at least 0.90. Model-based duration calculations to control PSA levels during off-treatment are provided. There are two pairs of models. In one pair, the Generalized Linear Model and Mixed Model are both used to analyze the variability of PSA at the individual patient level by using the variable “Patient ID” as a repeated measure. In the second pair, Patient ID is not used as a repeated measure but additional baseline variables are included to analyze the variability of PSA.
Informational Index And Its Applications In High Dimensional Data, Qingcong Yuan
Informational Index And Its Applications In High Dimensional Data, Qingcong Yuan
Theses and Dissertations--Statistics
We introduce a new class of measures for testing independence between two random vectors, which uses expected difference of conditional and marginal characteristic functions. By choosing a particular weight function in the class, we propose a new index for measuring independence and study its property. Two empirical versions are developed, their properties, asymptotics, connection with existing measures and applications are discussed. Implementation and Monte Carlo results are also presented.
We propose a two-stage sufficient variable selections method based on the new index to deal with large p small n data. The method does not require model specification and especially focuses …
Gamma/Hadron Separation For The Hawc Observatory, Michael J. Gerhardt
Gamma/Hadron Separation For The Hawc Observatory, Michael J. Gerhardt
Dissertations, Master's Theses and Master's Reports
The High-Altitude Water Cherenkov (HAWC) Observatory is a gamma-ray observatory sensitive to gamma rays from 100 GeV to 100 TeV with an instantaneous field of view of ~2 sr. It is located on the Sierra Negra plateau in Mexico at an elevation of 4,100 m and began full operation in March 2015. The purpose of the detector is to study relativistic particles that are produced by interstellar and intergalactic objects such as: pulsars, supernova remnants, molecular clouds, black holes and more. To achieve optimal angular resolution, energy reconstruction and cosmic ray background suppression for the extensive air showers detected by …
A Traders Guide To The Predictive Universe- A Model For Predicting Oil Price Targets And Trading On Them, Jimmie Harold Lenz
A Traders Guide To The Predictive Universe- A Model For Predicting Oil Price Targets And Trading On Them, Jimmie Harold Lenz
Doctor of Business Administration Dissertations
At heart every trader loves volatility; this is where return on investment comes from, this is what drives the proverbial “positive alpha.” As a trader, understanding the probabilities related to the volatility of prices is key, however if you could also predict future prices with reliability the world would be your oyster. To this end, I have achieved three goals with this dissertation, to develop a model to predict future short term prices (direction and magnitude), to effectively test this by generating consistent profits utilizing a trading model developed for this purpose, and to write a paper that anyone with …
Tutorial For Using The Center For High Performance Computing At The University Of Utah And An Example Using Random Forest, Stephen Barton
Tutorial For Using The Center For High Performance Computing At The University Of Utah And An Example Using Random Forest, Stephen Barton
All Graduate Plan B and other Reports, Spring 1920 to Spring 2023
Random Forests are very memory intensive machine learning algorithms and most computers would fail at building models from datasets with millions of observations. Using the Center for High Performance Computing (CHPC) at the University of Utah and an airline on-time arrival dataset with 7 million observations from the U.S. Department of Transportation Bureau of Transportation Statistics we built 316 models by adjusting the depth of the trees and randomness of each forest and compared the accuracy and time each took. Using this dataset we discovered that substantial restrictions to the size of trees, observations allowed for each tree, and variables …
Advanced Data Analysis - Lecture Notes, Erik B. Erhardt, Edward J. Bedrick, Ronald M. Schrader
Advanced Data Analysis - Lecture Notes, Erik B. Erhardt, Edward J. Bedrick, Ronald M. Schrader
Open Textbooks
Lecture notes for Advanced Data Analysis (ADA1 Stat 427/527 and ADA2 Stat 428/528), Department of Mathematics and Statistics, University of New Mexico, Fall 2016-Spring 2017. Additional material including RMarkdown templates for in-class and homework exercises, datasets, R code, and video lectures are available on the course websites: https://statacumen.com/teaching/ada1 and https://statacumen.com/teaching/ada2 .
Contents
I ADA1: Software
- 0 Introduction to R, Rstudio, and ggplot
II ADA1: Summaries and displays, and one-, two-, and many-way tests of means
- 1 Summarizing and Displaying Data
- 2 Estimation in One-Sample Problems
- 3 Two-Sample Inferences
- 4 Checking Assumptions
- 5 One-Way Analysis of Variance
III ADA1: Nonparametric, categorical, …