Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Social and Behavioral Sciences (1395)
- Statistical Theory (1191)
- Statistical Models (374)
- Applied Mathematics (362)
- Mathematics (341)
-
- Statistical Methodology (274)
- Data Science (194)
- Computer Sciences (171)
- Biostatistics (155)
- Engineering (155)
- Medicine and Health Sciences (151)
- Probability (150)
- Multivariate Analysis (143)
- Life Sciences (140)
- Business (127)
- Longitudinal Data Analysis and Time Series (122)
- Categorical Data Analysis (116)
- Economics (113)
- Law (102)
- Other Statistics and Probability (93)
- Education (91)
- Environmental Sciences (86)
- Artificial Intelligence and Robotics (81)
- Design of Experiments and Sample Surveys (77)
- Econometrics (76)
- Psychology (65)
- Numerical Analysis and Scientific Computing (58)
- Institution
-
- Wayne State University (1097)
- Wright State University (138)
- Utah State University (103)
- Cornell University Law School (75)
- California Polytechnic State University, San Luis Obispo (59)
-
- Air Force Institute of Technology (55)
- Old Dominion University (55)
- University of Kentucky (54)
- Montclair State University (53)
- University of Arkansas, Fayetteville (51)
- Southern Methodist University (46)
- Western Kentucky University (46)
- University of Nebraska - Lincoln (43)
- Central Bank of Nigeria (41)
- City University of New York (CUNY) (39)
- Virginia Commonwealth University (39)
- Claremont Colleges (36)
- Illinois State University (35)
- Kennesaw State University (35)
- University of Richmond (34)
- Georgia Southern University (31)
- Louisiana Tech University (29)
- Stephen F. Austin State University (25)
- University of New Mexico (25)
- East Tennessee State University (24)
- University of Nevada, Las Vegas (24)
- Prairie View A&M University (22)
- The University of Akron (19)
- Technological University Dublin (17)
- Michigan Technological University (16)
- Keyword
-
- Statistics (122)
- Empirical legal studies (55)
- Simulation (49)
- Machine learning (42)
- Regression (40)
-
- Bias (35)
- Logistic regression (33)
- Monte Carlo simulation (30)
- Bootstrap (29)
- Power (29)
- Bayesian (28)
- Machine Learning (27)
- Reliability (27)
- Confidence interval (26)
- Western Kentucky University (26)
- Mean squared error (24)
- Estimation (22)
- Maximum likelihood estimation (22)
- Missing data (22)
- Monte Carlo (22)
- Nonparametric (21)
- Robustness (21)
- Sample size (21)
- Type I error (21)
- Effect size (20)
- Multicollinearity (20)
- Statistical analysis (20)
- Confidence intervals (19)
- Enrollment (19)
- Pure sciences (19)
- Publication Year
- Publication
-
- Journal of Modern Applied Statistical Methods (1093)
- Mathematics and Statistics Faculty Publications (137)
- Theses and Dissertations (91)
- Cornell Law Faculty Publications (75)
- Electronic Theses and Dissertations (58)
-
- Department of Applied Mathematics and Statistics Faculty Scholarship and Creative Works (49)
- All Graduate Theses and Dissertations, Spring 1920 to Summer 2023 (46)
- All Graduate Plan B and other Reports, Spring 1920 to Spring 2023 (42)
- CBN Journal of Applied Statistics (JAS) (40)
- Graduate Theses and Dissertations (38)
- Department of Math & Statistics Faculty Publications (34)
- Theses and Dissertations--Statistics (32)
- College of Graduate Studies: Theses & Dissertations (29)
- SMU Data Science Review (29)
- Master's Theses (27)
- Mathematics & Statistics Theses & Dissertations (26)
- WKU Administration Documents (26)
- Annual Symposium on Biomathematics and Ecology Education and Research (25)
- Symposium of Student Scholars (24)
- Applications and Applied Mathematics: An International Journal (AAM) (22)
- Articles (22)
- Statistics (21)
- Publications and Research (20)
- Williams Honors College, Honors Research Projects (19)
- Department of Statistics: Dissertations, Theses, and Student Research (17)
- CMC Senior Theses (16)
- Dissertations, Master's Theses and Master's Reports (16)
- Mathematics Senior Capstone Papers (15)
- Statistical Science Theses and Dissertations (15)
- Doctoral Dissertations (14)
- Publication Type
- File Type
Articles 931 - 960 of 2919
Full-Text Articles in Applied Statistics
Comparative Clinical Outcomes Between Direct Oral Anticoagulants And Warfarin Among Elderly Patients With Non-Valvular Atrial Fibrillation In The Cms Medicare Population, Alpesh Amin, Oluwaseyi Dina, Allison Keshishian, Amol Dhamane, Anagha Nadkarni, Eric Carda, Cristina Russ, Lisa Rosenblatt, Jack Mardekian, Huseyin Yuce, Christine L. Baker
Comparative Clinical Outcomes Between Direct Oral Anticoagulants And Warfarin Among Elderly Patients With Non-Valvular Atrial Fibrillation In The Cms Medicare Population, Alpesh Amin, Oluwaseyi Dina, Allison Keshishian, Amol Dhamane, Anagha Nadkarni, Eric Carda, Cristina Russ, Lisa Rosenblatt, Jack Mardekian, Huseyin Yuce, Christine L. Baker
Publications and Research
Atrial fibrillation (AF) prevalence increases with age; > 80% of US adults with AF are aged ≥ 65 years. Compare the risk of stroke/systemic embolism (SE), major bleeding (MB), net clinical outcome (NCO), and major adverse cardiac events (MACE) among elderly non-valvular AF (NVAF) Medicare patients prescribed direct oral anticoagulants (DOACs) vs warfarin. NVAF patients aged ≥ 65 years who initiated DOACs (apixaban, dabigatran, and rivaroxaban) or warfarin were selected from 01JAN2013-31DEC2015 in CMS Medicare data. Propensity score matching was used to balance DOAC and warfarin cohorts. Cox proportional hazards models estimated the risk of stroke/SE, MB, NCO, and MACE. 37,525 …
Six-Month Outcome Of Transient Ischemic Attack And Its Mimics, Alireza Sadighi, Vida Abedi, Alia C. Stanciu, Nada El Andary, Mihai Banciu, Neil Holland, Ramin Zand
Six-Month Outcome Of Transient Ischemic Attack And Its Mimics, Alireza Sadighi, Vida Abedi, Alia C. Stanciu, Nada El Andary, Mihai Banciu, Neil Holland, Ramin Zand
Faculty Journal Articles
Background and Objective: Although the risk of recurrent cerebral ischemia is higher after a transient ischemic attack (TIA), there is limited data on the outcome of TIA mimics. The goal of this study is to compare the 6-month outcome of patients with negative and positive diffusion-weighted imaging (DWI) TIAs (DWI-neg TIA vs. DWI-pos TIA) and also TIA mimics.
Methods: We prospectively studied consecutive patients with an initial diagnosis of TIA in our tertiary stroke centers in a 2-year period. Every included patient had an initial magnetic resonance (MR) with DWI and one-, three-, and six-month follow-up visits. The primary outcome …
A Random Forests Approach To Assess Determinants Of Central Bank Independence, Maddalena Cavicchioli, Angeliki Papana, Ariadni Papana Dagiasis, Barbara Pistoresi
A Random Forests Approach To Assess Determinants Of Central Bank Independence, Maddalena Cavicchioli, Angeliki Papana, Ariadni Papana Dagiasis, Barbara Pistoresi
Journal of Modern Applied Statistical Methods
A non-parametric efficient statistical method, Random Forests, is implemented for the selection of the determinants of Central Bank Independence (CBI) among a large database of economic, political, and institutional variables for OECD countries. It permits ranking all the determinants based on their importance in respect to the CBI and does not impose a priori assumptions on potential nonlinear relationships in the data. Collinearity issues are resolved, because correlated variables can be simultaneously considered.
Maximum Likelihood Estimation For The Generalized Pareto Distribution And Goodness-Of-Fit Test With Censored Data, Minh H. Pham, Chris Tsokos, Bong-Jin Choi
Maximum Likelihood Estimation For The Generalized Pareto Distribution And Goodness-Of-Fit Test With Censored Data, Minh H. Pham, Chris Tsokos, Bong-Jin Choi
Journal of Modern Applied Statistical Methods
The generalized Pareto distribution (GPD) is a flexible parametric model commonly used in financial modeling. Maximum likelihood estimation (MLE) of the GPD was proposed by Grimshaw (1993). Maximum likelihood estimation of the GPD for censored data is developed, and a goodness-of-fit test is constructed to verify an MLE algorithm in R and to support the model-validation step. The algorithms were composed in R. Grimshaw’s algorithm outperforms functions available in the R package ‘gPdtest’. A simulation study showed the MLE method for censored data and the goodness-of-fit test are both reliable.
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes, Lauren Spirko-Burns, Karthik Devarajan
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes, Lauren Spirko-Burns, Karthik Devarajan
COBRA Preprint Series
One of the major goals in large-scale genomic studies is to identify genes with a prognostic impact on time-to-event outcomes which provide insight into the disease's process. With rapid developments in high-throughput genomic technologies in the past two decades, the scientific community is able to monitor the expression levels of tens of thousands of genes and proteins resulting in enormous data sets where the number of genomic features is far greater than the number of subjects. Methods based on univariate Cox regression are often used to select genomic features related to survival outcome; however, the Cox model assumes proportional hazards …
A Simulation Study Of Diagnostics For Bias In Non-Probability Samples, Philip S. Boonstra, Roderick Ja Little, Brady T. West, Rebecca R. Andridge, Fernanda Alvarado-Leiton
A Simulation Study Of Diagnostics For Bias In Non-Probability Samples, Philip S. Boonstra, Roderick Ja Little, Brady T. West, Rebecca R. Andridge, Fernanda Alvarado-Leiton
The University of Michigan Department of Biostatistics Working Paper Series
A non-probability sampling mechanism is likely to bias estimates of parameters with respect to a target population of interest. This bias poses a unique challenge when selection is 'non-ignorable', i.e. dependent upon the unobserved outcome of interest, since it is then undetectable and thus cannot be ameliorated. We extend a simulation study by Nishimura et al. [International Statistical Review, 84, 43--62 (2016)], adding a recently published statistic, the so-called 'standardized measure of unadjusted bias', which explicitly quantifies the extent of bias under the assumption that a specified amount of non-ignorable selection exists. Our findings suggest that this new …
Bayesian Approximation Techniques For Scale Parameter Of Laplace Distribution, Uzma Jan, S. P. Ahmad
Bayesian Approximation Techniques For Scale Parameter Of Laplace Distribution, Uzma Jan, S. P. Ahmad
Journal of Modern Applied Statistical Methods
The Bayesian estimation of the scale parameter of a Laplace Distribution is obtained using two approximation techniques, like Normal approximation and Tierney and Kadane (T-K) approximation, under different informative priors.
Can One Test Fit All? Responses To The Article “Striving For Simple But Effective Advice For Comparing The Central Tendency Of Two Populations” (Ruxton & Neuhäuser, 2018), Diep Nguyen, Eun Sook Kim, Yi-Hsin Chen
Can One Test Fit All? Responses To The Article “Striving For Simple But Effective Advice For Comparing The Central Tendency Of Two Populations” (Ruxton & Neuhäuser, 2018), Diep Nguyen, Eun Sook Kim, Yi-Hsin Chen
Journal of Modern Applied Statistical Methods
Responses to suggestions made by Ruxton & Neuhäuser (2018) regarding Nguyen et al. (2016) are given.
Associations Of Hemodynamics, Morphology, And Patient Characteristics With Aneurysm Rupture Stratified By Aneurysm Location, Felicitas J. Detmer, Bong Jae Chung, Carlos Jimenez, Farid Hamzei-Sichani, David Kallmes, Christopher Putman, Juan R. Cebral
Associations Of Hemodynamics, Morphology, And Patient Characteristics With Aneurysm Rupture Stratified By Aneurysm Location, Felicitas J. Detmer, Bong Jae Chung, Carlos Jimenez, Farid Hamzei-Sichani, David Kallmes, Christopher Putman, Juan R. Cebral
Department of Applied Mathematics and Statistics Faculty Scholarship and Creative Works
Purpose: The mechanisms of cerebral aneurysm rupture are not fully understood. We analyzed the associations of hemodynamics, morphology, and patient age and gender with aneurysm rupture stratifying by location. Methods: Using image-based models, 20 hemodynamic and 17 morphological parameters were compared in 1931 ruptured and unruptured aneurysms with univariate logistic regression. Rupture rates were compared between males and females as well as younger and older patients and bifurcation versus sidewall aneurysms for different aneurysm locations. Subsequently, associations between hemodynamics and morphology and patient as well as aneurysm characteristics were analyzed for aneurysms at five locations. Results: Compared to unruptured aneurysms, …
On The Conditional And Unconditional Type I Error Rates And Power Of Tests In Linear Models With Heteroscedastic Errors, Patrick J. Rosopa, Alice M. Brawley, Theresa P. Atkinson, Stephen A. Robertson
On The Conditional And Unconditional Type I Error Rates And Power Of Tests In Linear Models With Heteroscedastic Errors, Patrick J. Rosopa, Alice M. Brawley, Theresa P. Atkinson, Stephen A. Robertson
Journal of Modern Applied Statistical Methods
Preliminary tests for homoscedasticity may be unnecessary in general linear models. Based on Monte Carlo simulations, results suggest that when testing for differences between independent slopes, the unconditional use of weighted least squares regression and HC4 regression performed the best across a wide range of conditions.
Φ-Divergence Loss-Based Artificial Neural Network, R. L. Salamwade, D. M. Sakate, S. K. Mathur
Φ-Divergence Loss-Based Artificial Neural Network, R. L. Salamwade, D. M. Sakate, S. K. Mathur
Journal of Modern Applied Statistical Methods
Artificial Neural Networks (ANNs) can fit non-linear functions and recognize patterns better than several standard techniques. Performance of ANNs is measured by using loss functions. Phi-divergence estimator is generalization of maximum likelihood estimator and it possesses all its properties. A neural network is proposed which is trained using phi-divergence loss.
Robust Ancova, Curvature, And The Curse Of Dimensionality, Rand Wilcox
Robust Ancova, Curvature, And The Curse Of Dimensionality, Rand Wilcox
Journal of Modern Applied Statistical Methods
There is a substantial collection of robust analysis of covariance (ANCOVA) methods that effectively deals with non-normality, unequal population slope parameters, outliers, and heteroscedasticity. Some are based on the usual linear model and others are based on smoothers (nonparametric regression estimators). However, extant results are limited to one or two covariates. A minor goal here is to extend a recently-proposed method, based on the usual linear model, to situations where there are up to six covariates. The usual linear model might provide a poor approximation of the true regression surface. The main goal is to suggest a method, based on …
A Strategy For Using Bias And Rmse As Outcomes In Monte Carlo Studies In Statistics, Michael Harwell
A Strategy For Using Bias And Rmse As Outcomes In Monte Carlo Studies In Statistics, Michael Harwell
Journal of Modern Applied Statistical Methods
To help ensure important patterns of bias and accuracy are detected in Monte Carlo studies in statistics this paper proposes conditioning bias and root mean square error (RMSE) measures on estimated Type I and Type II error rates. A small Monte Carlo study is used to illustrate this argument.
A Robust Nonparametric Measure Of Effect Size Based On An Analog Of Cohen's D, Plus Inferences About The Median Of The Typical Difference, Rand Wilcox
Journal of Modern Applied Statistical Methods
The paper describes a nonparametric analog of Cohen's d, Q. It is established that a confidence interval for Q can be computed via a method for computing a confidence interval for the median of D = X1 − X2, which in turn is related to making inferences about P(X1 < X2).
Should We Give Up On Causality?, Tom Knapp
Should We Give Up On Causality?, Tom Knapp
Journal of Modern Applied Statistical Methods
No abstract provided.
Striving For Simple But Effective Advice For Comparing The Central Tendency Of Two Populations, Graeme Ruxton, Markus Neuhäuser
Striving For Simple But Effective Advice For Comparing The Central Tendency Of Two Populations, Graeme Ruxton, Markus Neuhäuser
Journal of Modern Applied Statistical Methods
Nguyen et al. (2016) offered advice to researchers in the commonly-encountered situation where they are interested in testing for a difference in central tendency between two populations. Their data and the available literature support very simple advice that strikes the best balance between ease of implementation, power and reliability. Specifically, apply Satterthwaite’s test, with preliminary ranking of the data if a strong deviation from normality is expected, or is suggested by visual inspection of the data. This simple guideline will serve well except when dealing with small samples of discrete data, when more sophisticated treatment may be required.
Logistic Regression: An Inferential Method For Identifying The Best Predictors, Rand Wilcox
Logistic Regression: An Inferential Method For Identifying The Best Predictors, Rand Wilcox
Journal of Modern Applied Statistical Methods
When dealing with a logistic regression model, there is a simple method for estimating the strength of the association between the jth covariate and the dependent variable when all covariates are entered into the model. There is the issue of determining whether the jth independent variable has a stronger or weaker association than the kth independent variable. This note describes a method for dealing with this issue that was found to perform reasonably well in simulations.
Data Analytics Pipeline For Rna Structure Analysis Via Shape, Quinn Nelson
Data Analytics Pipeline For Rna Structure Analysis Via Shape, Quinn Nelson
UNO Student Research and Creative Activity Fair
Coxsackievirus B3 (CVB3) is a cardiovirulent enterovirus from the family Picornaviridae. The RNA genome houses an internal ribosome entry site (IRES) in the 5’ untranslated region (5’UTR) that enables cap-independent translation. Ample evidence suggests that the structure of the 5’UTR is a critical element for virulence. We probe RNA structure in solution using base-specific modifying agents such as dimethyl sulfate as well as backbone targeting agents such as N-methylisatoic anhydride used in Selective 2’-Hydroxyl Acylation Analyzed by Primer Extension (SHAPE). We have developed a pipeline that merges and evaluates base-specific and SHAPE data together with statistical analyses that provides confidence …
Sustainable Energy Governance In South Tyrol (Italy): A Probabilistic Bipartite Network Model, Jessica Belest, Laura Secco, Elena Pisani, Alberto Caimo
Sustainable Energy Governance In South Tyrol (Italy): A Probabilistic Bipartite Network Model, Jessica Belest, Laura Secco, Elena Pisani, Alberto Caimo
Articles
At the national scale, almost all of the European countries have already achieved energy transition targets, while at the regional and local scales, there is still some potential to further push sustainable energy transitions. Regions and localities have the support of political, social, and economic actors who make decisions for meeting existing social, environmental and economic needs recognising local specificities.
These actors compose the sustainable energy governance that is fundamental to effectively plan and manage energy resources. In collaborative relationships, these actors share, save, and protect several kinds of resources, thereby making energy transitions deeper and more effective.
This research …
Session: 4 Multilinear Subspace Learning And Its Applications To Machine Learning, Randy Hoover, Kyle Caudle Dr., Karen Braman Dr.
Session: 4 Multilinear Subspace Learning And Its Applications To Machine Learning, Randy Hoover, Kyle Caudle Dr., Karen Braman Dr.
SDSU Data Science Symposium
Multi-dimensional data analysis has seen increased interest in recent years. With more and more data arriving as 2-dimensional arrays (images) as opposed to 1-dimensioanl arrays (signals), new methods for dimensionality reduction, data analysis, and machine learning have been pursued. Most notably have been the Canonical Decompositions/Parallel Factors (commonly referred to as CP) and Tucker decompositions (commonly regarded as a high order SVD: HOSVD). In the current research we present an alternate method for computing singular value and eigenvalue decompositions on multi-way data through an algebra of circulants and illustrate their application to two well-known machine learning methods: Multi-Linear Principal Component …
Predicting Unplanned Medical Visits Among Patients With Diabetes Using Machine Learning, Arielle Selya, Eric L. Johnson
Predicting Unplanned Medical Visits Among Patients With Diabetes Using Machine Learning, Arielle Selya, Eric L. Johnson
SDSU Data Science Symposium
Diabetes poses a variety of medical complications to patients, resulting in a high rate of unplanned medical visits, which are costly to patients and healthcare providers alike. However, unplanned medical visits by their nature are very difficult to predict. The current project draws upon electronic health records (EMR’s) of adult patients with diabetes who received care at Sanford Health between 2014 and 2017. Various machine learning methods were used to predict which patients have had an unplanned medical visit based on a variety of EMR variables (age, BMI, blood pressure, # of prescriptions, # of diagnoses on problem list, A1C, …
Nonparametric Depth And Quantile Regression For Functional Data, Joydeep Chowdhury, Probal Chaudhuri
Nonparametric Depth And Quantile Regression For Functional Data, Joydeep Chowdhury, Probal Chaudhuri
Journal Articles
We investigate nonparametric regression methods based on spatial depth and quantiles when the response and the covariate are both functions. As in classical quantile regression for finite dimensional data, regression techniques developed here provide insight into the influence of the functional covariate on different parts, like the center as well as the tails, of the conditional distribution of the functional response. Depth and quantile based nonparametric regression methods are useful to detect heteroscedasticity in functional regression. We derive the asymptotic behavior of the nonparametric depth and quantile regression estimates, which depend on the small ball probabilities in the covariate space. …
Pedestrian Safety -- Fundamental To A Walkable City, Joshua Herrera, Patrick Mcdevitt, Preeti Swaminathan, Raghuram Srinivas
Pedestrian Safety -- Fundamental To A Walkable City, Joshua Herrera, Patrick Mcdevitt, Preeti Swaminathan, Raghuram Srinivas
SMU Data Science Review
In this paper, we present a method to identify urban areas with a higher likelihood of pedestrian safety related events. Pedestrian safety related events are pedestrian-vehicle interactions that result in fatalities, injuries, accidents without injury, or near--misses between pedestrians and vehicles. To develop a solution to this problem of identifying likely event locations, we assemble data, primarily from the City of Cincinnati and Hamilton County, that include safety reports from a five year period, geographic information for these events, citizen survey of pedestrian reported concerns, non-emergency requests for service for any cause in the city, property values and public transportation …
Improving Vix Futures Forecasts Using Machine Learning Methods, James Hosker, Slobodan Djurdjevic, Hieu Nguyen, Robert Slater
Improving Vix Futures Forecasts Using Machine Learning Methods, James Hosker, Slobodan Djurdjevic, Hieu Nguyen, Robert Slater
SMU Data Science Review
The problem of forecasting market volatility is a difficult task for most fund managers. Volatility forecasts are used for risk management, alpha (risk) trading, and the reduction of trading friction. Improving the forecasts of future market volatility assists fund managers in adding or reducing risk in their portfolios as well as in increasing hedges to protect their portfolios in anticipation of a market sell-off event. Our analysis compares three existing financial models that forecast future market volatility using the Chicago Board Options Exchange Volatility Index (VIX) to six machine/deep learning supervised regression methods. This analysis determines which models provide best …
Ample Provision: A Preliminary Study Relating Budget Composition And High School Graduation Rates In Select Washington State Public School Districts, Gregory P. Gadow
Ample Provision: A Preliminary Study Relating Budget Composition And High School Graduation Rates In Select Washington State Public School Districts, Gregory P. Gadow
All Undergraduate Projects
How to allocate scarce resources for an optimal outcome is of keen interest to those who set the budgets in public education. Simply throwing money at schools is not enough; it is important that money is spent where it will do the most good. This study considers Washington State public school districts and examines how the share of per-student expenditures in seven budget categories relates to on-time high school graduation rates. It is an investigative study, exploring whether there is enough evidence to merit further, more in-depth research. Using budget and graduation information from academic years 1997-98 through 2016-17 for …
Comparative Analysis Of Students’ Performance Between Online And On Campus In An Introductory Statistics Course, Kendal Mcdonald
Comparative Analysis Of Students’ Performance Between Online And On Campus In An Introductory Statistics Course, Kendal Mcdonald
The Corinthian
In this research, we compare students’ performance in an online and on-campus introductory statistics and probability course at Georgia College. MyStatLab is the learning management system used in both the online and on-campus courses for homework and quizzes. The online data is produced by five summer courses between Summer 2014 to Summer 2017 and the on-campus data is produced from nine on-campus courses from Spring 2014, Spring 2016, and Spring 2017. For homework, the research compares the scores made between online and on-campus. For quizzes, we test if there is a difference between the scores and the number of attempts …
Step Away From Stepwise, Gary N. Smith
Step Away From Stepwise, Gary N. Smith
Pomona Economics
Stepwise regression is a popular data-mining tool that uses statistical significance to select the explanatory variables to be used in a multiple-regression model. A fundamental problem with stepwise regression is that some real explanatory variables that have causal effects on the dependent variable may happen to not be statistically significant, while nuisance variables may be coincidentally significant. As a result, the model may fit the data well in-sample, but do poorly out-of-sample. Many Big-Data researchers believe that, the larger the number of possible explanatory variables, the more useful is stepwise regression for selecting explanatory variables. The reality is that stepwise …
Controlling For Confounding Via Propensity Score Methods Can Result In Biased Estimation Of The Conditional Auc: A Simulation Study, Hadiza I. Galadima, Donna K. Mcclish
Controlling For Confounding Via Propensity Score Methods Can Result In Biased Estimation Of The Conditional Auc: A Simulation Study, Hadiza I. Galadima, Donna K. Mcclish
Community & Environmental Health Faculty Publications
In the medical literature, there has been an increased interest in evaluating association between exposure and outcomes using nonrandomized observational studies. However, because assignments to exposure are not random in observational studies, comparisons of outcomes between exposed and nonexposed subjects must account for the effect of confounders. Propensity score methods have been widely used to control for confounding, when estimating exposure effect. Previous studies have shown that conditioning on the propensity score results in biased estimation of conditional odds ratio and hazard ratio. However, research is lacking on the performance of propensity score methods for covariate adjustment when estimating the …
The Dark Sky Character Of Archaeological Landscapes: Cultural Meaning And Conservation Strategies, Frank Prendergast
The Dark Sky Character Of Archaeological Landscapes: Cultural Meaning And Conservation Strategies, Frank Prendergast
Book/Book Chapter
This paper presents the first ever study of light pollution at selected Irish prehistoric archaeological landscapes. The concepts of cosmology and landscape are first briefly described and followed by a summary of early human settlement of the island. Building on this, the extant corpus of early prehistoric megalithic burial tombs is illustrated to show their contrasting distribution patterns and typology. Analysis of tomb locations using nearest-neighbour statistical methods reveals evidence of intentional clustering. Further geo-statistical analysis identifies the geographical locations and the density ranking of these nucleated clusters - a feature especially evident in the passage tomb tradition on this …
Be Wary Of Black-Box Trading Algorithms, Gary N. Smith
Be Wary Of Black-Box Trading Algorithms, Gary N. Smith
Pomona Economics
Black-box algorithms now account for nearly a third of all U. S. stock trades. It is a mistake to think that these algorithms possess superhuman intelligence. In reality, computers do not have the common sense and wisdom that humans have accumulated by living. Trading algorithms are particularly dangerous because they are so efficient at discovering statistical patterns—but so utterly useless in judging whether the discovered patterns are meaningful.