Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Statistical Models (162)
- Medicine and Health Sciences (154)
- Applied Statistics (142)
- Statistical Methodology (141)
- Social and Behavioral Sciences (134)
-
- Life Sciences (98)
- Data Science (73)
- Longitudinal Data Analysis and Time Series (71)
- Statistical Theory (71)
- Biostatistics (67)
- Categorical Data Analysis (66)
- Computer Sciences (61)
- Business (60)
- Public Health (58)
- Medical Specialties (53)
- Dentistry (51)
- Design of Experiments and Sample Surveys (51)
- Economics (51)
- Other Statistics and Probability (48)
- Probability (48)
- Econometrics (45)
- Engineering (45)
- Genetics and Genomics (33)
- Artificial Intelligence and Robotics (32)
- Macroeconomics (32)
- Mathematics (32)
- Bioinformatics (31)
- Institution
-
- Loma Linda University (90)
- COBRA (78)
- Central Bank of Nigeria (29)
- University of Kentucky (21)
- Hunan Provincial Institute of Scientific and Technology Information (13)
-
- City University of New York (CUNY) (12)
- California Polytechnic State University, San Luis Obispo (10)
- Southern Methodist University (10)
- University of Nevada, Las Vegas (10)
- Michigan Technological University (9)
- Stephen F. Austin State University (9)
- Virginia Commonwealth University (9)
- West Virginia University (9)
- East Tennessee State University (7)
- University of Arkansas, Fayetteville (7)
- University of Louisville (7)
- Air Force Institute of Technology (6)
- Claremont Colleges (6)
- Kennesaw State University (6)
- Louisiana State University (6)
- Portland State University (6)
- Dartmouth College (5)
- Georgia Southern University (5)
- Murray State University (5)
- University of Nebraska - Lincoln (5)
- Duquesne University (4)
- Illinois State University (4)
- LSU New Orleans (4)
- Missouri State University (4)
- South Dakota State University (4)
- Keyword
-
- Classification (13)
- Statistics (12)
- Machine learning (11)
- Data mining (9)
- Prediction (9)
-
- Machine Learning (8)
- Regression (8)
- American Southeast (7)
- Caddo (7)
- Factor analysis (7)
- Archaeology (6)
- Clustering (6)
- Cross-validation (6)
- Forecasting (6)
- Gene expression (6)
- Logistic regression (6)
- Ceramics (5)
- Cluster analysis (5)
- GIS (5)
- Genetics (5)
- Geochemistry (5)
- MANOVA (5)
- Model selection (5)
- Modeling (5)
- Survival analysis (5)
- Time series (5)
- Bootstrap (4)
- COVID-19 (4)
- Chemometrics (4)
- INAA (4)
- Publication Year
- Publication
-
- Loma Linda University Electronic Theses, Dissertations & Projects (90)
- CBN Journal of Applied Statistics (JAS) (29)
- Electronic Theses and Dissertations (23)
- U.C. Berkeley Division of Biostatistics Working Paper Series (20)
- UW Biostatistics Working Paper Series (18)
-
- Harvard University Biostatistics Working Paper Series (17)
- Theses and Dissertations (17)
- Theses and Dissertations--Statistics (15)
- Journal of Scientific Information Research (13)
- COBRA Preprint Series (10)
- Dissertations, Master's Theses and Master's Reports (9)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (9)
- SMU Data Science Review (8)
- CRHR: Archaeology (7)
- Graduate Theses, Dissertations, and Problem Reports (ETD) (7)
- Statistics (6)
- Complex Systems Faculty Publications and Presentations (5)
- Dissertations, Theses, and Capstone Projects (5)
- Graduate Theses and Dissertations (5)
- Master's Theses (5)
- College of Graduate Studies: Theses & Dissertations (4)
- Dartmouth Scholarship (4)
- Dissertations (4)
- Dissertations and Theses (Open Access) (4)
- Graduate Theses/Dissertations (4)
- LSU Master's Theses (4)
- LSU New Orleans Theses and Dissertations (4)
- SDSU Data Science Symposium (4)
- UNLV Theses, Dissertations, Professional Papers, and Capstones (4)
- Williams Honors College, Honors Research Projects (4)
- Publication Type
- File Type
Articles 181 - 210 of 522
Full-Text Articles in Multivariate Analysis
Characterizing Uncertainty In Correlated Response Variables For Pareto Front Optimization, Peter A. Calhoun
Characterizing Uncertainty In Correlated Response Variables For Pareto Front Optimization, Peter A. Calhoun
Theses and Dissertations
Current research provides a method to incorporate uncertainty into Pareto front optimization by simulating additional response surface model parameters according to a Multivariate Normal Distribution (MVN). This research shows that analogous to the univariate case, the MVN understates uncertainty, leading to overconfident conclusions when variance is not known and there are few observations (less than 25-30 per response). This research builds upon current methods using simulated response surface model parameters that are distributed according to an Multivariate t-Distribution (MVT), which can be shown to produce a more accurate inference when variance is not known. The MVT better addresses uncertainty in …
Quantitative Model For Setting Manufacturer's Suggested Retail Price, Peter Byrd, Jonathan Knowles, Dmitry Andreev, Jacob Turner, Brian Mente, Laroux Wallace
Quantitative Model For Setting Manufacturer's Suggested Retail Price, Peter Byrd, Jonathan Knowles, Dmitry Andreev, Jacob Turner, Brian Mente, Laroux Wallace
SMU Data Science Review
In this paper, we present a quantitative approach to model the manufacturer’s suggested retail price (MSRP) for children’s doll- houses and establish relationships among key features that contribute most to establishing MSRP. Determination of the MSRP is a critical step in how consumers respond with their wallets when purchasing an item. KidKraft, a global leader in toys and juvenile products, sets MSRP subjectively using product experts. The process is arduous and time consuming requiring the focus of specialized resources and knowledge of the interaction between key attributes and their impact on consumer value. An accurate prediction of MSRP during the …
Theory Of Principal Components For Applications In Exploratory Crime Analysis And Clustering, Daniel Silva
Theory Of Principal Components For Applications In Exploratory Crime Analysis And Clustering, Daniel Silva
All Graduate Theses, Dissertations, and Other Capstone Projects
The purpose of this paper is to develop the theory of principal components analysis succinctly from the fundamentals of matrix algebra and multivariate statistics. Principal components analysis is sometimes used as a descriptive technique to explain the variance-covariance or correlation structure of a dataset. However, most often, it is used as a dimensionality reduction technique to visualize a high dimensional dataset in a lower dimensional space. Principal components analysis accomplishes this by using the first few principal components, provided that they account for a substantial proportion of variation in the original dataset. In the same way, the first few principal …
Joint Simulation Of Continuous And Categorical Variables For Mineral Resource Modeling And Recoverable Reserves Calculation, Sentle Augustinus Hlajoane
Joint Simulation Of Continuous And Categorical Variables For Mineral Resource Modeling And Recoverable Reserves Calculation, Sentle Augustinus Hlajoane
Dissertations, Master's Theses and Master's Reports
Spatial variability and uncertainty of continuous variables (grade) and categorical variables (rock-types) in mineral evaluation significantly impact the economics of mining projects. The conventional approach of simulating grades using deterministic rock- types is problematic since spatial variability, and uncertainty of grades at rock-type contacts are not well captured in deposits where the grade changes gradually between rock-types. Therefore, jointly simulating these variables can improve confidence (reduce uncertainty) in a resource model. Also, resource classification and recoverable reserve calculation can significantly improve the understanding of the deposit and its economic viability. This research utilized the Plural-Gaussian geostatistical simulation to jointly simulate …
Zero-Inflated Longitudinal Mixture Model For Stochastic Radiographic Lung Compositional Change Following Radiotherapy Of Lung Cancer, Viviana A. Rodríguez Romero
Zero-Inflated Longitudinal Mixture Model For Stochastic Radiographic Lung Compositional Change Following Radiotherapy Of Lung Cancer, Viviana A. Rodríguez Romero
Theses and Dissertations
Compositional data (CD) is mostly analyzed as relative data, using ratios of components, and log-ratio transformations to be able to use known multivariable statistical methods. Therefore, CD where some components equal zero represent a problem. Furthermore, when the data is measured longitudinally, observations are spatially related and appear to come from a mixture population, the analysis becomes highly complex. For this matter, a two-part model was proposed to deal with structural zeros in longitudinal CD using a mixed-effects model. Furthermore, the model has been extended to the case where the non-zero components of the vector might a two component mixture …
Nonparametric Analysis Of Clustered And Multivariate Data, Yue Cui
Nonparametric Analysis Of Clustered And Multivariate Data, Yue Cui
Theses and Dissertations--Statistics
In this dissertation, we investigate three distinct but interrelated problems for nonparametric analysis of clustered data and multivariate data in pre-post factorial design.
In the first project, we propose a nonparametric approach for one-sample clustered data in pre-post intervention design. In particular, we consider the situation where for some clusters all members are only observed at either pre or post intervention but not both. This type of clustered data is referred to us as partially complete clustered data. Unlike most of its parametric counterparts, we do not assume specific models for data distributions, intra-cluster dependence structure or variability, in effect …
Nonparametric Tests Of Lack Of Fit For Multivariate Data, Yan Xu
Nonparametric Tests Of Lack Of Fit For Multivariate Data, Yan Xu
Theses and Dissertations--Statistics
A common problem in regression analysis (linear or nonlinear) is assessing the lack-of-fit. Existing methods make parametric or semi-parametric assumptions to model the conditional mean or covariance matrices. In this dissertation, we propose fully nonparametric methods that make only additive error assumptions. Our nonparametric approach relies on ideas from nonparametric smoothing to reduce the test of association (lack-of-fit) problem into a nonparametric multivariate analysis of variance. A major problem that arises in this approach is that the key assumptions of independence and constant covariance matrix among the groups will be violated. As a result, the standard asymptotic theory is not …
Moment Kernels For T-Central Subspace, Weihang Ren
Moment Kernels For T-Central Subspace, Weihang Ren
Theses and Dissertations--Statistics
The T-central subspace allows one to perform sufficient dimension reduction for any statistical functional of interest. We propose a general estimator using a third moment kernel to estimate the T-central subspace. In particular, in this dissertation we develop sufficient dimension reduction methods for the central mean subspace via the regression mean function and central subspace via Fourier transform, central quantile subspace via quantile estimator and central expectile subsapce via expectile estima- tor. Theoretical results are established and simulation studies show the advantages of our proposed methods.
An Assessment Of Convergence In The Feeding Morphology Of Xiphactinus Audax And Megalops Atlanticus Using Landmark-Based Geometric Morphometrics, Edward Chase Shelburne
An Assessment Of Convergence In The Feeding Morphology Of Xiphactinus Audax And Megalops Atlanticus Using Landmark-Based Geometric Morphometrics, Edward Chase Shelburne
Master's Theses or Doctor of Nursing Practice
Convergence is an evolutionary phenomenon wherein distantly related organisms independently develop features or functional adaptations to overcome similar environmental constraints. Historically, convergence among organisms has been speculated or asserted with little rigorous or quantitative investigation. More recent advancements in systematics has allowed for the detection and study of convergence in a phylogenetic context, but this does little to elucidate convergent anatomical features in extinct taxa with poorly understood evolutionary histories. The purpose of this study is to investigate one potentially convergent system—the feeding structure of Xiphactinus audax (Teleostei: Ichthyodectiformes) and Megalops atlanticus (Teleostei: Elopiformes)—using a comparative anatomical approach to assess …
Process Based Analysis Of Fluvial Stratigraphic Record: Middle Pennsylvanian Allegheny Formation, North-Central Wv, Oluwasegun O. Abatan
Process Based Analysis Of Fluvial Stratigraphic Record: Middle Pennsylvanian Allegheny Formation, North-Central Wv, Oluwasegun O. Abatan
Graduate Theses, Dissertations, and Problem Reports (ETD)
Fluvial deposits represent some of the best hydrocarbon reservoirs, but the quality of fluvial reservoirs varies depending on the reservoir architecture, which is controlled by allogenic and autogenic processes. Allogenic controls, including paleoclimate, tectonics, and glacio-eustasy, have long been debated as dominant controls in the deposition of fluvial strata. However, recent research has questioned the validity of this cyclicity and may indicate major influence from autogenic controls. To further investigate allogenic controls on stratal order, I analyzed the facies architecture, geomorphology, paleohydrology, and the stratigraphic framework of the Middle Pennsylvanian Allegheny Formation (MPAF), a fluvial depositional system in the Appalachian …
Accounting For The Uncertainty Due To Chemicals Below The Detection Limit In Mixture Analysis, Paul M. Hargarten
Accounting For The Uncertainty Due To Chemicals Below The Detection Limit In Mixture Analysis, Paul M. Hargarten
Theses and Dissertations
Humans are exposed to multiple chemicals every day. Epidemiological studies have shown that chemical mixtures are associated with cancers, allergies, neurodevelopmental disorders, and other adverse health effects. To assess these associations, investigators are increasingly using chemical mixture approaches like weighted quantile sum (WQS) regression. In these studies, the research objectives are to determine whether a mixture of correlated chemicals is associated with an adverse health outcome and to identify the important chemicals. However, as experimental equipment measures each exposure to a chemical-specific detection limit, the exposures are unknown between zero and the detection limit. Indeed, the number of exposures below …
Three Essays On Health Economics And Policy Evaluation, Shishir Shakya
Three Essays On Health Economics And Policy Evaluation, Shishir Shakya
Graduate Theses, Dissertations, and Problem Reports (ETD)
This dissertation consists of three essays on the U.S. Health care policy. Each paragraph below refers to the three abstracts for the three chapters in this dissertation, respectively. I provide quantitative evidence on how much Prescription Drug Monitoring Programs (PDMPs) affects the retail opioid prescribing behaviors. Using the American Community Survey (ACS), I retrieve county-level high dimensional panel data set from 2010 to 2017. I employ three separate identification strategies: difference-in-difference, double selection post-LASSO, and spatial difference-in-difference. I compare how the retail opioid prescribing behaviors of counties, that are mandatory for prescribers to check the PDMP before prescribing controlled substances …
Inventory Models For Perishable Items Under Markdown Policy, Nurzahara Atika Kamaruzaman
Inventory Models For Perishable Items Under Markdown Policy, Nurzahara Atika Kamaruzaman
Student Works (2020-2029)
As expected, the demand for a fresh product depends on how fresh it is, therefore, it is important to take expiration date into consideration. Based on marketing and economic theory, several factors such as price, inventory level and advertisement play a crucial role in influencing the demand. Hence, we study the effect of these factors in influencing the demand in the inventory model. Since the demand for perishable product declines over time, markdown policy is offered to increase the demand and profit while reducing the inventory. Salvage value is incorporated to the deteriorating units. In this research, we extend previous …
Generalized Matrix Decomposition Regression: Estimation And Inference For Two-Way Structured Data, Yue Wang, Ali Shojaie, Tim Randolph, Jing Ma
Generalized Matrix Decomposition Regression: Estimation And Inference For Two-Way Structured Data, Yue Wang, Ali Shojaie, Tim Randolph, Jing Ma
UW Biostatistics Working Paper Series
Analysis of two-way structured data, i.e., data with structures among both variables and samples, is becoming increasingly common in ecology, biology and neuro-science. Classical dimension-reduction tools, such as the singular value decomposition (SVD), may perform poorly for two-way structured data. The generalized matrix decomposition (GMD, Allen et al., 2014) extends the SVD to two-way structured data and thus constructs singular vectors that account for both structures. While the GMD is a useful dimension-reduction tool for exploratory analysis of two-way structured data, it is unsupervised and cannot be used to assess the association between such data and an outcome of interest. …
Statistical Inference For Networks Of High-Dimensional Point Processes, Xu Wang, Mladen Kolar, Ali Shojaie
Statistical Inference For Networks Of High-Dimensional Point Processes, Xu Wang, Mladen Kolar, Ali Shojaie
UW Biostatistics Working Paper Series
Fueled in part by recent applications in neuroscience, high-dimensional Hawkes process have become a popular tool for modeling the network of interactions among multivariate point process data. While evaluating the uncertainty of the network estimates is critical in scientific applications, existing methodological and theoretical work have only focused on estimation. To bridge this gap, this paper proposes a high-dimensional statistical inference procedure with theoretical guarantees for multivariate Hawkes process. Key to this inference procedure is a new concentration inequality on the first- and second-order statistics for integrated stochastic processes, which summarizes the entire history of the process. We apply this …
Function Space Tensor Decomposition And Its Application In Sports Analytics, Justin Reising
Function Space Tensor Decomposition And Its Application In Sports Analytics, Justin Reising
Electronic Theses and Dissertations
Recent advancements in sports information and technology systems have ushered in a new age of applications of both supervised and unsupervised analytical techniques in the sports domain. These automated systems capture large volumes of data points about competitors during live competition. As a result, multi-relational analyses are gaining popularity in the field of Sports Analytics. We review two case studies of dimensionality reduction with Principal Component Analysis and latent factor analysis with Non-Negative Matrix Factorization applied in sports. Also, we provide a review of a framework for extending these techniques for higher order data structures. The primary scope of this …
Habitat Associations And Reproduction Of Fishes On The Northwestern Gulf Of Mexico Shelf Edge, Elizabeth Marie Keller
Habitat Associations And Reproduction Of Fishes On The Northwestern Gulf Of Mexico Shelf Edge, Elizabeth Marie Keller
LSU Doctoral Dissertations
Several of the northwestern Gulf of Mexico (GOM) shelf-edge banks provide critical hard bottom habitat for coral and fish communities, supporting a wide diversity of ecologically and economically important species. These sites may be fish aggregation and spawning sites and provide important habitat for fish growth and reproduction. Already designated as habitat areas of particular concern, many of these banks are also under consideration for inclusion in the expansion of the Flower Garden Banks National Marine Sanctuary. This project aimed to gain a more comprehensive understanding of the communities and fish species on shelf-edge banks by way of gonad histology, …
Classification Of Coronary Artery Disease In Non-Diabetic Patients Using Artificial Neural Networks, Demond Handley
Classification Of Coronary Artery Disease In Non-Diabetic Patients Using Artificial Neural Networks, Demond Handley
Annual Symposium on Biomathematics and Ecology Education and Research
No abstract provided.
Identifying Risk Factors Related To Premature Birth Through Binary Logistic And Proportional Odds Ordinal Logistic Regression, Clayton Elwood
Identifying Risk Factors Related To Premature Birth Through Binary Logistic And Proportional Odds Ordinal Logistic Regression, Clayton Elwood
Electronic Theses and Dissertations
Premature birth has been identified as the single greatest cause of death worldwide in children under the age of five. This thesis will implement binary logistic regression and proportional odds ordinal logistic regression to predict different levels of premature birth and identify associated risk factors. The models will be built from the Center for Disease Control and Prevention's 2014 Vital Statistics Natality Birth Data containing nearly 4 million live births within the United States. Odds ratios and confidence intervals on risk factors were produced utilizing binary logistic regression.
Optimal Design For A Causal Structure, Zaher Kmail
Optimal Design For A Causal Structure, Zaher Kmail
Department of Statistics: Dissertations, Theses, and Student Research
Linear models and mixed models are important statistical tools. But in many natural phenomena, there is more than one endogenous variable involved and these variables are related in a sophisticated way. Structural Equation Modeling (SEM) is often used to model the complex relationships between the endogenous and exogenous variables. It was first implemented in research to estimate the strength and direction of direct and indirect effects among variables and to measure the relative magnitude of each causal factor.
Historically, traditional optimal design theory focuses on univariate linear, nonlinear, and mixed models. There is no current literature on the subject of …
Taking Multiple Regression Analysis To Task: A Review Of Mindware: Tools For Smart Thinking, By Richard Nisbett (2015), Jason Makansi
Taking Multiple Regression Analysis To Task: A Review Of Mindware: Tools For Smart Thinking, By Richard Nisbett (2015), Jason Makansi
Numeracy
Richard Nisbett. 2015. Mindware: Tools for Smart Thinking.(New York, NY: Farrar, Strauss, and Giroux). 336 pp. ISBN: 9780374536244
Nisbett, a psychologist, may not achieve his stated goal of teaching readers to “effortlessly” extend their common sense when it comes to quantitative analysis applied to everyday issues, but his critique of multiple regression analysis (MRA) in the middle chapters of Mindware is worth attention from, and contemplation by, the QL/QR and Numeracy community. While in at least one other source, Nisbett’s critique has been called a “crusade” against MRA, what he really advocates is that it not be used as …
Implementation Of Multivariate Artificial Neural Networks Coupled With Genetic Algorithms For The Multi-Objective Property Prediction And Optimization Of Emulsion Polymers, David Chisholm
Master's Theses
Machine learning has been gaining popularity over the past few decades as computers have become more advanced. On a fundamental level, machine learning consists of the use of computerized statistical methods to analyze data and discover trends that may not have been obvious or otherwise observable previously. These trends can then be used to make predictions on new data and explore entirely new design spaces. Methods vary from simple linear regression to highly complex neural networks, but the end goal is similar. The application of these methods to material property prediction and new material discovery has been of high interest …
Analyzing Two-Year College Student Success Using Structural Equation Modeling, Jessica Taylor
Analyzing Two-Year College Student Success Using Structural Equation Modeling, Jessica Taylor
Graduate Theses, Dissertations, and Capstones
The goal of this study is to more fully understand the scope of community college student success using the principles of mindset, engagement, and college readiness. Using structural equation modeling ensures this study is able to measure the combined effects these concepts have on student success, group differences, and the combined model of student success. Findings suggest student success can be significantly impacted by self-belief and mindset behaviors that can outweigh the initial effect of academically under-prepared students. Groups included in this study are non-traditional students, minority populations, first generation students, and Pell eligible students.
Leveraging Reviews To Improve User Experience, Anthony Schams, Iram Bakhtiar, Cristina Stanley
Leveraging Reviews To Improve User Experience, Anthony Schams, Iram Bakhtiar, Cristina Stanley
SMU Data Science Review
In this paper, we will explore and present a method of finding characteristics of a restaurant using its reviews through machine learning algorithms. We begin by building models to predict the ratings of individual reviews using text and categorical features. This is to examine the efficacy of the algorithms to the task. Both XGBoost and logistic regression will be examined. With these models, our goal is then to identify key phrases in reviews that are correlated with positive and negative experience. Our analysis makes use of review data publicly made available by Yelp. Key bigrams extracted were non-specific to the …
A Systematic Assessment Of Socio-Economic Impacts Of Prolonged Episodic Volcano Crises, Justin Peers
A Systematic Assessment Of Socio-Economic Impacts Of Prolonged Episodic Volcano Crises, Justin Peers
Electronic Theses and Dissertations
Uncertainty surrounding volcanic activity can lead to socio-economic crises with or without an eruption as demonstrated by the post-1978 response to unrest of Long Valley Caldera (LVC), CA. Extensive research in physical sciences provides a foundation on which to assess direct impacts of hazards, but fewer resources have been dedicated towards understanding human responses to volcanic risk. To evaluate natural hazard risk issues at LVC, a multi-hazard, mail-based, household survey was conducted to compare perceptions of volcanic, seismic, and wildfire hazards. Impacts of volcanic activity on housing prices and businesses were examined at the county-level for three volcanoes with a …
Comparison Of Imputation Methods For Mixed Data Missing At Random, Kaitlyn Heidt
Comparison Of Imputation Methods For Mixed Data Missing At Random, Kaitlyn Heidt
Electronic Theses and Dissertations
A statistician's job is to produce statistical models. When these models are precise and unbiased, we can relate them to new data appropriately. However, when data sets have missing values, assumptions to statistical methods are violated and produce biased results. The statistician's objective is to implement methods that produce unbiased and accurate results. Research in missing data is becoming popular as modern methods that produce unbiased and accurate results are emerging, such as MICE in R, a statistical software. Using real data, we compare four common imputation methods, in the MICE package in R, at different levels of missingness. The …
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes, Lauren Spirko-Burns, Karthik Devarajan
Unified Methods For Feature Selection In Large-Scale Genomic Studies With Censored Survival Outcomes, Lauren Spirko-Burns, Karthik Devarajan
COBRA Preprint Series
One of the major goals in large-scale genomic studies is to identify genes with a prognostic impact on time-to-event outcomes which provide insight into the disease's process. With rapid developments in high-throughput genomic technologies in the past two decades, the scientific community is able to monitor the expression levels of tens of thousands of genes and proteins resulting in enormous data sets where the number of genomic features is far greater than the number of subjects. Methods based on univariate Cox regression are often used to select genomic features related to survival outcome; however, the Cox model assumes proportional hazards …
Predicting Unplanned Medical Visits Among Patients With Diabetes Using Machine Learning, Arielle Selya, Eric L. Johnson
Predicting Unplanned Medical Visits Among Patients With Diabetes Using Machine Learning, Arielle Selya, Eric L. Johnson
SDSU Data Science Symposium
Diabetes poses a variety of medical complications to patients, resulting in a high rate of unplanned medical visits, which are costly to patients and healthcare providers alike. However, unplanned medical visits by their nature are very difficult to predict. The current project draws upon electronic health records (EMR’s) of adult patients with diabetes who received care at Sanford Health between 2014 and 2017. Various machine learning methods were used to predict which patients have had an unplanned medical visit based on a variety of EMR variables (age, BMI, blood pressure, # of prescriptions, # of diagnoses on problem list, A1C, …
Nonparametric Depth And Quantile Regression For Functional Data, Joydeep Chowdhury, Probal Chaudhuri
Nonparametric Depth And Quantile Regression For Functional Data, Joydeep Chowdhury, Probal Chaudhuri
Journal Articles
We investigate nonparametric regression methods based on spatial depth and quantiles when the response and the covariate are both functions. As in classical quantile regression for finite dimensional data, regression techniques developed here provide insight into the influence of the functional covariate on different parts, like the center as well as the tails, of the conditional distribution of the functional response. Depth and quantile based nonparametric regression methods are useful to detect heteroscedasticity in functional regression. We derive the asymptotic behavior of the nonparametric depth and quantile regression estimates, which depend on the small ball probabilities in the covariate space. …
Estimation Of Multivariate Asset Models With Jumps, Angela Loregian, Laura Ballotta, Gianluca Gianluca Fusai, Marcos Fabricio Perez
Estimation Of Multivariate Asset Models With Jumps, Angela Loregian, Laura Ballotta, Gianluca Gianluca Fusai, Marcos Fabricio Perez
Business Faculty Publications
We propose a consistent and computationally efficient two-step methodology for the estimation of multidimensional non-Gaussian asset models built using Levy processes. The proposed framework allows for dependence between assets and different tail behaviors and jump structures for each asset. Our procedure can be applied to portfolios with a large number of assets as it is immune to estimation dimensionality problems. Simulations show good finite sample properties and significant efficiency gains. This method is especially relevant for risk management purposes such as, for example, the computation of portfolio Value at Risk and intra-horizon Value at Risk, as we show in detail …