Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- Changsha University of Science and Technology (570)
- COBRA (362)
- University of Kentucky (43)
- Central Bank of Nigeria (28)
- Southern Methodist University (27)
-
- Virginia Commonwealth University (23)
- Kennesaw State University (20)
- University of Denver (18)
- City University of New York (CUNY) (17)
- University of Arkansas, Fayetteville (16)
- University of Nebraska - Lincoln (16)
- Georgia Southern University (15)
- California Polytechnic State University, San Luis Obispo (14)
- Michigan Technological University (14)
- University of Louisville (13)
- University of Nevada, Las Vegas (13)
- Old Dominion University (12)
- Air Force Institute of Technology (11)
- Western Michigan University (11)
- Claremont Colleges (9)
- East Tennessee State University (9)
- The Texas Medical Center Library (9)
- The University of Akron (9)
- Bethel University (8)
- Clemson University (8)
- Rochester Institute of Technology (8)
- South Dakota State University (8)
- University of Central Florida (8)
- University of North Florida (8)
- West Virginia University (8)
- Keyword
-
- Road engineering (53)
- Statistics (50)
- Bridge engineering (31)
- Numerical simulation (31)
- Machine learning (18)
-
- Regression (18)
- Cable-stayed bridge (16)
- Causal inference (16)
- Prediction (16)
- Asphalt pavement (15)
- Simulation (15)
- Morgridge College of Education (14)
- Research Methods and Information Science (14)
- Research Methods and Statistics (14)
- Bootstrap (13)
- Missing data (13)
- Tunnel engineering (13)
- Machine Learning (12)
- Mechanical property (12)
- Model selection (12)
- Subgrade engineering (12)
- Classification (11)
- Genetics (11)
- Suspension bridge (11)
- Multiple testing (10)
- Psychology (10)
- Concrete (9)
- Cross-validation (9)
- Logistic regression (9)
- Stability (9)
- Publication Year
- Publication
-
- Journal of China & Foreign Highway (570)
- U.C. Berkeley Division of Biostatistics Working Paper Series (114)
- Harvard University Biostatistics Working Paper Series (74)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (59)
- UW Biostatistics Working Paper Series (56)
-
- Electronic Theses and Dissertations (41)
- Theses and Dissertations--Statistics (38)
- Theses and Dissertations (34)
- CBN Journal of Applied Statistics (JAS) (28)
- The University of Michigan Department of Biostatistics Working Paper Series (28)
- COBRA Preprint Series (24)
- Statistical Science Theses and Dissertations (16)
- Symposium of Student Scholars (15)
- College of Graduate Studies: Theses & Dissertations (14)
- Dissertations, Master's Theses and Master's Reports (14)
- Dissertations (13)
- Graduate Theses and Dissertations (13)
- Articles (11)
- SMU Data Science Review (10)
- Master's Theses (9)
- Williams Honors College, Honors Research Projects (9)
- All Dissertations (8)
- Dissertations and Theses (Open Access) (8)
- Psychology Student Works (8)
- UNF Graduate Theses and Dissertations (8)
- SDSU Data Science Symposium (7)
- Data Science and Data Mining (6)
- Publications and Research (6)
- Department of Statistics: Dissertations, Theses, and Student Research (5)
- Dissertations and Theses (5)
- Publication Type
- File Type
Articles 871 - 900 of 1562
Full-Text Articles in Statistics and Probability
Innovative Statistical Models In Cancer Immunotherapy Trial Design, Jing Wei
Innovative Statistical Models In Cancer Immunotherapy Trial Design, Jing Wei
Theses and Dissertations--Statistics
A challenge arising in cancer immunotherapy trial design is the presence of non-proportional hazards (NPH) patterns in survival curves. We considered three different NPH patterns caused by delayed treatment effect, cure rate and responder rate of treatment group in this dissertation. These three NPH patterns would violate the proportional hazard model assumption and ignoring any of them in an immunotherapy trial design will result in substantial loss of statistical power.
In this dissertation, four models to deal with NPH patterns are discussed. First, a piecewise proportional hazards model is proposed to incorporate delayed treatment effect into the trial design consideration. …
Comparing Various Robust Estimation Techniques In Regression Analysis, Tracy S. Morrison
Comparing Various Robust Estimation Techniques In Regression Analysis, Tracy S. Morrison
All Graduate Theses, Dissertations, and Other Capstone Projects
In regression analysis, the use of the ordinary least squares (OLS) method is inadvisable when dealing with outlier or extreme observations. As a result, we require a method of robust estimation in which the estimation value is not significantly affected by outlier or extreme observations. Four methods of estimation will be compared in this paper in order to determine the best estimation: the M estimation method, the Least Trimmed Square Estimator, the S-estimation method, and the MM estimation method in robust regression. We discover that the best method is the MM-estimation method in this study. The M-estimation method is an …
The Need To Incorporate Communities In Compartmental Models, Michael J. Kane, Owais Gilani
The Need To Incorporate Communities In Compartmental Models, Michael J. Kane, Owais Gilani
Faculty Journal Articles
Tian et al. provide a framework for assessing population- level interventions of disease outbreaks through the construction of counterfactuals in a large-scale, natural experiment assessing the efficacy of mild, but early interventions compared to delayed interventions. The technique is applied to the recent SARS-CoV-2 outbreak with the population of Shenzhen, China acting as the mild-but-early treatment group and a combination of several US counties resembling Shenzhen but enacting a delayed intervention acting as the control. To help further the development of this framework and identify an avenue for further enhancement, we focus on the use and potential limitations of compartmental …
An Evaluation Of Knot Placement Strategies For Spline Regression, William Klein
An Evaluation Of Knot Placement Strategies For Spline Regression, William Klein
CMC Senior Theses
Regression splines have an established value for producing quality fit at a relatively low-degree polynomial. This paper explores the implications of adopting new methods for knot selection in tandem with established methodology from the current literature. Structural features of generated datasets, as well as residuals collected from sequential iterative models are used to augment the equidistant knot selection process. From analyzing a simulated dataset and an application onto the Racial Animus dataset, I find that a B-spline basis paired with equally-spaced knots remains the best choice when data are evenly distributed, even when structural features of a dataset are known …
A Class Of Copula-Based Bivariate Poisson Time Series Models With Applications, Mohammed Alqawba, Dimuthu Fernando, Norou Diawara
A Class Of Copula-Based Bivariate Poisson Time Series Models With Applications, Mohammed Alqawba, Dimuthu Fernando, Norou Diawara
Mathematics & Statistics Faculty Publications
A class of bivariate integer-valued time series models was constructed via copula theory. Each series follows a Markov chain with the serial dependence captured using copula-based transition probabilities from the Poisson and the zero-inflated Poisson (ZIP) margins. The copula theory was also used again to capture the dependence between the two series using either the bivariate Gaussian or “t-copula” functions. Such a method provides a flexible dependence structure that allows for positive and negative correlation, as well. In addition, the use of a copula permits applying different margins with a complicated structure such as the ZIP distribution. Likelihood-based inference was …
Novel Statistical Analysis In The Context Of A Comprehensive Needs Assessment For Secondary Stem Recruitment, Norou Diawara, Sarah Ferguson, Melva Grant, Kumer Das
Novel Statistical Analysis In The Context Of A Comprehensive Needs Assessment For Secondary Stem Recruitment, Norou Diawara, Sarah Ferguson, Melva Grant, Kumer Das
Mathematics & Statistics Faculty Publications
There is a myriad of career opportunities stemming from science, technology, engineering, and mathematics (STEM) disciplines. In addition to careers in corporate settings, teaching is a viable career option for individuals pursuing degrees in STEM disciplines. With national shortages of secondary STEM teachers, efforts to recruit, train, and retain quality STEM teachers is greatly important. Prior to exploring ways to attract potential STEM teacher candidates to pursue teacher training programs, it is important to understand the perceived value that potential recruits place on STEM careers, disciplines, and the teaching profession. The purpose of this study was to explore students’ perceptions …
Addressing The Ecological Fallacy With Lagrangian Inference, Michael Schwob
Addressing The Ecological Fallacy With Lagrangian Inference, Michael Schwob
Calvert Undergraduate Research Awards
Most epidemiologists elect to use statistical models that use population-level data to make inference on the spread of some virus or disease. This has become commonplace in the fields of epidemiology and biostatistics since most data used to construct and verify epidemic models are recorded at the population-level. Obtaining inference from a population-level model may be beneficial in studying the spread of disease in a homogeneous population, but the use of such models to describe a heterogeneous population results in inadequate inference. The inaccuracy of these models is further amplified when one tries to make individual-level inference from these population-level …
Bayesian Experimental Design For Bayesian Hierarchical Models With Differential Equations For Ecological Applications, Rebecca Atanga
Bayesian Experimental Design For Bayesian Hierarchical Models With Differential Equations For Ecological Applications, Rebecca Atanga
Theses and Dissertations
Ecologists are interested in the composition of species in various ecosystems. Studying population dynamics can assist environmental managers in making better decisions for the environment. Traditionally, the sampling of species has been recorded on a regular time frequency. However, sampling can be an expensive process due to financial and physical constraints. In some cases the environments are threatening, and ecologists prefer to limit their time collecting data in the field. Rather than convenience sampling, a statistical approach is introduced to improve data collection methods for ecologists by studying the dynamics associated with populations of interest. Population models including the logistic …
Upper-Sided Ewma-Based Distribution-Specific Tolerance Limits, Owen Visser
Upper-Sided Ewma-Based Distribution-Specific Tolerance Limits, Owen Visser
UNF Graduate Theses and Dissertations
Tolerance limits are constructed from sample data to ascertain if a proportion of a process is within specification limits. There exists multiple methods of calculating the sample size requirements for tolerance limits under various assumptions. In this research, a distribution-specific algorithm that utilizes the exponentially weighted moving average technique (EWMA), first introduced by Sa and Razaila (2004), is reconstructed. The algorithm is used to calculate the required sample sizes for continuous construction of upper-sided tolerance limits. The sample sizes and intervals constructed from them are compared to three existing methods for various distributions. The distribution-specific algorithm was observed to reduce …
Statistical Methods In Genetic Studies, Cheng Gao
Statistical Methods In Genetic Studies, Cheng Gao
Dissertations, Master's Theses and Master's Reports
This dissertation includes three Chapters. A brief description of each chapter is organized as follows.
In Chapter 1, we proposed a new method, called MF-TOWmuT, for genome-wide association studies with multiple genetic variants and multiple phenotypes using family samples. MF-TOWmuT uses kinship matrix to account for sample relatedness. It is worth mentioning that in simulations, we considered hidden polygenic effects and varied the proportion of variance contributed by it to generate phenotypes. Simulation studies show that MF-TOWmuT can preserve the type I error rates and is more powerful than several existing methods in different simulation scenarios, MFTOWmuT is also quite …
Bayesian Semi-Supervised Keyphrase Extraction And Jackknife Empirical Likelihood For Assessing Heterogeneity In Meta-Analysis, Guanshen Wang
Bayesian Semi-Supervised Keyphrase Extraction And Jackknife Empirical Likelihood For Assessing Heterogeneity In Meta-Analysis, Guanshen Wang
Statistical Science Theses and Dissertations
This dissertation investigates: (1) A Bayesian Semi-supervised Approach to Keyphrase Extraction with Only Positive and Unlabeled Data, (2) Jackknife Empirical Likelihood Confidence Intervals for Assessing Heterogeneity in Meta-analysis of Rare Binary Events.
In the big data era, people are blessed with a huge amount of information. However, the availability of information may also pose great challenges. One big challenge is how to extract useful yet succinct information in an automated fashion. As one of the first few efforts, keyphrase extraction methods summarize an article by identifying a list of keyphrases. Many existing keyphrase extraction methods focus on the unsupervised setting, …
Improved Statistical Methods For Time-Series And Lifetime Data, Xiaojie Zhu
Improved Statistical Methods For Time-Series And Lifetime Data, Xiaojie Zhu
Statistical Science Theses and Dissertations
In this dissertation, improved statistical methods for time-series and lifetime data are developed. First, an improved trend test for time series data is presented. Then, robust parametric estimation methods based on system lifetime data with known system signatures are developed.
In the first part of this dissertation, we consider a test for the monotonic trend in time series data proposed by Brillinger (1989). It has been shown that when there are highly correlated residuals or short record lengths, Brillinger’s test procedure tends to have significance level much higher than the nominal level. This could be related to the discrepancy between …
Multi-Level Small Area Estimation Based On Calibrated Hierarchical Likelihood Approach Through Bias Correction With Applications To Covid-19 Data, Nirosha Rathnayake
Multi-Level Small Area Estimation Based On Calibrated Hierarchical Likelihood Approach Through Bias Correction With Applications To Covid-19 Data, Nirosha Rathnayake
Theses & Dissertations
Small area estimation (SAE) has been widely used in a variety of applications to draw estimates in geographic domains represented as a metropolitan area, district, county, or state. The direct estimation methods provide accurate estimates when the sample size of study participants within each area unit is sufficiently large, but it might not always be realistic to have large sample sizes of study participants when considering small geographical regions. Meanwhile, high dimensional socio-ecological data exist at the community level, providing an opportunity for model-based estimation by incorporating rich auxiliary information at the individual and area levels. Thus, it is critical …
Development Of An Effect Size To Classify The Magnitude Of Dif In Dichotomous And Polytomous Items, James D. Weese
Development Of An Effect Size To Classify The Magnitude Of Dif In Dichotomous And Polytomous Items, James D. Weese
Graduate Theses and Dissertations
A standardized effect size for the SIBTEST/POLYSIBTEST procedure is proposed, allowing for Differential Item Functioning (DIF) to be classified with a single set of DIF heuristics regardless of whether data are dichotomous or polytomous. This proposed standardized effect size accounts for both variability in responses and whether participants are included in the SIBTEST/POLYSIBTEST calculations. First, a new set of unstandardized effect size heuristics are established for dichotomous data that are more aligned with Educational Testing Service (ETS) standards using two and three parameter logistic (2PL and 3PL) models. Second, a standardized effect size is proposed and compared to other DIF …
Statistical Approaches Of Gene Set Analysis With Quantitative Trait Loci For High-Throughput Genomic Studies., Samarendra Das
Statistical Approaches Of Gene Set Analysis With Quantitative Trait Loci For High-Throughput Genomic Studies., Samarendra Das
Electronic Theses and Dissertations
Recently, gene set analysis has become the first choice for gaining insights into the underlying complex biology of diseases through high-throughput genomic studies, such as Microarrays, bulk RNA-Sequencing, single cell RNA-Sequencing, etc. It also reduces the complexity of statistical analysis and enhances the explanatory power of the obtained results. Further, the statistical structure and steps common to these approaches have not yet been comprehensively discussed, which limits their utility. Hence, a comprehensive overview of the available gene set analysis approaches used for different high-throughput genomic studies is provided. The analysis of gene sets is usually carried out based on …
Quantifying The Simultaneous Effect Of Socio-Economic Predictors And Build Environment On Spatial Crime Trends, Alfieri Daniel Ek
Quantifying The Simultaneous Effect Of Socio-Economic Predictors And Build Environment On Spatial Crime Trends, Alfieri Daniel Ek
Graduate Theses and Dissertations
Proper allocation of law enforcement agencies falls under the umbrella of risk terrainmodeling (Caplan et al., 2011, 2015; Drawve, 2016) that primarily focuses on crime prediction and prevention by spatially aggregating response and predictor variables of interest. Although mental health incidents demand resource allocation from law enforcement agencies and the city, relatively less emphasis has been placed on building spatial models for mental health incidents events. Analyzing spatial mental health events in Little Rock, AR over 2015 to 2018, we found evidence of spatial heterogeneity via Moran’s I statistic. A spatial modeling framework is then built using generalized linear models, …
On Simes’S Second Conjecture: An Extended Single-Step Simes Test Procedure For Multiple Testing, Matthew G. Hudson
On Simes’S Second Conjecture: An Extended Single-Step Simes Test Procedure For Multiple Testing, Matthew G. Hudson
Dissertations
One of the major concerns with multiple tests of significance is controlling the family wise error rate. Various methods have been developed to ensure that the false positive rate be maintained at some prespecified level. One of the most well know being the Bonferroni procedure. Simes presented an improved Bonferroni procedure for testing the global hypothesis that is more powerful and less conservative, especially with positively correlated tests. While Simes’s procedure is more powerful, it does not allow for making inferences on the individual hypotheses. However, the Simes procedure has since become the foundation of many p-value based multiple testing …
Exploring The Relationship Between Children’S Vocabulary And Their Understanding Of Cardinality: A Methodological Approach, Justin Slifer, Emily Carrigan, Kristin Walker, Marie Coppola
Exploring The Relationship Between Children’S Vocabulary And Their Understanding Of Cardinality: A Methodological Approach, Justin Slifer, Emily Carrigan, Kristin Walker, Marie Coppola
Honors Scholar Theses
Is there a relationship between vocabulary and children’s understanding of cardinality? Does the way in which we classify cardinality data as tested by the Give-a-Number task affect finding such a relationship? This thesis explored these questions using a methodological approach, by testing the relationship between children’s receptive vocabulary scores and Give-a-Number scores classified in two different ways, the traditional knower-level assessment, as well as by calculating the proportion of trials answered correctly. A significant correlation was found between participants’ receptive vocabulary scores and Give-a-Number scores using both manners of classification, independent of the children’s ages. The results were compared with …
Predicting Postoperative Delirium Risk For Intracranial Surgery: A Statistical Machine Learning Approach, Juliet Aygun, Alaina Bartfeld, Sahana Rayan
Predicting Postoperative Delirium Risk For Intracranial Surgery: A Statistical Machine Learning Approach, Juliet Aygun, Alaina Bartfeld, Sahana Rayan
The Journal of Purdue Undergraduate Research
No abstract provided.
Statistical Methods For Resolving Intratumor Heterogeneity With Single-Cell Dna Sequencing, Alexander Davis
Statistical Methods For Resolving Intratumor Heterogeneity With Single-Cell Dna Sequencing, Alexander Davis
Dissertations and Theses (Open Access)
Tumor cells have heterogeneous genotypes, which drives progression and treatment resistance. Such genetic intratumor heterogeneity plays a role in the process of clonal evolution that underlies tumor progression and treatment resistance. Single-cell DNA sequencing is a promising experimental method for studying intratumor heterogeneity, but brings unique statistical challenges in interpreting the resulting data. Researchers lack methods to determine whether sufficiently many cells have been sampled from a tumor. In addition, there are no proven computational methods for determining the ploidy of a cell, a necessary step in the determination of copy number. In this work, software for calculating probabilities from …
Lectures On Mathematical Computing With Python, Jay Gopalakrishnan
Lectures On Mathematical Computing With Python, Jay Gopalakrishnan
PDXOpen: Open Educational Resources
This open resource is a collection of class activities for use in undergraduate courses aimed at teaching mathematical computing, and computational thinking in general, using the python programming language. It was developed for a second-year course (MTH 271) revamped for a new undergraduate program in data science at Portland State University. The activities are designed to guide students' use of python modules effectively for scientific computation, data analysis, and visualization.
Adopt/Adapt
If you are an instructor adopting or adapting this open educational resource, please help us understand your use by filling out this form
Statistical Methodology To Establish A Benchmark For Evaluating Antimicrobial Resistance Genes Through Real Time Pcr Assay, Enakshy Dutta
Statistical Methodology To Establish A Benchmark For Evaluating Antimicrobial Resistance Genes Through Real Time Pcr Assay, Enakshy Dutta
Department of Statistics: Dissertations, Theses, and Student Research
Novel diagnostic tests are usually compared with gold standard tests for evaluating diagnostic accuracy. For assessing antimicrobial resistance (AMR) to bovine respiratory disease (BRD) pathogens, phenotypic broth microdilution method is used as gold standard (GS). The objective of the thesis is to evaluate the optimal cycle threshold (Ct) generated by real-time polymerase chain reaction (rtPCR) to genes that confer resistance that will translate to the phenotypic classification of AMR. Data from two different methodologies are assessed to identify Ct that will discriminate between resistance (R) and susceptibility (S). First, the receiver operating characteristic (ROC) curve was used to determine the …
Improving The Quality And Design Of Retrospective Clinical Outcome Studies That Utilize Electronic Health Records, Oliwier Dziadkowiec, Jeffery Durbin, Vignesh Jayaraman Muralidharan, Megan Novak, Brendon Cornett
Improving The Quality And Design Of Retrospective Clinical Outcome Studies That Utilize Electronic Health Records, Oliwier Dziadkowiec, Jeffery Durbin, Vignesh Jayaraman Muralidharan, Megan Novak, Brendon Cornett
HCA Healthcare Journal of Medicine
Electronic health records (EHRs) are an excellent source for secondary data analysis. Studies based on EHR-derived data, if designed properly, can answer previously unanswerable clinical research questions. In this paper we will highlight the benefits of large retrospective studies from secondary sources such as EHRs, examine retrospective cohort and case-control study design challenges, as well as methodological and statistical adjustment that can be made to overcome some of the inherent design limitations, in order to increase the generalizability, validity and reliability of the results obtained from these studies.
Italian Sociologists: A Community Of Disconnected Groups, Aliakbar Akbaritabar, Vincent Traag, Alberto Caimo, Flaminio Squazzoni
Italian Sociologists: A Community Of Disconnected Groups, Aliakbar Akbaritabar, Vincent Traag, Alberto Caimo, Flaminio Squazzoni
Articles
Examining coauthorship networks is key to study scientific collaboration patterns and structural characteristics of scientific communities. Here, we studied coauthorship networks of sociologists in Italy, using temporal and multi-level quantitative analysis. By looking at publications indexed in Scopus, we detected research communities among Italian sociologists. We found that Italian sociologists are fractured in many disconnected groups. The giant connected component of the Italian sociology could be split into five main groups with a mixture of three main disciplinary topics: sociology of culture and communication (present in two groups), economic sociology (present in three groups) and general sociology (present in three …
Working Children On Java Island 2017, Yuniarti
Working Children On Java Island 2017, Yuniarti
International Programs
Children's wellbeing has currently become a global concern as many of them are engaged in the labor force. A small area estimation (SAE) technique, EBLUP under Fey Herriot model, is employed to reveal their number in regencies of Java Island. Statistics have been disaggregated by geographical location (urban/rural) and gender. These statistics are required by the government as the basis for policy making.
Models For Data Analysis In Accelerated Reliability Growth, Cesar Alexander Ruiz Torres
Models For Data Analysis In Accelerated Reliability Growth, Cesar Alexander Ruiz Torres
Graduate Theses and Dissertations
This work develops new methodologies for analyzing accelerated testing data in the context of a reliability growth program for a complex multi-component system. Each component has multiple failure modes and the growth program consists of multiple test-fix stages with corrective actions applied at the end of each stage. The first group of methods considers time-to-failure data and test covariates for predicting the final reliability of the system. The time-to-failure of each failure mode is assumed to follow a Weibull distribution with rate parameter proportional to an acceleration factor. Acceleration factors are specific to each failure mode and test covariates. We …
Measuring Sexual Excitation And Sexual Inhibition In A Dutch-Speaking Sample, Malachi Willis
Measuring Sexual Excitation And Sexual Inhibition In A Dutch-Speaking Sample, Malachi Willis
Graduate Theses and Dissertations
Background: Individual differences in sexual excitation and sexual inhibition are important predictors of sexual functioning. Psychometric instruments for these aspects of sexual response were originally developed separately for men (Sexual Inhibition /Sexual Excitation Scales [SIS/SES]) and women (Sexual Excitation/Sexual Inhibition Inventory for Women [SESII-W]). These measures were then adapted to function similarly in samples comprising both men and women (Sexual Inhibition/Sexual Excitation Scales-Short Form [SIS/SES-SF] and Sexual Excitation/Sexual Inhibition Inventory for Women and Men [SESII-W/M], respectively). No published study to our knowledge has administered the SIS/SES and SESII-W/M questionnaires to a sample of both women and men. In the present …
Research In Short Term Actuarial Modeling, Elijah Howells
Research In Short Term Actuarial Modeling, Elijah Howells
Electronic Theses, Projects, and Dissertations
This paper covers mathematical methods used to conduct actuarial analysis in the short term, such as policy deductible analysis, maximum covered loss analysis, and mixtures of distributions. Assessment of a loss variable's distribution under the effect of a policy deductible, as well as one with an implemented maximum covered loss, and under both a policy deductible and maximum covered loss will also be covered. The derivation, meaning, and use of cost per loss and cost per payment will be discussed, as will those of an aggregate sum distribution, stop loss policy, and maximum likelihood estimation. For each topic, special cases …
A Study Of Cusum Statistics On Bitcoin Transactions, Ivan Perez
A Study Of Cusum Statistics On Bitcoin Transactions, Ivan Perez
Theses and Dissertations
In this thesis, our objective is to study the relationship between transaction price and volume in the BTC/USD Coinbase exchange. In the second chapter, we develop a consecutive CUSUM algorithm to detect instantaneous changes in the arrival rate of market orders. We begin by estimating a baseline rate using the assumption of a local time-homogeneous Poisson process. Our observations lead us to reject the plausibility of a time-homogeneous Poisson model on a more global scale by using a chi squared test. We thus proceed to use CUSUM-based alarms to detect consecutive upward and downward changes in the arrival rate of …
Sensitivity Analysis For Incomplete Data And Causal Inference, Heng Chen
Sensitivity Analysis For Incomplete Data And Causal Inference, Heng Chen
Statistical Science Theses and Dissertations
In this dissertation, we explore sensitivity analyses under three different types of incomplete data problems, including missing outcomes, missing outcomes and missing predictors, potential outcomes in \emph{Rubin causal model (RCM)}. The first sensitivity analysis is conducted for the \emph{missing completely at random (MCAR)} assumption in frequentist inference; the second one is conducted for the \emph{missing at random (MAR)} assumption in likelihood inference; the third one is conducted for one novel assumption, the ``sixth assumption'' proposed for the robustness of instrumental variable estimand in causal inference.