Open Access. Powered by Scholars. Published by Universities.®
- Discipline
- Keyword
-
- EM algorithm (4)
- Count data (3)
- Bayesian Analysis (2)
- Bootstrap (2)
- Data dispersion (2)
-
- EM Algorithm (2)
- Generalized linear models (2)
- Linear model (2)
- Multivariate Data (2)
- Normalization (2)
- ODE (2)
- Spatial-temporal process (2)
- Variable Screening (2)
- Variable Selection (2)
- Adjustment (1)
- Asymptotic distribution (1)
- Asymptotic properties (1)
- Autologistic regression models (1)
- Autoregressive models (1)
- Average Causal Effect (1)
- BRAiNS Study (1)
- Backfitting (1)
- Bayesian Adjustment for Confounding (1)
- Bayesian Inference (1)
- Bayesian analysis (1)
- Bayesian modeling (1)
- Beta Binomial Distribution (1)
- Big Data (1)
- Bilaterally contaminated normal model (1)
- Binary data (1)
Articles 31 - 42 of 42
Full-Text Articles in Statistical Models
Topics In Logistic Regression Analysis, Zhiheng Xie
Topics In Logistic Regression Analysis, Zhiheng Xie
Theses and Dissertations--Statistics
Discrete-time Markov chains have been used to analyze the transition of subjects from intact cognition to dementia with mild cognitive impairment and global impairment as intervening transient states, and death as competing risk. A multinomial logistic regression model is used to estimate the probability distribution in each row of the one-step transition matrix that correspond to the transient states. We investigate some goodness of fit tests for a multinomial distribution with covariates to assess the fit of this model to the data. We propose a modified chi-square test statistic and a score test statistic for the multinomial assumption in each …
Developing An Alternative Way To Analyze Nanostring Data, Shu Shen
Developing An Alternative Way To Analyze Nanostring Data, Shu Shen
Theses and Dissertations--Statistics
Nanostring technology provides a new method to measure gene expressions. It's more sensitive than microarrays and able to do more gene measurements than RT-PCR with similar sensitivity. This system produces counts for each target gene and tabulates them. Counts can be normalized by using an Excel macro or nSolver before analysis. Both methods rely on data normalization prior to statistical analysis to identify differentially expressed genes. Alternatively, we propose to model gene expressions as a function of positive controls and reference gene measurements. Simulations and examples are used to compare this model with Nanostring normalization methods. The results show that …
Statistical Inference On Dynamical Systems, Hongyuan Wang
Statistical Inference On Dynamical Systems, Hongyuan Wang
Theses and Dissertations--Statistics
The ordinary differential equation (ODE) is one representative and popular tool in modeling dynamical systems, which are widely implemented in physics, biology, economics, chemistry and biomedical sciences, etc. Because of the importance of dynamical systems in scientific studies, they are the main focuses of my dissertation.
The first chapter of the dissertation is introduction and literature review, which mainly focuses on numerical integration algorithms of ODEs that are difficult to solve analytically, as well as derivative-free optimization algorithms for the so-called inverse problem.
The second chapter is on the estimation method based on numerical solvers of differential equations. We start …
Statistical Methods For Environmental Exposure Data Subject To Detection Limits, Yuchen Yang
Statistical Methods For Environmental Exposure Data Subject To Detection Limits, Yuchen Yang
Theses and Dissertations--Statistics
In this dissertation, we develop unified and efficient nonparametric statistical methods for estimating and comparing environmental exposure distributions in presence of detection limits. In the first part, we propose a kernel-smoothed nonparametric estimator for the exposure distribution without imposing any independence assumption between the exposure level and detection limit. We show that the proposed estimator is consistent and asymptotically normal. Simulation studies demonstrate that the proposed estimator performs well in practical situations. A colon cancer study is provided for illustration. In the second part, we develop a class of test statistics to compare exposure distributions between two groups by using …
Improved Models For Differential Analysis For Genomic Data, Hong Wang
Improved Models For Differential Analysis For Genomic Data, Hong Wang
Theses and Dissertations--Statistics
This paper intend to develop novel statistical methods to improve genomic data analysis, especially for differential analysis. We considered two different data type: NanoString nCounter data and somatic mutation data. For NanoString nCounter data, we develop a novel differential expression detection method. The method considers a generalized linear model of the negative binomial family to characterize count data and allows for multi-factor design. Data normalization is incorporated in the model framework through data normalization parameters, which are estimated from control genes embedded in the nCounter system. For somatic mutation data, we develop beta-binomial model-based approaches to identify highly or lowly …
Continuous Time Multi-State Models For Interval Censored Data, Lijie Wan
Continuous Time Multi-State Models For Interval Censored Data, Lijie Wan
Theses and Dissertations--Statistics
Continuous-time multi-state models are widely used in modeling longitudinal data of disease processes with multiple transient states, yet the analysis is complex when subjects are observed periodically, resulting in interval censored data. Recently, most studies focused on modeling the true disease progression as a discrete time stationary Markov chain, and only a few studies have been carried out regarding non-homogenous multi-state models in the presence of interval-censored data. In this dissertation, several likelihood-based methodologies were proposed to deal with interval censored data in multi-state models.
Firstly, a continuous time version of a homogenous Markov multi-state model with backward transitions was …
New Results In Ell_1 Penalized Regression, Edward A. Roualdes
New Results In Ell_1 Penalized Regression, Edward A. Roualdes
Theses and Dissertations--Statistics
Here we consider penalized regression methods, and extend on the results surrounding the l1 norm penalty. We address a more recent development that generalizes previous methods by penalizing a linear transformation of the coefficients of interest instead of penalizing just the coefficients themselves. We introduce an approximate algorithm to fit this generalization and a fully Bayesian hierarchical model that is a direct analogue of the frequentist version. A number of benefits are derived from the Bayesian persepective; most notably choice of the tuning parameter and natural means to estimate the variation of estimates – a notoriously difficult task for the …
Genetic Association Testing Of Copy Number Variation, Yinglei Li
Genetic Association Testing Of Copy Number Variation, Yinglei Li
Theses and Dissertations--Statistics
Copy-number variation (CNV) has been implicated in many complex diseases. It is of great interest to detect and locate such regions through genetic association testings. However, the association testings are complicated by the fact that CNVs usually span multiple markers and thus such markers are correlated to each other. To overcome the difficulty, it is desirable to pool information across the markers. In this thesis, we propose a kernel-based method for aggregation of marker-level tests, in which first we obtain a bunch of p-values through association tests for every marker and then the association test involving CNV is based on …
Normal Mixture And Contaminated Model With Nuisance Parameter And Applications, Qian Fan
Normal Mixture And Contaminated Model With Nuisance Parameter And Applications, Qian Fan
Theses and Dissertations--Statistics
This paper intend to find the proper hypothesis and test statistic for testing existence of bilaterally contamination when there exists nuisance parameter. The test statistic is based on method of moments estimators. Union-Intersection test is used for testing if the distribution of population can be implemented by a bilaterally contaminated normal model with unknown variance. This paper also developed a hierarchical normal mixture model (HNM) and applied it to birth weight data. EM algorithm is employed for parameter estimation and a singular Bayesian information criterion (sBIC) is applied to choose the number components. We also proposed a singular flexible information …
Analysis Of Spatial Data, Xiang Zhang
Analysis Of Spatial Data, Xiang Zhang
Theses and Dissertations--Statistics
In many areas of the agriculture, biological, physical and social sciences, spatial lattice data are becoming increasingly common. In addition, a large amount of lattice data shows not only visible spatial pattern but also temporal pattern (see, Zhu et al. 2005). An interesting problem is to develop a model to systematically model the relationship between the response variable and possible explanatory variable, while accounting for space and time effect simultaneously.
Spatial-temporal linear model and the corresponding likelihood-based statistical inference are important tools for the analysis of spatial-temporal lattice data. We propose a general asymptotic framework for spatial-temporal linear models and …
Analysis Of Binary Data Via Spatial-Temporal Autologistic Regression Models, Zilong Wang
Analysis Of Binary Data Via Spatial-Temporal Autologistic Regression Models, Zilong Wang
Theses and Dissertations--Statistics
Spatial-temporal autologistic models are useful models for binary data that are measured repeatedly over time on a spatial lattice. They can account for effects of potential covariates and spatial-temporal statistical dependence among the data. However, the traditional parametrization of spatial-temporal autologistic model presents difficulties in interpreting model parameters across varying levels of statistical dependence, where its non-negative autocovariates could bias the realizations toward 1. In order to achieve interpretable parameters, a centered spatial-temporal autologistic regression model has been developed. Two efficient statistical inference approaches, expectation-maximization pseudo-likelihood approach (EMPL) and Monte Carlo expectation-maximization likelihood approach (MCEML), have been proposed. Also, Bayesian …
Stochastic Dynamics Of Gene Transcription, Yan Xie
Stochastic Dynamics Of Gene Transcription, Yan Xie
Theses and Dissertations--Statistics
Gene transcription in individual living cells is inevitably a stochastic and dynamic process. Little is known about how cells and organisms learn to balance the fidelity of transcriptional control and the stochasticity of transcription dynamics. In an effort to elucidate the contribution of environmental signals to this intricate balance, a Three State Model was recently proposed, and the transcription system was assumed to transit among three different functional states randomly.
In this work, we employ this model to demonstrate how the stochastic dynamics of gene transcription can be characterized by the three transition parameters. We compute the probability distribution of …