Open Access. Powered by Scholars. Published by Universities.®

Articles 1 - 17 of 17

Full-Text Articles in Longitudinal Data Analysis and Time Series

Feature Selection For Longitudinal Data By Using Sign Averages To Summarize Gene Expression Values Over Time, Suyan Tian, Chi Wang Mar 2019

Feature Selection For Longitudinal Data By Using Sign Averages To Summarize Gene Expression Values Over Time, Suyan Tian, Chi Wang

Biostatistics Faculty Publications

With the rapid evolution of high-throughput technologies, time series/longitudinal high-throughput experiments have become possible and affordable. However, the development of statistical methods dealing with gene expression profiles across time points has not kept up with the explosion of such data. The feature selection process is of critical importance for longitudinal microarray data. In this study, we proposed aggregating a gene’s expression values across time into a single value using the sign average method, thereby degrading a longitudinal feature selection process into a classic one. Regularized logistic regression models with pseudogenes (i.e., the sign average of genes across time as predictors) …


Bayesian Nonparametric Analysis Of Longitudinal Data With Non-Ignorable Non-Monotone Missingness, Yu Cao Jan 2019

Bayesian Nonparametric Analysis Of Longitudinal Data With Non-Ignorable Non-Monotone Missingness, Yu Cao

Theses and Dissertations

In longitudinal studies, outcomes are measured repeatedly over time, but in reality clinical studies are full of missing data points of monotone and non-monotone nature. Often this missingness is related to the unobserved data so that it is non-ignorable. In such context, pattern-mixture model (PMM) is one popular tool to analyze the joint distribution of outcome and missingness patterns. Then the unobserved outcomes are imputed using the distribution of observed outcomes, conditioned on missing patterns. However, the existing methods suffer from model identification issues if data is sparse in specific missing patterns, which is very likely to happen with a …


Association Analyses Of Repeated Measures On Triglyceride And High-Density Lipoprotein Levels: Insights From Gaw20, Saurabh Ghosh, David W. Fardo Sep 2018

Association Analyses Of Repeated Measures On Triglyceride And High-Density Lipoprotein Levels: Insights From Gaw20, Saurabh Ghosh, David W. Fardo

Biostatistics Faculty Publications

Background: The GAW20 group formed on the theme of methods for association analyses of repeated measures comprised 4sets of investigators. The provided “real” data set included genotypes obtained from a human whole-genome association study based on longitudinal measurements of triglycerides (TGs) and high-density lipoprotein in addition to methylation levels before and after administration of fenofibrate. The simulated data set contained 200 replications of methylation levels and posttreatment TGs, mimicking the real data set.

Results: The different investigators in the group focused on the statistical challenges unique to family-based association analyses of phenotypes measured longitudinally and applied a wide spectrum of …


Mediation Analysis With Time-Varying Exposures And Mediators, Tyler J. Vanderweele, Eric Tchetgen Tchetgen Mar 2014

Mediation Analysis With Time-Varying Exposures And Mediators, Tyler J. Vanderweele, Eric Tchetgen Tchetgen

Harvard University Biostatistics Working Paper Series

In this paper we consider mediation analysis when exposures and mediators vary over time. We give non-parametric identification results, discuss parametric implementation, and also provide a weighting approach to direct and indirect effects based on combining the results of two marginal structural models. We also discuss how our results give rise to a causal interpretation of the effect estimates produced from longitudinal structural equation models. When there are no time-varying confounders affected by prior exposure and mediator values, identification of direct and indirect effects is achieved by a longitudinal version of Pearl's mediation formula. When there are time-varying confounders affected …


Analysis Of Continuous Longitudinal Data With Arma(1, 1) And Antedependence Correlation Structures, Sirisha Mushti Apr 2013

Analysis Of Continuous Longitudinal Data With Arma(1, 1) And Antedependence Correlation Structures, Sirisha Mushti

Mathematics & Statistics Theses & Dissertations

Longitudinal or repeated measure data are common in biomedical and clinical trials. These data are often collected on individuals at scheduled times resulting in dependent responses. Inference methods for studying the behavior of responses over time as well as methods to study the association with certain risk factors or covariates taking into account the dependencies are of great importance. In this research we focus our study on the analysis of continuous longitudinal data. To model the dependencies of the responses over time, we consider appropriate correlation structures generated by the stationary and non-stationary time-series models. We develop new estimation procedures …


Canonical Correlation Analysis For Longitudinal Data, Raymond Mccollum Jan 2010

Canonical Correlation Analysis For Longitudinal Data, Raymond Mccollum

Mathematics & Statistics Theses & Dissertations

Data (multivariate data) on two sets of vectors commonly occur in applications. Statistical analysis of these data is usually done using a canonical correlation analysis (CCA). Occurrence of these data at multiple occasions or conditions leads to longitudinal multivariate data for a CCA. We address the problem of canonical correlation analysis on longitudinal data when the data have a Kronecker product covariance structure. Using structured correlation matrices we model the dependency of repeatedly observed data. Recent work of Srivastava, Nahtman, and von Rosen (2008) developed an iterative algorithm to determine the maximum likelihood estimate of the Kronecker product covariance structure …


Marginal Regression Modeling Under Irregular, Biased Sampling, Petra Buzkova, Thomas Lumley Sep 2005

Marginal Regression Modeling Under Irregular, Biased Sampling, Petra Buzkova, Thomas Lumley

UW Biostatistics Working Paper Series

In longitudinal studies observations are often obtained at continuous subject-specific times. Frequently the availability of outcome data may be related to the outcome measure or other covariates that are related to the outcome measure. Under such biased sampling designs unadjusted regression analysis yield biased estimates. Building on the work of Lin & Ying (2001) that integrates counting processes techniques with longitudinal data settings we propose a class of estimators that can handle biased sampling. We call those estimators ``inverse--intensity--rate--ratio--weighted'' (IIRR) estimators. Of major focus is a mean--response model where we examine the marginal effect of the covariate X at time …


Longitudinal Data Analysis For Generalized Linear Models Under Irregular, Biased Sampling: Situations With Follow-Up Dependent On Outcome Or Auxiliary Outcome-Related Variables, Petra Buzkova, Thomas Lumley Sep 2005

Longitudinal Data Analysis For Generalized Linear Models Under Irregular, Biased Sampling: Situations With Follow-Up Dependent On Outcome Or Auxiliary Outcome-Related Variables, Petra Buzkova, Thomas Lumley

UW Biostatistics Working Paper Series

In longitudinal studies, observations are often obtained at subject-specific observation times. Those times can be continuous times, not at a set of prespecified times. Frequently the observation times may be related to the outcome measure or other auxiliary variables that are related to the outcome measure but undesirable to condition upon in the regression model for outcome. Regression analysis unadjusted for such sampling designs yield biased estimates. Based on estimating equations, we propose a class of estimators in generalized linear regression models that can handle biased sampling under continuous observation times. We call those estimators ``inverse--intensity rate--ratio--weighted'' (IIRR) estimators. The …


Semiparametric Loglinear Regression For Longitudinal Measurements Subject To Irregular, Biased Follow-Up, Petra Buzkova, Thomas Lumley Sep 2005

Semiparametric Loglinear Regression For Longitudinal Measurements Subject To Irregular, Biased Follow-Up, Petra Buzkova, Thomas Lumley

UW Biostatistics Working Paper Series

We propose a method for analysis of loglinear regression models for longitudinal data that are subject to continuous and irregular follow-up. Frequently, if the follow-up is irregular, the availability of outcome data may be related to the outcome measure or other covariates that are related to the outcome measure. Under such biased sampling designs unadjusted regression analysis yield biased estimates. We examine the marginal association of the covariates X at time t and the logarithm of the mean of response Y at time t. We focus on semiparametric regression with unspecified baseline function of time. To predict the follow-up times …


G-Computation Estimation Of Nonparametric Causal Effects On Time-Dependent Mean Outcomes In Longitudinal Studies, Romain Neugebauer, Mark J. Van Der Laan Jul 2005

G-Computation Estimation Of Nonparametric Causal Effects On Time-Dependent Mean Outcomes In Longitudinal Studies, Romain Neugebauer, Mark J. Van Der Laan

U.C. Berkeley Division of Biostatistics Working Paper Series

Two approaches to Causal Inference based on Marginal Structural Models (MSM) have been proposed. They provide different representations of causal effects with distinct causal parameters. Initially, a parametric MSM approach to Causal Inference was developed: it relies on correct specification of a parametric MSM. Recently, a new approach based on nonparametric MSM was introduced. This later approach does not require the assumption of a correctly specified MSM and thus is more realistic if one believes that correct specification of a parametric MSM is unlikely in practice. However, this approach was described only for investigating causal effects on mean outcomes collected …


Statistical Analysis Of Longitudinal And Multivariate Discrete Data, Deepak Mav Apr 2005

Statistical Analysis Of Longitudinal And Multivariate Discrete Data, Deepak Mav

Mathematics & Statistics Theses & Dissertations

Correlated multivariate Poisson and binary variables occur naturally in medical, biological and epidemiological longitudinal studies. Modeling and simulating such variables is difficult because the correlations are restricted by the marginal means via Fréchet bounds in a complicated way. In this dissertation we will first discuss partially specified models and methods for estimating the regression and correlation parameters. We derive the asymptotic distributions of these parameter estimates. Using simulations based on extensions of the algorithm due to Sim (1993, Journal of Statistical Computation and Simulation, 47, pp. 1–10), we study the performance of these estimates using infeasibility, coverage probabilities of the …


Estimation Of Direct And Indirect Causal Effects In Longitudinal Studies, Mark J. Van Der Laan, Maya L. Petersen Aug 2004

Estimation Of Direct And Indirect Causal Effects In Longitudinal Studies, Mark J. Van Der Laan, Maya L. Petersen

U.C. Berkeley Division of Biostatistics Working Paper Series

The causal effect of a treatment on an outcome is generally mediated by several intermediate variables. Estimation of the component of the causal effect of a treatment that is mediated by a given intermediate variable (the indirect effect of the treatment), and the component that is not mediated by that intermediate variable (the direct effect of the treatment) is often relevant to mechanistic understanding and to the design of clinical and public health interventions. Under the assumption of no-unmeasured confounders, Robins & Greenland (1992) and Pearl (2000), develop two identifiability results for direct and indirect causal effects. They define an …


Marginal Modeling Of Multilevel Binary Data With Time-Varying Covariates, Diana Miglioretti, Patrick Heagerty Dec 2003

Marginal Modeling Of Multilevel Binary Data With Time-Varying Covariates, Diana Miglioretti, Patrick Heagerty

UW Biostatistics Working Paper Series

We propose and compare two approaches for regression analysis of multilevel binary data when clusters are not necessarily nested: a GEE method that relies on a working independence assumption coupled with a three-step method for obtaining empirical standard errors; and a likelihood-based method implemented using Bayesian computational techniques. Implications of time-varying endogenous covariates are addressed. The methods are illustrated using data from the Breast Cancer Surveillance Consortium to estimate mammography accuracy from a repeatedly screened population.


Efficient Semiparametric Marginal Estimation For Longitudinal/Clustered Data, Naisyin Wang, Raymond J. Carroll, Xihong Lin Sep 2003

Efficient Semiparametric Marginal Estimation For Longitudinal/Clustered Data, Naisyin Wang, Raymond J. Carroll, Xihong Lin

The University of Michigan Department of Biostatistics Working Paper Series

We consider marginal generalized semiparametric partially linear models for clustered data. Lin and Carroll (2001a) derived the semiparametric efficinet score funtion for this problem in the mulitvariate Gaussian case, but they were unable to contruct a semiparametric efficient estimator that actually achieved the semiparametric information bound. We propose such an estimator here and generalize the work to marginal generalized partially liner models. Asymptotic relative efficincies of the estimation or throughout are investigated. The finite sample performance of these estimators is evaluated through simulations and illustrated using a longtiudinal CD4 count data set. Both theoretical and numerical results indicate that properly …


Equivalent Kernels Of Smoothing Splines In Nonparametric Regression For Clustered/Longitudinal Data, Xihong Lin, Naisyin Wang, Alan H. Welsh, Raymond J. Carroll Sep 2003

Equivalent Kernels Of Smoothing Splines In Nonparametric Regression For Clustered/Longitudinal Data, Xihong Lin, Naisyin Wang, Alan H. Welsh, Raymond J. Carroll

The University of Michigan Department of Biostatistics Working Paper Series

We compare spline and kernel methods for clustered/longitudinal data. For independent data, it is well known that kernel methods and spline methods are essentially asymptotically equivalent (Silverman, 1984). However, the recent work of Welsh, et al. (2002) shows that the same is not true for clustered/longitudinal data. First, conventional kernel methods fail to account for the within- cluster correlation, while spline methods are able to account for this correlation. Second, kernel methods and spline methods were found to have different local behavior, with conventional kernels being local and splines being non-local. To resolve these differences, we show that a smoothing …


Histospline Method In Nonparametric Regression Models With Application To Clustered/Longitudinal Data, Raymond J. Carroll, Peter Hall, Tatiyana V. Apanasovich, Xihong Lin Sep 2003

Histospline Method In Nonparametric Regression Models With Application To Clustered/Longitudinal Data, Raymond J. Carroll, Peter Hall, Tatiyana V. Apanasovich, Xihong Lin

The University of Michigan Department of Biostatistics Working Paper Series

Kernel and smoothing methods for nonparametric function and curve estimation have been particularly successful in "standard" settings, where function values are observed subject to independent errors. However, when aspects of the function are known parametrically, or where the sampling scheme has significant structure, it can be quite difficult to adapt standard methods in such a way that they retain good statistical performance and continue to enjoy easy computability and good numerical properties. In particular, when using local linear modeling it is often awkward to both respect the sampling scheme and produce an estimator with good variance properties, without resorting to …


Double Robust Estimation In Longitudinal Marginal Structural Models, Zhuo Yu, Mark J. Van Der Laan Jun 2003

Double Robust Estimation In Longitudinal Marginal Structural Models, Zhuo Yu, Mark J. Van Der Laan

U.C. Berkeley Division of Biostatistics Working Paper Series

Consider estimation of causal parameters in a marginal structural model for the discrete intensity of the treatment specific counting process (e.g. hazard of a treatment specific survival time) based on longitudinal observational data on treatment, covariates and survival. We assume the sequential randomization assumption (SRA) on the treatment assignment mechanism and the so called experimental treatment assignment assumption which is needed to identify the causal parameters from the observed data distribution. Under SRA, the likelihood of the observed data structure factorizes in the auxiliary treatment mechanism and the partial likelihood consisting of the product over time of conditional distributions of …