Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Statistical Methodology (107)
- Statistical Theory (103)
- Medicine and Health Sciences (77)
- Public Health (67)
- Survival Analysis (67)
-
- Biostatistics (51)
- Epidemiology (45)
- Longitudinal Data Analysis and Time Series (34)
- Multivariate Analysis (32)
- Applied Mathematics (29)
- Numerical Analysis and Computation (29)
- Life Sciences (28)
- Genetics and Genomics (27)
- Genetics (23)
- Microarrays (22)
- Categorical Data Analysis (21)
- Clinical Epidemiology (20)
- Clinical Trials (20)
- Disease Modeling (14)
- Diseases (14)
- Design of Experiments and Sample Surveys (13)
- Bioinformatics (11)
- Computational Biology (11)
- Health Services Research (9)
- Medical Specialties (9)
- Laboratory and Basic Science Research (6)
- Applied Statistics (5)
- Keyword
-
- Causal inference (9)
- Prediction (8)
- Cross-validation (7)
- Gene expression (7)
- Model selection (7)
-
- Counting process (6)
- Estimating equation (6)
- Genetics (6)
- Semiparametric model (6)
- Air pollution (5)
- Longitudinal data (5)
- Censored data (4)
- Counterfactual (4)
- Generalized estimating equations (4)
- Hierarchical models (4)
- Mixture model (4)
- Multiple comparisons (4)
- Regression (4)
- Risk (4)
- Sensitivity (4)
- Survival analysis (4)
- Censoring (3)
- Classification (3)
- Clinical trials (3)
- Confounding (3)
- Likelihood (3)
- Linear regression (3)
- Loss function (3)
- MCMC (3)
- Marginal structural model (3)
- Publication Year
- Publication
-
- U.C. Berkeley Division of Biostatistics Working Paper Series (60)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (47)
- Harvard University Biostatistics Working Paper Series (43)
- The University of Michigan Department of Biostatistics Working Paper Series (40)
- UW Biostatistics Working Paper Series (31)
Articles 181 - 210 of 242
Full-Text Articles in Statistical Models
Semiparametric Estimation Of Time-Dependent: Roc Curves For Longitudinal Marker Data, Yingye Zheng, Patrick Heagerty
Semiparametric Estimation Of Time-Dependent: Roc Curves For Longitudinal Marker Data, Yingye Zheng, Patrick Heagerty
UW Biostatistics Working Paper Series
One approach to evaluating the strength of association between a longitudinal marker process and a key clinical event time is through predictive regression methods such as a time-dependent covariate hazard model. For example, a time-varying covariate Cox model specifies the instantaneous risk of the event as a function of the time-varying marker and additional covariates. In this manuscript we explore a second complementary approach which characterizes the distribution of the marker as a function of both the measurement time and the ultimate event time. Our goal is to flexibly extend the standard diagnostic accuracy concepts of sensitivity and specificity to …
Comparison Of The Inverse Probability Of Treatment Weighted (Iptw) Estimator With A Naïve Estimator In The Analysis Of Longitudinal Data With Time-Dependent Confounding: A Simulation Study, Thaddeus Haight, Romain Neugebauer, Ira B. Tager, Mark J. Van Der Laan
Comparison Of The Inverse Probability Of Treatment Weighted (Iptw) Estimator With A Naïve Estimator In The Analysis Of Longitudinal Data With Time-Dependent Confounding: A Simulation Study, Thaddeus Haight, Romain Neugebauer, Ira B. Tager, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
A simulation study was conducted to compare estimates from a naïve estimator, using standard conditional regression, and an IPTW (Inverse Probability of Treatment Weighted) estimator, to true causal parameters for a given MSM (Marginal Structural Model). The study was extracted from a larger epidemiological study (Longitudinal Study of Effects of Physical Activity and Body Composition on Functional Limitation in the Elderly, by Tager et. al [accepted, Epidemiology, September 2003]), which examined the causal effects of physical activity and body composition on functional limitation. The simulation emulated the larger study in terms of the exposure and outcome variables of interest-- physical …
Optimization Of Breast Cancer Screening Modalities, Yu Shen, Giovanni Parmigiani
Optimization Of Breast Cancer Screening Modalities, Yu Shen, Giovanni Parmigiani
Johns Hopkins University, Dept. of Biostatistics Working Papers
Mathematical models and decision analyses based on microsimulations have been shown to be useful in evaluating relative merits of various screening strategies in terms of cost and mortality reduction. Most investigations regarding the balance between mortality reduction and costs have focused on a single modality, mammography. A systematic evaluation of the relative expenses and projected benefit of combining clinical breast examination and mammography is not at present available. The purpose of this report is to provide methodologic details including assumptions and data used in the process of modeling for complex decision analyses, when searching for optimal breast cancer screening strategies …
Modeling The Incubation Period Of Anthrax, Ron Brookmeyer, Elizabeth Johnson, Sarah Barry
Modeling The Incubation Period Of Anthrax, Ron Brookmeyer, Elizabeth Johnson, Sarah Barry
Johns Hopkins University, Dept. of Biostatistics Working Papers
Models of the incubation period of anthrax are important to public health planners because they can be used to predict the delay before outbreaks are detected, the size of an outbreak and the duration of time that persons should remain on antibiotics to prevent disease. The difficulty is that there is little direct data about the incubation period in humans. The objective of this paper is to develop and apply models for the incubation period of anthrax. Mechanistic models that account for the biology of spore clearance and germination are developed based on a competing risks formulation. The models predict …
Unified Cross-Validation Methodology For Selection Among Estimators And A General Cross-Validated Adaptive Epsilon-Net Estimator: Finite Sample Oracle Inequalities And Examples, Mark J. Van Der Laan, Sandrine Dudoit
Unified Cross-Validation Methodology For Selection Among Estimators And A General Cross-Validated Adaptive Epsilon-Net Estimator: Finite Sample Oracle Inequalities And Examples, Mark J. Van Der Laan, Sandrine Dudoit
U.C. Berkeley Division of Biostatistics Working Paper Series
In Part I of this article we propose a general cross-validation criterian for selecting among a collection of estimators of a particular parameter of interest based on n i.i.d. observations. It is assumed that the parameter of interest minimizes the expectation (w.r.t. to the distribution of the observed data structure) of a particular loss function of a candidate parameter value and the observed data structure, possibly indexed by a nuisance parameter. The proposed cross-validation criterian is defined as the empirical mean over the validation sample of the loss function at the parameter estimate based on the training sample, averaged over …
Underestimation Of Standard Errors In Multi-Site Time Series Studies, Michael Daniels, Francesca Dominici, Scott L. Zeger
Underestimation Of Standard Errors In Multi-Site Time Series Studies, Michael Daniels, Francesca Dominici, Scott L. Zeger
Johns Hopkins University, Dept. of Biostatistics Working Papers
Multi-site time series studies of air pollution and mortality and morbidity have figured prominently in the literature as comprehensive approaches for estimating acute effects of air pollution on health. Hierarchical models are generally used to combine site-specific information and estimate pooled air pollution effects taking into account both within-site statistical uncertainty, and across-site heterogeneity.
Within a site, characteristics of time series data of air pollution and health (small pollution effects, missing data, highly correlated predictors, non linear confounding etc.) make modelling all sources of uncertainty challenging. One potential consequence is underestimation of the statistical variance of the site-specific effects to …
Estimating Predictors For Long- Or Short-Term Survivors, Lu Tian, Wei Wang, L. J. Wei
Estimating Predictors For Long- Or Short-Term Survivors, Lu Tian, Wei Wang, L. J. Wei
Harvard University Biostatistics Working Paper Series
No abstract provided.
Time-Series Studies Of Particulate Matter, Michelle L. Bell, Jonathan M. Samet, Francesca Dominici
Time-Series Studies Of Particulate Matter, Michelle L. Bell, Jonathan M. Samet, Francesca Dominici
Johns Hopkins University, Dept. of Biostatistics Working Papers
Studies of air pollution and human health have evolved from descriptive studies of the early phenomena of large increases in adverse health effects following extreme air pollution episodes, to time-series analyses and the development of sophisticated regression models. In fact, advanced statistical methods are necessary to address the many challenges inherent in the detection of a small pollution risk in the presence of many confounders. This paper reviews the history, methods, and findings of the time-series studies estimating health risks associated with short-term exposure to particulate matter, though much of the discussion is applicable to epidemiological studies of air pollution …
Smooth Quantile Ratio Estimation With Regression: Estimating Medical Expenditures For Smoking Attributable Diseases, Francesca Dominici, Scott L. Zeger
Smooth Quantile Ratio Estimation With Regression: Estimating Medical Expenditures For Smoking Attributable Diseases, Francesca Dominici, Scott L. Zeger
Johns Hopkins University, Dept. of Biostatistics Working Papers
In this paper we introduce a semi-parametric regression model for estimating the difference in the expected value of two positive and highly skewed random variables as a function of covariates. Our method extends Smooth Quantile Ratio Estimation (SQUARE), a novel estimator of the mean difference of two positive random variables, to a regression model.
The methodological development of this paper is motivated by a common problem in econometrics where we are interested in estimating the difference in the average expenditures between two populations, say with and without a disease, taking covariates into account. Let Y1 and Y2 be two positive …
A Corrected Pseudo-Score Approach For Additive Hazards Model With Longitudinal Covariates Measured With Error, Xiao Song, Yijian Huang
A Corrected Pseudo-Score Approach For Additive Hazards Model With Longitudinal Covariates Measured With Error, Xiao Song, Yijian Huang
UW Biostatistics Working Paper Series
In medical studies, it is often of interest to characterize the relationship between a time-to-event and covariates, not only time-independent but also time-dependent. Time-dependent covariates are generally measured intermittently and with error. Recent interests focus on the proportional hazards framework, with longitudinal data jointly modeled through a mixed effects model. However, approaches under this framework depend on the normality assumption of the error, and might encounter intractable numerical difficulties in practice. This motivates us to consider an alternative framework, that is, the additive hazards model, under which little has been done when time-dependent covariates are measured with error. We propose …
A Nonparametric Comparison Of Conditional Distributions With Nonnegligible Cure Fractions, Yi Li, Jin Feng
A Nonparametric Comparison Of Conditional Distributions With Nonnegligible Cure Fractions, Yi Li, Jin Feng
Harvard University Biostatistics Working Paper Series
No abstract provided.
Survival Analysis With Heterogeneous Covariate Measurement Error, Yi Li, Louise Ryan
Survival Analysis With Heterogeneous Covariate Measurement Error, Yi Li, Louise Ryan
Harvard University Biostatistics Working Paper Series
No abstract provided.
To Model Or Not To Model? Competing Modes Of Inference For Finite Population Sampling, Rod Little
To Model Or Not To Model? Competing Modes Of Inference For Finite Population Sampling, Rod Little
The University of Michigan Department of Biostatistics Working Paper Series
Finite population sampling is perhaps the only area of statistics where the primary mode of analysis is based on the randomization distribution, rather than on statistical models for the measured variables. This article reviews the debate between design and model-based inference. The basic features of the two approaches are illustrated using the case of inference about the mean from stratified random samples. Strengths and weakness of design-based and model-based inference for surveys are discussed. It is suggested that models that take into account the sample design and make weak parametric assumptions can produce reliable and efficient inferences in surveys settings. …
Loss Function Based Ranking In Two-Stage, Hierarchical Models, Rongheng Lin, Thomas A. Louis, Susan M. Paddock, Greg Ridgeway
Loss Function Based Ranking In Two-Stage, Hierarchical Models, Rongheng Lin, Thomas A. Louis, Susan M. Paddock, Greg Ridgeway
Johns Hopkins University, Dept. of Biostatistics Working Papers
Several authors have studied the performance of optimal, squared error loss (SEL) estimated ranks. Though these are effective, in many applications interest focuses on identifying the relatively good (e.g., in the upper 10%) or relatively poor performers. We construct loss functions that address this goal and evaluate candidate rank estimates, some of which optimize specific loss functions. We study performance for a fully parametric hierarchical model with a Gaussian prior and Gaussian sampling distributions, evaluating performance for several loss functions. Results show that though SEL-optimal ranks and percentiles do not specifically focus on classifying with respect to a percentile cut …
Joint Modeling And Estimation For Recurrent Event Processes And Failure Time Data, Chiung-Yu Huang, Mei-Cheng Wang
Joint Modeling And Estimation For Recurrent Event Processes And Failure Time Data, Chiung-Yu Huang, Mei-Cheng Wang
Johns Hopkins University, Dept. of Biostatistics Working Papers
Recurrent event data are commonly encountered in longitudinal follow-up studies related to biomedical science, econometrics, reliability, and demography. In many studies, recurrent events serve as important measurements for evaluating disease progression, health deterioration, or insurance risk. When analyzing recurrent event data, an independent censoring condition is typically required for the construction of statistical methods. Nevertheless, in some situations, the terminating time for observing recurrent events could be correlated with the recurrent event process and, as a result, the assumption of independent censoring is violated. In this paper, we consider joint modeling of a recurrent event process and a failure time …
Semi-Parametric Box-Cox Power Transformation Models For Censored Survival Observations, Tianxi Cai, Lu Tian, L. J. Wei
Semi-Parametric Box-Cox Power Transformation Models For Censored Survival Observations, Tianxi Cai, Lu Tian, L. J. Wei
Harvard University Biostatistics Working Paper Series
No abstract provided.
Unification Of Variance Components And Haseman-Elston Regression For Quantitative Trait Linkage Analysis, Wei-Min Chen, Karl W. Broman, Kung-Yee Liang
Unification Of Variance Components And Haseman-Elston Regression For Quantitative Trait Linkage Analysis, Wei-Min Chen, Karl W. Broman, Kung-Yee Liang
Johns Hopkins University, Dept. of Biostatistics Working Papers
Two of the major approaches for linkage analysis with quantitative traits in humans include variance components and Haseman-Elston regression. Previously, these have been viewed as quite separate methods. We describe a general model, fit by use of generalized estimating equations (GEE), for which the variance components and Haseman-Elston methods (including many of the extensions to the original Haseman-Elston method) are special cases, corresponding to different choices for a working covariance matrix. We also show that the regression-based test of Sham et al.(2002) is equivalent to a robust score statistic derived from our GEE approach. These results have several important implications. …
Smooth Quantile Ratio Estimation, Francesca Dominici, Leslie Cope, Daniel Q. Naiman, Scott L. Zeger
Smooth Quantile Ratio Estimation, Francesca Dominici, Leslie Cope, Daniel Q. Naiman, Scott L. Zeger
Johns Hopkins University, Dept. of Biostatistics Working Papers
In a study of health care expenditures attributable to smoking, we seek to compare the distribution of medical costs for persons with lung cancer or chronic obstructive pulmonary disease (cases) to those without (controls) using a national survey which includes hundreds of cases and thousands of controls. The distribution of costs is highly skewed toward larger values, making estimates of the mean from the smaller sample dependent on a small fraction of the biggest values. One approach to deal with the smaller sample is to rely on a simple parametric model such as the log-normal, but this makes the undesirable …
Hierarchical Bivariate Time Series Models: A Combined Analysis Of The Effects Of Particulate Matter On Morbidity And Mortality, Francesca Dominici, Antonella Zanobetti, Scott L. Zeger, Joel Schwartz, Jonathan M. Samet
Hierarchical Bivariate Time Series Models: A Combined Analysis Of The Effects Of Particulate Matter On Morbidity And Mortality, Francesca Dominici, Antonella Zanobetti, Scott L. Zeger, Joel Schwartz, Jonathan M. Samet
Johns Hopkins University, Dept. of Biostatistics Working Papers
In this paper we develop a hierarchical bivariate time series model to characterize the relationship between particulate matter less than 10 microns in aerodynamic diameter (PM10) and both mortality and hospital admissions for cardiovascular diseases. The model is applied to time series data on mortality and morbidity for 10 metropolitan areas in the United States from 1986 to 1993. We postulate that these time series should be related through a shared relationship with PM10.
At the first stage of the hierarchy, we fit two seemingly unrelated Poisson regression models to produce city-specific estimates of the log relative rates of mortality …
Statistical Inferences Based On Non-Smooth Estimating Functions, Lu Tian, Jun S. Liu, Mary Zhao, L. J. Wei
Statistical Inferences Based On Non-Smooth Estimating Functions, Lu Tian, Jun S. Liu, Mary Zhao, L. J. Wei
Harvard University Biostatistics Working Paper Series
No abstract provided.
On The Cox Model With Time-Varying Regression Coefficients, Lu Tian, David Zucker, L. J. Wei
On The Cox Model With Time-Varying Regression Coefficients, Lu Tian, David Zucker, L. J. Wei
Harvard University Biostatistics Working Paper Series
No abstract provided.
Nonparametric Estimation Of The Bivariate Recurrence Time Distribution, Chiung-Yu Huang, Mei-Cheng Wang
Nonparametric Estimation Of The Bivariate Recurrence Time Distribution, Chiung-Yu Huang, Mei-Cheng Wang
Johns Hopkins University, Dept. of Biostatistics Working Papers
This paper considers statistical models in which two different types of events, such as the diagnosis of a disease and the remission of the disease, occur alternately over time and are observed subject to right censoring. We propose nonparametric estimators for the joint distribution of bivariate recurrence times and the marginal distribution of the first recurrence time. In general, the marginal distribution of the second recurrence time cannot be estimated due to an identifiability problem, but a conditional distribution of the second recurrence time can be estimated non-parametrically. In literature, statistical methods have been developed to estimate the joint distribution …
A Nested Unsupervised Approach To Identifying Novel Molecular Subtypes, Elizabeth Garrett, Giovanni Parmigiani
A Nested Unsupervised Approach To Identifying Novel Molecular Subtypes, Elizabeth Garrett, Giovanni Parmigiani
Johns Hopkins University, Dept. of Biostatistics Working Papers
In classification problems arising in genomics research it is common to study populations for which a broad class assignment is known (say, normal versus diseased) and one seeks to find undiscovered subclasses within one or both of the known classes. Formally, this problem can be thought of as an unsupervised analysis nested within a supervised one. Here we take the view that the nested unsupervised analysis can successfully utilize information from the entire data set for constructing and/or selecting useful predictors. Specifically, we propose a mixture model approach to the nested unsupervised problem, where the supervised information is used to …
A Population Pharmacokinetic Model With Time-Dependent Covariates Measured With Errors, Lang Lil, Xihong Lin, Mort B. Brown, Suneel Gupta, Kyung-Hoon Lee
A Population Pharmacokinetic Model With Time-Dependent Covariates Measured With Errors, Lang Lil, Xihong Lin, Mort B. Brown, Suneel Gupta, Kyung-Hoon Lee
The University of Michigan Department of Biostatistics Working Paper Series
We propose a population pharmacokinetic (PK) model with time-dependent covariates measured with errors. This model is used to model S-oxybutynin's kinetics following an oral administration of Ditropan, and allows the distribution rate to depend on time-dependent covariates blood pressure and heart rate, which are measured with errors. We propose two two-step estimation methods: the second order two-step method with numerical solutions of differential equations (2orderND), and the second order two-step method with closed form approximate solutions of differential equations (2orderAD). The proposed methods are computationally easy and require fitting a linear mixed model at the first step and a nonlinear …
Stochastic Models Based On Molecular Hybridization Theory For Short Oligonucleotide Microarrays, Zhijin Wu, Richard Leblanc, Rafael A. Irizarry
Stochastic Models Based On Molecular Hybridization Theory For Short Oligonucleotide Microarrays, Zhijin Wu, Richard Leblanc, Rafael A. Irizarry
Johns Hopkins University, Dept. of Biostatistics Working Papers
High density oligonucleotide expression arrays are a widely used tool for the measurement of gene expression on a large scale. Affymetrix GeneChip arrays appear to dominate this market. These arrays use short oligonucleotides to probe for genes in an RNA sample. Due to optical noise, non-specific hybridization, probe-specific effects, and measurement error, ad-hoc measures of expression, that summarize probe intensities, can lead to imprecise and inaccurate results. Various researchers have demonstrated that expression measures based on simple statistical models can provide great improvements over the ad-hoc procedure offered by Affymetrix. Recently, physical models based on molecular hybridization theory, have been …
Efficient Semiparametric Marginal Estimation For Longitudinal/Clustered Data, Naisyin Wang, Raymond J. Carroll, Xihong Lin
Efficient Semiparametric Marginal Estimation For Longitudinal/Clustered Data, Naisyin Wang, Raymond J. Carroll, Xihong Lin
The University of Michigan Department of Biostatistics Working Paper Series
We consider marginal generalized semiparametric partially linear models for clustered data. Lin and Carroll (2001a) derived the semiparametric efficinet score funtion for this problem in the mulitvariate Gaussian case, but they were unable to contruct a semiparametric efficient estimator that actually achieved the semiparametric information bound. We propose such an estimator here and generalize the work to marginal generalized partially liner models. Asymptotic relative efficincies of the estimation or throughout are investigated. The finite sample performance of these estimators is evaluated through simulations and illustrated using a longtiudinal CD4 count data set. Both theoretical and numerical results indicate that properly …
Equivalent Kernels Of Smoothing Splines In Nonparametric Regression For Clustered/Longitudinal Data, Xihong Lin, Naisyin Wang, Alan H. Welsh, Raymond J. Carroll
Equivalent Kernels Of Smoothing Splines In Nonparametric Regression For Clustered/Longitudinal Data, Xihong Lin, Naisyin Wang, Alan H. Welsh, Raymond J. Carroll
The University of Michigan Department of Biostatistics Working Paper Series
We compare spline and kernel methods for clustered/longitudinal data. For independent data, it is well known that kernel methods and spline methods are essentially asymptotically equivalent (Silverman, 1984). However, the recent work of Welsh, et al. (2002) shows that the same is not true for clustered/longitudinal data. First, conventional kernel methods fail to account for the within- cluster correlation, while spline methods are able to account for this correlation. Second, kernel methods and spline methods were found to have different local behavior, with conventional kernels being local and splines being non-local. To resolve these differences, we show that a smoothing …
Histospline Method In Nonparametric Regression Models With Application To Clustered/Longitudinal Data, Raymond J. Carroll, Peter Hall, Tatiyana V. Apanasovich, Xihong Lin
Histospline Method In Nonparametric Regression Models With Application To Clustered/Longitudinal Data, Raymond J. Carroll, Peter Hall, Tatiyana V. Apanasovich, Xihong Lin
The University of Michigan Department of Biostatistics Working Paper Series
Kernel and smoothing methods for nonparametric function and curve estimation have been particularly successful in "standard" settings, where function values are observed subject to independent errors. However, when aspects of the function are known parametrically, or where the sampling scheme has significant structure, it can be quite difficult to adapt standard methods in such a way that they retain good statistical performance and continue to enjoy easy computability and good numerical properties. In particular, when using local linear modeling it is often awkward to both respect the sampling scheme and produce an estimator with good variance properties, without resorting to …
Measuring Treatment Effects Using Semiparametric Models, Zhuo Yu, Mark J. Van Der Laan
Measuring Treatment Effects Using Semiparametric Models, Zhuo Yu, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
In order to estimate the causal effect of treatments on an outcome of interest, one has to account for the effect of confounding factors which covary with the treatments and also contribute to the outcome of interest. In this paper, we use the semiparametric regression model to estimate the causal parameters. We assume the causal effect of the treatments can be described by the parametric component of the semiparametric regression model. Following the general methodology which was developed in van der Laan and Robins (2002) we give the orthogonal complement of the nuisance tangent space which identifies all the estimating …
A Varying-Coefficient Cox Model For The Effect Of Age At A Marker Event On Age At Menopause, Bin Nan, Xihong Lin, Lynda D. Lisabeth, Sioban D. Harlow
A Varying-Coefficient Cox Model For The Effect Of Age At A Marker Event On Age At Menopause, Bin Nan, Xihong Lin, Lynda D. Lisabeth, Sioban D. Harlow
The University of Michigan Department of Biostatistics Working Paper Series
. It is of recent interest in reproductive health research to investigate the validity of a marker event for the onset of menopausal transition and to estimate age at menopause using age at the marker event. We propose a varying coefficient Cox model to investigate the association between age at a marker event, denned as a specific bleeding pattern change, and age at menopause, where both events are subject to censoring and their association varies with age at the marker event. Estimation proceeds using the regression spline method. The proposed method is applied to the Tremin Trust Data to evaluate …