Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Medicine and Health Sciences (10)
- Biostatistics (8)
- Statistical Methodology (6)
- Statistical Theory (6)
- Statistical Models (5)
-
- Epidemiology (4)
- Life Sciences (4)
- Medical Specialties (4)
- Public Health (4)
- Clinical Epidemiology (3)
- Dentistry (3)
- Genetics (3)
- Genetics and Genomics (3)
- Longitudinal Data Analysis and Time Series (3)
- Microarrays (3)
- Bioinformatics (2)
- Clinical Trials (2)
- Computational Biology (2)
- Disease Modeling (2)
- Diseases (2)
- Radiology (2)
- Survival Analysis (2)
- Analytical, Diagnostic and Therapeutic Techniques and Equipment (1)
- Applied Mathematics (1)
- Biomedical Devices and Instrumentation (1)
- Biomedical Engineering and Bioengineering (1)
- Dental Materials (1)
- Institution
- Keyword
-
- Bootstrap (2)
- Classification (2)
- Colon cancer (2)
- Cross-validation (2)
- Functional genomics (2)
-
- Linear regression (2)
- Non-neoplastic mucosa (2)
- Prediction (2)
- Prognosis prediction (2)
- Receiver operating characteristic curve (2)
- Adjusted p-value (1)
- Causal effect; efficient influence curve; estimating function; prediction; variable importance; subject-specific variable importance (1)
- Causal inference (1)
- Cluster analysis (1)
- Clustered/longitudinal data; Generalized estimating equations; Generalized linear mixed models; Kernel method (1)
- Clustering (1)
- Correlation (1)
- Counterfactual (1)
- Differential expression (1)
- Differentiated effects; Heterogeneity; Linear mixed model; MCMCEM; Multiple outcomes data (1)
- Discriminant analysis (1)
- Double Robust (1)
- Effect erosion; HIV/AIDS clinical trials; Longitudinal outcomes; Mean response model; Time-varying coefficient; Treatment effectiveness duration (1)
- Ensemble methods (1)
- Errors in variables (1)
- Expected residual life (1)
- Factor analysis (1)
- Family-wise error rate (1)
- G-computation (1)
- Gene clustering (1)
- Publication
- Publication Type
Articles 1 - 22 of 22
Full-Text Articles in Multivariate Analysis
Autism And Parental Marital Satisfaction: The Role Of Adequacy Of Resources, Geneeta Kaliah Chambers
Autism And Parental Marital Satisfaction: The Role Of Adequacy Of Resources, Geneeta Kaliah Chambers
Loma Linda University Electronic Theses, Dissertations & Projects
The goal of the present study was to expand on the existing literature exploring families with children who have developmental disabilities, particularly autism. Previous studies have been constrained by univariate approaches that have failed to adequately capture the nuances of family functioning. Using an ecological/context approach, stemming from an ongoing research program conducted within a university-based treatment center, the present study attempted to improve on the conceptualization of interrelationships among family members and the role that contextual factors play within that dynamic. Specifically, the present study explored the influence of children’s level of autism on parents’ reports of their marital …
Accuracy Of The Newtom 3g™ In Measuring The Angle Of The Articular Eminence, Rehana Khan
Accuracy Of The Newtom 3g™ In Measuring The Angle Of The Articular Eminence, Rehana Khan
Loma Linda University Electronic Theses, Dissertations & Projects
The purpose of this study was to determine the accuracy of the Newtom 3G™ in determining the angulation of the articular eminence. The benefits of conducting this study were to provide additional uses for the standard records that are taken for the purposes of orthodontic treatment, as well as evaluate the Newtom 3G™ for accuracy in measuring the anatomy of the glenoid fossa. This study required 20 participants that volunteered to allow their records to be used. Records evaluated were the Newtom 3G™, impressions, and wax check bite registrations. The wax record was taken using the 'forced bite' technique to …
Data Adaptive Pathway Testing, Merrill D. Birkner, Alan E. Hubbard, Mark J. Van Der Laan
Data Adaptive Pathway Testing, Merrill D. Birkner, Alan E. Hubbard, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
A majority of diseases are caused by a combination of factors, for example, composite genetic mutation profiles have been found in many cases to predict a deleterious outcome. There are several statistical techniques that have been used to analyze these types of biological data. This article implements a general strategy which uses data adaptive regression methods to build a specific pathway model, thus predicting a disease outcome by a combination of biological factors and assesses the significance of this model, or pathway, by using a permutation based null distribution. We also provide several simulation comparisons with other techniques. In addition, …
Application Of A Variable Importance Measure Method To Hiv-1 Sequence Data, Merrill D. Birkner, Mark J. Van Der Laan
Application Of A Variable Importance Measure Method To Hiv-1 Sequence Data, Merrill D. Birkner, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
van der Laan (2005) proposed a method to construct variable importance measures and provided the respective statistical inference. This technique involves determining the importance of a variable in predicting an outcome. This method can be applied as an inverse probability of treatment weighted (IPTW) or double robust inverse probability of treatment weighted (DR-IPTW) estimator. A respective significance of the estimator is determined by estimating the influence curve and hence determining the corresponding variance and p-value. This article applies the van der Laan (2005) variable importance measures and corresponding inference to HIV-1 sequence data. In this data application, protease and reverse …
Estimating A Treatment Effect With Repeated Measurements Accounting For Varying Effectiveness Duration, Ying Qing Chen, Jingrong Yang, Su-Chun Cheng
Estimating A Treatment Effect With Repeated Measurements Accounting For Varying Effectiveness Duration, Ying Qing Chen, Jingrong Yang, Su-Chun Cheng
UW Biostatistics Working Paper Series
To assess treatment efficacy in clinical trials, certain clinical outcomes are repeatedly measured for same subject over time. They can be regarded as function of time. The difference in their mean functions between the treatment arms usually characterises a treatment effect. Due to the potential existence of subject-specific treatment effectiveness lag and saturation times, erosion of treatment effect in the difference may occur during the observation period of time. Instead of using ad hoc parametric or purely nonparametric time-varying coefficients in statistical modeling, we first propose to model the treatment effectiveness durations, which are the varying time intervals between the …
Modeling Differentiated Treatment Effects For Multiple Outcomes Data, Hongfei Guo, Karen Bandeen-Roche
Modeling Differentiated Treatment Effects For Multiple Outcomes Data, Hongfei Guo, Karen Bandeen-Roche
Johns Hopkins University, Dept. of Biostatistics Working Papers
Multiple outcomes data are commonly used to characterize treatment effects in medical research, for instance, multiple symptoms to characterize potential remission of a psychiatric disorder. Often either a global, i.e. symptom-invariant, treatment effect is evaluated. Such a treatment effect may over generalize the effect across the outcomes. On the other hand individual treatment effects, varying across all outcomes, are complicated to interpret, and their estimation may lose precision relative to a global summary. An effective compromise to summarize the treatment effect may be through patterns of the treatment effects, i.e. "differentiated effects." In this paper we propose a two-category model …
Semiparametric Estimation In General Repeated Measures Problems, Xihong Lin, Raymond J. Carroll
Semiparametric Estimation In General Repeated Measures Problems, Xihong Lin, Raymond J. Carroll
Harvard University Biostatistics Working Paper Series
This paper considers a wide class of semiparametric problems with a parametric part for some covariate effects and repeated evaluations of a nonparametric function. Special cases in our approach include marginal models for longitudinal/clustered data, conditional logistic regression for matched case-control studies, multivariate measurement error models, generalized linear mixed models with a semiparametric component, and many others. We propose profile-kernel and backfitting estimation methods for these problems, derive their asymptotic distributions, and show that in likelihood problems the methods are semiparametric efficient. While generally not true, with our methods profiling and backfitting are asymptotically equivalent. We also consider pseudolikelihood methods …
The Outcome Of Mta As A Root End Filling Material: A Long Term Evaluation, Christopher M. Sechrist
The Outcome Of Mta As A Root End Filling Material: A Long Term Evaluation, Christopher M. Sechrist
Loma Linda University Electronic Theses, Dissertations & Projects
Periradicular surgery is a viable option to save natural teeth when non-surgical treatment fails or when endodontic retreatment is not feasible or contraindicated. Laboratory and animal studies have demonstrated that MTA is biocompatible, provides an excellent seal against penetrating bacteria, and promotes hard tissue healing. The purpose of this study was to provide long term (>3 years) clinical evidence for its use as a root-end filling material in endodontics. The clinical records of 294 patients who had MTA used during endodontic treatment from 1996 to 2001 were reviewed. From these, 75 patients whose root end cavities had been filled …
Laser And Led Effects On The Proliferation Rate Of Periodontal Ligament Fibroblasts, Allen J. Job
Laser And Led Effects On The Proliferation Rate Of Periodontal Ligament Fibroblasts, Allen J. Job
Loma Linda University Electronic Theses, Dissertations & Projects
PURPOSE: To compare the effectiveness of a Gallium Aluminum Arsenide (GaAlAs) diode laser and a light emitting diode (LED) on periodontal ligament fibroblast cell proliferative rates.
METHODS and MATERIALS: PDLF obtained from freshly extracted permanent teeth were cultured under standard conditions until a subconfluent monolayer was present. The next section took 5 days to complete. On day 1, the initial cell concentration of 700 uL/cm2 was plated on 96-well assay plates and placed in a CO2 incubator at 37° C for 24 hours. On day 2, cell counts were first verified using hemocytometry then were irradiated using an …
Statistical Inference For Variable Importance, Mark J. Van Der Laan
Statistical Inference For Variable Importance, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
Many statistical problems involve the learning of an importance/effect of a variable for predicting an outcome of interest based on observing a sample of n independent and identically distributed observations on a list of input variables and an outcome. For example, though prediction/machine learning is, in principle, concerned with learning the optimal unknown mapping from input variables to an outcome from the data, the typical reported output is a list of importance measures for each input variable. The typical approach in prediction has been to learn the unknown optimal predictor from the data and derive, for each of the input …
Test Statistics Null Distributions In Multiple Testing: Simulation Studies And Applications To Genomics, Katherine S. Pollard, Merrill D. Birkner, Mark J. Van Der Laan, Sandrine Dudoit
Test Statistics Null Distributions In Multiple Testing: Simulation Studies And Applications To Genomics, Katherine S. Pollard, Merrill D. Birkner, Mark J. Van Der Laan, Sandrine Dudoit
U.C. Berkeley Division of Biostatistics Working Paper Series
Multiple hypothesis testing problems arise frequently in biomedical and genomic research, for instance, when identifying differentially expressed or co-expressed genes in microarray experiments. We have developed generally applicable resampling-based single-step and stepwise multiple testing procedures (MTP) for control of a broad class of Type I error rates, defined as tail probabilities and expected values for arbitrary functions of the numbers of false positives and rejected hypotheses (Dudoit and van der Laan, 2005; Dudoit et al., 2004a,b; Pollard and van der Laan, 2004; van der Laan et al., 2005, 2004a,b). As argued in the early article of Pollard and van der …
On Additive Regression Of Expectancy, Ying Qing Chen
On Additive Regression Of Expectancy, Ying Qing Chen
UW Biostatistics Working Paper Series
Regression models have been important tools to study the association between outcome variables and their covariates. The traditional linear regression models usually specify such an association by the expectations of the outcome variables as function of the covariates and some parameters. In reality, however, interests often focus on their expectancies characterized by the conditional means. In this article, a new class of additive regression models is proposed to model the expectancies. The model parameters carry practical implication, which may allow the models to be useful in applications such as treatment assessment, resource planning or short-term forecasting. Moreover, the new model …
New Statistical Paradigms Leading To Web-Based Tools For Clinical/Translational Science, Knut M. Wittkowski
New Statistical Paradigms Leading To Web-Based Tools For Clinical/Translational Science, Knut M. Wittkowski
COBRA Preprint Series
As the field of functional genetics and genomics is beginning to mature, we become confronted with new challenges. The constant drop in price for sequencing and gene expression profiling as well as the increasing number of genetic and genomic variables that can be measured makes it feasible to address more complex questions. The success with rare diseases caused by single loci or genes has provided us with a proof-of-concept that new therapies can be developed based on functional genomics and genetics.
Common diseases, however, typically involve genetic epistasis, genomic pathways, and proteomic pattern. Moreover, to better understand the underlying biologi-cal …
Prognosis Of Stage Ii Colon Cancer By Non-Neoplastic Mucosa Gene Expresssion Profiling, Alain Barrier, Sandrine Dudoit, Et Al.
Prognosis Of Stage Ii Colon Cancer By Non-Neoplastic Mucosa Gene Expresssion Profiling, Alain Barrier, Sandrine Dudoit, Et Al.
U.C. Berkeley Division of Biostatistics Working Paper Series
Aims. This study assessed the possibility to build a prognosis predictor, based on non-neoplastic mucosa microarray gene expression measures, in stage II colon cancer patients. Materials and Methods. Non-neoplastic colonic mucosa mRNA samples from 24 patients (10 with a metachronous metastasis, 14 with no recurrence) were profiled using the Affymetrix HGU133A GeneChip. The k-nearest neighbor method was used for prognosis prediction using microarray gene expression measures. Leave-one-out cross-validation was used to select the number of neighbors and number of informative genes to include in the predictor. Based on this information, a prognosis predictor was proposed and its accuracy estimated by …
Colon Cancer Prognosis Prediction By Gene Expression Profiling, Alain Barrier, Sandrine Dudoit, Et Al.
Colon Cancer Prognosis Prediction By Gene Expression Profiling, Alain Barrier, Sandrine Dudoit, Et Al.
U.C. Berkeley Division of Biostatistics Working Paper Series
Aims. This study assessed the possibility to build a prognosis predictor, based on microarray gene expression measures, in stage II and III colon cancer patients. Materials and Methods. Tumour (T) and non-neoplastic mucosa (NM) mRNA samples from 18 patients (9 with a recurrence, 9 with no recurrence) were profiled using the Affymetrix HGU133A GeneChip. The k-nearest neighbour method was used for prognosis prediction using T and NM gene expression measures. Six-fold cross-validation was applied to select the number of neighbours and the number of informative genes to include in the predictors. Based on this information, one T-based and one NM-based …
Causal Inference In Longitudinal Studies With History-Restricted Marginal Structural Models, Romain Neugebauer, Mark J. Van Der Laan, Ira B. Tager
Causal Inference In Longitudinal Studies With History-Restricted Marginal Structural Models, Romain Neugebauer, Mark J. Van Der Laan, Ira B. Tager
U.C. Berkeley Division of Biostatistics Working Paper Series
Causal Inference based on Marginal Structural Models (MSMs) is particularly attractive to subject-matter investigators because MSM parameters provide explicit representations of causal effects. We introduce History-Restricted Marginal Structural Models (HRMSMs) for longitudinal data for the purpose of defining causal parameters which may often be better suited for Public Health research. This new class of MSMs allows investigators to analyze the causal effect of a treatment on an outcome based on a fixed, shorter and user-specified history of exposure compared to MSMs. By default, the latter represents the treatment causal effect of interest based on a treatment history defined by the …
Survival Ensembles, Torsten Hothorn, Peter Buhlmann, Sandrine Dudoit, Annette M. Molinaro, Mark J. Van Der Laan
Survival Ensembles, Torsten Hothorn, Peter Buhlmann, Sandrine Dudoit, Annette M. Molinaro, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
We propose a unified and flexible framework for ensemble learning in the presence of censoring. For right-censored data, we introduce a random forest algorithm and a generic gradient boosting algorithm for the construction of prognostic models. The methodology is utilized for predicting the survival time of patients suffering from acute myeloid leukemia based on clinical and genetic covariates. Furthermore, we compare the diagnostic capabilities of the proposed censored data random forest and boosting methods applied to the recurrence free survival time of node positive breast cancer patients with previously published findings.
The Clustering Of Regression Models Method With Applications In Gene Expression Data, Li-Xuan Qin, Steven G. Self
The Clustering Of Regression Models Method With Applications In Gene Expression Data, Li-Xuan Qin, Steven G. Self
UW Biostatistics Working Paper Series
Identification of differentially expressed genes and clustering of genes are two important and complementary objectives addressed with gene expression data. For the differential expression question, many "per-gene" analytic methods have been proposed. These methods can generally be characterized as using a regression function to independently model the observations for each gene; various adjustments for multiplicity are then used to interpret the statistical significance of these per-gene regression models over the collection of genes analyzed. Motivated by this common structure of per-gene models, we propose a new model-based clustering method -- the clustering of regression models method, which groups genes that …
Insights Into Latent Class Analysis, Margaret S. Pepe, Holly Janes
Insights Into Latent Class Analysis, Margaret S. Pepe, Holly Janes
UW Biostatistics Working Paper Series
Latent class analysis is a popular statistical technique for estimating disease prevalence and test sensitivity and specificity. It is used when a gold standard assessment of disease is not available but results of multiple imperfect tests are. We derive analytic expressions for the parameter estimates in terms of the raw data, under the conditional independence assumption. These expressions indicate explicitly how observed two- and three-way associations between test results are used to infer disease prevalence and test operating characteristics. Although reasonable if the conditional independence model holds, the estimators have no basis when it fails. We therefore caution against using …
Standardizing Markers To Evaluate And Compare Their Performances, Margaret S. Pepe, Gary M. Longton
Standardizing Markers To Evaluate And Compare Their Performances, Margaret S. Pepe, Gary M. Longton
UW Biostatistics Working Paper Series
Introduction: Markers that purport to distinguish subjects with a condition from those without a condition must be evaluated rigorously for their classification accuracy. A single approach to statistically evaluating and comparing markers is not yet established.
Methods: We suggest a standardization that uses the marker distribution in unaffected subjects as a reference. For an affected subject with marker value Y, the standardized placement value is the proportion of unaffected subjects with marker values that exceed Y.
Results: We apply the standardization to two illustrative datasets. In patients with pancreatic cancer placement values calculated for the CA 19-9 marker are smaller …
Combining Predictors For Classification Using The Area Under The Roc Curve, Margaret S. Pepe, Tianxi Cai, Zheng Zhang, Gary M. Longton
Combining Predictors For Classification Using The Area Under The Roc Curve, Margaret S. Pepe, Tianxi Cai, Zheng Zhang, Gary M. Longton
UW Biostatistics Working Paper Series
No single biomarker for cancer is considered adequately sensitive and specific for cancer screening. It is expected that the results of multiple markers will need to be combined in order to yield adequately accurate classification. Typically the objective function that is optimized for combining markers is the likelihood function. In this paper we consider an alternative objective function -- the area under the empirical receiver operating characteristic curve (AUC). We note that it yields consistent estimates of parameters in a generalized linear model for the risk score but does not require specifying the link function. Like logistic regression it yields …
Cluster Analysis Of Genomic Data With Applications In R, Katherine S. Pollard, Mark J. Van Der Laan
Cluster Analysis Of Genomic Data With Applications In R, Katherine S. Pollard, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
In this paper, we provide an overview of existing partitioning and hierarchical clustering algorithms in R. We discuss statistical issues and methods in choosing the number of clusters, the choice of clustering algorithm, and the choice of dissimilarity matrix. In particular, we illustrate how the bootstrap can be employed as a statistical method in cluster analysis to establish the reproducibility of the clusters and the overall variability of the followed procedure. We also show how to visualize a clustering result by plotting ordered dissimilarity matrices in R. We present a new R package, hopach, which implements the hybrid clustering method, …