Open Access. Powered by Scholars. Published by Universities.®
- Institution
- Keyword
-
- Classification (4)
- Gene expression (4)
- Genetics (4)
- Bioinformatics (2)
- Bootstrap (2)
-
- Clustering (2)
- Comparative genomic hybridization (2)
- Cross-validation (2)
- Differential expression (2)
- Genomics (2)
- Microarray (2)
- Mixture models (2)
- Prediction (2)
- Survival analysis (2)
- (Quasi)separation (1)
- Adjust p value (1)
- Aging (1)
- Alcoholic liver disease (1)
- BLUPs; Kernel function; Model/variable selection; Nonparametric regression; Penalized likelihood; REML; Score test; Smoothing parameter; Support vector machines (1)
- Bayesian Additive Regression Trees (1)
- Bayesian Inference (1)
- Bayesian shrinkage priors (1)
- Bayesian trees (1)
- Beta mixture; DNA methylation; cancer; epigenetics; mixture model (1)
- Bilaterally contaminated normal model (1)
- Biomarker (1)
- Biomedical signal processing (1)
- Brain tumor (1)
- Cancer genomics (1)
- Case-control study (1)
- Publication Year
- Publication
-
- UW Biostatistics Working Paper Series (6)
- U.C. Berkeley Division of Biostatistics Working Paper Series (5)
- COBRA Preprint Series (4)
- Harvard University Biostatistics Working Paper Series (4)
- Dissertations and Theses (Open Access) (2)
-
- Electronic Theses and Dissertations (2)
- The University of Michigan Department of Biostatistics Working Paper Series (2)
- Theses and Dissertations--Statistics (2)
- Annual Symposium on Biomathematics and Ecology Education and Research (1)
- Bioconductor Project Working Papers (1)
- Graduate Theses and Dissertations (1)
- Theses and Dissertations (1)
- Publication Type
Articles 31 - 31 of 31
Full-Text Articles in Microarrays
Statistical Inference For Simultaneous Clustering Of Gene Expression Data, Katherine S. Pollard, Mark J. Van Der Laan
Statistical Inference For Simultaneous Clustering Of Gene Expression Data, Katherine S. Pollard, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
Current methods for analysis of gene expression data are mostly based on clustering and classification of either genes or samples. We offer support for the idea that more complex patterns can be identified in the data if genes and samples are considered simultaneously. We formalize the approach and propose a statistical framework for two-way clustering. A simultaneous clustering parameter is defined as a function of the true data generating distribution, and an estimate is obtained by applying this function to the empirical distribution. We illustrate that a wide range of clustering procedures, including generalized hierarchical methods, can be defined as …