Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Statistical Theory (103)
- Applied Statistics (71)
- Social and Behavioral Sciences (62)
- Statistical Methodology (50)
- Statistical Models (50)
-
- Medicine and Health Sciences (43)
- Public Health (33)
- Mathematics (27)
- Survival Analysis (27)
- Life Sciences (25)
- Applied Mathematics (23)
- Design of Experiments and Sample Surveys (19)
- Numerical Analysis and Computation (19)
- Genetics and Genomics (17)
- Microarrays (17)
- Epidemiology (16)
- Genetics (15)
- Computer Sciences (14)
- Longitudinal Data Analysis and Time Series (14)
- Multivariate Analysis (14)
- Categorical Data Analysis (11)
- Clinical Epidemiology (10)
- Clinical Trials (8)
- Diseases (8)
- Disease Modeling (7)
- Education (7)
- Health Services Research (7)
- Law (7)
- Institution
-
- COBRA (93)
- Wayne State University (52)
- Missouri University of Science and Technology (13)
- Brigham Young University (7)
- Cornell University Law School (7)
-
- Loma Linda University (7)
- New Jersey Institute of Technology (6)
- Portland State University (5)
- Claremont Colleges (3)
- Marquette University (3)
- Singapore Management University (3)
- Swarthmore College (3)
- University of Richmond (3)
- Western Michigan University (3)
- Wright State University (3)
- Air Force Institute of Technology (2)
- California Polytechnic State University, San Luis Obispo (2)
- Old Dominion University (2)
- Southern Illinois University Carbondale (2)
- University of Kentucky (2)
- WellBeing International (2)
- Chulalongkorn University (1)
- Department of Primary Industries and Regional Development, Western Australia (1)
- Edith Cowan University (1)
- Indiana State University (1)
- Kennesaw State University (1)
- Louisiana Tech University (1)
- Montclair State University (1)
- University of Denver (1)
- University of Massachusetts Boston (1)
- Keyword
-
- Bootstrap (6)
- Cross-validation (6)
- Gene expression (6)
- Counting process (5)
- Empirical legal studies (5)
-
- Prediction (5)
- Causal inference (4)
- Counterfactual (4)
- Data mining (4)
- False discovery rate (4)
- MCMC (4)
- Microarray (4)
- Missing data (4)
- Model selection (4)
- Multiple comparisons (4)
- Multiple testing (4)
- Adjusted p-value (3)
- Air pollution (3)
- Capital punishment (3)
- Censoring (3)
- Classification (3)
- Clinical trials (3)
- Confounding (3)
- Current status data (3)
- Death penalty (3)
- Differential expression (3)
- Double robust estimation (3)
- Estimating equation (3)
- G-computation estimation (3)
- Generalized family-wise error rate (3)
- Publication
-
- Journal of Modern Applied Statistical Methods (52)
- The University of Michigan Department of Biostatistics Working Paper Series (28)
- U.C. Berkeley Division of Biostatistics Working Paper Series (24)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (19)
- Mathematics and Statistics Faculty Research & Creative Works (13)
-
- UW Biostatistics Working Paper Series (13)
- Theses and Dissertations (9)
- Cornell Law Faculty Publications (7)
- Loma Linda University Electronic Theses, Dissertations & Projects (7)
- Theses (6)
- Harvard University Biostatistics Working Paper Series (5)
- Bioconductor Project Working Papers (4)
- Complex Systems Faculty Publications and Presentations (4)
- Department of Math & Statistics Faculty Publications (3)
- Dissertations (3)
- Mathematics & Statistics Faculty Works (3)
- Mathematics, Statistics and Computer Science Faculty Research and Publications (3)
- Research Collection School Of Economics (3)
- Articles and Preprints (2)
- Mathematics and Statistics Faculty Publications (2)
- Statistics (2)
- All Graduate Plan B and other Reports, Spring 1920 to Spring 2023 (1)
- All HMC Faculty Publications and Research (1)
- All-Inclusive List of Electronic Theses and Dissertations (1)
- CGU Faculty Publications and Research (1)
- Chulalongkorn University Theses and Dissertations (Chula ETD) (1)
- Department of Applied Mathematics and Statistics Faculty Scholarship and Creative Works (1)
- Dissertations and Theses (1)
- Doctoral Dissertations (1)
- Electrical & Computer Engineering Theses & Dissertations (1)
- Publication Type
Articles 91 - 120 of 235
Full-Text Articles in Statistics and Probability
A Hierarchical Multivariate Two-Part Model For Profiling Providers' Effects On Healthcare Charges, John W. Robinson, Scott L. Zeger, Christopher B. Forrest
A Hierarchical Multivariate Two-Part Model For Profiling Providers' Effects On Healthcare Charges, John W. Robinson, Scott L. Zeger, Christopher B. Forrest
Johns Hopkins University, Dept. of Biostatistics Working Papers
Procedures for analyzing and comparing healthcare providers' effects on health services delivery and outcomes have been referred to as provider profiling. In a typical profiling procedure, patient-level responses are measured for clusters of patients treated by providers that in turn, can be regarded as statistically exchangeable. Thus, a hierarchical model naturally represents the structure of the data. When provider effects on multiple responses are profiled, a multivariate model rather than a series of univariate models, can capture associations among responses at both the provider and patient levels. When responses are in the form of charges for healthcare services and sampled …
Building A Model Of Campus Diversity: An Organization-As-The-Individual Based Approach, Ian Spencer Ray
Building A Model Of Campus Diversity: An Organization-As-The-Individual Based Approach, Ian Spencer Ray
Electronic Theses and Dissertations
Diversity is a vital concept when engaging with modern social science research and among higher education practitioners. Despite the pervasive discourse surrounding diversity, equity, and inclusion (DEI) initiatives, quantification of diversity remains a challenge across the social sciences. This study proposes a new conceptual framework, applies the framework to the quantification of diversity and measurement of its impacts. Further, this study demonstrates the robustness of the new conceptual framework alongside the positive impacts of diversity on institutional outcomes.
The Organization-As-The-Individual (OATI) framework arises from an integration of the social and biological sciences. The framework argues that humans, as biological organisms, …
Estimation Of Direct And Indirect Causal Effects In Longitudinal Studies, Mark J. Van Der Laan, Maya L. Petersen
Estimation Of Direct And Indirect Causal Effects In Longitudinal Studies, Mark J. Van Der Laan, Maya L. Petersen
U.C. Berkeley Division of Biostatistics Working Paper Series
The causal effect of a treatment on an outcome is generally mediated by several intermediate variables. Estimation of the component of the causal effect of a treatment that is mediated by a given intermediate variable (the indirect effect of the treatment), and the component that is not mediated by that intermediate variable (the direct effect of the treatment) is often relevant to mechanistic understanding and to the design of clinical and public health interventions. Under the assumption of no-unmeasured confounders, Robins & Greenland (1992) and Pearl (2000), develop two identifiability results for direct and indirect causal effects. They define an …
Linear Life Expectancy Regression With Censored Data, Ying Qing Chen, Su-Chun Cheng
Linear Life Expectancy Regression With Censored Data, Ying Qing Chen, Su-Chun Cheng
U.C. Berkeley Division of Biostatistics Working Paper Series
Life expectancy, i.e., mean residual life function, has been of important practical and scientific interests to characterise the distribution of residual life. Regression models are often needed to model the association between life expectancy and its covariates. In this article, we consider a linear mean residual life model and further developed some inference procedures in presence of censoring. The new model and proposed inference procedure will be demonstrated by numerical examples and application to the well-known Stanford heart transplant data. Additional semiparametric efficiency calculation and information bound are also considered.
A Comparative Study Of Interrater Reliability Coefficients Obtained From Different Statistical Procedures Using Monte Carlo Simulation Techniques, Ebrima Nying
Dissertations
Reliability estimation is a key research component within the global area of educational assessment. The literature reports numerous studies using different statistical techniques for estimating reliability of educational measures. However, few have focused on the estimation of interrater reliability of performance assessment (Abedi, Baker, & Herl, 1995). Specifically, this study compared three different methods for estimating interrater reliability to determine if there are differences among these estimates as a function of: sample size, measurement scale, number of raters and the theoretical population reliability (rho). The three methods of estimation were the Intraclass Correlation (ICC(2, k )) (Shrout & Fleiss, 1979), …
Enhancements To Crisp Possibilistic Reconstructability Analysis, Anas Al-Rabadi, Martin Zwick
Enhancements To Crisp Possibilistic Reconstructability Analysis, Anas Al-Rabadi, Martin Zwick
Complex Systems Faculty Publications and Presentations
Modified Reconstructibility Analysis (MRA), a novel decomposition within the framework of set-theoretic (crisp possibilistic) Reconstructibility Analysis, is presented. It is shown that in some cases while 3-variable NPN-classified Boolean functions are not decomposable using Conventional Reconstructibility Analysis (CRA), they are decomposable using Modified Reconstructibility Analysis (MRA). Also, it is shown that whenever a decomposition of 3-variable NPN-classified Boolean functions exists in both MRA and CRA, MRA yields simpler or equal complexity decompositions. A comparison of the corresponding complexities for Ashenhurst-Curtis decompositions, and Modified Reconstructibility Analysis (MRA) is also presented. While both AC and MRA decompose some but …
Instructing Teachers Of Children With Disabilities Within The Church Of Jesus Christ Of Latter-Day Saints, Katie E. Sampson
Instructing Teachers Of Children With Disabilities Within The Church Of Jesus Christ Of Latter-Day Saints, Katie E. Sampson
Theses and Dissertations
This study investigates benefits of in-service training on LDS primary teachers' ability to state an objective, obtain and keep attention, use wait time, incorporate active participation, teach to the multiple intelligences, and employ positive behavior management techniques. Two groups of 30 viewed either a video-tape or read a handout. Pre and post surveys were used to determine mean gain.
Using an ANCOVA, comparisons were made of overall mean gain for each group. Results showed participants made a gain of approximately 1/2 point per question on a 4-point scale on the video and the handout (video gain = .6032 p<.01; handout gain = .6264 p<.01). The results of this study support the hypothesis that teachers receiving one in-service will increase their perception of their ability to teach students with special needs.
Non-Parametric Estimation Of Roc Curves In The Absence Of A Gold Standard, Xiao-Hua Zhou, Pete Castelluccio, Chuan Zhou
Non-Parametric Estimation Of Roc Curves In The Absence Of A Gold Standard, Xiao-Hua Zhou, Pete Castelluccio, Chuan Zhou
UW Biostatistics Working Paper Series
In evaluation of diagnostic accuracy of tests, a gold standard on the disease status is required. However, in many complex diseases, it is impossible or unethical to obtain such the gold standard. If an imperfect standard is used as if it were a gold standard, the estimated accuracy of the tests would be biased. This type of bias is called imperfect gold standard bias. In this paper we develop a maximum likelihood (ML) method for estimating ROC curves and their areas of ordinal-scale tests in the absence of a gold standard. Our simulation study shows the proposed estimates for the …
The Optimal Confidence Region For A Random Parameter, Hajime Uno, Lu Tian, L.J. Wei
The Optimal Confidence Region For A Random Parameter, Hajime Uno, Lu Tian, L.J. Wei
Harvard University Biostatistics Working Paper Series
Under a two-level hierarchical model, suppose that the distribution of the random parameter is known or can be estimated well. Data are generated via a fixed, but unobservable realization of this parameter. In this paper, we derive the smallest confidence region of the random parameter under a joint Bayesian/frequentist paradigm. On average this optimal region can be much smaller than the corresponding Bayesian highest posterior density region. The new estimation procedure is appealing when one deals with data generated under a highly parallel structure, for example, data from a trial with a large number of clinical centers involved or genome-wide …
Probabilistic Methodology For Record Linkage Determining Robustness Of Weights, Krista Peine Jensen
Probabilistic Methodology For Record Linkage Determining Robustness Of Weights, Krista Peine Jensen
Theses and Dissertations
Record linkage is the process that joins separately recorded pieces of information for a particular individual from one or more sources. To facilitate record linkage, a reliable computer based approach is ideal. In genealogical research computerized record linkage is useful in combing information for an individual across multiple censuses.
In creating a computerized method for linking censuse records it needs to be determined if weights calculated from one geographical area, can be used to link records from another geographical area. Research performed by Marcie Francis calculates field weights using census records from 1910 and 1920 for Ascension Parish Louisiana. These …
Hiring Practices For Graphic Designers In Utah County, Utah, Landon T. Densley
Hiring Practices For Graphic Designers In Utah County, Utah, Landon T. Densley
Theses and Dissertations
The purpose of this study was to show how hiring standards of evidence for graphic designers in Utah County compared with the national standards of evidence. The four major national standards of evidence for hiring graphic designers, identified by American Institute of Graphic Arts (AIGA) and Goldfarb, in order of importance are portfolio, recommendations, personality, and education. The data from this study revealed that Utah County employer's standards of evidence matched up closely to national standards of evidence, but the order of importance was slightly different because personality was ranked ahead of recommendations and education.
A Note On Empirical Likelihood Inference Of Residual Life Regression, Ying Qing Chen, Yichuan Zhao
A Note On Empirical Likelihood Inference Of Residual Life Regression, Ying Qing Chen, Yichuan Zhao
U.C. Berkeley Division of Biostatistics Working Paper Series
Mean residual life function, or life expectancy, is an important function to characterize distribution of residual life. The proportional mean residual life model by Oakes and Dasu (1990) is a regression tool to study the association between life expectancy and its associated covariates. Although semiparametric inference procedures have been proposed in the literature, the accuracy of such procedures may be low when the censoring proportion is relatively large. In this paper, the semiparametric inference procedures are studied with an empirical likelihood ratio method. An empirical likelihood confidence region is constructed for the regression parameters. The proposed method is further compared …
Large Prandtl Number Behavior Of The Boussinesq System Of Rayleigh-Bénard Convection, Xiaoming Wang
Large Prandtl Number Behavior Of The Boussinesq System Of Rayleigh-Bénard Convection, Xiaoming Wang
Mathematics and Statistics Faculty Research & Creative Works
We establish the validity of the infinite Prandtl number model as an approximation of the Boussinesq system at large Prandtl number on finite and infinite time interval, as well as in some statistical sense. © 2004 Elsevier Ltd. All rights reserved.
Semiparametric Quantitative-Trait-Locus Mapping: I. On Functional Growth Curves, Ying Qing Chen, Rongling Wu
Semiparametric Quantitative-Trait-Locus Mapping: I. On Functional Growth Curves, Ying Qing Chen, Rongling Wu
U.C. Berkeley Division of Biostatistics Working Paper Series
The genetic study of certain quantitative traits in growth curves as a function of time has recently been of major scientific interest to explore the developmental evolution processes of biological subjects. Various parametric approaches in the statistical literature have been proposed to study the quantitative-trait-loci (QTL) mapping of the growth curves as multivariate outcomes. In this article, we view the growth curves as functional quantitative traits and propose some semiparametric models to relax the strong parametric assumptions which may not be always practical in reality. Appropriate inference procedures are developed to estimate the parameters of interest which characterise the possible …
Semiparametric Quantitative-Trait-Locus Mapping: Ii. On Censored Age-At-Onset, Ying Qing Chen, Chengcheng Hu, Rongling Wu
Semiparametric Quantitative-Trait-Locus Mapping: Ii. On Censored Age-At-Onset, Ying Qing Chen, Chengcheng Hu, Rongling Wu
U.C. Berkeley Division of Biostatistics Working Paper Series
In genetic studies, the variation in genotypes may not only affect different inheritance patterns in qualitative traits, but may also affect the age-at-onset as quantitative trait. In this article, we use standard cross designs, such as backcross or F2, to propose some hazard regression models, namely, the additive hazards model in quantitative trait loci mapping for age-at-onset, although the developed method can be extended to more complex designs. With additive invariance of the additive hazards models in mixture probabilities, we develop flexible semiparametric methodologies in interval regression mapping without heavy computing burden. A recently developed multiple comparison procedures is adapted …
The Morphology Of Steve, Eugenie C. Scott, Nicholas J. Matzke, Glenn Branch, Steven Mccullagh
The Morphology Of Steve, Eugenie C. Scott, Nicholas J. Matzke, Glenn Branch, Steven Mccullagh
Faculty Articles
This report is part of Project Steve. Project Steve is, among other things, the first scientific analysis of the sex, geographic location, and body size of scientists named Steve. We performed this research for the best of all reasons: we discovered that we had lots of data. No scientist can resist the opportunity to analyze data, regardless of where that data came from or why it was gathered.
Oscillation Of Symplectic Dynamic Systems, Martin Bohner, Ondřej Došlý
Oscillation Of Symplectic Dynamic Systems, Martin Bohner, Ondřej Došlý
Mathematics and Statistics Faculty Research & Creative Works
We investigate oscillatory properties of a perturbed symplectic dynamic system on a time scale that is unbounded above. the unperturbed system is supposed to be Non oscillatory, and we give conditions on the perturbation matrix, which guarantee that the perturbed system becomes oscillatory. Examples illustrating the general results are given as well. © Australian Mathematical Society 2004.
Implicit Racial Attitudes Of Death Penalty Lawyers, Theodore Eisenberg, Sheri Lynn Johnson
Implicit Racial Attitudes Of Death Penalty Lawyers, Theodore Eisenberg, Sheri Lynn Johnson
Cornell Law Faculty Publications
Defense attorneys commonly suspect that the defendant's race plays a role in prosecutors' decisions to seek the death penalty, especially when the victim of the crime was white. When the defendant is convicted of the crime and sentenced to death, it is equally common for such attorneys to question the racial attitudes of the jury. These suspicions are not merely partisan conjectures; ample historical, statistical, and anecdotal evidence supports the inference that race matters in capital cases. Even the General Accounting Office of the United States concludes as much. Despite McCleskey v. Kemp, in which the United States Supreme Court …
Quantification And Visualization Of Ld Patterns And Identification Of Haplotype Blocks, Yan Wang, Sandrine Dudoit
Quantification And Visualization Of Ld Patterns And Identification Of Haplotype Blocks, Yan Wang, Sandrine Dudoit
U.C. Berkeley Division of Biostatistics Working Paper Series
Classical measures of linkage disequilibrium (LD) between two loci, based only on the joint distribution of alleles at these loci, present noisy patterns. In this paper, we propose a new distance-based LD measure, R, which takes into account multilocus haplotypes around the two loci in order to exploit information from neighboring loci. The LD measure R yields a matrix of pairwise distances between markers, based on the correlation between the lengths of shared haplotypes among chromosomes around these markers. Data analysis demonstrates that visualization of LD patterns through the R matrix reveals more deterministic patterns, with much less noise, than …
Quantitative Methods For Tracking Cognitive Change 3 Years After Coronary Artery Bypass Surgery, Sarah Barry, Scott L. Zeger, Ola A. Selnes, Maura A. Grega, Louis M. Borowicz, Jr., Guy M. Mckhann
Quantitative Methods For Tracking Cognitive Change 3 Years After Coronary Artery Bypass Surgery, Sarah Barry, Scott L. Zeger, Ola A. Selnes, Maura A. Grega, Louis M. Borowicz, Jr., Guy M. Mckhann
Johns Hopkins University, Dept. of Biostatistics Working Papers
Background: The analysis and interpretation of change in cognitive function test scores after Coronary Artery Bypass Grafting (CABG). Longitudinal studies with multiple outcomes present considerable statistical challenges. Application of hierarchical linear statistical models can estimate the effects of a surgical intervention on the time course of multiple biomarkers.
Methods: We use an "analyze then summarize" approach whereby we estimate the intervention effects separately for each cognitive test and then pool them, taking appropriate account of their statistical correlations. The model accounts for dropouts at follow-up, the chance of which may be related to past cognitive score, by implicitly imputing the …
New Estimating Methods For Surrogate Outcome Data, Bin Nan
New Estimating Methods For Surrogate Outcome Data, Bin Nan
The University of Michigan Department of Biostatistics Working Paper Series
Surrogate outcome data arise frequently in medical research. The true outcomes of interest are expensive or hard to ascertain, but measurements of surrogate outcomes (or more generally speaking, the correlates of the true outcomes) are usually available. In this paper we assume that the conditional expectation of the true outcome given covariates is known up to a finite dimensional parameter. When the true outcome is missing at random, the e±cient score function for the parameter in the conditional mean model has a simple form, which is similar to the generalized estimating functions. There is no integral equation involved as in …
Differential Expression With The Bioconductor Project, Anja Von Heydebreck, Wolfgang Huber, Robert Gentleman
Differential Expression With The Bioconductor Project, Anja Von Heydebreck, Wolfgang Huber, Robert Gentleman
Bioconductor Project Working Papers
A basic, yet challenging task in the analysis of microarray gene expression data is the identification of changes in gene expression that are associated with particular biological conditions. We discuss different approaches to this task and illustrate how they can be applied using software from the Bioconductor Project. A central problem is the high dimensionality of gene expression space, which prohibits a comprehensive statistical analysis without focusing on particular aspects of the joint distribution of the genes expression levels. Possible strategies are to do univariate gene-by-gene analysis, and to perform data-driven nonspecific filtering of genes before the actual statistical analysis. …
Asymptotic Results For Simultaneous Group Sequential Analysis Of Rank-Based And Weighted Kaplan-Meier Tests With Paired Survival Data In The Presence Of Censoring. Technical Report, Adin-Cristian Andrei, Susan Murray
Asymptotic Results For Simultaneous Group Sequential Analysis Of Rank-Based And Weighted Kaplan-Meier Tests With Paired Survival Data In The Presence Of Censoring. Technical Report, Adin-Cristian Andrei, Susan Murray
The University of Michigan Department of Biostatistics Working Paper Series
This research sequentially monitors paired survival differences using a new class of non-parametric tests based on functionals of standardized paired weighted log-rank (PWLR) and standardized paired weighted Kaplan-Meier (PWKM) tests. During a trial these tests may alternately assume the role of the more extreme statistic. By monitoring PEMAX, the maximum between the absolute values of the standardized PWLR and PWKM, one combines advantages of rank-based and non rank-based paired testing paradigms. Simulations show that monitoring treatment differences using PEMAX maintains type I error and is nearly as powerful as using the more advantageous of the two tests, in proportional hazards …
Bayesian Geostatistical Design, Peter J. Diggle, Soren Lophaven
Bayesian Geostatistical Design, Peter J. Diggle, Soren Lophaven
Johns Hopkins University, Dept. of Biostatistics Working Papers
This paper describes the use of model-based geostatistics for choosing the optimal set of sampling locations, collectively called the design, for a geostatistical analysis. Two types of design situations are considered. These are retrospective design, which concerns the addition of sampling locations to, or deletion of locations from, an existing design, and prospective design, which consists of choosing optimal positions for a new set of sampling locations. We propose a Bayesian design criterion which focuses on the goal of efficient spatial prediction whilst allowing for the fact that model parameter values are unknown. The results show that in this situation …
Nonparametric Methods For Analyzing Replication Origins In Genomewide Data, Debashis Ghosh
Nonparametric Methods For Analyzing Replication Origins In Genomewide Data, Debashis Ghosh
The University of Michigan Department of Biostatistics Working Paper Series
Due to the advent of high-throughput genomic technology, it has become possible to globally monitor cellular activities on a genomewide basis. With these new methods, scientists can begin to address important biological questions. One such question involves the identification of replication origins, which are regions in chromosomes where DNA replication is initiated. In addition, one hypothesis regarding replication origins is that their locations are non-random throughout the genome. In this article, we develop methods for identification of and cluster inference regarding replication origins involving genomewide expression data. We compare several nonparametric regression methods for the identification of replication origin locations. …
Mean Response Models Of Repeated Measurements In Presence Of Varying Effectiveness Onset, Ying Qing Chen, Su-Chun Cheng
Mean Response Models Of Repeated Measurements In Presence Of Varying Effectiveness Onset, Ying Qing Chen, Su-Chun Cheng
U.C. Berkeley Division of Biostatistics Working Paper Series
Repeated measurements are often collected over time to evaluate treatment efficacy in clinical trials. Most of the statistical models of the repeated measurements have been focusing on their mean response as function of time. These models usually assume that the treatment has persistent effect of constant additivity or multiplicity on the mean response functions throughout the observation period of time. In reality, however, such assumption may be confounded by the potential existence of the so-called effectiveness action onset, although they are often unobserved or difficult to obtain. Instead of including nonparametric time-varying coefficients in the mean response models, we propose …
Semiparametric Methods For Identification Of Tumor Progression Genes From Microarray Data, Debashis Ghosh, Arul Chinnaiyan
Semiparametric Methods For Identification Of Tumor Progression Genes From Microarray Data, Debashis Ghosh, Arul Chinnaiyan
The University of Michigan Department of Biostatistics Working Paper Series
The use of microarray data has become quite commonplace in medical and scientific experiments. We focus here on microarray data generated from cancer studies. It is potentially important for the discovery of biomarkers to identify genes whose expression levels correlate with tumor progression. In this article, we develop statistical procedures for the identification of such genes, which we term tumor progression genes. Two methods are considered in this paper. The first is use of a proportional odds procedure, combined with false discovery rate estimation techniques to adjust for the multiple testing problem. The second method is based on order-restricted estimation …
The False Discovery Rate: A Variable Selection Perspective, Debashis Ghosh, Wei Chen, Trivellore E. Raghuanthan
The False Discovery Rate: A Variable Selection Perspective, Debashis Ghosh, Wei Chen, Trivellore E. Raghuanthan
The University of Michigan Department of Biostatistics Working Paper Series
In many scientific and medical settings, large-scale experiments are generating large quantities of data that lead to inferential problems involving multiple hypotheses. This has led to recent tremendous interest in statistical methods regarding the false discovery rate (FDR). Several authors have studied the properties involving FDR in a univariate mixture model setting. In this article, we turn the problem on its side; in this manuscript, we show that FDR is a by-product of Bayesian analysis of variable selection problem for a hierarchical linear regression model. This equivalence gives many Bayesian insights as to why FDR is a natural quantity to …
Multiple Testing Methods For Chip-Chip High Density Oligonucleotide Array Data, Sunduz Keles, Mark J. Van Der Laan, Sandrine Dudoit, Simon E. Cawley
Multiple Testing Methods For Chip-Chip High Density Oligonucleotide Array Data, Sunduz Keles, Mark J. Van Der Laan, Sandrine Dudoit, Simon E. Cawley
U.C. Berkeley Division of Biostatistics Working Paper Series
Cawley et al. (2004) have recently mapped the locations of binding sites for three transcription factors along human chromosomes 21 and 22 using ChIP-Chip experiments. ChIP-Chip experiments are a new approach to the genome-wide identification of transcription factor binding sites and consist of chromatin (Ch) immunoprecipitation (IP) of transcription factor-bound genomic DNA followed by high density oligonucleotide hybridization (Chip) of the IP-enriched DNA. We investigate the ChIP-Chip data structure and propose methods for inferring the location of transcription factor binding sites from these data. The proposed methods involve testing for each probe whether it is part of a bound sequence …
Combining Predictors For Classification Using The Area Under The Roc Curve, Margaret S. Pepe, Tianxi Cai, Zheng Zhang
Combining Predictors For Classification Using The Area Under The Roc Curve, Margaret S. Pepe, Tianxi Cai, Zheng Zhang
UW Biostatistics Working Paper Series
We compare simple logistic regression with an alternative robust procedure for constructing linear predictors to be used for the two state classification task. Theoritical advantages of the robust procedure over logistic regression are: (i) although it assumes a generalized linear model for the dichotomous outcome variable, it does not require specification of the link function; (ii) it accommodates case-control designs even when the model is not logistic; and (iii) it yields sensible results even when the generalized linear model assumption fails to hold. Surprisingly, we find that the linear predictor derived from the logistic regression likelihood is very robust in …