Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- Universitas Indonesia (315)
- University of Kentucky (43)
- COBRA (26)
- Virginia Commonwealth University (18)
- University of South Carolina (13)
-
- SIT Graduate Institute/SIT Study Abroad (12)
- University of Louisville (12)
- Loma Linda University (11)
- The Texas Medical Center Library (11)
- Himmelfarb Health Sciences Library, The George Washington University (10)
- Old Dominion University (10)
- Dartmouth College (9)
- University of Arkansas, Fayetteville (8)
- University of Texas at El Paso (8)
- City University of New York (CUNY) (7)
- University of Nebraska - Lincoln (7)
- Illinois State University (6)
- Michigan Technological University (6)
- University at Albany, State University of New York (6)
- University of Nevada, Las Vegas (5)
- Walden University (5)
- West Virginia University (5)
- James Madison University (4)
- Purdue University (4)
- East Tennessee State University (3)
- LSU Health New Orleans (3)
- Liberty University (3)
- University of South Alabama (3)
- University of South Florida (3)
- Wayne State University (3)
- Keyword
-
- COVID-19 (22)
- Genetics (17)
- Knowledge (17)
- Humans (15)
- Indonesia (15)
-
- Machine learning (15)
- Obesity (12)
- Stunting (12)
- Adolescent (11)
- Bioinformatics (10)
- Physical activity (10)
- Female (9)
- Overweight (9)
- Behavior (8)
- Children (8)
- Diabetes mellitus (8)
- Exclusive breastfeeding (8)
- Algorithms (7)
- Attitude (7)
- Education (7)
- Elderly (7)
- GWAS (7)
- Gene expression (7)
- Genomics (7)
- Hypertension (7)
- Male (7)
- Pregnancy (7)
- Smoking (7)
- Biomarkers (6)
- Diabetes (6)
- Publication Year
- Publication
-
- Kesmas (315)
- Electronic Theses and Dissertations (15)
- Biostatistics Faculty Publications (14)
- Theses and Dissertations (14)
- Independent Study Project (ISP) Collection (12)
-
- Dissertations and Theses (Open Access) (11)
- Faculty Publications (11)
- Loma Linda University Electronic Theses, Dissertations & Projects (11)
- Graduate Theses and Dissertations (8)
- Open Access Theses & Dissertations (8)
- Epidemiology Faculty Publications (7)
- Harvard University Biostatistics Working Paper Series (7)
- Biology and Medicine Through Mathematics Conference (6)
- COBRA Preprint Series (6)
- Dissertations, Master's Theses and Master's Reports (6)
- Legacy Theses & Dissertations (2009 - 2024) (6)
- Annual Symposium on Biomathematics and Ecology Education and Research (5)
- Dartmouth Scholarship (5)
- Epidemiology and Environmental Health Faculty Publications (5)
- Faculty & Staff Scholarship (5)
- U.C. Berkeley Division of Biostatistics Working Paper Series (5)
- Walden Dissertations and Doctoral Studies (5)
- Dartmouth College Ph.D Dissertations (4)
- Open Access Dissertations (4)
- UW Biostatistics Working Paper Series (4)
- Internal Medicine Faculty Publications (3)
- Master's Theses (3)
- OES Theses and Dissertations (3)
- Publications and Research (3)
- Research Day (3)
- Publication Type
- File Type
Articles 601 - 630 of 644
Full-Text Articles in Biostatistics
Why Odds Ratio Estimates Of Gwas Are Almost Always Close To 1.0, Yutaka Yasui
Why Odds Ratio Estimates Of Gwas Are Almost Always Close To 1.0, Yutaka Yasui
COBRA Preprint Series
“Missing heritability” in genome-wide association studies (GWAS) refers to the seeming inability for GWAS data to capture the great majority of genetic causes of a disease in comparison to the known degree of heritability for the disease, in spite of GWAS’ genome-wide measures of genetic variations. This paper presents a simple mathematical explanation for this phenomenon, assuming that the heritability information exists in GWAS data. Specifically, it focuses on the fact that the great majority of association measures (in the form of odds ratios) from GWAS are consistently close to the value that indicates no association, explains why this occurs, …
Generalized Linear Latent Mixed Modeling Of Functional Independent Measures And Patient Outcomes, Maduranga Kasun Dassanayake
Generalized Linear Latent Mixed Modeling Of Functional Independent Measures And Patient Outcomes, Maduranga Kasun Dassanayake
Open Access Theses & Dissertations
The Functional Independent Measure (FIM) is one of the most widely accepted functional assessment measures used in the rehabilitation community. Past research studies have investigated the relationship between place of discharge, admission FIM scores or FIM difference scores, and patients' characteristics and found relationships between those variables. However, most of these studies fail to account for the multi-layered multidimensionality of the FIM and the measurement error associated with the FIM items. This study utilizes Generalized Linear Latent Mixed Models (GLLAMM) and Structural Equation Models (SEM) to assess which patient characteristics are associated with FIM difference scores and the structural relationship …
Association Between Chemical Constituents Of Particulate Matter And Cardiovascular And Respiratory Morbidities In Nys, Rena Jones
Legacy Theses & Dissertations (2009 - 2024)
Improved understanding of health risks from short- and long-term exposure to fine particulate matter (PM2.5) constituents may explain seasonal and geographic heterogeneity in PM2.5-health associations and inform control efforts targeting PM sources. Few studies have examined PM species health effects; most have been limited by their exposure assessments and modeling approaches. The goals of this project were to improve the PM exposure assessment and explore relationships between PM2.5 species and health in acute and chronic contexts.
Secondary Structure Prediction Of Long Rna Sequences Based On Inversion Excursions And A Modularized Mapreduce Framework, Daniel Tesfai Yehdego
Secondary Structure Prediction Of Long Rna Sequences Based On Inversion Excursions And A Modularized Mapreduce Framework, Daniel Tesfai Yehdego
Open Access Theses & Dissertations
Ribonucleic acid (RNA) molecules and their secondary structures play important roles in many biological processes including gene expression and regulation. The genomes of many viruses are also RNA molecules. Since secondary structures are crucial for RNA functionality, computational predictions of the RNA secondary structures have been widely studied. However, the tremendous demands on computer memory and computing time for complex secondary structures limit the capability of existing thermodynamically based algorithms for structure predictions to handling only short RNA sequences with a few hundred bases. One approach to overcome this limitation is by first cutting long RNA sequences into shorter, non-overlapping …
Estimation Of A Non-Parametric Variable Importance Measure Of A Continuous Exposure, Chambaz Antoine, Pierre Neuvial, Mark J. Van Der Laan
Estimation Of A Non-Parametric Variable Importance Measure Of A Continuous Exposure, Chambaz Antoine, Pierre Neuvial, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
We define a new measure of variable importance of an exposure on a continuous outcome, accounting for potential confounders. The exposure features a reference level x0 with positive mass and a continuum of other levels. For the purpose of estimating it, we fully develop the semi-parametric estimation methodology called targeted minimum loss estimation methodology (TMLE) [van der Laan & Rubin, 2006; van der Laan & Rose, 2011]. We cover the whole spectrum of its theoretical study (convergence of the iterative procedure which is at the core of the TMLE methodology; consistency and asymptotic normality of the estimator), practical implementation, simulation …
Gene By Bmi Interactions Influencing C-Reactive Protein Levels In European-Americans, Sarah Tudor
Gene By Bmi Interactions Influencing C-Reactive Protein Levels In European-Americans, Sarah Tudor
Dissertations and Theses (Open Access)
C-Reactive Protein (CRP) is a biomarker indicating tissue damage, inflammation, and infection. High-sensitivity CRP (hsCRP) is an emerging biomarker often used to estimate an individual’s risk for future coronary heart disease (CHD). hsCRP levels falling below 1.00 mg/l indicate a low risk for developing CHD, levels ranging between 1.00 mg/l and 3.00 mg/l indicate an elevated risk, and levels exceeding 3.00 mg/l indicate high risk. Multiple Genome-Wide Association Studies (GWAS) have identified a number of genetic polymorphisms which influence CRP levels. SNPs implicated in such studies have been found in or near genes of interest including: CRP, APOE, APOC, IL-6, …
A Novel Device For Cell-Cell Electrofusion, Justin T. Stewart
A Novel Device For Cell-Cell Electrofusion, Justin T. Stewart
USF Tampa Graduate Theses and Dissertations
Cell transplantation therapy is a potentially powerful tool and can be used to replace defective cells with healthy cells. This offers the possibility of alleviating the destructive symptoms for many diseases such as Parkinson's disease, Alzheimer's disease, stroke, spinal cord trauma, Type I diabetes and many more. While there are many diseases that could be positively impacted from cell transplantation therapy, the focus of this research is insulin dependent, Type I Diabetes.
The Islets of Langerhans are composed of various types of cells located in the pancreas and are responsible for a variety of biochemical functions. Specifically, the beta Islet …
Analysis Of Differential Gene Expression And Alternative Splicing In The Liver And Gastrointestinal Tract In The Lactating Rat, Antony Thomas Athippozhy
Analysis Of Differential Gene Expression And Alternative Splicing In The Liver And Gastrointestinal Tract In The Lactating Rat, Antony Thomas Athippozhy
University of Kentucky Doctoral Dissertations
Rat exon microarrays were utilized to detect changes in mRNA expression and alternative splicing in the liver, duodenum, jejunum, and ileum of the lactating rat when compared to age-matched virgin controls. Analysis of data at the level of gene expression revealed differential expression of genes involved in cholesterol biosynthesis in each tissue examined, suggesting increased Sterol Response Element Binding Protein activity. We also detected decreased mRNA from components of the T-cell signaling pathway in the jejunum and ileum. We characterized expression of solute carrier and adenosine triphosphate binding cassette proteins. In addition to characterizing genes by pathway, we have also …
Minimum Description Length Measures Of Evidence For Enrichment, Zhenyu Yang, David R. Bickel
Minimum Description Length Measures Of Evidence For Enrichment, Zhenyu Yang, David R. Bickel
COBRA Preprint Series
In order to functionally interpret differentially expressed genes or other discovered features, researchers seek to detect enrichment in the form of overrepresentation of discovered features associated with a biological process. Most enrichment methods treat the p-value as the measure of evidence using a statistical test such as the binomial test, Fisher's exact test or the hypergeometric test. However, the p-value is not interpretable as a measure of evidence apart from adjustments in light of the sample size. As a measure of evidence supporting one hypothesis over the other, the Bayes factor (BF) overcomes this drawback of the p-value but lacks …
Powerful Snp Set Analysis For Case-Control Genome Wide Association Studies, Michael C. Wu, Peter Kraft, Michael P. Epstein, Deanne M. Taylor, Stephen J. Chanock, David J. Hunter, Xihong Lin
Powerful Snp Set Analysis For Case-Control Genome Wide Association Studies, Michael C. Wu, Peter Kraft, Michael P. Epstein, Deanne M. Taylor, Stephen J. Chanock, David J. Hunter, Xihong Lin
Harvard University Biostatistics Working Paper Series
No abstract provided.
Survival Prediction For Brain Tumor Patients Using Gene Expression Data, Vinicius Bonato
Survival Prediction For Brain Tumor Patients Using Gene Expression Data, Vinicius Bonato
Dissertations and Theses (Open Access)
Brain tumor is one of the most aggressive types of cancer in humans, with an estimated median survival time of 12 months and only 4% of the patients surviving more than 5 years after disease diagnosis. Until recently, brain tumor prognosis has been based only on clinical information such as tumor grade and patient age, but there are reports indicating that molecular profiling of gliomas can reveal subgroups of patients with distinct survival rates. We hypothesize that coupling molecular profiling of brain tumors with clinical information might improve predictions of patient survival time and, consequently, better guide future treatment decisions. …
Anthropometric Parameters Of Under-Five Years Old Children With Different Dietary Habits In Ukambani Region : A Study In Eastern Rural Kenya, Hellen M. Ndiku
Anthropometric Parameters Of Under-Five Years Old Children With Different Dietary Habits In Ukambani Region : A Study In Eastern Rural Kenya, Hellen M. Ndiku
Loma Linda University Electronic Theses, Dissertations & Projects
The objective of this descriptive cross sectional study was to assess dietary intake and nutritional status of children under-five years in two rural sites of Eastern Kenya where the staple cereals may differ. A modified rapid, knowledge, practice and coverage (KPC) questionnaire and a 24-hr dietary recall form were used to collect the data. A total of 403 households were surveyed from four randomly selected divisions. This yielded 629 surrogate 24-hr dietary recalls of children < 5 years with 314 from Mwingi district and 315 from Makueni district (49 % boys and 51 % girls).
Statistical analysis was done using SPSS and SAS. Comparison of means was done using t- test and chi square was used for proportions. The 24-hr …
Joint Multiple Testing Procedures For Graphical Model Selection With Applications To Biological Networks, Houston N. Gilbert, Mark J. Van Der Laan, Sandrine Dudoit
Joint Multiple Testing Procedures For Graphical Model Selection With Applications To Biological Networks, Houston N. Gilbert, Mark J. Van Der Laan, Sandrine Dudoit
U.C. Berkeley Division of Biostatistics Working Paper Series
Gaussian graphical models have become popular tools for identifying relationships between genes when analyzing microarray expression data. In the classical undirected Gaussian graphical model setting, conditional independence relationships can be inferred from partial correlations obtained from the concentration matrix (= inverse covariance matrix) when the sample size n exceeds the number of parameters p which need to estimated. In situations where n < p, another approach to graphical model estimation may rely on calculating unconditional (zero-order) and first-order partial correlations. In these settings, the goal is to identify a lower-order conditional independence graph, sometimes referred to as a ‘0-1 graphs’. For either choice of graph, model selection may involve a multiple testing problem, in which edges in a graph are drawn only after rejecting hypotheses involving (saturated or lower-order) partial correlation parameters. Most multiple testing procedures applied in previously proposed graphical model selection algorithms rely on standard, marginal testing methods which do not take into account the joint distribution of the test statistics derived from (partial) correlations. We propose and implement a multiple testing framework useful when testing for edge inclusion during graphical model selection. Two features of our methodology include (i) a computationally efficient and asymptotically valid test statistics joint null distribution derived from influence curves for correlation-based parameters, and (ii) the application of empirical Bayes joint multiple testing procedures which can effectively control a variety of popular Type I error rates by incorpo- rating joint null distributions such as those described here (Dudoit and van der Laan, 2008). Using a dataset from Arabidopsis thaliana, we observe that the use of more sophisticated, modular approaches to multiple testing allows one to identify greater numbers of edges when approximating an undirected graphical model using a 0-1 graph. Our framework may also be extended to edge testing algorithms for other types of graphical models (e.g., for classical undirected, bidirected, and directed acyclic graphs).
Estimation And Testing For The Effect Of A Genetic Pathway On A Disease Outcome Using Logistic Kernel Machine Regression Via Logistic Mixed Models, Dawei Liu, Debashis Ghosh, Xihong Lin
Estimation And Testing For The Effect Of A Genetic Pathway On A Disease Outcome Using Logistic Kernel Machine Regression Via Logistic Mixed Models, Dawei Liu, Debashis Ghosh, Xihong Lin
Harvard University Biostatistics Working Paper Series
No abstract provided.
A Powerful And Flexible Multilocus Association Test For Quantitative Traits, Lydia Coulter Kwee, Dawei Liu, Xihong Lin, Debashis Ghosh, Michael P. Epstein
A Powerful And Flexible Multilocus Association Test For Quantitative Traits, Lydia Coulter Kwee, Dawei Liu, Xihong Lin, Debashis Ghosh, Michael P. Epstein
Harvard University Biostatistics Working Paper Series
No abstract provided.
The Expression Of Microrna Mir-107 Decreases Early In Alzheimer's Disease And May Accelerate Disease Progression Through Regulation Of Β-Site Amyloid Precursor Protein-Cleaving Enzyme 1, Wang-Xia Wang, Bernard W. Rajeev, Arnold J. Stromberg, Na Ren, Guiliang Tang, Qingwei Huang, Isidore Rigoutsos, Peter T. Nelson
The Expression Of Microrna Mir-107 Decreases Early In Alzheimer's Disease And May Accelerate Disease Progression Through Regulation Of Β-Site Amyloid Precursor Protein-Cleaving Enzyme 1, Wang-Xia Wang, Bernard W. Rajeev, Arnold J. Stromberg, Na Ren, Guiliang Tang, Qingwei Huang, Isidore Rigoutsos, Peter T. Nelson
Sanders-Brown Center on Aging Faculty Publications
MicroRNAs (miRNAs) are small regulatory RNAs that participate in posttranscriptional gene regulation in a sequence-specific manner. However, little is understood about the role(s) of miRNAs in Alzheimer's disease (AD). We used miRNA expression microarrays on RNA extracted from human brain tissue from the University of Kentucky Alzheimer's Disease Center Brain Bank with near-optimal clinicopathological correlation. Cases were separated into four groups: elderly nondemented with negligible AD-type pathology, nondemented with incipient AD pathology, mild cognitive impairment (MCI) with moderate AD pathology, and AD. Among the AD-related miRNA expression changes, miR-107 was exceptional because miR-107 levels decreased significantly even in patients with …
Assessment Of A Cgh-Based Genetic Instability, David A. Engler, Yiping Shen, J F. Gusella, Rebecca A. Betensky
Assessment Of A Cgh-Based Genetic Instability, David A. Engler, Yiping Shen, J F. Gusella, Rebecca A. Betensky
Harvard University Biostatistics Working Paper Series
No abstract provided.
Survival Analysis With Large Dimensional Covariates: An Application In Microarray Studies, David A. Engler, Yi Li
Survival Analysis With Large Dimensional Covariates: An Application In Microarray Studies, David A. Engler, Yi Li
Harvard University Biostatistics Working Paper Series
Use of microarray technology often leads to high-dimensional and low- sample size data settings. Over the past several years, a variety of novel approaches have been proposed for variable selection in this context. However, only a small number of these have been adapted for time-to-event data where censoring is present. Among standard variable selection methods shown both to have good predictive accuracy and to be computationally efficient is the elastic net penalization approach. In this paper, adaptation of the elastic net approach is presented for variable selection both under the Cox proportional hazards model and under an accelerated failure time …
Is A Basketball Free-Throw Sequence Nonrandom? A Group Exercise For Undergraduate Statistics Students, Stephen C. Adolph
Is A Basketball Free-Throw Sequence Nonrandom? A Group Exercise For Undergraduate Statistics Students, Stephen C. Adolph
All HMC Faculty Publications and Research
I describe a group exercise that I give to my undergraduate biostatistics class. The exercise involves analyzing a series of 200 consecutive basketball free-throw attempts to determine whether there is any evidence for sequential dependence in the probability of making a free-throw. The students are given the exercise before they have learned the appropriate statistical tests, so that they can come up with ideas on their own. Students spend a full class period working on the problem, with my guidance and hints. In the next class period, we discuss how each student group approached the problem. I then present several …
Semiparametric Regression Of Multi-Dimensional Genetic Pathway Data: Least Squares Kernel Machines And Linear Mixed Models, Dawei Liu, Xihong Lin, Debashis Ghosh
Semiparametric Regression Of Multi-Dimensional Genetic Pathway Data: Least Squares Kernel Machines And Linear Mixed Models, Dawei Liu, Xihong Lin, Debashis Ghosh
Harvard University Biostatistics Working Paper Series
No abstract provided.
Bounded Search For De Novo Identification Of Degenerate Cis-Regulatory Elements, Jonathan M. Carlson, Arijit Chakravarty, Radhika S. Khetani, Robert H. Gross
Bounded Search For De Novo Identification Of Degenerate Cis-Regulatory Elements, Jonathan M. Carlson, Arijit Chakravarty, Radhika S. Khetani, Robert H. Gross
Dartmouth Scholarship
The identification of statistically overrepresented sequences in the upstream regions of coregulated genes should theoretically permit the identification of potential cis-regulatory elements. However, in practice many cis-regulatory elements are highly degenerate, precluding the use of an exhaustive word-counting strategy for their identification. While numerous methods exist for inferring base distributions using a position weight matrix, recent studies suggest that the independence assumptions inherent in the model, as well as the inability to reach a global optimum, limit this approach.
Genome Scanning Methods For Comparing Sequences Between Groups, With Application To Hiv Vaccine Trials, Peter B. Gilbert, Chunyuan Wu, David V. Jobes
Genome Scanning Methods For Comparing Sequences Between Groups, With Application To Hiv Vaccine Trials, Peter B. Gilbert, Chunyuan Wu, David V. Jobes
UW Biostatistics Working Paper Series
Consider a placebo-controlled preventive HIV vaccine efficacy trial. An HIV amino acid sequence is measured from each volunteer who acquires HIV, and these sequences are aligned together with the reference HIV sequence represented in the vaccine. We develop genome scanning methods to identify HIV positions at which the amino acids in sequences from infected vaccine recipients tend to be more divergent from the corresponding reference amino acid than the amino acids in sequences from infected placebo recipients. We consider five two-sample test statistics, based on Euclidean, Mahalanobis, and Kullback-Leibler divergence measures. Weights are incorporated to reflect biological information contained in …
Multiple Tests Of Association With Biological Annotation Metadata, Sandrine Dudoit, Sunduz Keles, Mark J. Van Der Laan
Multiple Tests Of Association With Biological Annotation Metadata, Sandrine Dudoit, Sunduz Keles, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
We propose a general and formal statistical framework for the multiple tests of associations between known fixed features of a genome and unknown parameters of the distribution of variable features of this genome in a population of interest. The known fixed gene-annotation profiles, corresponding to the fixed features of the genome, may concern Gene Ontology (GO) annotation, pathway membership, regulation by particular transcription factors, nucleotide sequences, or protein sequences. The unknown gene-parameter profiles, corresponding to the variable features of the genome, may be, for example, regression coefficients relating genome-wide transcript levels or DNA copy numbers to possibly censored biological and …
Gpnn: Power Studies And Applications Of A Neural Network Method For Detecting Gene-Gene Interactions In Studies Of Human Disease, Alison A. Motsinger, Stephen L. Lee, George Mellick, Marylyn D. Ritchie
Gpnn: Power Studies And Applications Of A Neural Network Method For Detecting Gene-Gene Interactions In Studies Of Human Disease, Alison A. Motsinger, Stephen L. Lee, George Mellick, Marylyn D. Ritchie
Dartmouth Scholarship
The identification and characterization of genes that influence the risk of common, complex multifactorial disease primarily through interactions with other genes and environmental factors remains a statistical and computational challenge in genetic epidemiology. We have previously introduced a genetic programming optimized neural network (GPNN) as a method for optimizing the architecture of a neural network to improve the identification of gene combinations associated with disease risk. The goal of this study was to evaluate the power of GPNN for identifying high-order gene-gene interactions. We were also interested in applying GPNN to a real data analysis in Parkinson's disease.
Principal Component Analysis For Predicting Transcription-Factor Binding Motifs From Array-Derived Data, Yunlong Liu, Matthew P Vincenti, Hiroki Yokota
Principal Component Analysis For Predicting Transcription-Factor Binding Motifs From Array-Derived Data, Yunlong Liu, Matthew P Vincenti, Hiroki Yokota
Dartmouth Scholarship
The responses to interleukin 1 (IL-1) in human chondrocytes constitute a complex regulatory mechanism, where multiple transcription factors interact combinatorially to transcription-factor binding motifs (TFBMs). In order to select a critical set of TFBMs from genomic DNA information and an array-derived data, an efficient algorithm to solve a combinatorial optimization problem is required. Although computational approaches based on evolutionary algorithms are commonly employed, an analytical algorithm would be useful to predict TFBMs at nearly no computational cost and evaluate varying modelling conditions. Singular value decomposition (SVD) is a powerful method to derive primary components of a given matrix. Applying SVD …
A Pseudolikelihood Approach For Simultaneous Analysis Of Array Comparative Genomic Hybridizations (Acgh), David A. Engler, Gayatry Mohapatra, David N. Louis, Rebecca Betensky
A Pseudolikelihood Approach For Simultaneous Analysis Of Array Comparative Genomic Hybridizations (Acgh), David A. Engler, Gayatry Mohapatra, David N. Louis, Rebecca Betensky
Harvard University Biostatistics Working Paper Series
DNA sequence copy number has been shown to be associated with cancer development and progression. Array-based Comparative Genomic Hybridization (aCGH) is a recent development that seeks to identify the copy number ratio at large numbers of markers across the genome. Due to experimental and biological variations across chromosomes and across hybridizations, current methods are limited to analyses of single chromosomes. We propose a more powerful approach that borrows strength across chromosomes and across hybridizations. We assume a Gaussian mixture model, with a hidden Markov dependence structure, and with random effects to allow for intertumoral variation, as well as intratumoral clonal …
Application Of A Multiple Testing Procedure Controlling The Proportion Of False Positives To Protein And Bacterial Data, Merrill D. Birkner, Alan E. Hubbard, Mark J. Van Der Laan
Application Of A Multiple Testing Procedure Controlling The Proportion Of False Positives To Protein And Bacterial Data, Merrill D. Birkner, Alan E. Hubbard, Mark J. Van Der Laan
U.C. Berkeley Division of Biostatistics Working Paper Series
Simultaneously testing multiple hypotheses is important in high-dimensional biological studies. In these situations, one is often interested in controlling the Type-I error rate, such as the proportion of false positives to total rejections (TPPFP) at a specific level, alpha. This article will present an application of the E-Bayes/Bootstrap TPPFP procedure, presented in van der Laan et al. (2005), which controls the tail probability of the proportion of false positives (TPPFP), on two biological datasets. The two data applications include firstly, the application to a mass-spectrometry dataset of two leukemia subtypes, AML and ALL. The protein data measurements include intensity and …
Test Statistics Null Distributions In Multiple Testing: Simulation Studies And Applications To Genomics, Katherine S. Pollard, Merrill D. Birkner, Mark J. Van Der Laan, Sandrine Dudoit
Test Statistics Null Distributions In Multiple Testing: Simulation Studies And Applications To Genomics, Katherine S. Pollard, Merrill D. Birkner, Mark J. Van Der Laan, Sandrine Dudoit
U.C. Berkeley Division of Biostatistics Working Paper Series
Multiple hypothesis testing problems arise frequently in biomedical and genomic research, for instance, when identifying differentially expressed or co-expressed genes in microarray experiments. We have developed generally applicable resampling-based single-step and stepwise multiple testing procedures (MTP) for control of a broad class of Type I error rates, defined as tail probabilities and expected values for arbitrary functions of the numbers of false positives and rejected hypotheses (Dudoit and van der Laan, 2005; Dudoit et al., 2004a,b; Pollard and van der Laan, 2004; van der Laan et al., 2005, 2004a,b). As argued in the early article of Pollard and van der …
New Statistical Paradigms Leading To Web-Based Tools For Clinical/Translational Science, Knut M. Wittkowski
New Statistical Paradigms Leading To Web-Based Tools For Clinical/Translational Science, Knut M. Wittkowski
COBRA Preprint Series
As the field of functional genetics and genomics is beginning to mature, we become confronted with new challenges. The constant drop in price for sequencing and gene expression profiling as well as the increasing number of genetic and genomic variables that can be measured makes it feasible to address more complex questions. The success with rare diseases caused by single loci or genes has provided us with a proof-of-concept that new therapies can be developed based on functional genomics and genetics.
Common diseases, however, typically involve genetic epistasis, genomic pathways, and proteomic pattern. Moreover, to better understand the underlying biologi-cal …
Geographic Variation In The Morphology Of Crotalus Horridus (Serpentes: Viperidae), John Robert Allsteadt
Geographic Variation In The Morphology Of Crotalus Horridus (Serpentes: Viperidae), John Robert Allsteadt
Biological Sciences Theses & Dissertations
The Timber Rattlesnake (Crotalus horridus) occurs in discontinuous populations throughout the eastern and central United States. The species exhibits high levels of polymorphism in morphological traits, especially in coloration and pattern. Previous studies recognized either distinct northern and southern subspecies or three regional morphs (northern, southern, and western), but conflicting data sets and limited geographic sampling of previous studies have left the relationships among those regional variants unclear. In this study, univariate and multivariate statistics, together with a geographic information system, were used to analyze geographic variation in 36 morphological characters recorded from 2,420 specimens of C. horridus …