Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- The Texas Medical Center Library (131)
- COBRA (113)
- Augustana College (42)
- University of Kentucky (40)
- University of Nebraska - Lincoln (35)
-
- Dartmouth College (26)
- Virginia Commonwealth University (22)
- University of Nebraska at Omaha (19)
- City University of New York (CUNY) (18)
- University of Connecticut (16)
- Louisiana State University (14)
- University of New Mexico (14)
- Michigan Technological University (12)
- University of Arkansas, Fayetteville (12)
- University of Louisville (12)
- Dordt University (11)
- University of Nebraska Medical Center (11)
- Kennesaw State University (10)
- Loyola University Chicago (10)
- The University of Southern Mississippi (10)
- Western University (10)
- California Polytechnic State University, San Luis Obispo (9)
- Munster Technological University (9)
- West Virginia University (9)
- Clemson University (8)
- Illinois State University (8)
- University of Missouri, St. Louis (8)
- Wayne State University (8)
- Mississippi State University (7)
- University of Montana (7)
- Keyword
-
- Bioinformatics (132)
- Humans (58)
- Genetics (55)
- Genome (49)
- Genomics (46)
-
- Annotation (38)
- Meiothermus ruber (38)
- Gene expression (26)
- GENI-ACT (25)
- Machine learning (20)
- Evolution (19)
- Transcriptome (19)
- Cancer (17)
- Animals (15)
- Algorithms (14)
- Computational biology (14)
- Genome-Wide Association Study (14)
- Phylogenetics (14)
- Population genetics (14)
- Cancer genomics (13)
- Female (13)
- Polymorphism, Single Nucleotide (13)
- Transcriptomics (13)
- Biology (11)
- Epigenetics (11)
- Male (11)
- Polymorphism (11)
- RNA-seq (11)
- Biomarkers (10)
- Computational Biology (10)
- Publication Year
- Publication
-
- Faculty, Staff and Student Publications (72)
- Dissertations and Theses (Open Access) (59)
- Meiothermus ruber Genome Analysis Project (40)
- Theses and Dissertations (31)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (27)
-
- Harvard University Biostatistics Working Paper Series (24)
- Electronic Theses and Dissertations (19)
- COBRA Preprint Series (17)
- Dartmouth Scholarship (16)
- Dissertations, Theses, and Capstone Projects (13)
- Master's Theses (13)
- Graduate Theses and Dissertations (12)
- UW Biostatistics Working Paper Series (12)
- Dissertations, Master's Theses and Master's Reports (11)
- Faculty Work Comprehensive List (11)
- Honors Scholar Theses (11)
- LSU Doctoral Dissertations (11)
- Theses & Dissertations (11)
- Bioconductor Project Working Papers (10)
- Interdisciplinary Informatics Faculty Publications (10)
- UPenn Biostatistics Working Papers (10)
- Biochemistry Publications (9)
- Dartmouth College Ph.D Dissertations (9)
- Bioinformatics Faculty Publications (8)
- Biology ETDs (8)
- Graduate Theses, Dissertations, and Problem Reports (ETD) (8)
- Theses and Dissertations--Biology (8)
- Annual Symposium on Biomathematics and Ecology Education and Research (7)
- Dissertations (7)
- Theses (7)
- Publication Type
- File Type
Articles 841 - 870 of 904
Full-Text Articles in Genetics and Genomics
Group Scad Regression Analysis For Microarray Time Course Gene Expression Data, Lifeng Wang, Guang Chen, Hongzhe Li Phd
Group Scad Regression Analysis For Microarray Time Course Gene Expression Data, Lifeng Wang, Guang Chen, Hongzhe Li Phd
UPenn Biostatistics Working Papers
Since many important biological systems or processes are dynamic systems, it is important to study the gene expression patterns over time in a genomic scale in order to capture the dynamic behavior of gene expression. Microarray technologies have made it possible to measure the gene expression levels of essentially all the genes during a given biological process. In order to determine the transcriptional factors involved in gene regulation during a given biological process, we propose to develop a functional response model with varying coefficients in order to model the transcriptional effects on gene expression levels and to develop a group …
Trab: Testing Whether Mutation Frequencies Are Above An Unknown Background, Giovanni Parmigiani, Sining Chen, Victor E. Velculescu
Trab: Testing Whether Mutation Frequencies Are Above An Unknown Background, Giovanni Parmigiani, Sining Chen, Victor E. Velculescu
Johns Hopkins University, Dept. of Biostatistics Working Papers
To rigorously determine whether a gene or a population of genes have alterations that are involved in carcinogenesis requires comparison of the prevalence of identified changes to the background mutation frequency present in tumor DNA. To facilitate this task, we develop a testing approach and the associated R library, called TRAB, that evaluates whether the frequency of somatic mutation is higher than an unknown, but estimable, background. We test the null hypothesis that the frequency belongs to background population of frequencies against the alternative hypothesis that the frequency is higher. Background mutation frequencies are themselves allowed to be variable. TRAB …
Optimized Cross-Study Analysis Of Microarray-Based Predictors, Xiaogang Zhong, Luigi Marchionni, Leslie Cope, Edwin S. Iversen, Elizabeth S. Garrett-Mayer, Edward Gabrielson, Giovanni Parmigiani
Optimized Cross-Study Analysis Of Microarray-Based Predictors, Xiaogang Zhong, Luigi Marchionni, Leslie Cope, Edwin S. Iversen, Elizabeth S. Garrett-Mayer, Edward Gabrielson, Giovanni Parmigiani
Johns Hopkins University, Dept. of Biostatistics Working Papers
Background: Microarray-based gene expression analysis is widely used in cancer research to discover molecular signatures for cancer classification and prediction. In addition to numerous independent profiling projects, a number of investigators have analyzed multiple published data sets for purposes of cross-study validation. However, the diverse microarray platforms and technical approaches make direct comparisons across studies difficult, and without means to identify aberrant data patterns, less than optimal. To address this issue, we previously developed an integrative correlation approach to systematically address agreement of gene expression measurements across studies, providing a basis for cross-study validation analysis. Here we generalize this methodology …
Improving Gsea For Analysis Of Biologic Pathways For Differential Gene Expression Across A Binary Phenotype , Irina Dinu, John D. Potter, Thomas Mueller, Qi Liu, Adeniyi J. Adewale, Gian S. Jhangri, Gunilla Einecke, Konrad S. Famulski, Philip Halloran, Yutaka Yasui
Improving Gsea For Analysis Of Biologic Pathways For Differential Gene Expression Across A Binary Phenotype , Irina Dinu, John D. Potter, Thomas Mueller, Qi Liu, Adeniyi J. Adewale, Gian S. Jhangri, Gunilla Einecke, Konrad S. Famulski, Philip Halloran, Yutaka Yasui
COBRA Preprint Series
Gene-set analysis evaluates the expression of biological pathways, or a priori defined gene sets, rather than that of single genes, in association with a binary phenotype, and is of great biologic interest in many DNA microarray studies. Gene Set Enrichment Analysis (GSEA) has been applied widely as a tool for gene-set analyses. We describe here some critical problems with GSEA and propose an alternative method by extending the single-gene analysis method, Significance Analysis of Microarray (SAM), to gene-set analyses (SAM-GS). Specifically, we illustrate, in a simulation study, that GSEA gives statistical significance to gene sets that have no gene associated …
The Plant Structure Ontology, A Unified Vocabulary Of Anatomy And Morphology Of A Flowering Plant, Katica Ilic, Elizabeth Kellogg, Pankaj Jaiswal, Felipe Zapata, Peter Stevens, Leszek Vincent, Shulamit Avraham, Leonore Reiser, Anuradha Pujar, Martin Sachs, Noah Whitman, Susan Mccouch, Mary Schaeffer, Doreen Ware, Lincoln Stein, Seung Rhee
The Plant Structure Ontology, A Unified Vocabulary Of Anatomy And Morphology Of A Flowering Plant, Katica Ilic, Elizabeth Kellogg, Pankaj Jaiswal, Felipe Zapata, Peter Stevens, Leszek Vincent, Shulamit Avraham, Leonore Reiser, Anuradha Pujar, Martin Sachs, Noah Whitman, Susan Mccouch, Mary Schaeffer, Doreen Ware, Lincoln Stein, Seung Rhee
Biology Department Faculty Works
Formal description of plant phenotypes and standardized annotation of gene expression and protein localization data require uniform terminology that accurately describes plant anatomy and morphology. This facilitates cross species comparative studies and quantitative comparison of phenotypes and expression patterns. A major drawback is variable terminology that is used to describe plant anatomy and morphology in publications and genomic databases for different species. The same terms are sometimes applied to different plant structures in different taxonomic groups. Conversely, similar structures are named by their species-specific terms. To address this problem, we created the Plant Structure Ontology (PSO), the first generic ontological …
Semiparametric Regression Of Multi-Dimensional Genetic Pathway Data: Least Squares Kernel Machines And Linear Mixed Models, Dawei Liu, Xihong Lin, Debashis Ghosh
Semiparametric Regression Of Multi-Dimensional Genetic Pathway Data: Least Squares Kernel Machines And Linear Mixed Models, Dawei Liu, Xihong Lin, Debashis Ghosh
Harvard University Biostatistics Working Paper Series
No abstract provided.
Penalized Likelihood And Bayesian Methods For Sparse Contingency Tables: An Analysis Of Alternative Splicing In Full-Length Cdna Libraries, Corinne Dahinden, Giovanni Parmigiani, Mark C. Emerick, Peter Buhlmann
Penalized Likelihood And Bayesian Methods For Sparse Contingency Tables: An Analysis Of Alternative Splicing In Full-Length Cdna Libraries, Corinne Dahinden, Giovanni Parmigiani, Mark C. Emerick, Peter Buhlmann
Johns Hopkins University, Dept. of Biostatistics Working Papers
We develop methods to perform model selection and parameter estimation in loglinear models for the analysis of sparse contingency tables to study the interaction of two or more factors. Typically, datasets arising from so-called full-length cDNA libraries, in the context of alternatively spliced genes, lead to such sparse contingency tables. Maximum Likelihood estimation of log-linear model coefficients fails to work because of zero cell entries. Therefore new methods are required to estimate the coefficients and to perform model selection. Our suggestions include computationally efficient penalization (Lasso-type) approaches as well as Bayesian methods using MCMC. We compare these procedures in a …
Multiple Testing With An Empirical Alternative Hypothesis, James E. Signorovitch
Multiple Testing With An Empirical Alternative Hypothesis, James E. Signorovitch
Harvard University Biostatistics Working Paper Series
An optimal multiple testing procedure is identified for linear hypotheses under the general linear model, maximizing the expected number of false null hypotheses rejected at any significance level. The optimal procedure depends on the unknown data-generating distribution, but can be consistently estimated. Drawing information together across many hypotheses, the estimated optimal procedure provides an empirical alternative hypothesis by adapting to underlying patterns of departure from the null. Proposed multiple testing procedures based on the empirical alternative are evaluated through simulations and an application to gene expression microarray data. Compared to a standard multiple testing procedure, it is not unusual for …
Estimating Genome-Wide Copy Number Using Allele Specific Mixture Models, Wenyi Wang , Benilton Caravalho, Nate Miller, Jonathan Pevsner, Aravinda Chakravarti, Rafael A. Irizarry
Estimating Genome-Wide Copy Number Using Allele Specific Mixture Models, Wenyi Wang , Benilton Caravalho, Nate Miller, Jonathan Pevsner, Aravinda Chakravarti, Rafael A. Irizarry
Johns Hopkins University, Dept. of Biostatistics Working Papers
Genomic changes such as copy number alterations are thought to be one of the major underlying causes of human phenotypic variation among normal and disease subjects [23,11,25,26,5,4,7,18]. These include chromosomal regions with so-called copy number alterations: instead of the expected two copies, a section of the chromosome for a particular individual may have zero copies (homozygous deletion), one copy (hemizygous deletions), or more than two copies (amplifications). The canonical example is Down syndrome which is caused by an extra copy of chromosome 21. Identification of such abnormalities in smaller regions has been of great interest, because it is believed to …
Exploration Of Distributional Models For A Novel Intensity-Dependent Normalization , Nicola Lama, Patrizia Boracchi, Elia Mario Biganzoli
Exploration Of Distributional Models For A Novel Intensity-Dependent Normalization , Nicola Lama, Patrizia Boracchi, Elia Mario Biganzoli
COBRA Preprint Series
Currently used gene intensity-dependent normalization methods, based on regression smoothing techniques, usually approach the two problems of location bias detrending and data re-scaling without taking into account the censoring characteristic of certain gene expressions produced by experiment measurement constraints or by previous normalization steps. Moreover, the bias vs variance balance control of normalization procedures is not often discussed but left to the user's experience. Here an approximate maximum likelihood procedure to fit a model smoothing the dependences of log-fold gene expression differences on average gene intensities is presented. Central tendency and scaling factor were modeled by means of B-splines smoothing …
Structural Inference In Transition Measurement Error Models For Longitudinal Data, Wenqin Pan, Xihong Lin, Donglin Zeng
Structural Inference In Transition Measurement Error Models For Longitudinal Data, Wenqin Pan, Xihong Lin, Donglin Zeng
Harvard University Biostatistics Working Paper Series
No abstract provided.
Estimation In Semiparametric Transition Measurement Error Models For Longitudinal Data, Wenqin Pan, Donglin Zeng, Xihong Lin
Estimation In Semiparametric Transition Measurement Error Models For Longitudinal Data, Wenqin Pan, Donglin Zeng, Xihong Lin
Harvard University Biostatistics Working Paper Series
No abstract provided.
Nonparametric Regression Using Local Kernel Estimating Equations For Correlated Failure Time Data, Zhangsheng Yu, Xihong Lin
Nonparametric Regression Using Local Kernel Estimating Equations For Correlated Failure Time Data, Zhangsheng Yu, Xihong Lin
Harvard University Biostatistics Working Paper Series
No abstract provided.
Causal Inference In Hybrid Intervention Trials Involving Treatment Choice, Qi Long, Rod Little, Xihong Lin
Causal Inference In Hybrid Intervention Trials Involving Treatment Choice, Qi Long, Rod Little, Xihong Lin
Harvard University Biostatistics Working Paper Series
No abstract provided.
A Comparison Of Methods For Estimating The Causal Effect Of A Treatment In Randomized Clinical Trials Subject To Noncompliance, Rod Little, Qi Long, Xihong Lin
A Comparison Of Methods For Estimating The Causal Effect Of A Treatment In Randomized Clinical Trials Subject To Noncompliance, Rod Little, Qi Long, Xihong Lin
Harvard University Biostatistics Working Paper Series
No abstract provided.
Group Additive Regression Models For Genomic Data Analysis, Yihui Luan, Hongzhe Li
Group Additive Regression Models For Genomic Data Analysis, Yihui Luan, Hongzhe Li
UPenn Biostatistics Working Papers
One important problem in genomic research is to identify genomic features such as gene expression data or DNA single nucleotide polymorphisms (SNPs) that are related to clinical phenotypes. Often these genomic data can be naturally divided into biologically meaningful groups such as genes belonging to the same pathways or SNPs within genes. In this paper, we propose group additive regression models and a group gradient descent boosting procedure for identifying groups of genomic features that are related to clinical phenotypes. Our simulation results show that by dividing the variables into appropriate groups, we can obtain better identification of the group …
Extensions To Gene Set Enrichment, Zhen Jiang, Robert Gentleman
Extensions To Gene Set Enrichment, Zhen Jiang, Robert Gentleman
Bioconductor Project Working Papers
Motivation: Gene Set Enrichment Analysis (GSEA) has been developed recently to capture moderate but coordinated changes in the expression of sets of functionally related genes. We propose number of extensions to GSEA, which uses different statistics to describe the association between genes and phenotype of interest. We make use of dimension reduction procedures, such as principle component analysis to identify gene sets containing coordinated genes. We also address the problem of overlapping among gene sets in this paper.
Results: We applied our methods to the data come from a clinical trial in acute lymphoblastic leukemia (ALL) [1]. We identified interesting …
Fdr And Bayesian Multiple Comparisons Rules, Peter Muller, Giovanni Parmigiani, Kenneth Rice
Fdr And Bayesian Multiple Comparisons Rules, Peter Muller, Giovanni Parmigiani, Kenneth Rice
Johns Hopkins University, Dept. of Biostatistics Working Papers
We discuss Bayesian approaches to multiple comparison problems, using a decision theoretic perspective to critically compare competing approaches. We set up decision problems that lead to the use of FDR-based rules and generalizations. Alternative definitions of the probability model and the utility function lead to different rules and problem-specific adjustments. Using a loss function that controls realized FDR we derive an optimal Bayes rule that is a variation of the Benjamini and Hochberg (1995) procedure. The cutoff is based on increments in ordered posterior probabilities instead of ordered p- values. Throughout the discussion we take a Bayesian perspective. In particular, …
Exploration, Normalization, And Genotype Calls Of High Density Oligonucleotide Snp Array Data, Benilton Carvalho, Terence P. Speed, Rafael A. Irizarry
Exploration, Normalization, And Genotype Calls Of High Density Oligonucleotide Snp Array Data, Benilton Carvalho, Terence P. Speed, Rafael A. Irizarry
Johns Hopkins University, Dept. of Biostatistics Working Papers
In most microarray technologies, a number of critical steps are required to convert raw intensity measurements into the data relied upon by data analysts, biologists and clinicians. These data manipulations, referred to as preprocessing, can influence the quality of the ultimate measurements. In the last few years, the high-throughput measurement of gene expression is the most popular application of microarray technology. For this application, various groups have demonstrated that the use of modern statistical methodology can substantially improve accuracy and precision of gene expression measurements, relative to ad-hoc procedures introduced by designers and manufacturers of the technology. Currently, other applications …
Multivariate Analysis And Visualization Of Splicing Correlations In Single-Gene Transcriptomes, Mark C. Emerick, Giovanni Parmigiani, William S. Agnew
Multivariate Analysis And Visualization Of Splicing Correlations In Single-Gene Transcriptomes, Mark C. Emerick, Giovanni Parmigiani, William S. Agnew
Johns Hopkins University, Dept. of Biostatistics Working Papers
Through ‘combinatorial splicing’, RNA metabolism may create enormous structural diversity in the proteome. Functional interactions among multiple alternative domains can have a disproportionate impact on the phenotype, requiring integrated RNA-level regulation of molecular composition. Splicing correlations within molecules expressed from a single gene, where these effects would be greatest, provide valuable clues to functional relationships and targets for splicing regulation. We present tools to visualize complex splicing patterns in full-length cDNA libraries. Developmental changes in pair-wise correlations are presented vectorially in ‘clock plots’ and linkage grids. Higher-order correlations are assessed via a loglinear model and Monte Carlo analysis with an …
Plasq: A Generalized Linear Model-Based Procedure To Determine Allelic Dosage Ini Cancer Cells From Snp Array Data, Thomas Laframboise, David P. Harrington, Barbara A. Weir
Plasq: A Generalized Linear Model-Based Procedure To Determine Allelic Dosage Ini Cancer Cells From Snp Array Data, Thomas Laframboise, David P. Harrington, Barbara A. Weir
Harvard University Biostatistics Working Paper Series
No abstract provided.
Bounded Search For De Novo Identification Of Degenerate Cis-Regulatory Elements, Jonathan M. Carlson, Arijit Chakravarty, Radhika S. Khetani, Robert H. Gross
Bounded Search For De Novo Identification Of Degenerate Cis-Regulatory Elements, Jonathan M. Carlson, Arijit Chakravarty, Radhika S. Khetani, Robert H. Gross
Dartmouth Scholarship
The identification of statistically overrepresented sequences in the upstream regions of coregulated genes should theoretically permit the identification of potential cis-regulatory elements. However, in practice many cis-regulatory elements are highly degenerate, precluding the use of an exhaustive word-counting strategy for their identification. While numerous methods exist for inferring base distributions using a position weight matrix, recent studies suggest that the independence assumptions inherent in the model, as well as the inability to reach a global optimum, limit this approach.
In Vitro Expression And Purification Of Class I Mhc Molecules, Loi Cheng
In Vitro Expression And Purification Of Class I Mhc Molecules, Loi Cheng
Honors Scholar Theses
The major histocompatibility complex (MHC) is a gene family responsible for many critical functions of the immune system in most vertebrates. The MHC consists of three classes differentiated by their structure and function, and MHC class I encodes antigen binding proteins as well as chaperone and accessory proteins such as tapasin. The purpose of this project is to reconstitute several human MHC class I molecules in their peptide-filled and peptide-deficient forms, and to purify these proteins for biochemical study. The expressed proteins include wild type and mutant variants of the fusion protein human leukocyte antigen HLA-B*0801-fos, and human beta-2-microglobulin (β2m). …
Selecting 'Significant' Differentially Expressed Genes From The Combined Perspective Of The Null And The Alternative, Beatrijs Moerkerke, Els Goetghebeur
Selecting 'Significant' Differentially Expressed Genes From The Combined Perspective Of The Null And The Alternative, Beatrijs Moerkerke, Els Goetghebeur
Harvard University Biostatistics Working Paper Series
No abstract provided.
Feature-Level Exploration Of The Choe Et Al. Affymetrix Genechip Control Dataset, Rafael A. Irizarry, Leslie Cope, Zhijin Wu
Feature-Level Exploration Of The Choe Et Al. Affymetrix Genechip Control Dataset, Rafael A. Irizarry, Leslie Cope, Zhijin Wu
Johns Hopkins University, Dept. of Biostatistics Working Papers
We describe why the Choe et al. control dataset should not be used to assess GeneChip expression measures.
Genome Scanning Methods For Comparing Sequences Between Groups, With Application To Hiv Vaccine Trials, Peter B. Gilbert, Chunyuan Wu, David V. Jobes
Genome Scanning Methods For Comparing Sequences Between Groups, With Application To Hiv Vaccine Trials, Peter B. Gilbert, Chunyuan Wu, David V. Jobes
UW Biostatistics Working Paper Series
Consider a placebo-controlled preventive HIV vaccine efficacy trial. An HIV amino acid sequence is measured from each volunteer who acquires HIV, and these sequences are aligned together with the reference HIV sequence represented in the vaccine. We develop genome scanning methods to identify HIV positions at which the amino acids in sequences from infected vaccine recipients tend to be more divergent from the corresponding reference amino acid than the amino acids in sequences from infected placebo recipients. We consider five two-sample test statistics, based on Euclidean, Mahalanobis, and Kullback-Leibler divergence measures. Weights are incorporated to reflect biological information contained in …
2^K Factorials In Blocks Of Size 2, With Application To Two-Color Microarray Experiments, Kathleen F. Kerr
2^K Factorials In Blocks Of Size 2, With Application To Two-Color Microarray Experiments, Kathleen F. Kerr
UW Biostatistics Working Paper Series
When a two-level design must be run in blocks of size two, there is a unique blocking scheme that enables estimation of all the main effects. Unfortunately this design does not enable estimation of any two-factor interactions. When the experimental goal is to estimate all main effects and two-factor interactions, it is necessary to combine replicates of the experiment that use different blocking schemes. In this paper we identify such designs for up to eight factors that enable estimation of all main effects and two-factor interactions with the fewest number of replications. In addition, we give a construction for general …
Nonparametric Pathway-Based Regression Models For Analysis Of Genomic Data, Zhi Wei, Hongzhe Li
Nonparametric Pathway-Based Regression Models For Analysis Of Genomic Data, Zhi Wei, Hongzhe Li
UPenn Biostatistics Working Papers
High-throughout genomic data provide an opportunity for identifying pathways and genes that are related to various clinical phenotypes. Besides these genomic data, another valuable source of data is the biological knowledge about genes and pathways that might be related to the phenotypes of many complex diseases. Databases of such knowledge are often called the metadata. In microarray data analysis, such metadata are currently explored in post hoc ways by gene set enrichment analysis but have hardly been utilized in the modeling step. We propose to develop and evaluate a pathway-based gradient descent boosting procedure for nonparametric pathways-based regression(NPR) analysis to …
Visualizing Genomic Data, Robert Gentleman, Florian Hahne, Wolfgang Huber
Visualizing Genomic Data, Robert Gentleman, Florian Hahne, Wolfgang Huber
Bioconductor Project Working Papers
The advent of experimental techniques capable of probing biomolecules and cells at high levels of resolution has led to a rapid change in the methods used for the analysis of experimental molecular biology data. In this article we give an overview over visualization techniques and methods that can be used to assess various aspects of genomic data.
Gpnn: Power Studies And Applications Of A Neural Network Method For Detecting Gene-Gene Interactions In Studies Of Human Disease, Alison A. Motsinger, Stephen L. Lee, George Mellick, Marylyn D. Ritchie
Gpnn: Power Studies And Applications Of A Neural Network Method For Detecting Gene-Gene Interactions In Studies Of Human Disease, Alison A. Motsinger, Stephen L. Lee, George Mellick, Marylyn D. Ritchie
Dartmouth Scholarship
The identification and characterization of genes that influence the risk of common, complex multifactorial disease primarily through interactions with other genes and environmental factors remains a statistical and computational challenge in genetic epidemiology. We have previously introduced a genetic programming optimized neural network (GPNN) as a method for optimizing the architecture of a neural network to improve the identification of gene combinations associated with disease risk. The goal of this study was to evaluate the power of GPNN for identifying high-order gene-gene interactions. We were also interested in applying GPNN to a real data analysis in Parkinson's disease.