Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- The Texas Medical Center Library (131)
- COBRA (113)
- Augustana College (42)
- University of Kentucky (40)
- University of Nebraska - Lincoln (35)
-
- Dartmouth College (26)
- Virginia Commonwealth University (22)
- University of Nebraska at Omaha (19)
- City University of New York (CUNY) (18)
- University of Connecticut (16)
- Louisiana State University (14)
- University of New Mexico (14)
- Michigan Technological University (12)
- University of Arkansas, Fayetteville (12)
- University of Louisville (12)
- Dordt University (11)
- University of Nebraska Medical Center (11)
- Kennesaw State University (10)
- Loyola University Chicago (10)
- The University of Southern Mississippi (10)
- Western University (10)
- California Polytechnic State University, San Luis Obispo (9)
- Clemson University (9)
- Munster Technological University (9)
- West Virginia University (9)
- Illinois State University (8)
- University of Missouri, St. Louis (8)
- Wayne State University (8)
- Mississippi State University (7)
- University of Montana (7)
- Keyword
-
- Bioinformatics (132)
- Humans (58)
- Genetics (55)
- Genome (49)
- Genomics (46)
-
- Annotation (38)
- Meiothermus ruber (38)
- Gene expression (26)
- GENI-ACT (25)
- Machine learning (20)
- Evolution (19)
- Transcriptome (19)
- Cancer (17)
- Animals (15)
- Algorithms (14)
- Computational biology (14)
- Genome-Wide Association Study (14)
- Phylogenetics (14)
- Population genetics (14)
- Cancer genomics (13)
- Female (13)
- Polymorphism, Single Nucleotide (13)
- Transcriptomics (13)
- Epigenetics (12)
- Biology (11)
- Male (11)
- Polymorphism (11)
- RNA-seq (11)
- Biomarkers (10)
- Computational Biology (10)
- Publication Year
- Publication
-
- Faculty, Staff and Student Publications (72)
- Dissertations and Theses (Open Access) (59)
- Meiothermus ruber Genome Analysis Project (40)
- Theses and Dissertations (31)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (27)
-
- Harvard University Biostatistics Working Paper Series (24)
- Electronic Theses and Dissertations (19)
- COBRA Preprint Series (17)
- Dartmouth Scholarship (16)
- Dissertations, Theses, and Capstone Projects (13)
- Master's Theses (13)
- Graduate Theses and Dissertations (12)
- UW Biostatistics Working Paper Series (12)
- Dissertations, Master's Theses and Master's Reports (11)
- Faculty Work Comprehensive List (11)
- Honors Scholar Theses (11)
- LSU Doctoral Dissertations (11)
- Theses & Dissertations (11)
- Bioconductor Project Working Papers (10)
- Interdisciplinary Informatics Faculty Publications (10)
- UPenn Biostatistics Working Papers (10)
- Biochemistry Publications (9)
- Dartmouth College Ph.D Dissertations (9)
- Bioinformatics Faculty Publications (8)
- Biology ETDs (8)
- Graduate Theses, Dissertations, and Problem Reports (ETD) (8)
- Theses and Dissertations--Biology (8)
- All Dissertations (7)
- Annual Symposium on Biomathematics and Ecology Education and Research (7)
- Dissertations (7)
- Publication Type
- File Type
Articles 631 - 660 of 905
Full-Text Articles in Genetics and Genomics
Comparing Meiothermus Ruber And Myxococcus Xanthus In The Purine Metabolism Pathway, Linnea J. Ritchie, Dr. Lori Scott
Comparing Meiothermus Ruber And Myxococcus Xanthus In The Purine Metabolism Pathway, Linnea J. Ritchie, Dr. Lori Scott
Meiothermus ruber Genome Analysis Project
This project is part of the Meiothermus ruber genome analysis project, which uses the bioinformatics tools associated with the Guiding Education through Novel Investigation – Annotation Collaboration Toolkit (GENI-ACT) to predict gene function. I investigated the biological functions of Mrub_1053 Mrub_2281 and Mrub_2299. I predicted that Mrub_1053 and Mrub_2281 (DNA coordinates 1053364..1054359 on the forward strand and 2333172..2334113 on the forward strand respectively) encodes the enzyme phosphoribose-1-pyrophosphate synthetase (PRS) which is the first step of the purine synthesis pathway (KEGG). I also predicted that Mrub_2299 (DNA coordinates: 2352378..2353775 on the forward strand) encodes for Phosphoribosyl pyrophosphate (PRPP) amidotransferase, which is …
Valine Biosynthesis: Mrub_2994 Is Orthologous To E. Coli B3770 And Mrub_1844 Is Orthologous To E. Coli B3771, Bennett A. Hartmann, Dr. Lori Scott
Valine Biosynthesis: Mrub_2994 Is Orthologous To E. Coli B3770 And Mrub_1844 Is Orthologous To E. Coli B3771, Bennett A. Hartmann, Dr. Lori Scott
Meiothermus ruber Genome Analysis Project
This project is part of the Meiothermus ruber genome analysis project, which uses the bioinformatics tools associated with the Guiding Education through Novel Investigation – Annotation Collaboration Toolkit (GENI-ACT) to predict gene function. We investigated the biological function of the genes Mrub_2994 and Mrub_1844. We predict that Mrub_1884 encodes the enzyme dihydroxy-acid dehydratase (DNA coordinates 1901362..1903026 on the forward strand), which is the third step of the valine biosynthesis pathway (KEGG map number 00290). It catalyzes the conversion of 2,3-dihydroxy-3methylbutanoate to 3-methyl-2-oxobutanoate. The E. coli K12 MG1655 ortholog is predicted to be b3771, which has the gene identifier ilvD. …
Bioinformatic Comparison Of Genes In The Leucine Biosynthesis Pathway Of Escherichia Coli To Meiothermus Ruber, Isaac D. Schmied, Benjamin T. Ryan, Dr. Lori Scott
Bioinformatic Comparison Of Genes In The Leucine Biosynthesis Pathway Of Escherichia Coli To Meiothermus Ruber, Isaac D. Schmied, Benjamin T. Ryan, Dr. Lori Scott
Meiothermus ruber Genome Analysis Project
We predict that Mrub_1905 and Mrub_1906 encode the enzyme 2-isopropylmalate synthase (Mrub_1906 DNA coordinates complement(1965044..1966603) Mrub_1905 DNA coordinates complement(1963455..1965041)), which is the first step of the leucine biosynthesis pathway (KEGG map number 00290). It catalyzes the conversion of (2S)-2-isopropylmalate to 2-isopropylmaleate. The E. coli K12 MG1655 ortholog is predicted to be b0074, which has the gene identifier leuA. We predict that Mrub_1846 encodes the enzyme 3-isopropylmalate dehydrogenase (DNA coordinates complement(1903909..1904961)), which is the third step of the leucine biosynthesis pathway (KEGG map number 00290). It catalyzes the conversion of (2R,3S)-3-isopropylmalate to (2S)-2-isopropyl-3-oxosuccinate. The E. coli K12 MG1655 ortholog is predicted …
Riboflavin Metabolism: A Study To See If Mrub_1256 Is Orthologous To E. Coli B0415, And If Mrub_1254 Is Orthologous To E. Coli B1662, Anish Sora Reddy, Dr. Lori Scott
Riboflavin Metabolism: A Study To See If Mrub_1256 Is Orthologous To E. Coli B0415, And If Mrub_1254 Is Orthologous To E. Coli B1662, Anish Sora Reddy, Dr. Lori Scott
Meiothermus ruber Genome Analysis Project
This project is part of the Meiothermus ruber genome analysis project, which uses the bioinformatics tools associated with the Guiding Education through Novel Investigation –Annotation Collaboration Toolkit (GENI-ACT) to predict gene function. We investigated the biological function of the genes Mrub_1256 and Mrub_1254. We predict that Mrub_1256 encodes the enzyme 6,7-dimethyl-8-ribityllumazine synthase (Dna Coordinates 1282509..1282982 forward strand), which is part of the Riboflavin Metabolism pathway (KEGG map number 00740). It catalyzes the conversion of 3,4-Dihydroxy-2-butanone-4-phosphate or 5-amino-6-ribityl-aminouracil to Quinone. The E. coli K12 MG1655 ortholog is predicted to be E. coli b0415, which has the gene identifier …
The Meiothermus Ruber Mrub_2572 Gene Is An Ortholog Of The Escherichia Coli Pyre B3642 Gene And The Meiothermus Ruber Mrub_2071 Gene Is An Ortholog Of The Escherichia Coli Pyrf B1281 Gene, Cale J. Mccormick, Dr. Lori Scott
The Meiothermus Ruber Mrub_2572 Gene Is An Ortholog Of The Escherichia Coli Pyre B3642 Gene And The Meiothermus Ruber Mrub_2071 Gene Is An Ortholog Of The Escherichia Coli Pyrf B1281 Gene, Cale J. Mccormick, Dr. Lori Scott
Meiothermus ruber Genome Analysis Project
This project is part of the Meiothermus ruber genome analysis project, which uses the bioinformatics tools associated with the Guiding Education through Novel Investigation –Annotation Collaboration Toolkit (GENI-ACT) to predict gene function. We investigated the biological function of the genes Mrub_2572 and Mrub_2071. We predict that Mrub_2572 encodes the enzyme orotate phosphoribosyltransferase (DNA coordinates 2617545..2618096 on the forward strand), which is the 5th step of the UMP biosynthesis pathway (KEGG map number 00240). It catalyzes the conversion of orotate + PRPP to orotidine 5’-phosphate. The E. coli K12 MG1655 ortholog is predicted to be b3642, which has the gene …
Bioinformatics Indicates That Meiothermus Ruber Genes Mrub_1710 And Mrub_1712 Are Homologs Of The Escherichia Coli Genes B2903 (P-Protein; Glycine Dehydrogenase) And B2905 (T-Protein; Aminomethyltransferase), Respectively, Malory J. Groen, Dr. Lori Scott
Bioinformatics Indicates That Meiothermus Ruber Genes Mrub_1710 And Mrub_1712 Are Homologs Of The Escherichia Coli Genes B2903 (P-Protein; Glycine Dehydrogenase) And B2905 (T-Protein; Aminomethyltransferase), Respectively, Malory J. Groen, Dr. Lori Scott
Meiothermus ruber Genome Analysis Project
This project is part of the Meiothermus ruber genome analysis project, which uses the bioinformatics tools associated with the Guiding Education through Novel Investigation – Annotation Collaboration Toolkit (GENI-ACT) to predict gene function. We investigated the biological function of the genes Mrub_1710 and Mrub_1712. We predict that Mrub_1710 encodes the enzyme glycine dehydrogenase (DNA coordinates 3046168.. 3049041 on the reverse strand), which is the first step of the glycine degradation pathway (KEGG map number 00260). It catalyzes the conversion of glycine to S-Amino-methyldihydro-lipoylprotein. The E. coli K12 MG1655 ortholog is predicted to be b2903, which has the gene identifier gcvP. …
E. Coli B3639 And B3634 Are Orthologs Of Mrub_2047 And Mrub_1372, Rong Zheng, Dr. Lori Scott
E. Coli B3639 And B3634 Are Orthologs Of Mrub_2047 And Mrub_1372, Rong Zheng, Dr. Lori Scott
Meiothermus ruber Genome Analysis Project
This project is part of the Meiothermus ruber genome analysis project, which uses the bioinformatics tools associated with the Guiding Education through Novel Investigation –Annotation Collaboration Toolkit (GENI-ACT) to predict gene function. We investigated the biological function of the genes Mrub_2047 and Mrub_1372. We predict that Mrub_2047 encodes the enzyme fused 4'-phosphopantothenoylcysteine decarboxylase/phosphopantothenoylcysteine synthetase, FMN-binding (DNA coordinates 2083590..2084816 on the forward strand), which is the first and the second steps of the CoA biosynthesis pathway (KEGG map number 00770). It catalyzes the conversion of (R)-4’-phosphopantothenate to (R)-4’-phosphopantothenoyl-L-cysteine and the conversion of (R)-4’-phosphopantothenoyl-L-cysteine to 4’-phosphopantetheine. The E. coli K12 MG1655 ortholog …
Mrub_0258 Gene Is An Ortholog Of The B4226 Gene (Ppa) Found In Escherichia Coli; Mrub_1198 Gene Is An Ortholog Of The B2501 Gene (Ppk) Found In Escherichia Coli;, Brandon M. Wills, Dr. Lori Scott
Mrub_0258 Gene Is An Ortholog Of The B4226 Gene (Ppa) Found In Escherichia Coli; Mrub_1198 Gene Is An Ortholog Of The B2501 Gene (Ppk) Found In Escherichia Coli;, Brandon M. Wills, Dr. Lori Scott
Meiothermus ruber Genome Analysis Project
This project is part of the Meiothermus ruber genome analysis project, which uses the bioinformatics tools associated with the Guiding Education through Novel Investigation –Annotation Collaboration Toolkit (GENI-ACT) to predict gene function. We investigated the biological function of the genes Mrub_0258 and Mrub_1198. We predict that Mrub_0258 encodes the enzyme inorganic pyrophosphatase (226403..226942), which is indirectly involved with the oxidative phosphorylation pathway (KEGG map number 00190). It catalyzes the conversion of the diphosphate ions made by Mrub_1198 into two orthophosphate ions, which can then be used by ATP synthase to produce energy. The E. coli K12 MG1655 ortholog is predicted …
Progress In Pathogen Detection By Whole-Genome Sequencing, Chung Wong
Progress In Pathogen Detection By Whole-Genome Sequencing, Chung Wong
Chemistry & Biochemistry Faculty Works
No abstract provided.
Hpcnmf: A High-Performance Toolbox For Non-Negative Matrix Factorization, Karthik Devarajan, Guoli Wang
Hpcnmf: A High-Performance Toolbox For Non-Negative Matrix Factorization, Karthik Devarajan, Guoli Wang
COBRA Preprint Series
Non-negative matrix factorization (NMF) is a widely used machine learning algorithm for dimension reduction of large-scale data. It has found successful applications in a variety of fields such as computational biology, neuroscience, natural language processing, information retrieval, image processing and speech recognition. In bioinformatics, for example, it has been used to extract patterns and profiles from genomic and text-mining data as well as in protein sequence and structure analysis. While the scientific performance of NMF is very promising in dealing with high dimensional data sets and complex data structures, its computational cost is high and sometimes could be critical for …
Models For Hsv Shedding Must Account For Two Levels Of Overdispersion, Amalia Magaret
Models For Hsv Shedding Must Account For Two Levels Of Overdispersion, Amalia Magaret
UW Biostatistics Working Paper Series
We have frequently implemented crossover studies to evaluate new therapeutic interventions for genital herpes simplex virus infection. The outcome measured to assess the efficacy of interventions on herpes disease severity is the viral shedding rate, defined as the frequency of detection of HSV on the genital skin and mucosa. We performed a simulation study to ascertain whether our standard model, which we have used previously, was appropriately considering all the necessary features of the shedding data to provide correct inference. We simulated shedding data under our standard, validated assumptions and assessed the ability of 5 different models to reproduce the …
Expression Of Zinc Fingers And Homeoboxes 2 (Zhx2) And Zhx2 Target Genes In Multiple Tissues Of Wild-Type And Zhx2 Knockout Mice, Minen Al-Kafajy
Expression Of Zinc Fingers And Homeoboxes 2 (Zhx2) And Zhx2 Target Genes In Multiple Tissues Of Wild-Type And Zhx2 Knockout Mice, Minen Al-Kafajy
Theses and Dissertations--Microbiology, Immunology, and Molecular Genetics
The Spear lab has had a long-standing interest in gene regulation in the liver during development and disease. Several years ago, these studies identified a novel transcriptional regulator called Zinc fingers and homeoboxes 2 (Zhx2), which is a member of a small family that includes Zhx1 and Zhx3. All Zhx proteins contain two amino-terminal C2-H2 zinc fingers and four or five carboxy-terminal homeodomains. Previous studies indicate that Zhx proteins can form homodimers and heterodimers with each other.
Zhx2 regulates numerous hepatic genes, including alpha-fetoprotein (AFP) and H19. Genes controlling lipid and cholesterol homeostasis are also regulated by …
A Mechanistic Study Of An Ipsc Model For Leigh’S Disease Caused By Mtdna Mutataion (8993 T>G), John P. Galdun
A Mechanistic Study Of An Ipsc Model For Leigh’S Disease Caused By Mtdna Mutataion (8993 T>G), John P. Galdun
Theses and Dissertations
Mitochondrial diseases encompass a broad range of devastating disorders that typically affect tissues with high-energy requirements. These disorders have been difficult to diagnose and research because of the complexity of mitochondrial genetics, and the large variability seen among patient populations. We have devised and carried out a mechanistic study to generate a cell based model for Leigh’s disease caused by mitochondrial DNA mutation 8993 T>G. Leigh’s disease is a multi-organ system disorder that depends heavily on the mutation burden seen within various tissues. Using new reprogramming and sequencing technologies, we were able to show that Leigh’s disease patient fibroblasts …
Power Analysis In Applied Linear Regression For Cell Type-Specific Differential Expression Detection, Edmund Glass
Power Analysis In Applied Linear Regression For Cell Type-Specific Differential Expression Detection, Edmund Glass
Theses and Dissertations
The goal of many human disease-oriented studies is to detect molecular mechanisms different between healthy controls and patients. Yet, commonly used gene expression measurements from any tissues suffer from variability of cell composition. This variability hinders the detection of differentially expressed genes and is often ignored. However, this variability may actually be advantageous, as heterogeneous gene expression measurements coupled with cell counts may provide deeper insights into the gene expression differences on the cell type-specific level. Published computational methods use linear regression to estimate cell type-specific differential expression. Yet, they do not consider many artifacts hidden in high-dimensional gene expression …
Genomic Comparisons And Genome Architecture Of Divergent Trypanosoma Species, Katie Bradwell
Genomic Comparisons And Genome Architecture Of Divergent Trypanosoma Species, Katie Bradwell
Theses and Dissertations
Virulent Trypanosoma cruzi, and the non-pathogenic Trypanosoma conorhini and Trypanosoma rangeli are protozoan parasites with divergent lifestyles. T. cruzi and T. rangeli are endemic to Latin America, whereas T. conorhini is tropicopolitan. Reduviid bug vectors spread these parasites to mammalian hosts, within which T. rangeli and T. conorhini replicate extracellularly, while T. cruzi has intracellular stages. Firstly, this work compares the genomes of these parasites to understand their differing phenotypes. Secondly, genome architecture of T. cruzi is examined to address the effect of a complex hybridization history, polycistronic transcription, and genome plasticity on this organism, and study its highly …
Differential Gene Expression Of Minnesota (Mn) Hygienic Honeybees (Apis Mellifera) Performing Hygienic Behavior, Eric Northrup
Differential Gene Expression Of Minnesota (Mn) Hygienic Honeybees (Apis Mellifera) Performing Hygienic Behavior, Eric Northrup
All Graduate Theses, Dissertations, and Other Capstone Projects
Hygienic behavior is the ability to remove dead and diseased brood from the comb early as to limit the detrimental impact of the parasite or pathogen. Minnesota (MN) Hygienic bees are generalists of hygienic behavior with the ability to remove several brood infected with several pathogens including the Varroa mite. This study explored the mechanisms of MN Hygienic behavior by comparing the transcriptome of MN Hygienic bee brains to non-hygienic bee brains via cDNA microarray. The results suggest that the brains of MN Hygienic bees may have a greater number of dendritic connections or are more sensitive to neurotransmitters. Quantitative …
Resolving Gnetum Evolutionary History, Angela Mcfadden
Resolving Gnetum Evolutionary History, Angela Mcfadden
All Master's Theses
Gnetum are non-flowering seed plants of the tropics, indigenous to South America, Africa, and Asia. This group of about 40 species is fascinating to botanists because it shares distinctive morphological characteristics with flowering plants, such as broad leaves, woody stems, and flower-like strobili. There are still questions surrounding the relationships within the genus of Gnetum. With that in mind, I focused my work on generating phylogenetic hypotheses, using two molecular data sets: a concatenation of over 60 different chloroplast genes (66,815 base pairs), and the whole chloroplast genome (128,772 base pairs). This allowed me to compare the two phylogenies …
Modeling The Mechanism Underlying Environmental And Genetic Determinants Of Gene Expression And Complex Traits, Gregory Alan Moyerbrailean
Modeling The Mechanism Underlying Environmental And Genetic Determinants Of Gene Expression And Complex Traits, Gregory Alan Moyerbrailean
Wayne State University Dissertations
Advances in next-generation sequencing technologies and functional genomics strategies have allowed researchers to identify both common and rare genetic variation, to deeply profile gene expression, and even to determine regions of active gene transcription.
While these technologies and strategies have contributed greatly to our understanding of complex traits and diseases, there are many biological questions and analytical issues to be addressed.
Genome-wide association studies (GWAS) have successfully identified large numbers of genetic variants associated with complex traits and diseases. However, in many cases the mechanistic link between the phenotype and associated variant remains unclear. This may be because most variants …
System Genetic Analysis Of Mechanisms Underlying Excessive Alcohol Consumption, Maren L. Smith
System Genetic Analysis Of Mechanisms Underlying Excessive Alcohol Consumption, Maren L. Smith
Theses and Dissertations
Increased alcohol consumption over time is one of the characteristic symptoms of Alcohol Use Disorder (AUD). The molecular mechanisms underlying this escalation in intake is still the subject of study. However, the mesocortical and mesolimbic dopamine pathways, and the extended amygdala, because of their involvement in reward and reinforcement are believed to play key roles in these behavioral changes. Multiple gene expression studies have shown that alcohol affects the expression of thousands of genes in the brain. The studies discussed in this document use the systems biology technique of co-expression network analysis to attempt to find
patterns within genome-wide expression …
A Pipeline For Creation Of Genome-Scale Metabolic Reconstructions, Shaun W. Norris
A Pipeline For Creation Of Genome-Scale Metabolic Reconstructions, Shaun W. Norris
Theses and Dissertations
The decreasing costs of next generation sequencing technologies and the increasing speeds at which they work have lead to an abundance of 'omic datasets. The need for tools and methods to analyze, annotate, and model these datasets to better understand biological systems is growing. Here we present a novel software pipeline to reconstruct the metabolic model of an organism in silico starting from its genome sequence and a novel compilation of biological databases to better serve the generation of metabolic models. We validate these methods using five Gardnerella vaginalis strains and compare the gene annotation results to NCBI and the …
Population Genetic Structure Of Necturus Maculosus In Central And Eastern Kentucky, Mason Owen Murphy
Population Genetic Structure Of Necturus Maculosus In Central And Eastern Kentucky, Mason Owen Murphy
Theses and Dissertations--Biology
Population structure is influenced by extrinsic factors, such as landscape architecture and dispersal barriers. Lotic network architecture is known to constrain ecological, demographic and evolutionary processes, including population genetic structure. I assessed the population structure of a widespread aquatic salamander, Necturus maculosus, across three river basins in central and eastern Kentucky. I examined the role of network architecture, anthropogenic barriers, and spatial scale on patterns of population structure. I also provided a review of N. maculosus capture methods and offer an improved trap design. I identified significant structuring between the combined Licking/Kinniconick basin and the Kentucky River basin, with …
Deep Models For Brain Em Image Segmentation: Novel Insights And Improved Performance, Ahmed Fakhry, Hanchuan Peng, Shuiwang Ji
Deep Models For Brain Em Image Segmentation: Novel Insights And Improved Performance, Ahmed Fakhry, Hanchuan Peng, Shuiwang Ji
Computer Science Faculty Publications
Motivation: Accurate segmentation of brain electron microscopy (EM) images is a critical step in dense circuit reconstruction. Although deep neural networks (DNNs) have been widely used in a number of applications in computer vision, most of these models that proved to be effective on image classification tasks cannot be applied directly to EM image segmentation, due to the different objectives of these tasks. As a result, it is desirable to develop an optimized architecture that uses the full power of DNNs and tailored specifically for EM image segmentation.
Results: In this work, we proposed a novel design of DNNs for …
Clusters Of Alpha Satellite On Human Chromosome 21 Are Dispersed Far Onto The Short Arm And Lack Ancient Layers, William Ziccardi, Chongjian Zhao, Valery Shepelev, Lev Uralsky, Ivan Alexandrov, Tatyana Andreeva, Evgeny Rogaev, Christopher Bun, Emily Miller, Catherine Putonti, Jeffrey Doering
Clusters Of Alpha Satellite On Human Chromosome 21 Are Dispersed Far Onto The Short Arm And Lack Ancient Layers, William Ziccardi, Chongjian Zhao, Valery Shepelev, Lev Uralsky, Ivan Alexandrov, Tatyana Andreeva, Evgeny Rogaev, Christopher Bun, Emily Miller, Catherine Putonti, Jeffrey Doering
Bioinformatics Faculty Publications
Human alpha satellite (AS) sequence domains that currently function as centromeres are typically flanked by layers of evolutionarily older AS that presumably represent the remnants of earlier primate centromeres. Studies on several human chromosomes reveal that these older AS arrays are arranged in an age gradient, with the oldest arrays farthest from the functional centromere and arrays progressively closer to the centromere being progressively younger. The organization of AS on human chromosome 21 (HC21) has not been well-characterized. We have used newly available HC21 sequence data and an HC21p YAC map to determine the size, organization, and location of the …
Neuronal Insult Either By Exposure To Lead Or By Direct Neuronal Damage Cause Genome-Wide Changes In Dna Methylation And Histone 3 Lysine 36 Trimethylation, Arko Sen
Wayne State University Dissertations
Prenatal and postnatal exposure to pervasive neuro-toxicants such as Lead (Pb) has been reported to causes extensive and diverse changes in the epigenetic profile. Among epigenetic modification, DNA methylation (5mC) is perhaps the most widely studied and has been proposed to be potential early biomarkers for Pb toxicity. Several studies have demonstrated the association between Pb-exposure and 5mC. However most of these studies are restricted to looking at a specific set of target genes or repetitive elements. Therefore, one of the main objectives of our study was to use an unbiased genome-wide approach to look at Pb-exposure associated changes in …
Finding Function In The Unknown, Kelly Boyd, Emma Highland, Amanda Misch, Amber Hu, Sushma Reddy, Catherine Putonti
Finding Function In The Unknown, Kelly Boyd, Emma Highland, Amanda Misch, Amber Hu, Sushma Reddy, Catherine Putonti
Bioinformatics Faculty Publications
Through high-throughput RNA sequencing (RNAseq), transcriptomes for a single cell, tissue, or organism(s) can be ascertained at a high resolution. While a number of bioinformatic tools have been developed for transcriptome analyses, significant challenges exist for studies of non-model organisms. Without a reference sequence available, raw reads must first be assembled de novo followed by the tedious task of BLAST searches and data mining for functional information. We have created a pipeline, PyRanger, to automate this process. The pipeline includes functionality to assess a single transcriptome and also facilitate comparative transcriptomic studies.
Identifying Gene-Gene Interactions That Are Highly Associated With Body Mass Index Using Quantitative Multifactor Dimensionality Reduction (Qmdr), Rishika De, Shefali S. Verma, Fotios Drenos, Emily R. Holzinger
Identifying Gene-Gene Interactions That Are Highly Associated With Body Mass Index Using Quantitative Multifactor Dimensionality Reduction (Qmdr), Rishika De, Shefali S. Verma, Fotios Drenos, Emily R. Holzinger
Dartmouth Scholarship
Despite heritability estimates of 40–70% for obesity, less than 2% of its variation is explained by Body Mass Index (BMI) associated loci that have been identified so far. Epistasis, or gene-gene interactions are a plausible source to explain portions of the missing heritability of BMI. Using genotypic data from 18,686 individuals across five study cohorts – ARIC, CARDIA, FHS, CHS, MESA – we filtered SNPs (Single Nucleotide Polymorphisms) using two parallel approaches. SNPs were filtered either on the strength of their main effects of association with BMI, or on the number of knowledge sources supporting a specific SNP-SNP interaction in …
The Importance Of Physicochemical Characteristics And Nonlinear Classifiers In Determining Hiv-1 Protease Specificity, Timmy Manning, Paul Walsh
The Importance Of Physicochemical Characteristics And Nonlinear Classifiers In Determining Hiv-1 Protease Specificity, Timmy Manning, Paul Walsh
Department of Biological Sciences Publications
This paper reviews recent research relating to the application of bioinformatics approaches to determining HIV-1 protease specificity, outlines outstanding issues, and presents a new approach to addressing these issues. Leading machine learning theory for the problem currently suggests that the direct encoding of the physicochemical properties of the amino acid substrates is not required for optimal performance. A number of amino acid encoding approaches which incorporate potentially relevant physicochemical properties of the substrate are identified, and are evaluated using a nonlinear task decomposition based neuroevolution algorithm. The results are evaluated, and compared against a recent benchmark set on a nonlinear …
A Survey Of The Common Loon (Gavia Immer) Genome Reveals Patterns Of Natural Selection, Zach G. Gayk
A Survey Of The Common Loon (Gavia Immer) Genome Reveals Patterns Of Natural Selection, Zach G. Gayk
All NMU Master's Theses
With rapid advances in Next-Generation Sequencing technology, comparative genomics has become a viable method for studying the adaptation of species to their environment at the genome level. I investigated this in common loons (Gavia immer)—for which molecular adaptation has not been characterized—by finding signatures of positive selection as evidence for genomic adaptation.
I used Illumina short read sequencing data from a single female common loon to produce a fragmented assembly of the common loon (Gavia immer) genome. The resulting assembly had a contig N50 of 814 bp, a total length of 767,326,331 bp, and 45.7 % …
Apply Data Clustering To Gene Expression Data, Abdullah Jameel Abualhamayl Mr.
Apply Data Clustering To Gene Expression Data, Abdullah Jameel Abualhamayl Mr.
Electronic Theses, Projects, and Dissertations
Data clustering plays an important role in effective analysis of gene expression. Although DNA microarray technology facilitates expression monitoring, several challenges arise when dealing with gene expression datasets. Some of these challenges are the enormous number of genes, the dimensionality of the data, and the change of data over time. The genetic groups which are biologically interlinked can be identified through clustering. This project aims to clarify the steps to apply clustering analysis of genes involved in a published dataset. The methodology for this project includes the selection of the dataset representation, the selection of gene datasets, Similarity Matrix Selection, …
A Polyglot Approach To Bioinformatics Data Integration: A Phylogenetic Analysis Of Hiv-1, Steven Reisman, Thomas Hatzopoulos, Konstantin Laufer, George K. Thiruvathukal, Catherine Putonti
A Polyglot Approach To Bioinformatics Data Integration: A Phylogenetic Analysis Of Hiv-1, Steven Reisman, Thomas Hatzopoulos, Konstantin Laufer, George K. Thiruvathukal, Catherine Putonti
Bioinformatics Faculty Publications
As sequencing technologies continue to drop in price and increase in throughput, new challenges emerge for the management and accessibility of genomic sequence data. We have developed a pipeline for facilitating the storage, retrieval, and subsequent analysis of molecular data, integrating both sequence and metadata. Taking a polyglot approach involving multiple languages, libraries, and persistence mechanisms, sequence data can be aggregated from publicly available and local repositories. Data are exposed in the form of a RESTful web service, formatted for easy querying, and retrieved for downstream analyses. As a proof of concept, we have developed a resource for annotated HIV-1 …