Open Access. Powered by Scholars. Published by Universities.®

Computational Biology Commons™

Open Access. Powered by Scholars. Published by Universities.®

Medicine and Health Sciences

Institution
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 121 - 150 of 156

Full-Text Articles in Computational Biology

Hpcnmf: A High-Performance Toolbox For Non-Negative Matrix Factorization, Karthik Devarajan, Guoli Wang Feb 2016

Hpcnmf: A High-Performance Toolbox For Non-Negative Matrix Factorization, Karthik Devarajan, Guoli Wang

COBRA Preprint Series

Non-negative matrix factorization (NMF) is a widely used machine learning algorithm for dimension reduction of large-scale data. It has found successful applications in a variety of fields such as computational biology, neuroscience, natural language processing, information retrieval, image processing and speech recognition. In bioinformatics, for example, it has been used to extract patterns and profiles from genomic and text-mining data as well as in protein sequence and structure analysis. While the scientific performance of NMF is very promising in dealing with high dimensional data sets and complex data structures, its computational cost is high and sometimes could be critical for …


Models For Hsv Shedding Must Account For Two Levels Of Overdispersion, Amalia Magaret Jan 2016

Models For Hsv Shedding Must Account For Two Levels Of Overdispersion, Amalia Magaret

UW Biostatistics Working Paper Series

We have frequently implemented crossover studies to evaluate new therapeutic interventions for genital herpes simplex virus infection. The outcome measured to assess the efficacy of interventions on herpes disease severity is the viral shedding rate, defined as the frequency of detection of HSV on the genital skin and mucosa. We performed a simulation study to ascertain whether our standard model, which we have used previously, was appropriately considering all the necessary features of the shedding data to provide correct inference. We simulated shedding data under our standard, validated assumptions and assessed the ability of 5 different models to reproduce the …


Leveraging Global Gene Expression Patterns To Predict Expression Of Unmeasured Genes, James Rudd, René A. Zelaya, Eugene Demidenko, Ellen L. Goode, Casey S. Greene S. Greene, Jennifer A. Doherty Dec 2015

Leveraging Global Gene Expression Patterns To Predict Expression Of Unmeasured Genes, James Rudd, René A. Zelaya, Eugene Demidenko, Ellen L. Goode, Casey S. Greene S. Greene, Jennifer A. Doherty

Dartmouth Scholarship

BackgroundLarge collections of paraffin-embedded tissue represent a rich resource to test hypotheses based on gene expression patterns; however, measurement of genome-wide expression is cost-prohibitive on a large scale. Using the known expression correlation structure within a given disease type (in this case, high grade serous ovarian cancer; HGSC), we sought to identify reduced sets of directly measured (DM) genes which could accurately predict the expression of a maximized number of unmeasured genes.


Meta-Analysis Of Dna Methylation And Expression In Liver Cancer Patients, Daniele Todorov, Barbara Stefanska, Katarzyna Lubecka-Pietruszewska Aug 2015

Meta-Analysis Of Dna Methylation And Expression In Liver Cancer Patients, Daniele Todorov, Barbara Stefanska, Katarzyna Lubecka-Pietruszewska

The Summer Undergraduate Research Fellowship (SURF) Symposium

Hepatocellular carcinoma (HCC), the most common liver cancer, is the second leading cause of cancer-related death worldwide. HCC is often diagnosed at late stages, for which there are no effective chemotherapies. Biomarkers unique to HCC patients could be used to detect HCC early and improve treatment. In the present project, we have performed a meta-analysis to compare the gene-specific DNA methylation and gene expression patterns of HCC patients as reported by four independent studies. Our goal was to discover the strongest changes that distinguish HCC from normal tissue. The relationship between methylation and expression in HCC was examined and genes …


Functional Analysis Of Synthetic Gene Circuits Controlling A Protein Pump In Yeast, Junchen Diao Aug 2015

Functional Analysis Of Synthetic Gene Circuits Controlling A Protein Pump In Yeast, Junchen Diao

Dissertations and Theses (Open Access)

Synthetic biology aims to build biological devices to understand living systems and explore new applications. Synthetic gene circuits such as genetic switches, oscillators and logic gates are at the core of many synthetic biology applications. These gene circuits often include a sensor/regulator protein capable to detect small molecules and then transduce them into a regulatory signal to generate measurable output. Similar signal transduction networks are also abundant in nature. However, in many natural and engineered scenarios, the output also affects the regulator/sensor protein. How such interactions between the regulator/sensor and the output affect synthetic gene circuit function has not been …


Investigation Of Genetic Alterations In Emt Suppressor, Dear1, Through Pan-Cancer Analysis And Ultra-Deep Targeted Sequencing In Ductal Carcinoma In Situ, Jacquelyn Reuther May 2015

Investigation Of Genetic Alterations In Emt Suppressor, Dear1, Through Pan-Cancer Analysis And Ultra-Deep Targeted Sequencing In Ductal Carcinoma In Situ, Jacquelyn Reuther

Dissertations and Theses (Open Access)

Ductal carcinoma in situ (DCIS) is thought to be one of the earliest pre-invasive form of and non-obligate precursor to invasive ductal carcinoma (IDC). There is an urgent need to identify predictive and prognostic biomarkers for breast cancers with a heightened risk of progression from DCIS to IDC. Our laboratory has previously discovered a novel TRIM family member, DEAR1 (Ductal Epithelium Associated Ring Chromosome 1, annotated as TRIM62) within chromosome 1p35.1, that is mutated and homozygously deleted in breast cancer and whose expression is downregulated/lost in DCIS. Previous work has shown that DEAR1 is a novel tumor suppressor …


Genetics Of Obesity In Starr County, Texas Mexican Americans, Heather M. Highland May 2015

Genetics Of Obesity In Starr County, Texas Mexican Americans, Heather M. Highland

Dissertations and Theses (Open Access)

Currently, over two-thirds of Americans are classified as over-weight or obese. Obesity increases risk for many other diseases including type 2 diabetes, heart disease, stroke, and cancer, making obesity the largest public health problem in America and most other Westernized nations. Hispanics have a higher rate of both obesity and type 2 diabetes, making them a particularly interesting population in which to study obesity. For the last 33 years, the Starr County Health Studies has collected an array of phenotypes and biological samples from residents of Starr County, along Texas-Mexico border. This study includes 825 subjects who were not known …


Systems Level Analysis Of Systemic Sclerosis Shows A Network Of Immune And Profibrotic Pathways Connected With Genetic Polymorphisms, J. Matthew Mahoney, Jaclyn Taroni, Viktor Martyanov, Tammara A. A. Wood, Casey S. Greene, Patricia A. Pioli, Monique E. Hinchcliff, Michael L. Whitfield Jan 2015

Systems Level Analysis Of Systemic Sclerosis Shows A Network Of Immune And Profibrotic Pathways Connected With Genetic Polymorphisms, J. Matthew Mahoney, Jaclyn Taroni, Viktor Martyanov, Tammara A. A. Wood, Casey S. Greene, Patricia A. Pioli, Monique E. Hinchcliff, Michael L. Whitfield

Dartmouth Scholarship

Systemic sclerosis (SSc) is a rare systemic autoimmune disease characterized by skin and organ fibrosis. The pathogenesis of SSc and its progression are poorly understood. The SSc intrinsic gene expression subsets (inflammatory, fibroproliferative, normal-like, and limited) are observed in multiple clinical cohorts of patients with SSc. Analysis of longitudinal skin biopsies suggests that a patient's subset assignment is stable over 6-12 months. Genetically, SSc is multi-factorial with many genetic risk loci for SSc generally and for specific clinical manifestations. Here we identify the genes consistently associated with the intrinsic subsets across three independent cohorts, show the relationship between these genes …


Genomic Characterization Of Polyps In Familial Adenomatous Polyposis Patients And Identification Of Candidate Chemopreventive Drugs, Francis A. San Lucas Aug 2014

Genomic Characterization Of Polyps In Familial Adenomatous Polyposis Patients And Identification Of Candidate Chemopreventive Drugs, Francis A. San Lucas

Dissertations and Theses (Open Access)

Familial adenomatous polyposis (FAP) is an autosomal dominant disease characterized by APC germline mutations and the development of hundreds to thousands of premalignant adenomas in the gastrointestinal tract at a young age. If left untreated, these patients inevitably develop colon cancer (CRC) and small bowel tumors. We performed exome sequencing of samples from 12 FAP patients to characterize adenomas and to identify candidate genes of adenoma development that may serve as potential targets for chemoprevention drug development. From each patient, a blood and at least one polyp were sequenced with a total of 25 polyps analyzed. In some cases, normal …


Structural Features Of The Pseudomonas Fluorescens Biofilm Adhesin Lapa Required For Lapg-Dependent Cleavage, Biofilm Formation, And Cell Surface Localization, Chelsea D. Boyd, T. Jarrod Smith, Sofiane El-Kirat-Chatel, Peter D. Newell, Yves F. Dufrêne, George A. O'Toole May 2014

Structural Features Of The Pseudomonas Fluorescens Biofilm Adhesin Lapa Required For Lapg-Dependent Cleavage, Biofilm Formation, And Cell Surface Localization, Chelsea D. Boyd, T. Jarrod Smith, Sofiane El-Kirat-Chatel, Peter D. Newell, Yves F. Dufrêne, George A. O'Toole

Dartmouth Scholarship

The localization of the LapA protein to the cell surface is a key step required by Pseudomonas fluorescens Pf0-1 to irreversibly attach to a surface and form a biofilm. LapA is a member of a diverse family of predicted bacterial adhesins, and although lacking a high degree of sequence similarity, family members do share common predicted domains. Here, using mutational analysis, we determine the significance of each domain feature of LapA in relation to its export and localization to the cell surface and function in biofilm formation. Our previous work showed that the N terminus of LapA is required for …


The Association Between The Il-1 Pathway, Isaac C. Wun May 2014

The Association Between The Il-1 Pathway, Isaac C. Wun

Dissertations and Theses (Open Access)

Cutaneous malignant melanoma (CMM) is a potentially lethal malignancy that warrants attention and further research, as it is known to that there is an increasing rate of incidence in theUnited States, and it is also known that exposure to UV light is its most crucial risk factor, and family history of melanoma is also an important risk factor. Melanoma is an aggressive and lethal cancer in humans. There are an estimated new 132,000 melanoma cases annually worldwide, and the trend has doubled in the past 20 years. However, attempts to treat melanoma have encountered considerable resistance and remained ineffective. The …


Integrated Assessment Of Predicted Mhc Binding And Cross-Conservation With Self Reveals Patterns Of Viral Camouflage, Lu He, Anne S. De Groot, Andres H. Gutierrez, William D. Martin, Lenny Moise, Chris Bailey-Kellogg Mar 2014

Integrated Assessment Of Predicted Mhc Binding And Cross-Conservation With Self Reveals Patterns Of Viral Camouflage, Lu He, Anne S. De Groot, Andres H. Gutierrez, William D. Martin, Lenny Moise, Chris Bailey-Kellogg

Dartmouth Scholarship

Immune recognition of foreign proteins by T cells hinges on the formation of a ternary complex sandwiching a constituent peptide of the protein between a major histocompatibility complex (MHC) molecule and a T cell receptor (TCR). Viruses have evolved means of "camouflaging" themselves, avoiding immune recognition by reducing the MHC and/or TCR binding of their constituent peptides. Computer-driven T cell epitope mapping tools have been used to evaluate the degree to which articular viruses have used this means of avoiding immune response, but most such analyses focus on MHC-facing ‘agretopes'. Here we set out a new means of evaluating the …


Computational Model For Survey And Trend Analysis Of Patients With Endometriosis : A Decision Aid Tool For Ebm, Salvo Reina, Vito Reina, Franco Ameglio, Mauro Costa, Alessandro Fasciani Feb 2014

Computational Model For Survey And Trend Analysis Of Patients With Endometriosis : A Decision Aid Tool For Ebm, Salvo Reina, Vito Reina, Franco Ameglio, Mauro Costa, Alessandro Fasciani

COBRA Preprint Series

Endometriosis is increasingly collecting worldwide attention due to its medical complexity and social impact. The European community has identified this as a “social disease”. A large amount of information comes from scientists, yet several aspects of this pathology and staging criteria need to be clearly defined on a suitable number of individuals. In fact, available studies on endometriosis are not easily comparable due to a lack of standardized criteria to collect patients’ informations and scarce definitions of symptoms. Currently, only retrospective surgical stadiation is used to measure pathology intensity, while the Evidence Based Medicine (EBM) requires shareable methods and correct …


A Systems Biology Approach To Detect Eqtls Associated With Mirna And Mrna Co-Expression Networks In The Nucleus Accumbens Of Chronic Alcoholic Patients, Mohammed Mamdani Jan 2014

A Systems Biology Approach To Detect Eqtls Associated With Mirna And Mrna Co-Expression Networks In The Nucleus Accumbens Of Chronic Alcoholic Patients, Mohammed Mamdani

Theses and Dissertations

Alcohol Dependence (AD) is a chronic substance use disorder with moderate heritability (60%). Linkage and genome-wide association studies (GWAS) have implicated a number of loci; however, the molecular mechanisms underlying AD are unclear. Advances in systems biology allow genome-wide expression data to be integrated with genetic data to detect expression quantitative trait loci (eQTL), polymorphisms that regulate gene expression levels, influence phenotypes and are significantly enriched among validated genetic signals for many commonly studied traits including AD.

We integrated genome-wide mRNA and miRNA expression data with genotypic data from the nucleus accumbens (NAc), a major addiction-related brain region, of 36 …


Rna-Sequencing Applications: Gene Expression Quantification And Methylator Phenotype Identification, Guoshuai Cai Aug 2013

Rna-Sequencing Applications: Gene Expression Quantification And Methylator Phenotype Identification, Guoshuai Cai

Dissertations and Theses (Open Access)

My dissertation focuses on two aspects of RNA sequencing technology. The first is the methodology for modeling the overdispersion inherent in RNA-seq data for differential expression analysis. This aspect is addressed in three sections. The second aspect is the application of RNA-seq data to identify the CpG island methylator phenotype (CIMP) by integrating datasets of mRNA expression level and DNA methylation status.

Section 1: The cost of DNA sequencing has reduced dramatically in the past decade. Consequently, genomic research increasingly depends on sequencing technology. However it remains elusive how the sequencing capacity influences the accuracy of mRNA expression measurement. We …


Transcription Factor Binding Profiles Reveal Cyclic Expression Of Human Protein-Coding Genes And Non-Coding Rnas, Chao Cheng, Matthew Ung, Gavin D. Grant, Michael L. Whitfield Jul 2013

Transcription Factor Binding Profiles Reveal Cyclic Expression Of Human Protein-Coding Genes And Non-Coding Rnas, Chao Cheng, Matthew Ung, Gavin D. Grant, Michael L. Whitfield

Dartmouth Scholarship

Cell cycle is a complex and highly supervised process that must proceed with regulatory precision to achieve successful cellular division. Despite the wide application, microarray time course experiments have several limitations in identifying cell cycle genes. We thus propose a computational model to predict human cell cycle genes based on transcription factor (TF) binding and regulatory motif information in their promoters. We utilize ENCODE ChIP-seq data and motif information as predictors to discriminate cell cycle against non-cell cycle genes. Our results show that both the trans- TF features and the cis- motif features are predictive of cell cycle genes, and …


Chapter 11: Genome-Wide Association Studies, William S. Bush, Jason H. Moore Dec 2012

Chapter 11: Genome-Wide Association Studies, William S. Bush, Jason H. Moore

Dartmouth Scholarship

Genome-wide association studies (GWAS) have evolved over the last ten years into a powerful tool for investigating the genetic architecture of human disease. In this work, we review the key concepts underlying GWAS, including the architecture of common diseases, the structure of common human genetic variation, technologies for capturing genetic information, study designs, and the statistical methods used for data analysis. We also look forward to the future beyond GWAS.


An Integrated Bioinformatics And Computational Biology Approach Identifies New Bh3-Only Protein Candidates, Robert G. Hawley, Yuzhong Chen, Irene Riz, Chen Zeng May 2012

An Integrated Bioinformatics And Computational Biology Approach Identifies New Bh3-Only Protein Candidates, Robert G. Hawley, Yuzhong Chen, Irene Riz, Chen Zeng

Anatomy and Regenerative Biology Faculty Publications

FoxD4L1 is a forkhead transcription factor that expands the neural ectoderm by down-regulating genes that promote the onset of neural differentiation and up-regulating genes that maintain proliferative neural precursors in an immature state. We previously demonstrated that binding of Grg4 to an Eh-1 motif enhances the ability of FoxD4L1 to down-regulate target neural genes but does not account for all of its repressive activity. Herein we analyzed the protein sequence for additional interaction motifs and secondary structure. Eight conserved motifs were identified in the C-terminal region of fish and frog proteins. Extending the analysis to mammals identified a high scoring …


Dynamic Herbal Monographs For A Digital World, Niamh O'Brien Jan 2012

Dynamic Herbal Monographs For A Digital World, Niamh O'Brien

Theses

Post analysis of a worldwide survey of Medical Herbalists, 93% of respondents were in favour of an online system which could update monographs dynamically. 63% of respondents suggested that some current monographs are out of date and lack certain practicalities in areas such as : Interactions, Dosage and Safety. Research into gaining optimal responses from surveys led to a 78% response rate. Survey analysis resulted in a modem up-to-date monograph template being created and each of the aforementioned information systems tested against same. Testing involved the generation of XML, HTML, PHP and OWL languages for encoding documents to allow for …


Efficient And Robust Rna-Seq Process For Cultured Bacteria And Complex Community Transcriptomes, Georgia Giannoukos, Dawn M. Ciulla, Katherine Huang, Brian J. Haas, Jacques Izard, Joshua Z. Levin, Jonathan Livny, Ashlee M. Earl, Dirk Gevers, Doyle V. Ward, Chad Nusbaum, Bruce W. Birren, Andreas Gnirke Jan 2012

Efficient And Robust Rna-Seq Process For Cultured Bacteria And Complex Community Transcriptomes, Georgia Giannoukos, Dawn M. Ciulla, Katherine Huang, Brian J. Haas, Jacques Izard, Joshua Z. Levin, Jonathan Livny, Ashlee M. Earl, Dirk Gevers, Doyle V. Ward, Chad Nusbaum, Bruce W. Birren, Andreas Gnirke

Department of Food Science and Technology: Faculty Publications

We have developed a process for transcriptome analysis of bacterial communities that accommodates both intact and fragmented starting RNA and combines efficient rRNA removal with strand-specific RNA-seq. We applied this approach to an RNA mixture derived from three diverse cultured bacterial species and to RNA isolated from clinical stool samples. The resulting expression profiles were highly reproducible, enriched up to 40-fold for non-rRNA transcripts, and correlated well with profiles representing undepleted total RNA.


Optimization Algorithms For Functional Deimmunization Of Therapeutic Proteins, Andrew S. Parker, Wei Zheng, Karl E. Griswold, Chris Bailey-Kellogg Apr 2010

Optimization Algorithms For Functional Deimmunization Of Therapeutic Proteins, Andrew S. Parker, Wei Zheng, Karl E. Griswold, Chris Bailey-Kellogg

Dartmouth Scholarship

To develop protein therapeutics from exogenous sources, it is necessary to mitigate the risks of eliciting an anti-biotherapeutic immune response. A key aspect of the response is the recognition and surface display by antigen-presenting cells of epitopes, short peptide fragments derived from the foreign protein. Thus, developing minimal-epitope variants represents a powerful approach to deimmunizing protein therapeutics. Critically, mutations selected to reduce immunogenicity must not interfere with the protein's therapeutic activity.


Advancing Epidemiological Science Through Computational Modeling: A Review With Novel Examples, Scott M. Duke-Sylvester, Eli N. Perencevich, Jon P. Furuno, Leslie A. Real, Holly Gaff Jan 2008

Advancing Epidemiological Science Through Computational Modeling: A Review With Novel Examples, Scott M. Duke-Sylvester, Eli N. Perencevich, Jon P. Furuno, Leslie A. Real, Holly Gaff

Biological Sciences Faculty Publications

Computational models have been successfully applied to a wide variety of research areas including infectious disease epidemiology. Especially for questions that are difficult to examine in other ways, computational models have been used to extend the range of epidemiological issues that can be addressed, advance theoretical understanding of disease processes and help identify specific intervention strategies. We explore each of these contributions to epidemiology research through discussion and examples. We also describe in detail models for raccoon rabies and methicillin-resis-tant Staphylococcus aureus, drawn from our own research, to further illustrate the role of computation in epidemiological modeling.


Power Boosting In Genome-Wide Studies Via Methods For Multivariate Outcomes, Mary J. Emond Feb 2007

Power Boosting In Genome-Wide Studies Via Methods For Multivariate Outcomes, Mary J. Emond

UW Biostatistics Working Paper Series

Whole-genome studies are becoming a mainstay of biomedical research. Examples include expression array experiments, comparative genomic hybridization analyses and large case-control studies for detecting polymorphism/disease associations. The tactic of applying a regression model to every locus to obtain test statistics is useful in such studies. However, this approach ignores potential correlation structure in the data that could be used to gain power, particularly when a Bonferroni correction is applied to adjust for multiple testing. In this article, we propose using regression techniques for misspecified multivariate outcomes to increase statistical power over independence-based modeling at each locus. Even when the outcome …


Semiparametric Regression Of Multi-Dimensional Genetic Pathway Data: Least Squares Kernel Machines And Linear Mixed Models, Dawei Liu, Xihong Lin, Debashis Ghosh Nov 2006

Semiparametric Regression Of Multi-Dimensional Genetic Pathway Data: Least Squares Kernel Machines And Linear Mixed Models, Dawei Liu, Xihong Lin, Debashis Ghosh

Harvard University Biostatistics Working Paper Series

No abstract provided.


Structural Inference In Transition Measurement Error Models For Longitudinal Data, Wenqin Pan, Xihong Lin, Donglin Zeng Aug 2006

Structural Inference In Transition Measurement Error Models For Longitudinal Data, Wenqin Pan, Xihong Lin, Donglin Zeng

Harvard University Biostatistics Working Paper Series

No abstract provided.


Estimation In Semiparametric Transition Measurement Error Models For Longitudinal Data, Wenqin Pan, Donglin Zeng, Xihong Lin Aug 2006

Estimation In Semiparametric Transition Measurement Error Models For Longitudinal Data, Wenqin Pan, Donglin Zeng, Xihong Lin

Harvard University Biostatistics Working Paper Series

No abstract provided.


Nonparametric Regression Using Local Kernel Estimating Equations For Correlated Failure Time Data, Zhangsheng Yu, Xihong Lin Aug 2006

Nonparametric Regression Using Local Kernel Estimating Equations For Correlated Failure Time Data, Zhangsheng Yu, Xihong Lin

Harvard University Biostatistics Working Paper Series

No abstract provided.


Causal Inference In Hybrid Intervention Trials Involving Treatment Choice, Qi Long, Rod Little, Xihong Lin Aug 2006

Causal Inference In Hybrid Intervention Trials Involving Treatment Choice, Qi Long, Rod Little, Xihong Lin

Harvard University Biostatistics Working Paper Series

No abstract provided.


A Comparison Of Methods For Estimating The Causal Effect Of A Treatment In Randomized Clinical Trials Subject To Noncompliance, Rod Little, Qi Long, Xihong Lin Aug 2006

A Comparison Of Methods For Estimating The Causal Effect Of A Treatment In Randomized Clinical Trials Subject To Noncompliance, Rod Little, Qi Long, Xihong Lin

Harvard University Biostatistics Working Paper Series

No abstract provided.


Circadian Rhythmicity By Autocatalysis, Arun Mehra, Christian I. Hong, Mi Shi, Jennifer J. Loros, Jay C. Dunlap, Peter Ruoff Jul 2006

Circadian Rhythmicity By Autocatalysis, Arun Mehra, Christian I. Hong, Mi Shi, Jennifer J. Loros, Jay C. Dunlap, Peter Ruoff

Dartmouth Scholarship

The temperature compensated in vitro oscillation of cyanobacterial KaiC phosphorylation, the first example of a thermodynamically closed system showing circadian rhythmicity, only involves the three Kai proteins (KaiA, KaiB, and KaiC) and ATP. In this paper, we describe a model in which the KaiA- and KaiB-assisted autocatalytic phosphorylation and dephosphorylation of KaiC are the source for circadian rhythmicity. This model, based upon autocatalysis instead of transcription-translation negative feedback, shows temperature-compensated circadian limit-cycle oscillations with KaiC phosphorylation profiles and has period lengths and rate constant values that are consistent with experimental observations.