An intelligent data-centric approach toward identification of conserved motifs in protein sequences

Kathryn Dempsey, Benjamin Currall, Richard Hallworth, Hesham Ali

Research output: Chapter in Book/Report/Conference proceedingConference contribution

4 Scopus citations

Abstract

The continued integration of the computational and biological sciences has revolutionized genomic and proteomic studies. However, efficient collaboration between these fields requires the creation of shared standards. A common problem arises when biological input does not properly fit the expectations of the algorithm, which can result in misinterpretation of the output. This potential confounding of input/output is a drawback especially when regarding motif finding software. Here we propose a method for improving output by selecting input based upon evolutionary distance, domain architecture, and known function. This method improved detection of both known and unknown motifs in two separate case studies. By standardizing input considerations, both biologists and bioinformaticians can better interpret and design the evolving sophistication of bioinformatic software.

Original languageEnglish (US)
Title of host publication2010 ACM International Conference on Bioinformatics and Computational Biology, ACM-BCB 2010
Pages398-401
Number of pages4
DOIs
StatePublished - 2010
Event2010 ACM International Conference on Bioinformatics and Computational Biology, ACM-BCB 2010 - Niagara Falls, NY, United States
Duration: Aug 2 2010Aug 4 2010

Publication series

Name2010 ACM International Conference on Bioinformatics and Computational Biology, ACM-BCB 2010

Conference

Conference2010 ACM International Conference on Bioinformatics and Computational Biology, ACM-BCB 2010
Country/TerritoryUnited States
CityNiagara Falls, NY
Period8/2/108/4/10

Keywords

  • Intelligent tools
  • Motif finding
  • Prestin
  • Protein sequences

ASJC Scopus subject areas

  • Biomedical Engineering
  • Health Information Management

Fingerprint

Dive into the research topics of 'An intelligent data-centric approach toward identification of conserved motifs in protein sequences'. Together they form a unique fingerprint.

Cite this