Mutation Miner (CPI 2005)

Introduction

Biological researchers today have access to vast amounts of exponentially growing research data in a structured form within several publicly accessible databases. A large proportion of salient information is however still hidden within individual research papers, since costly manual database curation efforts are overwhelmed by the scale of new information being generated. In the domain of protein engineering, critical units of information required from the literature include: the identity of the mutated protein, the identity and position of wild type residues that are mutated, the identity of the resulting mutant residues and the impacts of the mutations on functional properties of the proteins.
Mutation Miner is a system designed to automate the extraction of mutations and textual annotations describing the impacts of mutations on protein properties (mutation annotations) from full text scientific literature. Furthermore, the system retrieves and carries out bioinformatic analyses on mutated sequences providing the mapped coordinates of mutants on a selected structure. Integration of multiple formatted mutation annotations with associated residue coordinates facilitates their rendering with structure visualization tools. We describe the architecture and tools that support Mutation Miner (Text mining-NLP, Sequence Analysis, Structure Visualization) and present performance evaluations that demonstrate the feasibility of this approach.

Mutation Miner Poster at CPI 2005

Reference

Christopher J. O. Baker, René Witte, Ashwin Bhat Gurpur, and Vladislav Ryzhikov, Mutation Miner. 5th International Conference of the Canadian Proteomics Initiative (CPI 2005), May 13-14, 2005, Toronto, Ontario, Canada.

Bibtex entry (also for download):

@InProceedings{BWBR_CPI2005,
  author = 	 {Christopher J. O. Baker and Ren{\'e} Witte and 
                  Ashwin Bhat Gurpur and Vladislav Ryzhikov},
  title = 	 {{Mutation Miner}},
  booktitle =	 {5th International Conference of the 
                  Canadian Proteomics Initiative (CPI 2005)},
  year =	 {2005},
  address =	 {Toronto, Ontario, Canada},
  month =	 {May 13--14}
}

This poster was also presented at the Knowledge-based Bioinformatics Workshop (KBB), September 21st-23rd, 2005, Montréal, Québec, Canada.

Software

For downloading our open source software, please refer to the successor project, Open Mutation Miner (OMM).