Unknown

Dataset Information

0

A text-mining system for extracting metabolic reactions from full-text articles.


ABSTRACT: BACKGROUND:Increasingly biological text mining research is focusing on the extraction of complex relationships relevant to the construction and curation of biological networks and pathways. However, one important category of pathway - metabolic pathways - has been largely neglected.Here we present a relatively simple method for extracting metabolic reaction information from free text that scores different permutations of assigned entities (enzymes and metabolites) within a given sentence based on the presence and location of stemmed keywords. This method extends an approach that has proved effective in the context of the extraction of protein-protein interactions. RESULTS:When evaluated on a set of manually-curated metabolic pathways using standard performance criteria, our method performs surprisingly well. Precision and recall rates are comparable to those previously achieved for the well-known protein-protein interaction extraction task. CONCLUSIONS:We conclude that automated metabolic pathway construction is more tractable than has often been assumed, and that (as in the case of protein-protein interaction extraction) relatively simple text-mining approaches can prove surprisingly effective. It is hoped that these results will provide an impetus to further research and act as a useful benchmark for judging the performance of more sophisticated methods that are yet to be developed.

SUBMITTER: Czarnecki J 

PROVIDER: S-EPMC3475109 | biostudies-literature | 2012 Jul

REPOSITORIES: biostudies-literature

altmetric image

Publications

A text-mining system for extracting metabolic reactions from full-text articles.

Czarnecki Jan J   Nobeli Irene I   Smith Adrian M AM   Shepherd Adrian J AJ  

BMC bioinformatics 20120723


<h4>Background</h4>Increasingly biological text mining research is focusing on the extraction of complex relationships relevant to the construction and curation of biological networks and pathways. However, one important category of pathway - metabolic pathways - has been largely neglected.Here we present a relatively simple method for extracting metabolic reaction information from free text that scores different permutations of assigned entities (enzymes and metabolites) within a given sentence  ...[more]

Similar Datasets

| S-EPMC3667078 | biostudies-literature
| S-EPMC3534466 | biostudies-literature
| S-EPMC7251675 | biostudies-literature
| S-EPMC3441580 | biostudies-literature
| S-EPMC3179973 | biostudies-literature
| S-EPMC4874549 | biostudies-literature
| S-EPMC6602571 | biostudies-literature
| S-EPMC6825414 | biostudies-literature