Header

UZH-Logo

Maintenance Infos

Detecting Protein-Protein Interactions in Biomedical Literature Using a Parser


Schneider, Gerold (2009). Detecting Protein-Protein Interactions in Biomedical Literature Using a Parser. In: Clematide, Simon; Klenner, Manfred; Volk, Martin. Searching Answers. Münster: MV Verlag, 109-118.

Abstract

We describe the task of automatically detecting interactions between proteins in biomedical literature. We use a syntactic parser, a corpus annotated for proteins, and manual decisions as training material. After automatically parsing the GENIA corpus, which is manually annotated for proteins, all syntactic paths between proteins are extracted. These syntactic paths are manually disambiguated between meaningful paths and irrelevant paths. Meaningful paths are paths that express an interaction between the syntactically connected proteins, irrelevant paths are paths that do not convey any interaction. The resource created by these manual decisions is used in two ways. First, words that appear frequently inside a meaningful path are learnt using simple machine learning. Second, these resources are applied to the task of automatically detecting interactions between proteins in biomedical literature.

Abstract

We describe the task of automatically detecting interactions between proteins in biomedical literature. We use a syntactic parser, a corpus annotated for proteins, and manual decisions as training material. After automatically parsing the GENIA corpus, which is manually annotated for proteins, all syntactic paths between proteins are extracted. These syntactic paths are manually disambiguated between meaningful paths and irrelevant paths. Meaningful paths are paths that express an interaction between the syntactically connected proteins, irrelevant paths are paths that do not convey any interaction. The resource created by these manual decisions is used in two ways. First, words that appear frequently inside a meaningful path are learnt using simple machine learning. Second, these resources are applied to the task of automatically detecting interactions between proteins in biomedical literature.

Statistics

Altmetrics

Downloads

91 downloads since deposited on 23 Dec 2009
4 downloads since 12 months
Detailed statistics

Additional indexing

Item Type:Book Section, not refereed, original work
Communities & Collections:06 Faculty of Arts > Institute of Computational Linguistics
06 Faculty of Arts > English Department
Dewey Decimal Classification:000 Computer science, knowledge & systems
820 English & Old English literatures
410 Linguistics
Uncontrolled Keywords:IR, NLP, text mining, parsing, biomedicine
Language:English
Date:2009
Deposited On:23 Dec 2009 13:29
Last Modified:14 Sep 2016 13:40
Publisher:MV Verlag
ISBN:978-3-642-00381-3
Funders:Swiss National Science Fund, Grant 100014-118396/1
Related URLs:http://www.recherche-portal.ch/primo_library/libweb/action/search.do?fn=search&mode=Advanced&vid=ZAD&vl%28186672378UI0%29=isbn&vl%281UI0%29=contains&vl%28freeText0%29=978-3-642-00381-3

Download

Download PDF  'Detecting Protein-Protein Interactions in Biomedical Literature Using a Parser'.
Preview
Filetype: PDF
Size: 1MB