Cascaded classifiers for confidence-based chemical named entity recognition

15 years 2 months ago

Download www.aclweb.org

Chemical named entities represent an important facet of biomedical text. We have developed a system to use character-based ngrams, Maximum Entropy Markov Models and rescoring to recognise chemical names and other such entities, and to make confidence estimates for the extracted entities. An adjustable threshold allows the system to be tuned to high precision or high recall. At a threshold set for balanced precision and recall, we were able to extract named entities at an F score of 80.7% from chemistry papers and om PubMed abstracts. Furthermore, we were able to achieve 57.6% and 60.3% recall at 95% precision, and 58.9% and 49.1% precision at 90% recall. These results show that chemical named entities can be extracted with good performance, and that the properties of the extraction can be tuned to suit the demands of the task.

Peter Corbett, Ann A. Copestake

Real-time Traffic

BMCBI 2008 | Chemical | Entities | Maximum Entropy Markov |

claim paper

Added	09 Dec 2010
Updated	09 Dec 2010
Type	Journal
Year	2008
Where	BMCBI
Authors	Peter Corbett, Ann A. Copestake

Sciweavers

Cascaded classifiers for confidence-based chemical named entity recognition

BMCBI 2008 | Chemical | Entities | Maximum Entropy Markov |

Explore & Download

Productivity Tools

Document Tools

Image Tools

Sciweavers