4.6 Article

Large-scale biomedical concept recognition: an evaluation of current automatic annotators and their parameters

期刊

BMC BIOINFORMATICS
卷 15, 期 -, 页码 -

出版社

BMC
DOI: 10.1186/1471-2105-15-59

关键词

-

资金

  1. NIH [2T15LM009451]
  2. NSF [0965616]
  3. National ICT Australia (NICTA)
  4. Department of Broadband, Communications and the Digital Economy
  5. Australian Research Council through the ICT Centre of Excellence program
  6. Direct For Biological Sciences
  7. Div Of Biological Infrastructure [0965616] Funding Source: National Science Foundation

向作者/读者索取更多资源

Background: Ontological concepts are useful for many different biomedical tasks. Concepts are difficult to recognize in text due to a disconnect between what is captured in an ontology and how the concepts are expressed in text. There are many recognizers for specific ontologies, but a general approach for concept recognition is an open problem. Results: Three dictionary-based systems (MetaMap, NCBO Annotator, and ConceptMapper) are evaluated on eight biomedical ontologies in the Colorado Richly Annotated Full-Text (CRAFT) Corpus. Over 1,000 parameter combinations are examined, and best-performing parameters for each system-ontology pair are presented. Conclusions: Baselines for concept recognition by three systems on eight biomedical ontologies are established (F-measures range from 0.14-0.83). Out of the three systems we tested, ConceptMapper is generally the best-performing system; it produces the highest F-measure of seven out of eight ontologies. Default parameters are not ideal for most systems on most ontologies; by changing parameters F-measure can be increased by up to 0.4. Not only are best performing parameters presented, but suggestions for choosing the best parameters based on ontology characteristics are presented.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.6
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据