☆ 4.2 Article

Arabic Named Entity Recognition: A Feature-Driven Study

IEEE TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING (2009)

Journal

IEEE TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING

Volume 17, Issue 5, Pages 926-934

Publisher

IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC

DOI: 10.1109/TASL.2009.2019927

Keywords

Arabic; machine learning comparison; named entity recognition; natural language processing (NLP)

Funding

MCyT [TIN2006-15265-C06-04, AECID-PCI B/017961/08]
Defense Advanced Research Projects Agency (DARPA) [HR0011-06-C-0023]

Ask authors/readers for more resources

Protocol

Community support

Reagent

Community support

Abstract

The Named Entity Recognition task aims at identifying and classifying Named Entities within an open-domain text. This task has been garnering significant attention recently as it has been shown to help improve the performance of many Natural Language Processing applications. In this paper, we investigate the impact of using different sets of features in three discriminative machine learning frameworks, namely, support vector machines, maximum entropy and conditional random fields for the task of Named Entity Recognition. Our language of interest is Arabic. We explore lexical, contextual and morphological features and nine data-sets of different genres and annotations. We measure the impact of the different features in isolation and incrementally combine them in order to evaluate the robustness to noise of each approach. We achieve the highest performance using a combination of 15 features in conditional random fields using Broadcast News data (F(beta=1) = 83.34).

Arabic Named Entity Recognition: A Feature-Driven Study

Journal

IEEE TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING

Publisher

IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC

Keywords

Categories

Funding

Ask authors/readers for more resources

Protocol

Reagent

Authors

I am an author on this paper

Reviews

Primary Rating

Secondary Ratings

Novelty

Significance

Scientific rigor

Rate this paper

Recommended

Arabic Named Entity Recognition: A Feature-Driven Study

Journal

IEEE TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING

Publisher

IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC

Keywords

Categories

Funding

Ask authors/readers for more resources

Protocol

Reagent

Authors

I am an author on this paper

Reviews

Primary Rating

Secondary Ratings

Novelty

Significance

Scientific rigor

Rate this paper

Recommended

Export Citation

Share Paper