Arabic Part of Speech Tagging

被引:0
|
作者
Mohamed, Emad [1 ]
Kuebler, Sandra [1 ]
机构
[1] Indiana Univ, Dept Linguist, Bloomington, IN 47405 USA
关键词
D O I
暂无
中图分类号
H [语言、文字];
学科分类号
05 ;
摘要
Arabic is a morphologically rich language, which presents a challenge for part of speech tagging. In this paper, we compare two novel methods for POS tagging of Arabic without the use of gold standard word segmentation but with the full POS tagset of the Penn Arabic Treebank. The first approach uses complex tags that describe full words and does not require any word segmentation. The second approach is segmentation-based, using a machine learning segmenter. In this approach, the words are first segmented, then the segments are annotated with POS tags. Because of the word-based approach, we evaluate full word accuracy rather than segment accuracy. Word-based POS tagging yields better results than segment-based tagging (93.93% vs. 93.41%). Word based tagging also gives the best results on known words, the segmentation-based approach gives better results on unknown words. Combining both methods results in a word accuracy of 94.37%, which is very close to the result obtained by using gold standard segmentation (94.91%).
引用
收藏
页码:2537 / 2543
页数:7
相关论文
共 50 条
  • [41] Block and memory based part of speech tagging
    Computer Human Interaction and Intelligent Information Processing Laboratory, Institute of Software, Chinese Acad. of Sci., Beijing 100080, China
    [J]. Jilin Daxue Xuebao (Gongxueban), 2006, 4 (560-563):
  • [42] Part-of-Speech Tagging for Azerbaijani Language
    Mammadov, Samir
    Rustamov, Samir
    Mustafali, Ali
    Sadigov, Ziyaddin
    Mollayev, Rasim
    Mammadov, Zamir
    [J]. 2018 IEEE 12TH INTERNATIONAL CONFERENCE ON APPLICATION OF INFORMATION AND COMMUNICATION TECHNOLOGIES (AICT), 2018, : 40 - 45
  • [43] Part-of-Speech Tagging by Latent Analogy
    Bellegarda, Jerome R.
    [J]. IEEE JOURNAL OF SELECTED TOPICS IN SIGNAL PROCESSING, 2010, 4 (06) : 985 - 993
  • [44] Corpus based part-of-speech tagging
    Lv, Chengyao
    Liu, Huihua
    Dong, Yuanxing
    Chen, Yunliang
    [J]. INTERNATIONAL JOURNAL OF SPEECH TECHNOLOGY, 2016, 19 (03) : 647 - 654
  • [45] Domain adaptation in part-of-speech tagging
    Institute of Exact and Natural Sciences, Federal University of Pará , Pará, Brazil
    不详
    [J]. Emerging Applic. of Nat. Lang. Proc.: Concepts and New Res., (52-72):
  • [46] Part of Speech Tagging with Naive Bayes Methods
    Cretulescu, R.
    David, A.
    Morariu, D.
    Vintan, L.
    [J]. 2014 18TH INTERNATIONAL CONFERENCE SYSTEM THEORY, CONTROL AND COMPUTING (ICSTCC), 2014, : 446 - 451
  • [47] Part-of-speech tagging without training
    Bressan, S
    Indradjaja, LS
    [J]. INTELLIGENCE IN COMMUNICATION SYSTEMS, 2004, 3283 : 112 - 119
  • [48] Voted approach for part of speech tagging in Bengali
    Department of Computational Linguistics, University of Heidelberg, 1m Neuenheimer Feld 325, 69120 Heidelberg, Germany
    不详
    不详
    [J]. PACLIC 23 - Proc. 23rd Pacific Asia Conf. Lang. Inf. Comput., 2009, (120-129):
  • [49] Semi-supervised Part-of-speech Tagging in Speech Applications
    Dufour, Richard
    Favre, Benoit
    [J]. 11TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION 2010 (INTERSPEECH 2010), VOLS 1-2, 2010, : 1373 - 1376
  • [50] Toward An Efficient Arabic Part of Speech Tagger
    Abdelali, Ahmed
    Elhadj, Yahya O. Mohamed
    Bouziane, Rachid
    [J]. 2013 ACS INTERNATIONAL CONFERENCE ON COMPUTER SYSTEMS AND APPLICATIONS (AICCSA), 2013,