An Adaptive Harmony Search Part-of-Speech tagger for Square Hmong Corpus

被引:1
|
作者
Kang, Di -Wen [1 ]
Ye, Shao-Qiang [2 ,4 ]
Ahmad, Sharifah Zarith Rahmah Syed [2 ]
Mo, Li-Ping [3 ]
Qin, Feng [2 ]
Zhou, Pan [1 ]
机构
[1] Jishou Univ, Sch Commun & Elect Engn, Jishou 416000, Peoples R China
[2] Univ Teknol Malaysia, Fac Comp, Skudai 80310, Johor, Malaysia
[3] Jishou Univ, Coll Comp Sci & Engn, Jishou, Hunan, Peoples R China
[4] Hunan Appl Technol Univ, Coll Informat & Engn, Changde 415000, Hunan, Peoples R China
关键词
Harmony Search Algorithm; Low-resource language; Optimization; Part-of-Speech tagging; Unknown words; ALGORITHM;
D O I
10.21123/bsj.2024.9694
中图分类号
O [数理科学和化学]; P [天文学、地球科学]; Q [生物科学]; N [自然科学总论];
学科分类号
07 ; 0710 ; 09 ;
摘要
Data -driven models perform poorly on part -of -speech tagging problems with the square Hmong language, a low -resource corpus. This paper designs a weight evaluation function to reduce the influence of unknown words. It proposes an improved harmony search algorithm utilizing the roulette and local evaluation strategies for handling the square Hmong part -of -speech tagging problem. The experiment shows that the average accuracy of the proposed model is 6%, 8% more than HMM and BiLSTM-CRF models, respectively. Meanwhile, the average F1 of the proposed model is also 6%, 3% more than HMM and BiLSTM-CRF models, respectively.
引用
收藏
页码:622 / 632
页数:11
相关论文
共 50 条
  • [21] Developing a robust part-of-speech tagger for biomedical text
    Tsuruoka, Y
    Tateishi, Y
    Kim, JD
    Ohta, T
    McNaught, J
    Ananiadou, S
    Tsujii, J
    ADVANCES IN INFORMATICS, PROCEEDINGS, 2005, 3746 : 382 - 392
  • [22] Part-of-Speech Tagger for Malay Social Media Texts
    Ariffin, Siti Noor Allia Noor
    Tiun, Sabrina
    GEMA ONLINE JOURNAL OF LANGUAGE STUDIES, 2018, 18 (04): : 124 - 142
  • [23] Part-of-Speech Tagger Based on Maximum Entropy Model
    Huang Heyan
    Zhang Xiaofei
    2009 2ND IEEE INTERNATIONAL CONFERENCE ON COMPUTER SCIENCE AND INFORMATION TECHNOLOGY, VOL 3, 2009, : 26 - 29
  • [24] A morphology-system and part-of-speech tagger for German
    Lezius, W
    Rapp, R
    Wettler, M
    NATURAL LANGUAGE PROCESSING AND SPEECH TECHNOLOGY: RESULTS OF THE 3RD KONVENS CONFERENCE, 1996, : 369 - 378
  • [25] Adding Morphological Information to a Connectionist Part-Of-Speech Tagger
    Zamora-Martinez, Francisco
    Jose Castro-Bleda, Maria
    Espana-Boquera, Salvador
    Tortajada-Velert, Salvador
    CURRENT TOPICS IN ARTIFICIAL INTELLIGENCE, 2010, 5988 : 191 - +
  • [26] Corpus based part-of-speech tagging
    Lv, Chengyao
    Liu, Huihua
    Dong, Yuanxing
    Chen, Yunliang
    INTERNATIONAL JOURNAL OF SPEECH TECHNOLOGY, 2016, 19 (03) : 647 - 654
  • [27] Building an Indonesian Rule-Based Part-of-Speech Tagger
    Rashel, Fam
    Luthfi, Andry
    Dinakaramani, Arawinda
    Manurung, Ruli
    PROCEEDINGS OF THE 2014 INTERNATIONAL CONFERENCE ON ASIAN LANGUAGE PROCESSING (IALP 2014), 2014, : 70 - 73
  • [28] Bayesian reinforcement for a probabilistic neural net Part-of-Speech tagger
    Maragoudakis, M
    Ganchev, T
    Fakotakis, N
    TEXT, SPEECH AND DIALOGUE, PROCEEDINGS, 2004, 3206 : 137 - 145
  • [29] A Supervised Part-Of-Speech Tagger for the Greek Language of the Social Web
    Nikiforos, Maria Nefeli
    Kermanidis, Katia Lida
    PROCEEDINGS OF THE 12TH INTERNATIONAL CONFERENCE ON LANGUAGE RESOURCES AND EVALUATION (LREC 2020), 2020, : 3861 - 3867
  • [30] Arabic part-of-speech tagger based support vectors machines
    Yousif, Jabar Hassan
    Sembok, Tengku Mohd Tengku
    INTERNATIONAL SYMPOSIUM OF INFORMATION TECHNOLOGY 2008, VOLS 1-4, PROCEEDINGS: COGNITIVE INFORMATICS: BRIDGING NATURAL AND ARTIFICIAL KNOWLEDGE, 2008, : 2084 - +