Second-Order Text Matching Algorithm for Agricultural Text

被引:0
|
作者
Sun, Xiaoyang [1 ]
Song, Yunsheng [1 ,2 ]
Huang, Jianing [1 ]
机构
[1] Shandong Agr Univ, Sch Informat Sci & Engn, Tai An 271018, Peoples R China
[2] Minist Agr & Rural Affairs, Key Lab Huang Huai Hai Smart Agr Technol, Tai An 271018, Peoples R China
来源
APPLIED SCIENCES-BASEL | 2024年 / 14卷 / 16期
关键词
natural language processing; deep learning; text matching; agriculture text;
D O I
10.3390/app14167012
中图分类号
O6 [化学];
学科分类号
0703 ;
摘要
Text matching promotes the research and application of deep understanding of text information, and it provides the basis for information retrieval, recommendation systems and natural language processing by exploring the similar structures in text data. Owning to the outstanding performance and automatically extract text features for the target, the methods based-pre-training models gradually become the mainstream. However, such models usually suffer from the disadvantages of slow retrieval speed and low running efficiency. On the other hand, previous text matching algorithms have mainly focused on horizontal domain research, and there are relatively few vertical domain algorithms for agricultural text, which need to be further investigated. To address this issue, a second-order text matching algorithm has been developed. This paper first obtains a large amount of text about typical agricultural crops and constructs a database by using web crawlers and querying relevant textbooks, etc. Then BM25 algorithm is used to generate a candidate set and BERT model is used to filter the optimal match based on the candidate set. Experiments have shown that the Precision@1 of this second-order algorithm can reach 88.34% on the dataset constructed in this paper, and the average time to match a piece of text is only 2.02 s. Compared with BERT model and BM25 algorithm, there is an increase of 8.81% and 13.73% in Precision@1 respectively. In terms of the average time required for matching a text, it is 55.2 s faster than BERT model and only 2 s slower than BM25 algorithm. It can improve the efficiency and accuracy of agricultural information retrieval, agricultural decision support, agricultural market analysis, etc., and promote the sustainable development of agriculture.
引用
下载
收藏
页数:20
相关论文
共 50 条
  • [21] Matching wildcards:: An algorithm -: A simple wildcard text-matching algorithm in a single while loop
    Krauss, Kirk J.
    DR DOBBS JOURNAL, 2008, 33 (09): : 37 - 39
  • [22] Method for Retrieving Digital Agricultural Text Information Based on Local Matching
    Song, Yue
    Wang, Minjuan
    Gao, Wanlin
    SYMMETRY-BASEL, 2020, 12 (07):
  • [23] Text and text history of the second Book of Esra
    Rösel, M
    ZEITSCHRIFT FUR DIE ALTTESTAMENTLICHE WISSENSCHAFT, 2005, 117 (01): : 144 - 145
  • [24] A UNIFICATION ALGORITHM FOR SECOND-ORDER MONADIC TERMS
    FARMER, WM
    ANNALS OF PURE AND APPLIED LOGIC, 1988, 39 (02) : 131 - 174
  • [25] Normalized table-matching algorithm as approach to text categorization
    Taeho Jo
    Soft Computing, 2015, 19 : 839 - 849
  • [26] An Optimal Algorithm for Matching String Patterns in Large Text Databases
    Kumar, K. S. M. V.
    Raju, S. Viswanadha
    Govardha, Ka.
    INTERNATIONAL JOURNAL OF COMPUTER SCIENCE AND NETWORK SECURITY, 2013, 13 (06): : 31 - 40
  • [27] A second-order parareal algorithm for fractional PDEs
    Wu, Shu-Lin
    JOURNAL OF COMPUTATIONAL PHYSICS, 2016, 307 : 280 - 290
  • [28] Efficient algorithm for second-order reliability analysis
    Der Kiureghian, Armen
    De Stefano, Mario
    Journal of Engineering Mechanics, 1991, 117 (12) : 2904 - 2923
  • [29] A Second-Order Proximal Algorithm for Consensus Optimization
    Wu, Xuyang
    Qu, Zhihai
    Lu, Jie
    IEEE TRANSACTIONS ON AUTOMATIC CONTROL, 2021, 66 (04) : 1864 - 1871
  • [30] A Second-Order Algorithm for the Distance Of A Point to An Epigraph
    Zhao, Wenhui
    Li, Ruan
    Gao, Yan
    2018 IEEE 4TH INTERNATIONAL CONFERENCE ON CONTROL SCIENCE AND SYSTEMS ENGINEERING (ICCSSE 2018), 2018, : 6 - 9