Scientific text citation analysis using CNN features and ensemble learning model

被引:0
|
作者
Alnowaiser, Khaled [1 ]
机构
[1] Prince Sattam Bin Abdulaziz Univ, Coll Comp Engn & Sci, Dept Comp Engn, Al Kharj, Saudi Arabia
来源
PLOS ONE | 2024年 / 19卷 / 05期
关键词
BIBLIOMETRICS; CONTEXT; COUNTS; INDEX;
D O I
10.1371/journal.pone.0302304
中图分类号
O [数理科学和化学]; P [天文学、地球科学]; Q [生物科学]; N [自然科学总论];
学科分类号
07 ; 0710 ; 09 ;
摘要
Citation illustrates the link between citing and cited documents. Different aspects of achievements like the journal's impact factor, author's ranking, and peers' judgment are analyzed using citations. However, citations are given the same weight for determining these important metrics. However academics contend that not all citations can ever have equal weight. Predominantly, such rankings are based on quantitative measures and the qualitative aspect is completely ignored. For a fair evaluation, qualitative evaluation of citations is needed in addition to quantitative ones. Many existing works that use qualitative evaluation consider binary class and categorize citations as important or unimportant. This study considers multi-class tasks for citation sentiments on imbalanced data and presents a novel framework for sentiment analysis in in-text citations of research articles. In the proposed technique, features are retrieved using a convolutional neural network (CNN), and classification is performed using a voting classifier that combines Logistic Regression (LR) and Stochastic Gradient Descent (SGD). The class imbalance problem is handled by the synthetic minority oversampling technique (SMOTE). Extensive experiments are performed in comparison with the proposed approach using SMOTE-generated data and machine learning models by term frequency (TF), and term frequency-inverse document frequency (TF-IDF) to evaluate the efficacy of the proposed approach for citation analysis. It is found that the proposed voting classifier using CNN features achieves an accuracy, precision, recall, and F1 score of 0.99 for all. This work not only advances the field of sentiment analysis in academic citations but also underscores the importance of incorporating qualitative aspects in evaluating the impact and sentiments conveyed through citations.
引用
收藏
页数:19
相关论文
共 50 条
  • [1] Sentiment Analysis for Arabic Text Using Ensemble Learning
    Al-Saqqa, Samar
    Obeid, Nadim
    Awajan, Arafat
    [J]. 2018 IEEE/ACS 15TH INTERNATIONAL CONFERENCE ON COMPUTER SYSTEMS AND APPLICATIONS (AICCSA), 2018,
  • [2] Predicting numeric ratings for Google apps using text features and ensemble learning
    Umer, Muhammad
    Ashraf, Imran
    Mehmood, Arif
    Ullah, Saleem
    Choi, Gyu Sang
    [J]. ETRI JOURNAL, 2021, 43 (01) : 95 - 108
  • [3] A scientific citation recommendation model integrating network and text representations
    Tianshuang Qiu
    Chuanming Yu
    Yunci Zhong
    Lu An
    Gang Li
    [J]. Scientometrics, 2021, 126 : 9199 - 9221
  • [4] A scientific citation recommendation model integrating network and text representations
    Qiu, Tianshuang
    Yu, Chuanming
    Zhong, Yunci
    An, Lu
    Li, Gang
    [J]. SCIENTOMETRICS, 2021, 126 (11) : 9199 - 9221
  • [5] DEEP ENSEMBLE LEARNING MODEL BASED ON COVARIANCE POOLING OF MULTI-LAYER CNN FEATURES
    Akodad, Sara
    Bombrun, Lionel
    Puscasu, Maria
    Xia, Junshi
    Germain, Christian
    Berthoumieu, Yannick
    [J]. 2022 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, ICIP, 2022, : 1081 - 1085
  • [6] Scientific papers citation analysis using textual features and SMOTE resampling techniques
    Umer, Muhammad
    Sadiq, Saima
    Missen, Malik Muhammad Saad
    Hameed, Zahid
    Aslam, Zahid
    Siddique, Muhammad Abubakar
    Nappi, Michele
    [J]. PATTERN RECOGNITION LETTERS, 2021, 150 : 250 - 257
  • [7] Scientific Text Sentiment Analysis using Machine Learning Techniques
    Raza, Hassan
    Faizan, M.
    Hamza, Ahsan
    Mushtaq, Ahmed
    Akhtar, Naeem
    [J]. INTERNATIONAL JOURNAL OF ADVANCED COMPUTER SCIENCE AND APPLICATIONS, 2019, 10 (12) : 157 - 165
  • [8] Measuring Scientific Knowledge Flows by Deploying Citation Context Analysis using Machine Learning Approach on PLoS ONE Full Text
    Saeed-Ul Hassan
    Akram, Anam
    Asghar, Awais
    Aljohani, Naif Radi
    [J]. 16TH INTERNATIONAL CONFERENCE ON SCIENTOMETRICS & INFORMETRICS (ISSI 2017), 2017, : 322 - 333
  • [9] An ensemble learning model for driver drowsiness detection and accident prevention using the behavioral features analysis
    Sharanabasappa
    Nandyal, Suvarna
    [J]. INTERNATIONAL JOURNAL OF INTELLIGENT COMPUTING AND CYBERNETICS, 2022, 15 (02) : 224 - 244
  • [10] Sentiment analysis using a deep ensemble learning model
    Başarslan, Muhammet Sinan
    Kayaalp, Fatih
    [J]. Multimedia Tools and Applications, 2024, 83 (14) : 42207 - 42231