HiPool: Modeling Long Documents Using Graph Neural Networks

被引:0
|
作者
Li, Irene R. [1 ]
Feng, Aosong [2 ]
Radev, Dragomir [2 ]
Ying, Rex [2 ]
机构
[1] Univ Tokyo, Tokyo, Japan
[2] Yale Univ, New Haven, CT USA
关键词
D O I
暂无
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Encoding long sequences in Natural Language Processing (NLP) is a challenging problem. Though recent pretraining language models achieve satisfying performances in many NLP tasks, they are still restricted by a pre-defined maximum length, making them challenging to be extended to longer sequences. So some recent works utilize hierarchies to model long sequences. However, most of them apply sequential models for upper hierarchies, suffering from long dependency issues. In this paper, we alleviate these issues through a graph-based method. We first chunk the sequence with a fixed length to model the sentence-level information. We then leverage graphs to model intraand cross-sentence correlations with a new attention mechanism. Additionally, due to limited standard benchmarks for long document classification (LDC), we propose a new challenging benchmark, totaling six datasets with up to 53k samples and 4034 average tokens' length. Evaluation shows our model surpasses competitive baselines by 2.6% in F1 score, and 4.8% on the longest sequence dataset. Our method is shown to outperform hierarchical sequential models with better performance and scalability, especially for longer sequences.
引用
下载
收藏
页码:161 / 171
页数:11
相关论文
共 50 条
  • [11] Graph Neural Networks and Representation Embedding for Table Extraction in PDF Documents
    Gemelli, Andrea
    Vivoli, Emanuele
    Marinai, Simone
    2022 26TH INTERNATIONAL CONFERENCE ON PATTERN RECOGNITION (ICPR), 2022, : 1719 - 1726
  • [12] Web documents categorization using neural networks
    Corrêa, RF
    Ludermir, TB
    NEURAL INFORMATION PROCESSING, 2004, 3316 : 758 - 762
  • [13] Three-Dimensional Structural Geological Modeling Using Graph Neural Networks
    Michael Hillier
    Florian Wellmann
    Boyan Brodaric
    Eric de Kemp
    Ernst Schetselaar
    Mathematical Geosciences, 2021, 53 : 1725 - 1749
  • [14] On Wavelet based Modeling of Neural Networks using Graph-theoretic Approach
    Bhosale, B.
    20TH INTERNATIONAL CONGRESS ON MODELLING AND SIMULATION (MODSIM2013), 2013, : 712 - 718
  • [15] Three-Dimensional Structural Geological Modeling Using Graph Neural Networks
    Hillier, Michael
    Wellmann, Florian
    Brodaric, Boyan
    de Kemp, Eric
    Schetselaar, Ernst
    MATHEMATICAL GEOSCIENCES, 2021, 53 (08) : 1725 - 1749
  • [16] Using Neural Networks to Detect Emotions in Documents
    Kumar, Keshav
    Mittra, Yash
    PROCEEDINGS OF SECOND INTERNATIONAL CONFERENCE ON ADVANCES IN COMPUTER ENGINEERING AND COMMUNICATION SYSTEMS, ICACECS 2021, 2022, : 191 - 198
  • [17] Learning Long-Term Dependencies Using Layered Graph Neural Networks
    Bandinelli, Niccolo
    Bianchini, Monica
    Scarselli, Franco
    2010 INTERNATIONAL JOINT CONFERENCE ON NEURAL NETWORKS IJCNN 2010, 2010,
  • [18] xNet: Modeling Network Performance With Graph Neural Networks
    Huang, Sijiang
    Wei, Yunze
    Peng, Lingfeng
    Wang, Mowei
    Hui, Linbo
    Liu, Peng
    Du, Zongpeng
    Liu, Zhenhua
    Cui, Yong
    IEEE-ACM TRANSACTIONS ON NETWORKING, 2024, 32 (02) : 1753 - 1767
  • [19] Passage Retrieval on Structured Documents Using Graph Attention Networks
    Albarede, Lucas
    Mulhem, Philippe
    Goeuriot, Lorraine
    Le Pape-Gardeux, Claude
    Marie, Sylvain
    Chardin-Segui, Trinidad
    ADVANCES IN INFORMATION RETRIEVAL, PT II, 2022, 13186 : 13 - 21
  • [20] Graph-based Recommendation using Graph Neural Networks
    Dossena, Marco
    Irwin, Christopher
    Portinale, Luigi
    2022 21ST IEEE INTERNATIONAL CONFERENCE ON MACHINE LEARNING AND APPLICATIONS, ICMLA, 2022, : 1769 - 1774