HiPool: Modeling Long Documents Using Graph Neural Networks

被引:0
|
作者
Li, Irene R. [1 ]
Feng, Aosong [2 ]
Radev, Dragomir [2 ]
Ying, Rex [2 ]
机构
[1] Univ Tokyo, Tokyo, Japan
[2] Yale Univ, New Haven, CT USA
关键词
D O I
暂无
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Encoding long sequences in Natural Language Processing (NLP) is a challenging problem. Though recent pretraining language models achieve satisfying performances in many NLP tasks, they are still restricted by a pre-defined maximum length, making them challenging to be extended to longer sequences. So some recent works utilize hierarchies to model long sequences. However, most of them apply sequential models for upper hierarchies, suffering from long dependency issues. In this paper, we alleviate these issues through a graph-based method. We first chunk the sequence with a fixed length to model the sentence-level information. We then leverage graphs to model intraand cross-sentence correlations with a new attention mechanism. Additionally, due to limited standard benchmarks for long document classification (LDC), we propose a new challenging benchmark, totaling six datasets with up to 53k samples and 4034 average tokens' length. Evaluation shows our model surpasses competitive baselines by 2.6% in F1 score, and 4.8% on the longest sequence dataset. Our method is shown to outperform hierarchical sequential models with better performance and scalability, especially for longer sequences.
引用
收藏
页码:161 / 171
页数:11
相关论文
共 50 条
  • [1] Modeling Graph Neural Networks and Dynamic Role Sorting for Argument Extraction in Documents
    Zhang, Qingchuan
    Chen, Hongxi
    Cai, Yuanyuan
    Dong, Wei
    Liu, Peng
    [J]. APPLIED SCIENCES-BASEL, 2023, 13 (16):
  • [2] Modeling TCP Performance using Graph Neural Networks
    Jaeger, Benedikt
    Helm, Max
    Schwegmann, Lars
    Carle, Georg
    [J]. PROCEEDINGS OF THE 1ST INTERNATIONAL WORKSHOP ON GRAPH NEURAL NETWORKING, GNNET 2022, 2022, : 18 - 23
  • [3] Graph Neural Networks for Metasurface Modeling
    Khoram, Erfan
    Wu, Zhicheng
    Qu, Yurui
    Zhou, Ming
    Yu, Zongfu
    [J]. ACS PHOTONICS, 2023, 10 (04): : 892 - 899
  • [4] Foundations and Modeling of Dynamic Networks Using Dynamic Graph Neural Networks: A Survey
    Skarding, Joakim
    Gabrys, Bogdan
    Musial, Katarzyna
    [J]. IEEE ACCESS, 2021, 9 : 79143 - 79168
  • [5] Dating Documents using Graph Convolution Networks
    Vashishth, Shikhar
    Dasgupta, Shib Sankar
    Ray, Swayambhu Nath
    Talukdar, Partha
    [J]. PROCEEDINGS OF THE 56TH ANNUAL MEETING OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS (ACL), VOL 1, 2018, : 1605 - 1615
  • [6] Using graph neural networks for wall modeling in compressible anisothermal flows
    Dupuy, Dorian
    Odier, Nicolas
    Lapeyre, Corentin
    [J]. DATA-CENTRIC ENGINEERING, 2024, 5
  • [7] Modeling Human Innate Immune Response Using Graph Neural Networks
    Henna, Shagufta
    [J]. IEEE ACCESS, 2021, 9 : 167117 - 167127
  • [8] Graph-based knowledge tracing: Modeling student proficiency using graph neural networks
    Nakagawa, Hiromi
    Iwasawa, Yusuke
    Matsuo, Yutaka
    [J]. WEB INTELLIGENCE, 2021, 19 (1-2) : 87 - 102
  • [9] Timing Macro Modeling with Graph Neural Networks
    Chang, Kevin Kai-Chun
    Chiang, Chun-Yao
    Lee, Pei-Yu
    Jiang, Iris Hui-Ru
    [J]. PROCEEDINGS OF THE 59TH ACM/IEEE DESIGN AUTOMATION CONFERENCE, DAC 2022, 2022, : 1219 - 1224
  • [10] Modeling IoT Equipment With Graph Neural Networks
    Zhang, Weishan
    Zhang, Yafei
    Xu, Liang
    Zhou, Jiehan
    Liu, Yan
    Guis, Mu
    Liu, Xin
    Yang, Su
    [J]. IEEE ACCESS, 2019, 7 : 32754 - 32764