CoType: Joint Extraction of Typed Entities and Relations with Knowledge Bases

被引:198
|
作者
Ren, Xiang [1 ]
Wu, Zeqiu [1 ]
He, Wenqi [1 ]
Qu, Meng [1 ]
Voss, Clare R. [2 ]
Ji, Heng [3 ]
Abdelzaher, Tarek F. [1 ]
Han, Jiawei [1 ]
机构
[1] Univ Illinois, Urbana, IL 61801 USA
[2] Army Res Lab, Computat & Informat Sci Directorate, Adelphi, MD USA
[3] Rensselaer Polytech Inst, Comp Sci Dept, Troy, NY USA
基金
美国国家科学基金会;
关键词
D O I
10.1145/3038912.3052708
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Extracting entities and relations for types of interest from text is important for understanding massive text corpora. Traditionally, systems of entity relation extraction have relied on human-annotated corpora for training and adopted an incremental pipeline. Such systems require additional human expertise to be ported to a new domain, and are vulnerable to errors cascading down the pipeline. In this paper, we investigate joint extraction of typed entities and relations with labeled data heuristically obtained from knowledge bases (i.e., distant supervision). As our algorithm for type labeling via distant supervision is context-agnostic, noisy training data poses unique challenges for the task. We propose a novel domain-independent framework, called COTYPE, that runs a data-driven text segmentation algorithm to extract entity mentions, and jointly embeds entity mentions, relation mentions, text features and type labels into two low-dimensional spaces (for entity and relation mentions respectively), where, in each space, objects whose types are close will also have similar representations. COTYPE, then using these learned embeddings, estimates the types of test (unlinkable) mentions. We formulate a joint optimization problem to learn embeddings from text corpora and knowledge bases, adopting a novel partial-label loss function for noisy labeled data and introducing an object "translation" function to capture the cross-constraints of entities and relations on each other. Experiments on three public datasets demonstrate the effectiveness of COTYPE across different domains (e.g., news, biomedical), with an average of 25% improvement in F1 score compared to the next best method.
引用
收藏
页码:1015 / 1024
页数:10
相关论文
共 50 条
  • [1] Reinforcement Learning for Joint Extraction of Entities and Relations
    Liu, Wenpeng
    Cao, Yanan
    Liu, Yanbing
    Hu, Yue
    Tan, Jianlong
    ARTIFICIAL NEURAL NETWORKS AND MACHINE LEARNING - ICANN 2018, PT II, 2018, 11140 : 263 - 272
  • [2] Joint Extraction of Entities and Relations in the News Domain
    Li, Zeyu
    Ma, Haoyang
    Lv, Yingying
    Shen, Hao
    ADVANCED DATA MINING AND APPLICATIONS (ADMA 2022), PT I, 2022, 13725 : 81 - 92
  • [3] A Decoupling and Aggregating Framework for Joint Extraction of Entities and Relations
    Wang, Yao
    Liu, Xin
    Kong, Weikun
    Yu, Hai-Tao
    Racharak, Teeradaj
    Kim, Kyoung-Sook
    Nguyen, Le Minh
    IEEE ACCESS, 2024, 12 : 103313 - 103328
  • [4] Sequence to sequence learning for joint extraction of entities and relations
    Liang, Zeyu
    Du, Junping
    Neurocomputing, 2022, 501 : 480 - 488
  • [5] Sequence to sequence learning for joint extraction of entities and relations
    Liang, Zeyu
    Du, Junping
    NEUROCOMPUTING, 2022, 501 : 480 - 488
  • [6] Investigating LSTMs for Joint Extraction of Opinion Entities and Relations
    Katiyar, Arzoo
    Cardie, Claire
    PROCEEDINGS OF THE 54TH ANNUAL MEETING OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS, VOL 1, 2016, : 919 - 929
  • [7] A Text-Generated Method to Joint Extraction of Entities and Relations
    E, Haihong
    Xiao, Siqi
    Song, Meina
    APPLIED SCIENCES-BASEL, 2019, 9 (18):
  • [8] Joint Extraction of Entities and Relations Based on a Novel Graph Scheme
    Wang, Shaolei
    Zhang, Yue
    Che, Wanxiang
    Liu, Ting
    PROCEEDINGS OF THE TWENTY-SEVENTH INTERNATIONAL JOINT CONFERENCE ON ARTIFICIAL INTELLIGENCE, 2018, : 4461 - 4467
  • [9] A joint extraction model of entities and relations based on relation decomposition
    Gao, Chen
    Zhang, Xuan
    Liu, Hui
    Yun, Wei
    Jiang, Jiahao
    INTERNATIONAL JOURNAL OF MACHINE LEARNING AND CYBERNETICS, 2022, 13 (07) : 1833 - 1845
  • [10] A joint extraction model of entities and relations based on relation decomposition
    Chen Gao
    Xuan Zhang
    Hui Liu
    Wei Yun
    Jiahao Jiang
    International Journal of Machine Learning and Cybernetics, 2022, 13 : 1833 - 1845