Learning various length dependence by dual recurrent neural networks

被引：3

作者：

Zhang, Chenpeng ^{[1
]}

Li, Shuai ^{[2
]}

Ye, Mao ^{[1
]}

Zhu, Ce ^{[2
]}

Li, Xue ^{[3
]}

机构：

[1] Univ Elect Sci & Technol China, Sch Comp Sci & Engn, Chengdu 611731, Peoples R China

[2] Univ Elect Sci & Technol China, Sch Informat & Commun Engn, Chengdu 611731, Peoples R China

[3] Univ Queensland, Sch Informat Technol & Elect Engn, Brisbane, Qld 4072, Australia

来源：

NEUROCOMPUTING | 2021年 / 466卷

基金：

中国国家自然科学基金; 国家重点研发计划;

关键词：

Sequence learning; Recurrent neural networks; Long-term; Dependence separating;

D O I：

10.1016/j.neucom.2021.09.043

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

Recurrent neural networks (RNNs) are widely used as a memory model for sequence-related problems. Many variants of RNN have been proposed to solve the gradient problems of training RNNs and process long sequences. Although some classical models have been proposed, capturing long-term dependence while responding to short-term changes remains a challenge. To address this problem, we propose a new model named Dual Recurrent Neural Networks (DuRNN). The DuRNN consists of two parts to learn the short-term dependence and progressively learn the long-term dependence. The first part is a recurrent neural network with constrained full recurrent connections to deal with short-term dependence in sequence and generate short-term memory. Another part is a recurrent neural network with independent recurrent connections which helps to learn long-term dependence and generate long-term memory. A selection mechanism is added between two parts to transfer the needed long-term information to the independent neurons. Multiple modules can be stacked to form a multi-layer model for better performance. Our contributions are: 1) a new recurrent model developed based on the divide-and-conquer strategy to learn long and short-term dependence separately, and 2) a selection mechanism to enhance the separating and learning of different temporal scales of dependence. Both theoretical analysis and extensive experiments are conducted to validate the performance of our model. Experimental results indicate that the proposed DuRNN model can handle not only very long sequences (over 5,000 time steps), but also short sequences very well. (c) 2021 Elsevier B.V. All rights reserved.

引用

页码：1 / 15

页数：15

共 50 条

[1] Learning Contextual Dependence With Convolutional Hierarchical Recurrent Neural Networks
Zuo, Zhen
Shuai, Bing
Wang, Gang
Liu, Xiao
Wang, Xingxing
Wang, Bing
Chen, Yushi
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2016, 25 (07) : 2983 - 2996
[2] LEARNING IN RECURRENT NEURAL NETWORKS
WHITE, H
MATHEMATICAL SOCIAL SCIENCES, 1991, 22 (01) : 102 - 103
[3] Dual recurrent neural networks using partial linear dependence for multivariate time series
Park, Hyungjin
Lee, Geonseok
Lee, Kichun
EXPERT SYSTEMS WITH APPLICATIONS, 2022, 208
[4] Minimum Description Length Recurrent Neural Networks
Lan, Nur
Geyer, Michal
Chemla, Emmanuel
Katzir, Roni
TRANSACTIONS OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS, 2022, 10 : 785 - 799
[5] Recurrent Neural Networks With Finite Memory Length
Long, Dingkun
Zhang, Richong
Mao, Yongyi
IEEE ACCESS, 2019, 7 : 12511 - 12520
[6] Learning Queuing Networks by Recurrent Neural Networks
Garbi, Giulio
Incerto, Emilio
Tribastone, Mirco
PROCEEDINGS OF THE ACM/SPEC INTERNATIONAL CONFERENCE ON PERFORMANCE ENGINEERING (ICPE'20), 2020, : 56 - 66
[7] Learning with recurrent neural networks - Conclusion
不详
LEARNING WITH RECURRENT NEURAL NETWORKS, 2000, 254 : 133 - 135
[8] Bayesian learning for recurrent neural networks
Crucianu, M
Boné, R
de Beauville, JPA
NEUROCOMPUTING, 2001, 36 (01) : 235 - 242
[9] Learning with recurrent neural networks - Introduction
Hammer, B
LEARNING WITH RECURRENT NEURAL NETWORKS, 2000, 254 : 1 - +
[10] LEARNING COMPACT RECURRENT NEURAL NETWORKS
Lu, Zhiyun
Sindhwani, Vikas
Sainath, Tara N.
2016 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING PROCEEDINGS, 2016, : 5960 - 5964

← 1 2 3 4 5 →