Keeping Models Consistent between Pretraining and Translation for Low-Resource Neural Machine Translation

被引：3

作者：

Zhang, Wenbo ^{[1
,2
,3
]}

Li, Xiao ^{[1
,2
,3
]}

Yang, Yating ^{[1
,2
,3
]}

Dong, Rui ^{[1
,2
,3
]}

Luo, Gongxu ^{[1
,2
,3
]}

机构：

[1] Chinese Acad Sci, Xinjiang Tech Inst Phys & Chem, Urumqi 830011, Peoples R China

[2] Univ Chinese Acad Sci, Beijing 100049, Peoples R China

[3] Xinjiang Lab Minor Speech & Language Informat Pro, Urumqi 830011, Peoples R China

来源：

FUTURE INTERNET | 2020年 / 12卷 / 12期

基金：

中国国家自然科学基金;

关键词：

low-resource neural machine translation; monolingual data; pretraining; transformer;

D O I：

10.3390/fi12120215

中图分类号：

TP [自动化技术、计算机技术];

学科分类号：

0812 ;

摘要：

Recently, the pretraining of models has been successfully applied to unsupervised and semi-supervised neural machine translation. A cross-lingual language model uses a pretrained masked language model to initialize the encoder and decoder of the translation model, which greatly improves the translation quality. However, because of a mismatch in the number of layers, the pretrained model can only initialize part of the decoder's parameters. In this paper, we use a layer-wise coordination transformer and a consistent pretraining translation transformer instead of a vanilla transformer as the translation model. The former has only an encoder, and the latter has an encoder and a decoder, but the encoder and decoder have exactly the same parameters. Both models can guarantee that all parameters in the translation model can be initialized by the pretrained model. Experiments on the Chinese-English and English-German datasets show that compared with the vanilla transformer baseline, our models achieve better performance with fewer parameters when the parallel corpus is small.

引用

页码：1 / 13

页数：13

共 50 条

[21] The neural machine translation models for the low-resource Kazakh-English language pair
Karyukin, Vladislav
Rakhimova, Diana
Karibayeva, Aidana
Turganbayeva, Aliya
Turarbek, Asem
PEERJ COMPUTER SCIENCE, 2023, 9
[22] Revisiting Back-Translation for Low-Resource Machine Translation Between Chinese and Vietnamese
Li, Hongzheng
Sha, Jiu
Shi, Can
IEEE ACCESS, 2020, 8 (08) : 119931 - 119939
[23] A Joint Back-Translation and Transfer Learning Method for Low-Resource Neural Machine Translation
Luo, Gong-Xu
Yang, Ya-Ting
Dong, Rui
Chen, Yan-Hong
Zhang, Wen-Bo
MATHEMATICAL PROBLEMS IN ENGINEERING, 2020, 2020 (2020)
[24] Rethinking the Exploitation of Monolingual Data for Low-Resource Neural Machine Translation
Pang, Jianhui
Yang, Baosong
Wong, Derek Fai
Wan, Yu
Liu, Dayiheng
Chao, Lidia Sam
Xie, Jun
COMPUTATIONAL LINGUISTICS, 2023, 50 (01) : 25 - 47
[25] A Diverse Data Augmentation Strategy for Low-Resource Neural Machine Translation
Li, Yu
Li, Xiao
Yang, Yating
Dong, Rui
INFORMATION, 2020, 11 (05)
[26] Semantic Perception-Oriented Low-Resource Neural Machine Translation
Wu, Nier
Hou, Hongxu
Li, Haoran
Chang, Xin
Jia, Xiaoning
MACHINE TRANSLATION, CCMT 2021, 2021, 1464 : 51 - 62
[27] An Analysis of Massively Multilingual Neural Machine Translation for Low-Resource Languages
Mueller, Aaron
Nicolai, Garrett
McCarthy, Arya D.
Lewis, Dylan
Wu, Winston
Yarowsky, David
PROCEEDINGS OF THE 12TH INTERNATIONAL CONFERENCE ON LANGUAGE RESOURCES AND EVALUATION (LREC 2020), 2020, : 3710 - 3718
[28] Incremental Domain Adaptation for Neural Machine Translation in Low-Resource Settings
Kalimuthu, Marimuthu
Barz, Michael
Sonntag, Daniel
FOURTH ARABIC NATURAL LANGUAGE PROCESSING WORKSHOP (WANLP 2019), 2019, : 1 - 10
[29] Benchmarking Neural and Statistical Machine Translation on Low-Resource African Languages
Duh, Kevin
McNamee, Paul
Post, Matt
Thompson, Brian
PROCEEDINGS OF THE 12TH INTERNATIONAL CONFERENCE ON LANGUAGE RESOURCES AND EVALUATION (LREC 2020), 2020, : 2667 - 2675
[30] Towards a Low-Resource Neural Machine Translation for Indigenous Languages in Canada
Ngoc Tan Le
Sadat, Fatiha
TRAITEMENT AUTOMATIQUE DES LANGUES, 2021, 62 (03): : 39 - 63

← 1 2 3 4 5 →