Predicting and using target length in neural machine translation

Z Yang, Y Gao, W Wang, H Ney�- …�of the 1st Conference of the Asia�…, 2020 - aclanthology.org
Proceedings of the 1st Conference of the Asia-Pacific Chapter of the�…, 2020aclanthology.org
Attention-based encoder-decoder models have achieved great success in neural machine
translation tasks. However, the lengths of the target sequences are not explicitly predicted in
these models. This work proposes length prediction as an auxiliary task and set up a sub-
network to obtain the length information from the encoder. Experimental results show that
the length prediction sub-network brings improvements over the strong baseline system and
that the predicted length can be used as an alternative to length normalization during�…
Abstract
Attention-based encoder-decoder models have achieved great success in neural machine translation tasks. However, the lengths of the target sequences are not explicitly predicted in these models. This work proposes length prediction as an auxiliary task and set up a sub-network to obtain the length information from the encoder. Experimental results show that the length prediction sub-network brings improvements over the strong baseline system and that the predicted length can be used as an alternative to length normalization during decoding.
aclanthology.org
Showing the best result for this search. See all results