跳到主要导航 跳到搜索 跳到主要内容

Multi-channel encoder for neural machine translation

  • Baidu Inc

科研成果: 书/报告/会议事项章节会议稿件同行评审

摘要

Attention-based Encoder-Decoder has the effective architecture for neural machine translation (NMT), which typically relies on recurrent neural networks (RNN) to build the blocks that will be lately called by attentive reader during the decoding process. This design of encoder yields relatively uniform composition on source sentence, despite the gating mechanism employed in encoding RNN. On the other hand, we often hope the decoder to take pieces of source sentence at varying levels suiting its own linguistic structure: for example, we may want to take the entity name in its raw form while taking an idiom as a perfectly composed unit. Motivated by this demand, we propose Multi-channel Encoder (MCE), which enhances encoding components with different levels of composition. More specifically, in addition to the hidden state of encoding RNN, MCE takes 1) the original word embedding for raw encoding with no composition, and 2) a particular design of external memory in Neural Turing Machine (NTM) for more complex composition, while all three encoding strategies are properly blended during decoding. Empirical study on Chinese-English translation shows that our model can improve by 6.52 BLEU points upon a strong open source NMT system: DL4MT 1 . On the WMT14 English-French task, our single shallow system achieves BLEU=38.8, comparable with the state-of-the-art deep models.

源语言英语
主期刊名32nd AAAI Conference on Artificial Intelligence, AAAI 2018
出版商AAAI press
4962-4969
页数8
ISBN(电子版)9781577358008
出版状态已出版 - 2018
已对外发布
活动32nd AAAI Conference on Artificial Intelligence, AAAI 2018 - New Orleans, 美国
期限: 2 2月 20187 2月 2018

出版系列

姓名32nd AAAI Conference on Artificial Intelligence, AAAI 2018

会议

会议32nd AAAI Conference on Artificial Intelligence, AAAI 2018
国家/地区美国
New Orleans
时期2/02/187/02/18

学术指纹

探究 'Multi-channel encoder for neural machine translation' 的科研主题。它们共同构成独一无二的学术指纹。

引用此