| 英文摘要 |
This paper proposes a Taiwanese-Mandarin neural machine translation system trained by all available Taiwanese corpora and translation datasets. The unit of translation can be either Chinese words or characters. Embedding will be either self-trained or pre-trained. And some strategies will be proposed to handle OOV output. The final Mandarin-to-Taiwanese outperforms all known Taiwanese MT systems. The best BLEU score evaluated on news articles is 75.02, and the best score on literature articles is 38.13. We also built the first Taiwanese-to-Mandarin NMT system in the world, which achieves BLEU scores of 73.38 and 35.32 on those two genres of articles. |