Specializing Multi-domain NMT via Penalizing Low Mutual Information

Lee, Jiyoung, Kim, Hantae, Cho, Hyunchang, Choi, Edward, Park, Cheonbok

Oct-23-2022–arXiv.org Artificial Intelligence

Multi-domain Neural Machine Translation (NMT) trains a single model with multiple domains. It is appealing because of its efficacy in handling multiple domains within one model. An ideal multi-domain NMT should learn distinctive domain characteristics simultaneously, however, grasping the domain peculiarity is a non-trivial task. In this paper, we investigate domain-specific information through the lens of mutual information (MI) and propose a new objective that penalizes low MI to become higher. Our method achieved the state-of-the-art performance among the current competitive multi-domain NMT models. Also, we empirically show our objective promotes low MI to be higher resulting in domain-specialized multi-domain NMT.

artificial intelligence, machine translation, natural language, (2 more...)

arXiv.org Artificial Intelligence

Oct-23-2022

arXiv.org PDF

Add feedback

Genre:
- Research Report (0.40)

Technology:
- Information Technology > Artificial Intelligence > Natural Language > Machine Translation (0.87)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found