{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,23]],"date-time":"2026-02-23T20:12:24Z","timestamp":1771877544798,"version":"3.50.1"},"reference-count":0,"publisher":"World Scientific Pub Co Pte Lt","issue":"07n08","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Int. J. Artif. Intell. Tools"],"published-print":{"date-parts":[[2020,12]]},"abstract":"<jats:p> Neural Machine Translation (NMT) model has become the mainstream technology in machine translation. The supervised neural machine translation model trains with abundant of sentence-level parallel corpora. But for low-resources language or dialect with no such corpus available, it is difficult to achieve good performance. Researchers began to focus on unsupervised neural machine translation (UNMT) that monolingual corpus as training data. UNMT need to construct the language model (LM) which learns semantic information from the monolingual corpus. This paper focuses on the pre-training of LM in unsupervised machine translation and proposes a pre-training method, NER-MLM (named entity recognition masked language model). Through performing NER, the proposed method can obtain better semantic information and language model parameters with better training results. In the unsupervised machine translation task, the BLEU scores on the WMT\u201916 English\u2013French, English\u2013German, data sets are 35.30, 27.30 respectively. To the best of our knowledge, this is the highest results in the field of UNMT reported so far. <\/jats:p>","DOI":"10.1142\/s0218213020400217","type":"journal-article","created":{"date-parts":[[2020,11,30]],"date-time":"2020-11-30T04:24:36Z","timestamp":1606710276000},"page":"2040021","source":"Crossref","is-referenced-by-count":10,"title":["Language Model Pre-training Method in Machine Translation Based on Named Entity Recognition"],"prefix":"10.1142","volume":"29","author":[{"given":"Zhen","family":"Li","sequence":"first","affiliation":[{"name":"Information System Engineering College, PLA Strategic Support Force Information Engineering University, 93 Hightech Zone, Zhengzhou, 450000, China"}]},{"given":"Dan","family":"Qu","sequence":"additional","affiliation":[{"name":"Information System Engineering College, PLA Strategic Support Force Information Engineering University, 93 Hightech Zone, Zhengzhou, 450000, China"}]},{"given":"Chaojie","family":"Xie","sequence":"additional","affiliation":[{"name":"Zhengzhou Xinda Institute of Advanced Technology, 93 Hightech Zone, Zhengzhou, 450000, China"}]},{"given":"Wenlin","family":"Zhang","sequence":"additional","affiliation":[{"name":"Information System Engineering College, PLA Strategic Support Force Information Engineering University, 93 Hightech Zone, Zhengzhou, 450000, China"}]},{"given":"Yanxia","family":"Li","sequence":"additional","affiliation":[{"name":"Foreign Languages College, PLA Strategic Support Force Information Engineering University, 93 Hightech Zone, Zhengzhou, 450000, China"}]}],"member":"219","published-online":{"date-parts":[[2020,11,30]]},"container-title":["International Journal on Artificial Intelligence Tools"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.worldscientific.com\/doi\/pdf\/10.1142\/S0218213020400217","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2020,11,30]],"date-time":"2020-11-30T04:24:55Z","timestamp":1606710295000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.worldscientific.com\/doi\/abs\/10.1142\/S0218213020400217"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,11,30]]},"references-count":0,"journal-issue":{"issue":"07n08","published-print":{"date-parts":[[2020,12]]}},"alternative-id":["10.1142\/S0218213020400217"],"URL":"https:\/\/doi.org\/10.1142\/s0218213020400217","relation":{},"ISSN":["0218-2130","1793-6349"],"issn-type":[{"value":"0218-2130","type":"print"},{"value":"1793-6349","type":"electronic"}],"subject":[],"published":{"date-parts":[[2020,11,30]]}}}