{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,12,21]],"date-time":"2025-12-21T06:24:31Z","timestamp":1766298271947,"version":"3.41.0"},"reference-count":42,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2018,2,13]],"date-time":"2018-02-13T00:00:00Z","timestamp":1518480000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc-sa\/4.0\/"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Asian Low-Resour. Lang. Inf. Process."],"published-print":{"date-parts":[[2018,9,30]]},"abstract":"<jats:p>We propose a new method for inducing a phrase-based translation model from a pair of unrelated monolingual corpora. Our method is able to deal with phrases of arbitrary length and to find phrase pairs that are useful for statistical machine translation, without requiring large parallel or comparable corpora. First, our method generates phrase pairs through coupling source and target phrases separately collected from respective monolingual data. Then, for each phrase pair, we compute features using the monolingual data and a small quantity of parallel sentences. Finally, incorrect phrase pairs are pruned, and a phrase table is made using the remaining phrase pairs. In our experiments on French--Japanese and Spanish--Japanese translation tasks under low-resource conditions, we observe that incorporating a phrase table induced by our method to the machine translation system leads to large improvements in translation quality. Furthermore, we show that a phrase table induced by our method can also be useful in a wide range of configurations, including configurations where we have already access to large parallel corpora and configurations where only small monolingual corpora are available.<\/jats:p>","DOI":"10.1145\/3168054","type":"journal-article","created":{"date-parts":[[2018,2,13]],"date-time":"2018-02-13T15:40:40Z","timestamp":1518536440000},"page":"1-25","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Phrase Table Induction Using Monolingual Data for Low-Resource Statistical Machine Translation"],"prefix":"10.1145","volume":"17","author":[{"given":"Benjamin","family":"Marie","sequence":"first","affiliation":[{"name":"National Institute of Information and Communications Technology, Kyoto, Japan"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3799-8428","authenticated-orcid":false,"given":"Atsushi","family":"Fujita","sequence":"additional","affiliation":[{"name":"National Institute of Information and Communications Technology, Kyoto, Japan"}]}],"member":"320","published-online":{"date-parts":[[2018,2,13]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D16-1250"},{"volume-title":"Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201912)","year":"2012","author":"Cherry Colin","key":"e_1_2_1_2_1"},{"volume-title":"Proceedings of the International Conference on Language Resources and Evaluation (LREC\u201916)","year":"2016","author":"Chu Chenhui","key":"e_1_2_1_3_1"},{"volume-title":"Proceedings of the Annual Conference of the Association for Computational Linguistics: Human Language Technologies (ACL-HLT\u201911)","author":"Clark Jonathan H.","key":"e_1_2_1_4_1"},{"volume-title":"Proceedings of the Conference of the Association for Computational Linguistics (ACL\u201907)","year":"2007","author":"Cohn Trevor","key":"e_1_2_1_5_1"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D15-1131"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D17-1148"},{"volume-title":"Proceedings of the Annual Conference of the Association for Computational Linguistics: Human Language Technologies (ACL-HLT\u201911)","year":"2011","author":"Jagadeesh Jagarlamudi Hal","key":"e_1_2_1_8_1"},{"volume-title":"Proceedings of the International Joint Conferences on Artificial Intelligence Organization (IJCAI\u201915)","year":"2015","author":"Dong Meiping","key":"e_1_2_1_9_1"},{"volume-title":"Proceedings of the Conference on Empirical Methods in Natural Language Processing and the Conference on Natural Language Learning (EMNLP-CoNLL\u201912)","year":"2012","author":"Dou Qing","key":"e_1_2_1_10_1"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D16-1136"},{"volume-title":"Proceedings of the 3rd Workshop on Very Large Corpora.","year":"1995","author":"Fung Pascale","key":"e_1_2_1_12_1"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.3115\/1118037.1118046"},{"volume-title":"Proceedings of the Conference of the Association for Computational Linguistics: Human Language Technologies (ACL-HLT\u201908)","author":"Haghighi A.","key":"e_1_2_1_14_1"},{"volume-title":"Proceedings of the International Conference on Language Resources and Evaluation (LREC\u201916)","year":"2016","author":"Han Jingyi","key":"e_1_2_1_15_1"},{"volume-title":"Proceedings of the Conference of the Association for Computational Linguistics (ACL\u201913)","year":"2013","author":"Heafield Kenneth","key":"e_1_2_1_16_1"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1017\/S1351324916000139"},{"volume-title":"Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201913)","year":"2013","author":"Irvine Ann","key":"e_1_2_1_18_1"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.3115\/v1\/W14-1617"},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1017\/S1351324916000127"},{"volume-title":"Proceedings of Conference on Empirical Methods in Natural Language Processing and the Conference on Natural Language Learning (EMNLP-CoNLL\u201907)","year":"2007","author":"Johnson Howard","key":"e_1_2_1_21_1"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.5555\/2380816.2380835"},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.5555\/1557769.1557821"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.3115\/1118627.1118629"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/W17-3204"},{"volume-title":"Exploiting similarities among languages for machine translation. CoRR abs\/1309.4168","year":"2013","author":"Mikolov Tomas","key":"e_1_2_1_26_1"},{"volume-title":"Proceedings of the Annual Conference on Neural Information Processing Systems (NIPS\u201913)","year":"2013","author":"Mikolov Tomas","key":"e_1_2_1_27_1"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.1551-6709.2010.01106.x"},{"volume-title":"Proceedings of the Conference of the Association for Computational Linguistics (ACL\u201912)","year":"2012","author":"Nuhn Malte","key":"e_1_2_1_29_1"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.3115\/1073083.1073135"},{"volume-title":"Proceedings of the International Conference on Computational Linguistics (COLING\u201916)","year":"2016","author":"Passban Peyman","key":"e_1_2_1_31_1"},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.3115\/981658.981709"},{"volume-title":"Proceedings of the Annual Conference of the Association for Computational Linguistics: Human Language Technologies (ACL-HLT\u201911)","year":"2011","author":"Ravi Sujith","key":"e_1_2_1_33_1"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.3115\/v1\/P14-1064"},{"volume-title":"Proceedings of the Conference of the Association for Computational Linguistics (ACL\u201913)","author":"Socher Richard","key":"e_1_2_1_35_1"},{"volume-title":"Proceedings of the Conference on Empirical Methods in Natural Language Processing (EMNLP\u201913)","year":"2013","author":"Socher Richard","key":"e_1_2_1_36_1"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P16-1157"},{"volume-title":"Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201907)","year":"2007","author":"Utiyama Masao","key":"e_1_2_1_38_1"},{"volume-title":"Proceedings of the Conference of the Association for Computational Linguistics (ACL\u201916)","year":"2016","author":"Vuli\u0107 Ivan","key":"e_1_2_1_39_1"},{"volume-title":"Proceedings of the Conference of the Association for Computational Linguistics (ACL\u201915)","year":"2015","author":"Vuli\u0107 Ivan","key":"e_1_2_1_40_1"},{"key":"e_1_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.5555\/3013558.3013583"},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.3115\/v1\/N15-1176"}],"container-title":["ACM Transactions on Asian and Low-Resource Language Information Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3168054","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3168054","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T02:26:08Z","timestamp":1750213568000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3168054"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,2,13]]},"references-count":42,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2018,9,30]]}},"alternative-id":["10.1145\/3168054"],"URL":"https:\/\/doi.org\/10.1145\/3168054","relation":{},"ISSN":["2375-4699","2375-4702"],"issn-type":[{"type":"print","value":"2375-4699"},{"type":"electronic","value":"2375-4702"}],"subject":[],"published":{"date-parts":[[2018,2,13]]},"assertion":[{"value":"2017-06-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2017-11-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2018-02-13","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}