{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2022,3,29]],"date-time":"2022-03-29T22:04:21Z","timestamp":1648591461640},"reference-count":7,"publisher":"Walter de Gruyter GmbH","issue":"1","license":[{"start":{"date-parts":[[2017,5,24]],"date-time":"2017-05-24T00:00:00Z","timestamp":1495584000000},"content-version":"unspecified","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by-nc-nd\/4.0"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2017,5,24]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p> In both Chinese and Dzongkha languages, the greatest challenge is to identify the word boundaries because there are no word delimiters as it is in English and other Western languages. Therefore, preprocessing and word segmentation is the first step in Dzongkha language processing, such as translation, spell-checking, and information retrieval. Research on Chinese word segmentation was conducted long time ago. Therefore, it is relatively mature, but the Dzongkha word segmentation has been less studied by researchers. In the paper, we have investigated this major problem in Dzongkha language processing using a probabilistic approach for selecting valid segments with probability being computed on the basis of the corpus.<\/jats:p>","DOI":"10.1515\/acss-2017-0008","type":"journal-article","created":{"date-parts":[[2017,6,13]],"date-time":"2017-06-13T10:01:20Z","timestamp":1497348080000},"page":"61-65","source":"Crossref","is-referenced-by-count":1,"title":["Analysing the Methods of Dzongkha Word Segmentation"],"prefix":"10.1515","volume":"21","author":[{"given":"Parshu Ram","family":"Dhungyel","sequence":"first","affiliation":[{"name":"College of Science and Technology, Royal University of Bhutan, Thimpu , Bhutan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"J\u0101nis","family":"Grundspe\u0146\u0137is","sequence":"additional","affiliation":[{"name":"Riga Technical University, Riga , Latvia"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"374","published-online":{"date-parts":[[2017,6,13]]},"reference":[{"key":"2021040803592340250_j_acss-2017-0008_ref_001_w2aab2b8b8b1b7b1ab1ab1Aa","unstructured":"[1] G. van Driem, \u201cLanguage Policy in Bhutan,\u201d presented at conference \u201cBhutan: a traditional order and the forces of change\u201d, London, UK, March 1993."},{"key":"2021040803592340250_j_acss-2017-0008_ref_002_w2aab2b8b8b1b7b1ab1ab2Aa","unstructured":"[2] D. C. Chhoeden et al., \u201cDzongkha Text-to-Speech Synthesis System - Phase II,\u201d in Conference on Human Language Technology for Development, May 2-5, 2011, Alexandria, Egypt, pp. 148-153."},{"key":"2021040803592340250_j_acss-2017-0008_ref_003_w2aab2b8b8b1b7b1ab1ab3Aa","unstructured":"[3] S. Norbu et al., \u201cDzongkha Word Segmentation,\u201d in 8th Workshop on Asian Language Resources, August 21-22, 2010, Beijing, China, pp. 95-102."},{"key":"2021040803592340250_j_acss-2017-0008_ref_004_w2aab2b8b8b1b7b1ab1ab4Aa","unstructured":"[4] H. Liu et al., \u201cTibetan Word Segmentation as Syllable Tagging Using Conditional Random Field,\u201d in 25th Pacific Asia Conference on Language, Information and Computation, December 16-18, 2011, Singapore, pp. 168-177."},{"key":"2021040803592340250_j_acss-2017-0008_ref_005_w2aab2b8b8b1b7b1ab1ab5Aa","unstructured":"[5] C. Chungku, J. Rabgay, and P. Choejey, \u201cDzongkha Text Corpus,\u201d in Conference on Human Language Technology for Development, May 2-5, 2011, Alexandria, Egypt, pp. 34-38."},{"key":"2021040803592340250_j_acss-2017-0008_ref_006_w2aab2b8b8b1b7b1ab1ab6Aa","doi-asserted-by":"crossref","unstructured":"[6] G. Andrew, \u201cA Hybrid Markov\/Semi-Markov Conditional Random Field for Sequence Segmentation,\u201d in Conference on Empirical Methods in Natural Language Processing, July 22-23, 2006, Sydney, Australia, pp. 465-472.","DOI":"10.3115\/1610075.1610140"},{"key":"2021040803592340250_j_acss-2017-0008_ref_007_w2aab2b8b8b1b7b1ab1ab7Aa","unstructured":"[7] C. Chungku, J. Rabgay, and G. Faa\u00df, \u201cBuilding NLP resources for Dzongkha:A Tagset and A Tagged Corpus,\u201d in 8th Workshop on Asian Language Resources, 21-22 August 2010, Beijing, China, pp. 103-110."}],"container-title":["Applied Computer Systems"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/content.sciendo.com\/view\/journals\/acss\/21\/1\/article-p61.xml","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.sciendo.com\/article\/10.1515\/acss-2017-0008","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2021,4,8]],"date-time":"2021-04-08T13:06:07Z","timestamp":1617887167000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.sciendo.com\/article\/10.1515\/acss-2017-0008"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2017,5,24]]},"references-count":7,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2017,6,13]]},"published-print":{"date-parts":[[2017,5,24]]}},"alternative-id":["10.1515\/acss-2017-0008"],"URL":"https:\/\/doi.org\/10.1515\/acss-2017-0008","relation":{},"ISSN":["2255-8691"],"issn-type":[{"value":"2255-8691","type":"electronic"}],"subject":[],"published":{"date-parts":[[2017,5,24]]}}}