{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,4]],"date-time":"2026-05-04T00:27:33Z","timestamp":1777854453157,"version":"3.51.4"},"reference-count":30,"publisher":"SAGE Publications","issue":"3","license":[{"start":{"date-parts":[[2012,3,15]],"date-time":"2012-03-15T00:00:00Z","timestamp":1331769600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/journals.sagepub.com\/page\/policies\/text-and-data-mining-license"}],"content-domain":{"domain":["journals.sagepub.com"],"crossmark-restriction":true},"short-container-title":["Journal of Information Science"],"published-print":{"date-parts":[[2012,6]]},"abstract":"<jats:p>A vector space model (VSM) composed of selected important features is a common way to represent documents, including patent documents. Patent documents have some special characteristics that make it difficult to apply traditional feature selection methods directly: (a) it is difficult to find common terms for patent documents in different categories; and (b) the class label of a patent document is hierarchical rather than flat. Hence, in this article we propose a new approach that includes a hierarchical feature selection (HFS) algorithm which can be used to select more representative features with greater discriminative ability to present a set of patent documents with hierarchical class labels. The performance of the proposed method is evaluated through application to two documents sets with 2400 and 9600 patent documents, where we extract candidate terms from their titles and abstracts. The experimental results reveal that a VSM whose features are selected by a proportional selection process gives better coverage, while a VSM whose features are selected with a weighted-summed selection process gives higher accuracy.<\/jats:p>","DOI":"10.1177\/0165551512437635","type":"journal-article","created":{"date-parts":[[2012,3,15]],"date-time":"2012-03-15T20:27:35Z","timestamp":1331843255000},"page":"222-233","update-policy":"https:\/\/doi.org\/10.1177\/sage-journals-update-policy","source":"Crossref","is-referenced-by-count":8,"title":["Vector space model for patent documents with hierarchical class labels"],"prefix":"10.1177","volume":"38","author":[{"given":"Yen-Liang","family":"Chen","sequence":"first","affiliation":[{"name":"National Central University, ROC"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yu-Ting","family":"Chiu","sequence":"additional","affiliation":[{"name":"National Central University, ROC"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"179","published-online":{"date-parts":[[2012,3,15]]},"reference":[{"key":"bibr1-0165551512437635","unstructured":"Baeza-Yates R, Ribeiro-Neto B. Modern information retrieval. Wokingham, UK: Addison-Wesley, 1999, pp. 163\u2013190."},{"key":"bibr2-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1017\/CBO9780511809071.014"},{"key":"bibr3-0165551512437635","unstructured":"Tan PN, Steinbach M, Kumar V. Introduction to data mining. USA: Addison Wesley, 2005, pp. 19\u201395."},{"key":"bibr4-0165551512437635","unstructured":"WIPO. \u2018Frequently Asked Questions about the International Patent Classification (IPC): What is the IPC?\u2019, http:\/\/www.wipo.int\/classifications\/ipc\/en\/faq\/index.html#G1 (2010, accessed May 2011)."},{"key":"bibr5-0165551512437635","doi-asserted-by":"publisher","DOI":"10.4018\/978-1-59904-373-9.ch012"},{"key":"bibr6-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1016\/0306-4573(88)90021-0"},{"key":"bibr7-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1016\/0306-4573(89)90100-3"},{"issue":"1","key":"bibr8-0165551512437635","doi-asserted-by":"crossref","first-page":"19","DOI":"10.21248\/jlcl.20.2005.68","volume":"20","author":"Hotho A","year":"2005","journal-title":"LDV-Forum GLDV Journal for Computational Linguistics and Language Technology"},{"key":"bibr9-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2007.190740"},{"key":"bibr10-0165551512437635","volume-title":"The text mining handbook: advanced approaches in analyzing unstructured data","author":"Feldman R","year":"2007"},{"key":"bibr11-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2006.11.011"},{"key":"bibr12-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1109\/FSKD.2010.5569326"},{"key":"bibr13-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1145\/361219.361220"},{"key":"bibr14-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1145\/312624.312682"},{"key":"bibr15-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1016\/j.knosys.2009.11.010"},{"key":"bibr16-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2010.06.001"},{"key":"bibr17-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1108\/02635570810847608"},{"key":"bibr18-0165551512437635","volume-title":"Introduction to modern information retrieval","author":"Salton G","year":"1983"},{"key":"bibr19-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1145\/175235.175243"},{"issue":"2","key":"bibr20-0165551512437635","first-page":"248","volume":"8","author":"Lai KK","year":"2006","journal-title":"The Journal of American Academy of Business"},{"key":"bibr21-0165551512437635","unstructured":"Wikipedia, United States Patent and Trademark Office, 30 September 2011, http:\/\/en.wikipedia.org\/wiki\/USPTO."},{"key":"bibr22-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2006.11.006"},{"key":"bibr23-0165551512437635","first-page":"87","volume-title":"Working notes for the AAAI-98 workshop on learning for text categorization","author":"Larkey LS","year":"1998"},{"key":"bibr24-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1145\/313238.313304"},{"key":"bibr25-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1145\/945546.945547"},{"key":"bibr26-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2007.02.002"},{"key":"bibr27-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2006.01.013"},{"issue":"4","key":"bibr28-0165551512437635","first-page":"413","volume":"23","author":"Trappey AJC","year":"2006","journal-title":"Journal of Management"},{"key":"bibr29-0165551512437635","unstructured":"Han J, Kamber M, Pei J.Data mining: concepts and techniques. 2nd ed. San Francisco, CA: Morgan Kaufmann, 2006, pp. 89\u201390."},{"key":"bibr30-0165551512437635","doi-asserted-by":"publisher","DOI":"10.1177\/0165551510368620"}],"container-title":["Journal of Information Science"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/0165551512437635","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/full-xml\/10.1177\/0165551512437635","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/0165551512437635","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,29]],"date-time":"2026-04-29T23:08:19Z","timestamp":1777504099000},"score":1,"resource":{"primary":{"URL":"https:\/\/journals.sagepub.com\/doi\/10.1177\/0165551512437635"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2012,3,15]]},"references-count":30,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2012,6]]}},"alternative-id":["10.1177\/0165551512437635"],"URL":"https:\/\/doi.org\/10.1177\/0165551512437635","relation":{},"ISSN":["0165-5515","1741-6485"],"issn-type":[{"value":"0165-5515","type":"print"},{"value":"1741-6485","type":"electronic"}],"subject":[],"published":{"date-parts":[[2012,3,15]]}}}