{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,8,2]],"date-time":"2025-08-02T18:08:44Z","timestamp":1754158124360,"version":"3.41.2"},"reference-count":46,"publisher":"Emerald","issue":"4","license":[{"start":{"date-parts":[[2009,8,7]],"date-time":"2009-08-07T00:00:00Z","timestamp":1249603200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.emerald.com\/insight\/site-policies"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2009,8,7]]},"abstract":"<jats:sec><jats:title content-type=\"abstract-heading\">Purpose<\/jats:title><jats:p>The purpose of this paper is to develop a new summarisation approach, namely structure\u2010preserving and query\u2010biased summarisation, to improve the effectiveness of web searching. During web searching, one aid for users is the document summaries provided in the search results. However, the summaries provided by current search engines have limitations in directing users to relevant documents.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Design\/methodology\/approach<\/jats:title><jats:p>The proposed system consists of two stages: document structure analysis and summarisation. In the first stage, a rule\u2010based approach is used to identify the sectional hierarchies of web documents. In the second stage, query\u2010biased summaries are created, making use of document structure both in the summarisation process and in the output summaries.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Findings<\/jats:title><jats:p>In structural processing, about 70 per cent accuracy in identifying document sectional hierarchies is obtained. The summarisation method is tested on a task\u2010based evaluation method using English and Turkish document collections. The results show that the proposed method is a significant improvement over both unstructured query\u2010biased summaries and Google snippets in terms of f\u2010measure.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Practical implications<\/jats:title><jats:p>The proposed summarisation system can be incorporated into search engines. The structural processing technique also has applications in other information systems, such as browsing, outlining and indexing documents.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Originality\/value<\/jats:title><jats:p>In the literature on summarisation, the effects of query\u2010biased techniques and document structure are considered in only a few works and are researched separately. The research reported here differs from traditional approaches by combining these two aspects in a coherent framework. The work is also the first automatic summarisation study for Turkish targeting web search.<\/jats:p><\/jats:sec>","DOI":"10.1108\/14684520910985684","type":"journal-article","created":{"date-parts":[[2009,10,5]],"date-time":"2009-10-05T10:28:02Z","timestamp":1254738482000},"page":"696-719","source":"Crossref","is-referenced-by-count":4,"title":["Structure\u2010preserving and query\u2010biased document summarisation for web searching"],"prefix":"10.1108","volume":"33","author":[{"given":"F.","family":"Canan Pembe","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tunga","family":"G\u00fcng\u00f6r","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"140","reference":[{"key":"key2022012420061005600_b1","unstructured":"Alam, H., Kumar, A., Nakamura, M., Rahman, F., Tarnikova, Y. and Wilcox, C. (2003), \u201cStructured and unstructured document summarisation: design of a commercial summariser using lexical chains\u201d, Proceedings of the Seventh International Conference on Document Analysis and Recognition, Edinburgh, pp. 1147\u20101152."},{"key":"key2022012420061005600_b2","doi-asserted-by":"crossref","unstructured":"Amini, M.R., Tombros, A., Usunier, N. and Lalmas, M. (2007), \u201cLearning based summarisation of XML documents\u201d, Journal of Information Retrieval, Vol. 10 No. 3, pp. 233\u201055.","DOI":"10.1007\/s10791-006-9017-1"},{"key":"key2022012420061005600_b3","unstructured":"Baeza\u2010Yates, R. and Ribeiro\u2010Neto, B. (1999), Modern Information Retrieval, Addison\u2010Wesley, New York, NY."},{"key":"key2022012420061005600_b4","unstructured":"Barzilay, R. and Elhadad, M. (1999), \u201cUsing lexical chains for text summarisation\u201d, in Mani, I. and Maybury, M.T. (Eds), Advances in Automatic Text Summarisation, MIT Press, Cambridge, MA, pp. 111\u201021."},{"key":"key2022012420061005600_b5","unstructured":"Berry, M.W. and Browne, M. (1999), Understanding Search Engines: Mathematical Modeling and Text Retrieval, SIAM, Philadelphia, PA."},{"key":"key2022012420061005600_b6","doi-asserted-by":"crossref","unstructured":"Broder, A. and Henzinger, M. (2002), \u201cAlgorithmic aspects of information retrieval on the web\u201d, in Abello, J., Pardalos, P.M. and Resende, M.G.C. (Eds), Handbook of Massive Data Sets, Kluwer Academic, Dordrecht.","DOI":"10.1007\/978-1-4615-0005-6_1"},{"key":"key2022012420061005600_b7","doi-asserted-by":"crossref","unstructured":"Buyukkokten, O., Garcia\u2010Molina, H. and Paepcke, A. (2001), \u201cAccordion summarisation for end\u2010game browsing on PDAs and cellular phones\u201d, Proceedings of the SIGCHI Conference on Human Factors in Computing Systems, Seattle, WA, pp. 213\u2010220.","DOI":"10.1145\/365024.365102"},{"key":"key2022012420061005600_b8","doi-asserted-by":"crossref","unstructured":"Can, F., Kocberber, S., Balcik, E., Kaynak, C., Ocalan, H.C. and Vursavas, O.M. (2008), \u201cInformation retrieval on Turkish texts\u201d, Journal of the American Society for Information Science and Technology, Vol. 59 No. 3, pp. 407\u201021.","DOI":"10.1002\/asi.20750"},{"key":"key2022012420061005600_b9","doi-asserted-by":"crossref","unstructured":"Chaudhuri, B.B. (2006), Digital Document Processing: Major Directions and Recent Advances, Springer, London.","DOI":"10.1007\/978-1-84628-726-8"},{"key":"key2022012420061005600_b10","doi-asserted-by":"crossref","unstructured":"Chen, Y., Ma, W.Y. and Zhang, H.J. (2003), \u201cDetecting web page structure for adaptive viewing on small form factor devices\u201d, Proceedings of the 12th International World Wide Web Conference, Budapest, pp. 225\u2010233.","DOI":"10.1145\/775152.775184"},{"key":"key2022012420061005600_b11","unstructured":"Chowdhury, G.G. (Ed.) (2006), Introduction to Modern Information Retrieval, Facet, London."},{"key":"key2022012420061005600_b12","unstructured":"Chung, C.Y., Gertz, M. and Sundarsan, N. (2002), \u201cReverse engineering for web data: from visual to semantic structures\u201d, Proceedings of the 18th International Conference on Data Engineering, San Jose, CA, pp. 53\u201063."},{"key":"key2022012420061005600_b13","doi-asserted-by":"crossref","unstructured":"Feng, J., Haffner, P. and Gilbert, M. (2005), \u201cA learning approach to discovering web page semantic structures\u201d, Proceedings of the Eighth International Conference on Document Analysis and Recognition, Seoul, pp. 1055\u20101059.","DOI":"10.1109\/ICDAR.2005.19"},{"key":"key2022012420061005600_b14","doi-asserted-by":"crossref","unstructured":"Grossman, D.A. and Frieder, O. (2004), Information Retrieval: Algorithms and Heuristics, Springer, Dordrecht.","DOI":"10.1007\/978-1-4020-3005-5"},{"key":"key2022012420061005600_b15","doi-asserted-by":"crossref","unstructured":"Guo, Y. and Stylios, G. (2005), \u201cAn intelligent summarisation system based on cognitive psychology\u201d, Information Sciences, Vol. 174 Nos 1\u20102, pp. 1\u201036.","DOI":"10.1016\/j.ins.2004.08.004"},{"key":"key2022012420061005600_b16","doi-asserted-by":"crossref","unstructured":"Gupta, S., Kaiser, G.E., Grimm, P., Chiang, M.F. and Starren, J. (2005), \u201cAutomating content extraction of HTML documents\u201d, World Wide Web, Vol. 8 No. 2, pp. 179\u2010224.","DOI":"10.1007\/s11280-004-4873-3"},{"key":"key2022012420061005600_b17","doi-asserted-by":"crossref","unstructured":"Hobson, S.P., Dorr, B.J., Monz, C. and Schwartz, R. (2007), \u201cTask\u2010based evaluation of text summarisation using relevance prediction\u201d, Information Processing and Management, Vol. 43 No. 6, pp. 1482\u201099.","DOI":"10.1016\/j.ipm.2007.01.002"},{"key":"key2022012420061005600_b18","doi-asserted-by":"crossref","unstructured":"Ingwersen, P. and J\u00e4rvelin, K. (2005), The Turn: Integration of Information Seeking and Retrieval in Context, Springer, Dordrecht.","DOI":"10.1145\/1113343.1113351"},{"key":"key2022012420061005600_b19","doi-asserted-by":"crossref","unstructured":"Jackson, P. and Moulinier, I. (2007), Natural Language Processing for Online Applications: Text Retrieval, Extraction and Categorization, John Benjamins, Amsterdam.","DOI":"10.1075\/nlp.5"},{"key":"key2022012420061005600_b20","doi-asserted-by":"crossref","unstructured":"Jansen, B.J. and Spink, A. (2005), \u201cAn analysis of web searching by European AlltheWeb.com users\u201d, Information Processing and Management, Vol. 41 No. 2, pp. 361\u201081.","DOI":"10.1016\/S0306-4573(03)00067-0"},{"key":"key2022012420061005600_b21","doi-asserted-by":"crossref","unstructured":"Kowalski, G.J. and Maybury, M.T. (2002), Information Storage and Retrieval Systems: Theory and Implementation, Kluwer Academic, Boston, MA.","DOI":"10.1007\/b116174"},{"key":"key2022012420061005600_b22","doi-asserted-by":"crossref","unstructured":"Liang, S.F., Devlin, S. and Tait, J. (2007), \u201cInvestigating sentence weighting components for automatic summarisation\u201d, Information Processing and Management, Vol. 43 No. 1, pp. 146\u201053.","DOI":"10.1016\/j.ipm.2006.05.010"},{"key":"key2022012420061005600_b23","unstructured":"Mani, I. (2001), Automatic Summarisation, John Benjamins, Amsterdam."},{"key":"key2022012420061005600_b24","unstructured":"Mani, I. and Maybury, M.T. (Eds) (1999), Advances in Automatic Text Summarisation, MIT Press, Cambridge, MA."},{"key":"key2022012420061005600_b25","unstructured":"Marcu, D. (1999), \u201cDiscourse trees are good indicators of importance in text\u201d, in Mani, I. and Maybury, M.T. (Eds), Advances in Automatic Text Summarisation, MIT Press, Cambridge, MA, pp. 123\u201036."},{"key":"key2022012420061005600_b26","doi-asserted-by":"crossref","unstructured":"Marcu, D. (2000), The Theory and Practice of Discourse Parsing and Summarisation, MIT Press, Cambridge, MA.","DOI":"10.7551\/mitpress\/6754.001.0001"},{"key":"key2022012420061005600_b27","doi-asserted-by":"crossref","unstructured":"Markey, K. (2007), \u201cTwenty\u2010five years of end\u2010user searching, part 1: research findings\u201d, Journal of the American Society for Information Science and Technology, Vol. 58 No. 8, pp. 1071\u201081.","DOI":"10.1002\/asi.20462"},{"key":"key2022012420061005600_b28","doi-asserted-by":"crossref","unstructured":"Maynard, D., Bontcheva, K., Saggion, H., Cunningham, H. and Hamza, O. (2002), \u201cUsing a text engineering framework to build an extendable and portable IE\u2010based summarisation system\u201d, Proceedings of the ACL Workshop on Text Summarisation, Philadelphia, PA, pp. 19\u201026.","DOI":"10.3115\/1118162.1118165"},{"key":"key2022012420061005600_b29","unstructured":"Moens, M.\u2010F. (2002), Automatic Indexing and Abstracting of Document Texts, Kluwer Academic, Boston, MA."},{"key":"key2022012420061005600_b30","unstructured":"Montgomery, D.C. (2001), Design and Analysis of Experiments, John Wiley, New York, NY."},{"key":"key2022012420061005600_b31","unstructured":"Mukherjee, S., Yang, G., Tan, W. and Ramakrishnan, I.V. (2003), \u201cAutomatic discovery of semantic structures in HTML documents\u201d, Proceedings of the Seventh International Conference on Document Analysis and Recognition, Edinburgh, pp. 245\u2010249."},{"key":"key2022012420061005600_b32","unstructured":"Niyogi, D. and Srihari, N. (1995), \u201cKnowledge\u2010based derivation of document logical structure\u201d, Proceedings of the Third International Conference on Document Analysis and Recognition, Montreal, pp. 472\u2010475."},{"key":"key2022012420061005600_b34","doi-asserted-by":"crossref","unstructured":"Otterbacher, J., Radev, D. and Kareem, O. (2006), \u201cNews to go: hierarchical text summarisation for mobile devices\u201d, Proceedings of 29th Annual ACM SIGIR Conference on Research and Development in Information Retrieval, Seattle, WA, pp. 589\u2010596.","DOI":"10.1145\/1148170.1148271"},{"key":"key2022012420061005600_b35","doi-asserted-by":"crossref","unstructured":"Porter, M. (1980), \u201cAn algorithm for suffix stripping\u201d, Program, Vol. 14 No. 3, pp. 130\u20107.","DOI":"10.1108\/eb046814"},{"key":"key2022012420061005600_b36","unstructured":"Raggett, D., Le Hors, A. and Jacobs, I. (1999), HTML 4.01 specification, available at: www.w3.org\/TR\/html401\/ (accessed 15 February 2008)."},{"key":"key2022012420061005600_b37","unstructured":"Sparck\u2010Jones, K. (1999), \u201cAutomatic summarising: factors and directions\u201d, in Mani, I. and Maybury, M.T. (Eds), Advances in Automatic Text Summarisation, MIT Press, Cambridge, MA, pp. 1\u201012."},{"key":"key2022012420061005600_b38","doi-asserted-by":"crossref","unstructured":"Szl\u00e1vik, Z., Tombros, A. and Lalmas, M. (2006), \u201cInvestigating the use of summarisation for interactive XML retrieval\u201d, Proceedings of the 2006 ACM Symposium on Applied Computing, Dijon, pp. 1068\u20101072.","DOI":"10.1145\/1141277.1141529"},{"key":"key2022012420061005600_b39","doi-asserted-by":"crossref","unstructured":"Tombros, A. and Sanderson, M. (1998), \u201cAdvantages of query biased summaries in information retrieval\u201d, Proceedings of the 28th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval, Salvador, pp. 2\u201010.","DOI":"10.1145\/290941.290947"},{"key":"key2022012420061005600_b40","doi-asserted-by":"crossref","unstructured":"Vadrevu, S., Gelgi, F. and Davulcu, H. (2007), \u201cInformation extraction from web pages using presentation regularities and domain knowledge\u201d, World Wide Web, Vol. 10 No. 2, pp. 157\u201079.","DOI":"10.1007\/s11280-007-0021-1"},{"key":"key2022012420061005600_b33","unstructured":"Van Oostendorp, H. and De Mul, S. (1996), Cognitive Aspects of Electronic Text Processing, Ablex, Norwood, NJ."},{"key":"key2022012420061005600_b41","doi-asserted-by":"crossref","unstructured":"Varadarajan, R. and Hristidis, V. (2005), \u201cStructure\u2010based query\u2010specific document summarisation\u201d, Proceedings of the 14th ACM International Conference on Information and Knowledge Management, Bremen, pp. 231\u2010232.","DOI":"10.1145\/1099554.1099602"},{"key":"key2022012420061005600_b42","doi-asserted-by":"crossref","unstructured":"White, R.W., Jose, J.M. and Ruthven, I. (2003), \u201cA task\u2010oriented study on the influencing effects of query\u2010biased summarisation in web searching\u201d, Information Processing and Management, Vol. 39 No. 5, pp. 707\u201033.","DOI":"10.1016\/S0306-4573(02)00033-X"},{"key":"key2022012420061005600_b43","doi-asserted-by":"crossref","unstructured":"Xue, Y., Hu, Y., Xin, G., Song, R., Shi, S., Cao, Y., Lin, C.Y. and Li, H. (2007), \u201cWeb page title extraction and its application\u201d, Information Processing and Management, Vol. 43 No. 5, pp. 1332\u201047.","DOI":"10.1016\/j.ipm.2006.11.007"},{"key":"key2022012420061005600_b44","doi-asserted-by":"crossref","unstructured":"Yang, C.C. and Wang, F.L. (2008), \u201cHierarchical summarisation of large documents\u201d, Journal of the American Society for Information Science and Technology, Vol. 59 No. 6, pp. 887\u2010902.","DOI":"10.1002\/asi.20781"},{"key":"key2022012420061005600_b45","unstructured":"Yang, Y. and Zhang, H.J. (2001), \u201cHTML page analysis based on visual cues\u201d, Proceedings of the Sixth International Conference on Document Analysis and Recognition, Seattle, WA, pp. 859\u2010864."},{"key":"key2022012420061005600_b46","doi-asserted-by":"crossref","unstructured":"Yeh, J.Y., Ke, H.R., Yang, W.P. and Meng, I.H. (2005), \u201cText summarisation using a trainable summariser and latent semantic analysis\u201d, Information Processing and Management, Vol. 41 No. 1, pp. 75\u201095.","DOI":"10.1016\/j.ipm.2004.04.003"}],"container-title":["Online Information Review"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/www.emeraldinsight.com\/doi\/full-xml\/10.1108\/14684520910985684","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/14684520910985684\/full\/xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/14684520910985684\/full\/html","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,7,25]],"date-time":"2025-07-25T00:41:18Z","timestamp":1753404078000},"score":1,"resource":{"primary":{"URL":"http:\/\/www.emerald.com\/oir\/article\/33\/4\/696-719\/314682"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2009,8,7]]},"references-count":46,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2009,8,7]]}},"alternative-id":["10.1108\/14684520910985684"],"URL":"https:\/\/doi.org\/10.1108\/14684520910985684","relation":{},"ISSN":["1468-4527"],"issn-type":[{"type":"print","value":"1468-4527"}],"subject":[],"published":{"date-parts":[[2009,8,7]]}}}