{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,20]],"date-time":"2025-10-20T18:15:28Z","timestamp":1760984128833,"version":"3.41.2"},"reference-count":34,"publisher":"Emerald","issue":"4","license":[{"start":{"date-parts":[[2011,8,9]],"date-time":"2011-08-09T00:00:00Z","timestamp":1312848000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.emerald.com\/insight\/site-policies"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2011,8,9]]},"abstract":"<jats:sec><jats:title content-type=\"abstract-heading\">Purpose<\/jats:title><jats:p>Video summarisation is one of the most active fields in content\u2010based video retrieval research. A new video summarisation scheme is proposed by this paper based on socially generated temporal tags.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Design\/methodology\/approach<\/jats:title><jats:p>To capture users' collaborative tagging activities the proposed scheme maintains video bookmarks, which contain some temporal or positional information about videos, such as relative time codes or byte offsets. For each video all the video bookmarks collected from users are then statistically analysed in order to extract some meaningful key frames (the video equivalent of keywords), which collectively constitute the summary of the video.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Findings<\/jats:title><jats:p>Compared with traditional video summarisation methods that use low\u2010level audio\u2010visual features, the proposed method is based on users' high\u2010level collaborative activities, and thus can produce semantically more important summaries than existing methods.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Research limitations\/implications<\/jats:title><jats:p>It is assumed that the video frames around the bookmarks inserted by users are informative and representative, and therefore can be used as good sources for summarising videos.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Originality\/value<\/jats:title><jats:p>Folksonomy, commonly called collaborative tagging, is a Web 2.0 method for users to freely annotate shared information resources with keywords. It has mostly been used for collaboratively tagging photos (Flickr), web site bookmarks (Del.icio.us), or blog posts (Technorati), but has never been applied to the field of automatic video summarisation. It is believed that this is the first attempt to utilise users' high\u2010level collaborative tagging activities, instead of low\u2010level audio\u2010visual features, for video summarisation.<\/jats:p><\/jats:sec>","DOI":"10.1108\/14684521111161981","type":"journal-article","created":{"date-parts":[[2011,8,20]],"date-time":"2011-08-20T07:10:10Z","timestamp":1313824210000},"page":"653-668","source":"Crossref","is-referenced-by-count":9,"title":["Video summarisation based on collaborative temporal tags"],"prefix":"10.1108","volume":"35","author":[{"given":"Min","family":"Gyo Chung","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Taehyung (George)","family":"Wang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Phillip C.\u2010Y.","family":"Sheu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"140","reference":[{"key":"key2022012119470392600_b1","doi-asserted-by":"crossref","unstructured":"Abdollahian, G. and Delp, E.J. (2009), \u201cUser generated video annotation using geo\u2010tagged image databases\u201d, Proceedings of the IEEE International Conference on Multimedia and Expo, IEEE, Los Alamitos, CA, pp. 610\u201013.","DOI":"10.1109\/ICME.2009.5202570"},{"key":"key2022012119470392600_b2","doi-asserted-by":"crossref","unstructured":"Ames, M. and Naaman, M. (2007), \u201cWhy we tag: motivations for annotation in mobile and online media\u201d, Proceedings of the Conference on Human Factors in Computing Systems (CHI 2007), ACM Press, New York, NY, pp. 971\u201080.","DOI":"10.1145\/1240624.1240772"},{"key":"key2022012119470392600_b3","doi-asserted-by":"crossref","unstructured":"Angus, E., Thelwall, M. and Stuart, D. (2008), \u201cGeneral patterns of tag usage among university groups in Flickr\u201d, Online Information Review, Vol. 32 No. 1, pp. 89\u2010101.","DOI":"10.1108\/14684520810866001"},{"key":"key2022012119470392600_b4","unstructured":"Avila, S., Luz, A., Araujo, A.A. and Cord, M. (2008), \u201cVSUMM: an approach for automatic video summarization and quantitative evaluation\u201d, Proceedings of XXI Brazilian Symposium on Computer Graphics and Image Processing, IEEE, Los Alamitos, CA, pp. 103\u201010."},{"key":"key2022012119470392600_b5","doi-asserted-by":"crossref","unstructured":"Barbieri, M., Agnihotri, L. and Dimitrova, N. (2003), \u201cVideo summarization: methods and landscape\u201d, Proceedings of the Society of Photographic Instrumentation Engineers Internet Multimedia Management Systems IV Conference, Vol. 5242, SPIE, Bellingham, WA, pp. 1\u201013.","DOI":"10.1117\/12.515733"},{"key":"key2022012119470392600_b6","unstructured":"Bateman, S., Brooks, C. and Brusilovsky, P. (2007), \u201cApplying collaborative tagging to e\u2010learning\u201d, Proceedings of WWW 2007 Workshop on Tagging and Metadata for Social Information Organization, available at: www2007.org\/workshops\/paper_56.pdf (accessed 14 April 2011)."},{"key":"key2022012119470392600_b7","doi-asserted-by":"crossref","unstructured":"Cernekova, Z., Pitas, I. and Nikou, C. (2006), \u201cInformation theory\u2010based shot cut\/fade detection and video summarization\u201d, IEEE Transactions on Circuits and Systems for Video Technology, Vol. 16 No. 1, pp. 82\u201091.","DOI":"10.1109\/TCSVT.2005.856896"},{"key":"key2022012119470392600_b8","doi-asserted-by":"crossref","unstructured":"Chippendale, P., Zanin, M. and Andreatta, C. (2009), \u201cCollective photography\u201d, Proceedings of the Conference for Visual Media Production, IEEE, Los Alamitos, CA, pp. 188\u201094.","DOI":"10.1109\/CVMP.2009.30"},{"key":"key2022012119470392600_b9","doi-asserted-by":"crossref","unstructured":"Gao, Y., Wang, W. and Yong, J. (2008), \u201cA video summarization tool using two\u2010level redundancy detection for personal video recorders\u201d, IEEE Transactions on Consumer Electronics, Vol. 54 No. 2, pp. 521\u20106.","DOI":"10.1109\/TCE.2008.4560124"},{"key":"key2022012119470392600_b10","doi-asserted-by":"crossref","unstructured":"Golder, S. and Huberman, B. (2006), \u201cUsage patterns of collaborative tagging systems\u201d, Journal of Information Science, Vol. 32 No. 2, pp. 198\u2010208.","DOI":"10.1177\/0165551506062337"},{"key":"key2022012119470392600_b11","doi-asserted-by":"crossref","unstructured":"Halpin, H., Robu, V. and Shepherd, H. (2007), \u201cThe complex dynamics of collaborative tagging\u201d, Proceedings of the WWW 2007 Conference, ACM Press, New York, NY, pp. 211\u201020.","DOI":"10.1145\/1242572.1242602"},{"key":"key2022012119470392600_b12","doi-asserted-by":"crossref","unstructured":"Huang, C.\u2010H., Kung, H.T. and Su, C.\u2010Y. (2008), \u201cUse of content tags in managing advertisements for online videos\u201d, Proceedings of 10th IEEE Conference on E\u2010Commerce Technology and the 5th IEEE Conference on Enterprise Computing, E\u2010Commerce and E\u2010Services, IEEE, Los Alamitos, CA, pp. 249\u201054.","DOI":"10.1109\/CECandEEE.2008.89"},{"key":"key2022012119470392600_b13","unstructured":"Huang, M., Mahajan, A.B. and DeMenthon, D.F. (2004), Automatic Performance Evaluation for Video Summarization, Technical Report UMIACS\u2010TR\u20102004\u201047, University of Maryland, College Park, MD, available at: http:\/\/lampsrv02.umiacs.umd.edu\/pubs\/TechReports\/LAMP_114\/LAMP_114.pdf (accessed 14 April 2011)."},{"key":"key2022012119470392600_b14","doi-asserted-by":"crossref","unstructured":"Hyndman, R., Koehler, A., Ord, J. and Snyder, R. (2008), Forecasting with Exponential Smoothing: The State Space Approach, Springer\u2010Verlag, Berlin.","DOI":"10.1007\/978-3-540-71918-2"},{"key":"key2022012119470392600_b15","unstructured":"John, A. and Seligmann, D. (2006), \u201cCollaborative tagging and expertise in the enterprise\u201d, Proceedings of WWW 2006 Workshop on Collaborative Web Tagging, available at: http:\/\/citeseerx.ist.psu.edu\/viewdoc\/summary?doi=10.1.1.134.296 (accessed 14 April 2011)."},{"key":"key2022012119470392600_b16","doi-asserted-by":"crossref","unstructured":"Kim, S.H., Ay, S.A. and Zimmermann, R. (2010), \u201cDesign and implementation of geo\u2010tagged video search framework\u201d, Journal of Visual Communication and Image Representation, Vol. 21 No. 8, pp. 773\u201086.","DOI":"10.1016\/j.jvcir.2010.07.004"},{"key":"key2022012119470392600_b17","doi-asserted-by":"crossref","unstructured":"Koelstra, S., Muhl, C. and Patras, I. (2009), \u201cEEG analysis for implicit tagging of video data\u201d, Proceedings of the 3rd International Conference on Affective Computing and Intelligent Interaction, IEEE, New York, NY, pp. 1\u20106.","DOI":"10.1109\/ACII.2009.5349482"},{"key":"key2022012119470392600_b18","doi-asserted-by":"crossref","unstructured":"Macgregor, G. and McCulloch, E. (2006), \u201cCollaborative tagging as a knowledge organization and resource discovery tool\u201d, Library Review, Vol. 55 No. 5, pp. 291\u2010300.","DOI":"10.1108\/00242530610667558"},{"key":"key2022012119470392600_b19","unstructured":"Marchetti, A., Tesconi, M. and Ronzano, F. (2007), \u201cSemKey: semantic collaborative tagging system\u201d, Proceedings of WWW 2007 Workshop on Tagging and Metadata for Social Information Organization, available at: www2007.org\/workshops\/paper_45.pdf (accessed 14 April 2011)."},{"key":"key2022012119470392600_b20","doi-asserted-by":"crossref","unstructured":"Marlow, C., Naaman, M., Boyd, D. and Davis, M. (2006), \u201cTagging paper, taxonomy, Flickr, academic article to read\u201d, Proceedings of the 17th Conference on Hypertext and Hypermedia (HT 2006), ACM Press, New York, NY, pp. 31\u20109.","DOI":"10.1145\/1149941.1149949"},{"key":"key2022012119470392600_b21","unstructured":"Mathes, A. (2004), \u201cFolksonomies \u2013 cooperative classification and communication through shared metadata\u201d, Computer Mediated Communication (LIS590CMC), University of Illinois, Urbana\u2010Champaign, IL, available at: www.adammathes.com\/academic\/computer\u2010mediated\u2010communication\/folksonomies.html (accessed 14 April 2011)."},{"key":"key2022012119470392600_b22","doi-asserted-by":"crossref","unstructured":"Min, H.\u2010S., Choi, J., De Neve, W., Ro, Y.M. and Plataniotis, K.N. (2009), \u201cSemantic annotation of personal video content using an image folksonomy\u201d, Proceedings of 16th IEEE International Conference on Image Processing, IEEE, Los Alamitos, CA, pp. 257\u201060.","DOI":"10.1109\/ICIP.2009.5413429"},{"key":"key2022012119470392600_b23","doi-asserted-by":"crossref","unstructured":"Money, A. and Agius, H. (2008), \u201cVideo summarization: a conceptual framework and survey of the state of the art\u201d, Journal of Visual Communication and Image Representation, Vol. 19 No. 2, pp. 121\u201043.","DOI":"10.1016\/j.jvcir.2007.04.002"},{"key":"key2022012119470392600_b24","doi-asserted-by":"crossref","unstructured":"Ngo, C., Ma, Y. and Zhang, H. (2005), \u201cVideo summarization and scene detection by graph modeling\u201d, IEEE Transactions on Circuits and Systems for Video Technology, Vol. 15 No. 2, pp. 296\u2010305.","DOI":"10.1109\/TCSVT.2004.841694"},{"key":"key2022012119470392600_b25","doi-asserted-by":"crossref","unstructured":"Pantic, M. and Vinciarelli, A. (2009), \u201cImplicit human\u2010centered tagging\u201d, IEEE Signal Processing Magazine, Vol. 26 No. 6, pp. 173\u201080.","DOI":"10.1109\/MSP.2009.934186"},{"key":"key2022012119470392600_b26","doi-asserted-by":"crossref","unstructured":"Paredes, R., Ulges, A. and Breuel, T. (2009), \u201cFast discriminative linear models for scalable video tagging\u201d, Proceedings of the 2009 International Conference on Machine Learning and Applications, IEEE, Los Alamitos, CA, pp. 571\u20106.","DOI":"10.1109\/ICMLA.2009.68"},{"key":"key2022012119470392600_b27","doi-asserted-by":"crossref","unstructured":"Ren, J. and Jiang, J. (2009), \u201cHierarchical modeling and adaptive clustering for real\u2010time summarization of rush videos\u201d, IEEE Transactions on Multimedia, Vol. 11 No. 5, pp. 906\u201017.","DOI":"10.1109\/TMM.2009.2021782"},{"key":"key2022012119470392600_b28","doi-asserted-by":"crossref","unstructured":"Scharcanski, J. and Gaviao, W. (2006), \u201cHierarchical summarization of diagnostic hysteroscopy videos\u201d, Proceedings of IEEE International Conference on Image Processing, IEEE, Los Alamitos, CA, pp. 129\u201032.","DOI":"10.1109\/ICIP.2006.312376"},{"key":"key2022012119470392600_b29","doi-asserted-by":"crossref","unstructured":"Siersdorfer, S., Pedro, J.S. and Sanderson, M. (2009), \u201cAutomatic video tagging using content redundancy\u201d, Proceedings of the 32nd International ACM SIGIR Conference on Research and Development in Information Retrieval, ACM, New York, NY, pp. 395\u2010402.","DOI":"10.1145\/1571941.1572010"},{"key":"key2022012119470392600_b30","doi-asserted-by":"crossref","unstructured":"Taskiran, C.M. (2006), \u201cEvaluation of automatic video summarization systems\u201d, Proceedings of the Society of Photographic Instrumentation Engineers Multimedia Content Analysis, Management and Retrieval Conference, 6073, SPIE, Bellingham, WA, pp. 178\u201087.","DOI":"10.1117\/12.655744"},{"key":"key2022012119470392600_b31","doi-asserted-by":"crossref","unstructured":"Truong, B.T. and Venkatesh, S. (2007), \u201cVideo abstraction: a systematic review and classification\u201d, ACM Transactions on Multimedia Computing, Communications, and Applications, Vol. 3 No. 1, available at: www.computing.edu.au\/\u223csvetha\/publications\/2006\/journals\/BaTu_Tommcat_2006.pdf (accessed 14 April 2011).","DOI":"10.1145\/1198302.1198305"},{"key":"key2022012119470392600_b32","unstructured":"Vander Wal, T. (2007), \u201cFolksonomy coinage and definition\u201d, available at: www.vanderwal.net\/folksonomy.html (accessed 30 April 2010)."},{"key":"key2022012119470392600_b33","doi-asserted-by":"crossref","unstructured":"Yamamoto, D., Masuda, T., Ohira, S. and Nagao, K. (2008), \u201cVideo scene annotation based on web social activities\u201d, IEEE Multimedia, Vol. 15 No. 3, pp. 22\u201032.","DOI":"10.1109\/MMUL.2008.67"},{"key":"key2022012119470392600_b34","doi-asserted-by":"crossref","unstructured":"Zeng, D. and Li, H. (2008), \u201cHow useful are tags? An empirical analysis of collaborative tagging for web page recommendation\u201d, Lecture Notes in Computer Science, Vol. 5075, pp. 320\u201030.","DOI":"10.1007\/978-3-540-69304-8_32"}],"container-title":["Online Information Review"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/www.emeraldinsight.com\/doi\/full-xml\/10.1108\/14684521111161981","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/14684521111161981\/full\/xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/14684521111161981\/full\/html","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,7,25]],"date-time":"2025-07-25T00:41:43Z","timestamp":1753404103000},"score":1,"resource":{"primary":{"URL":"http:\/\/www.emerald.com\/oir\/article\/35\/4\/653-668\/447314"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2011,8,9]]},"references-count":34,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2011,8,9]]}},"alternative-id":["10.1108\/14684521111161981"],"URL":"https:\/\/doi.org\/10.1108\/14684521111161981","relation":{},"ISSN":["1468-4527"],"issn-type":[{"type":"print","value":"1468-4527"}],"subject":[],"published":{"date-parts":[[2011,8,9]]}}}