{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,21]],"date-time":"2026-03-21T02:18:52Z","timestamp":1774059532625,"version":"3.50.1"},"reference-count":43,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2015,3,2]],"date-time":"2015-03-02T00:00:00Z","timestamp":1425254400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2015,3,2]]},"abstract":"<jats:p>Given the current profusion of devices for viewing media, video content created at one aspect ratio is often viewed on displays with different aspect ratios. Many previous solutions address this problem by retargeting or resizing the video, but a more general solution would re-edit the video for the new display. Our method employs the three primary editing operations: pan, cut, and zoom. We let viewers implicitly reveal what is important in a video by tracking their gaze as they watch the video. We present an algorithm that optimizes the path of a cropping window based on the collected eyetracking data, finds places to cut, and computes the size of the cropping window. We present results on a variety of video clips, including close-up and distant shots, and stationary and moving cameras. We conduct two experiments to evaluate our results. First, we eyetrack viewers on the result videos generated by our algorithm, and second, we perform a subjective assessment of viewer preference. These experiments show that viewer gaze patterns are similar on our result videos and on the original video clips, and that viewers prefer our results to an optimized crop-and-warp algorithm.<\/jats:p>","DOI":"10.1145\/2699644","type":"journal-article","created":{"date-parts":[[2015,3,3]],"date-time":"2015-03-03T14:08:19Z","timestamp":1425391699000},"page":"1-12","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":40,"title":["Gaze-Driven Video Re-Editing"],"prefix":"10.1145","volume":"34","author":[{"given":"Eakta","family":"Jain","sequence":"first","affiliation":[{"name":"Carnegie Mellon University, Disney Research Pittsburgh, and University of Florida, Pittsburgh, PA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yaser","family":"Sheikh","sequence":"additional","affiliation":[{"name":"Carnegie Mellon University, Pittsburgh, PA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ariel","family":"Shamir","sequence":"additional","affiliation":[{"name":"Disney Research Pittsburgh and Interdisciplinary Center Israel, Pittsburgh, PA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jessica","family":"Hodgins","sequence":"additional","affiliation":[{"name":"Carnegie Mellon University and Disney Research Pittsburgh, Pittsburgh, PA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2015,3,2]]},"reference":[{"key":"e_1_2_2_1_1","first-page":"1","article-title":"Ultra-low cost eyetracking as an high information throughput alternative to BMIS","volume":"12","author":"Abbot W.","year":"2011","unstructured":"W. Abbot and F. Aldo . 2011 . Ultra-low cost eyetracking as an high information throughput alternative to BMIS . BMC Neurosci. 12 , 1 . W. Abbot and F. Aldo. 2011. Ultra-low cost eyetracking as an high information throughput alternative to BMIS. BMC Neurosci. 12,1.","journal-title":"BMC Neurosci."},{"key":"e_1_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1743666.1743685"},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1276377.1276390"},{"key":"e_1_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.tins.2011.02.003"},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/2077451.2077453"},{"key":"e_1_2_2_6_1","volume-title":"Proceedings of the International Conference on Pattern Recognition (ICPR'08)","author":"Chamaret C.","unstructured":"C. Chamaret and O. Le Meur . 2008. Attention-based video reframing: Validation using eye-tracking . In Proceedings of the International Conference on Pattern Recognition (ICPR'08) . C. Chamaret and O. Le Meur. 2008. Attention-based video reframing: Validation using eye-tracking. In Proceedings of the International Conference on Pattern Recognition (ICPR'08)."},{"key":"e_1_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/566654.566650"},{"key":"e_1_2_2_8_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR'08)","author":"Deselaers T.","unstructured":"T. Deselaers , P. Dreuw , and H. Ney . 2008. Pan, zoom, scan -- Time coherent, trained automatic video cropping . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR'08) . 1--8. T. Deselaers, P. Dreuw, and H. Ney. 2008. Pan, zoom, scan -- Time coherent, trained automatic video cropping. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR'08). 1--8."},{"key":"e_1_2_2_9_1","volume-title":"On Film Editing","author":"Dmytryk E.","unstructured":"E. Dmytryk . 1984. On Film Editing . Focal Press . E. Dmytryk. 1984. On Film Editing. Focal Press."},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1167\/10.10.28"},{"key":"e_1_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/1291233.1291255"},{"key":"e_1_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.3758\/BF03203630"},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/358669.358692"},{"key":"e_1_2_2_14_1","unstructured":"J. D. Foley A. Van Dam S. K. Feiner and J. F. Hughes. 1996. Computer Graphics Principles and Practice 2nd Ed. Addison-Wesley.   J. D. Foley A. Van Dam S. K. Feiner and J. F. Hughes. 1996. Computer Graphics Principles and Practice 2 nd Ed. Addison-Wesley."},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.compbiomed.2006.08.018"},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/2355598.2355603"},{"key":"e_1_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/2338676.2338688"},{"key":"e_1_2_2_18_1","unstructured":"T. Judd F. Durand and A. Torralba. 2012. A benchmark of computational models of saliency to predict human fixations. Tech. rep. MITCSAIL-TR-2012-001 Massachusetts Institute of Technology. http:\/\/dspace.mit.edu\/handle\/1721.1\/68590.  T. Judd F. Durand and A. Torralba. 2012. A benchmark of computational models of saliency to predict human fixations. Tech. rep. MITCSAIL-TR-2012-001 Massachusetts Institute of Technology. http:\/\/dspace.mit.edu\/handle\/1721.1\/68590."},{"key":"e_1_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/2632284"},{"key":"e_1_2_2_20_1","volume-title":"Shot by Shot. Michael Wiese Productions","author":"Katz S. D.","unstructured":"S. D. Katz . 1991. Shot by Shot. Michael Wiese Productions , Focal Press . S. D. Katz. 1991. Shot by Shot. Michael Wiese Productions, Focal Press."},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11042-006-0076-5"},{"key":"e_1_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11042-010-0717-6"},{"key":"e_1_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/1618452.1618472"},{"key":"e_1_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/1180639.1180702"},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.1467-8659.2009.01616.x"},{"key":"e_1_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1007\/s12559-010-9074-z"},{"key":"e_1_2_2_27_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR'10)","author":"Niu Y.","unstructured":"Y. Niu , F. Liu , X. Li , and M. Gleicher . 2010. Warp propagation for video resizing . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR'10) . 537--544. Y. Niu, F. Liu, X. Li, and M. Gleicher. 2010. Warp propagation for video resizing. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR'10). 537--544."},{"key":"e_1_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/1360612.1360615"},{"key":"e_1_2_2_29_1","unstructured":"D. Rudoy D. B. Goldman E. Shechtman and L. Zelnik-Manor. 2012. Crowdsourcing gaze data collection. http:\/\/arxiv.org\/abs\/1204. 3367.  D. Rudoy D. B. Goldman E. Shechtman and L. Zelnik-Manor. 2012. Crowdsourcing gaze data collection. http:\/\/arxiv.org\/abs\/1204. 3367."},{"key":"e_1_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/1124772.1124886"},{"key":"e_1_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/1665817.1665828"},{"key":"e_1_2_2_32_1","doi-asserted-by":"crossref","first-page":"1","DOI":"10.16910\/jemr.2.2.6","article-title":"Edit blindness: The relationship between attention and global change blindness in dynamic scenes","volume":"2","author":"Smith T. J.","year":"2008","unstructured":"T. J. Smith and J. M. Henderson . 2008 . Edit blindness: The relationship between attention and global change blindness in dynamic scenes . J. Eye Movement Res. 2 , 2, 1 -- 17 . T. J. Smith and J. M. Henderson. 2008. Edit blindness: The relationship between attention and global change blindness in dynamic scenes. J. Eye Movement Res. 2, 2, 1--17.","journal-title":"J. Eye Movement Res."},{"key":"e_1_2_2_33_1","volume-title":"Proceedings of the Workshop on Dynamical Vision at the International Conference on Computer Vision (ICCV'07)","author":"Tao C.","unstructured":"C. Tao , J. Jia , and H. Sun . 2007. Active window oriented dynamic video retargeting . In Proceedings of the Workshop on Dynamical Vision at the International Conference on Computer Vision (ICCV'07) . C. Tao, J. Jia, and H. Sun. 2007. Active window oriented dynamic video retargeting. In Proceedings of the Workshop on Dynamical Vision at the International Conference on Computer Vision (ICCV'07)."},{"key":"e_1_2_2_34_1","volume-title":"Proceedings of the IEEE Conference on Multimedia and Expo (ICME'04)","author":"Wang J.","unstructured":"J. Wang , M. J. T. Reinders , R. L. Lagendijk , J. Lindenberg , and M. S. Kankanhalli . 2004. Video content representation on tiny devices . In Proceedings of the IEEE Conference on Multimedia and Expo (ICME'04) . 1711--1714. J. Wang, M. J. T. Reinders, R. L. Lagendijk, J. Lindenberg, and M. S. Kankanhalli. 2004. Video content representation on tiny devices. In Proceedings of the IEEE Conference on Multimedia and Expo (ICME'04). 1711--1714."},{"key":"e_1_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/1618452.1618473"},{"key":"e_1_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/2010324.1964983"},{"key":"e_1_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/1778765.1778827"},{"key":"e_1_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1145\/1409060.1409071"},{"key":"e_1_2_2_39_1","unstructured":"Wikipedia. 2015. http:\/\/en.wikipedia.org\/wiki\/pan_and_scan.  Wikipedia. 2015. http:\/\/en.wikipedia.org\/wiki\/pan_and_scan."},{"key":"e_1_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/1873951.1873991"},{"key":"e_1_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/1873951.1874113"},{"key":"e_1_2_2_42_1","unstructured":"J. Young. 2008. Sydney Pollack dies at 73. Variety May 26.  J. Young. 2008. Sydney Pollack dies at 73. Variety May 26."},{"key":"e_1_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1167\/12.6.22"}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2699644","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2699644","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T06:16:59Z","timestamp":1750227419000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2699644"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2015,3,2]]},"references-count":43,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2015,3,2]]}},"alternative-id":["10.1145\/2699644"],"URL":"https:\/\/doi.org\/10.1145\/2699644","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"value":"0730-0301","type":"print"},{"value":"1557-7368","type":"electronic"}],"subject":[],"published":{"date-parts":[[2015,3,2]]},"assertion":[{"value":"2012-12-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2014-09-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2015-03-02","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}