{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,5]],"date-time":"2026-03-05T20:07:54Z","timestamp":1772741274140,"version":"3.50.1"},"reference-count":30,"publisher":"Association for Computing Machinery (ACM)","issue":"6","license":[{"start":{"date-parts":[[2015,11,2]],"date-time":"2015-11-02T00:00:00Z","timestamp":1446422400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2015,11,4]]},"abstract":"<jats:p>We introduce JumpCut, a new mask transfer and interpolation method for interactive video cutout. Given a source frame for which a foreground mask is already available, we compute an estimate of the foreground mask at another, typically non-successive, target frame. Observing that the background and foreground regions typically exhibit different motions, we leverage these differences by computing two separate nearest-neighbor fields (split-NNF) from the target to the source frame. These NNFs are then used to jointly predict a coherent labeling of the pixels in the target frame. The same split-NNF is also used to aid a novel edge classifier in detecting silhouette edges (S-edges) that separate the foreground from the background. A modified level set method is then applied to produce a clean mask, based on the pixel labels and the S-edges computed by the previous two steps. The resulting mask transfer method may also be used for coherently interpolating the foreground masks between two distant source frames. Our results demonstrate that the proposed method is significantly more accurate than the existing state-of-the-art on a wide variety of video sequences. Thus, it reduces the required amount of user effort, and provides a basis for an effective interactive video object cutout tool.<\/jats:p>","DOI":"10.1145\/2816795.2818105","type":"journal-article","created":{"date-parts":[[2015,10,27]],"date-time":"2015-10-27T12:36:39Z","timestamp":1445949399000},"page":"1-10","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":101,"title":["JumpCut"],"prefix":"10.1145","volume":"34","author":[{"given":"Qingnan","family":"Fan","sequence":"first","affiliation":[{"name":"Shandong University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fan","family":"Zhong","sequence":"additional","affiliation":[{"name":"Shandong University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Dani","family":"Lischinski","sequence":"additional","affiliation":[{"name":"The Hebrew University of Jerusalem"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Daniel","family":"Cohen-Or","sequence":"additional","affiliation":[{"name":"Tel Aviv University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Baoquan","family":"Chen","sequence":"additional","affiliation":[{"name":"Shandong University"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2015,11,2]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/1015706.1015764"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1531326.1531376"},{"key":"e_1_2_1_3_1","volume-title":"Proc. ECCV, Springer-Verlag","author":"Bai X.","unstructured":"Bai , X. , Wang , J. , and Sapiro , G . 2010. Dynamic color flow: A motion-adaptive color model for object segmentation in video . In Proc. ECCV, Springer-Verlag , vol. V , 617--630. Bai, X., Wang, J., and Sapiro, G. 2010. Dynamic color flow: A motion-adaptive color model for object segmentation in video. In Proc. ECCV, Springer-Verlag, vol. V, 617--630."},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2014.2359374"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/1531326.1531330"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cviu.2007.09.014"},{"key":"e_1_2_1_7_1","volume-title":"Proc. ECCV'04","author":"Brox T.","unstructured":"Brox , T. , Bruhn , A. , Papenberg , N. , and Weickert , J . 2004. High accuracy optical flow estimation based on a theory for warping . In Proc. ECCV'04 . Springer, 25--36. Brox, T., Bruhn, A., Papenberg, N., and Weickert, J. 2004. High accuracy optical flow estimation based on a theory for warping. In Proc. ECCV'04. Springer, 25--36."},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1007979827043"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.316"},{"key":"e_1_2_1_10_1","volume-title":"Proc. CVPR","volume":"2","author":"Chuang Y.-Y.","year":"2001","unstructured":"Chuang , Y.-Y. , Curless , B. , Salesin , D. H. , and Szeliski , R . 2001. A bayesian approach to digital matting . In Proc. CVPR 2001 , IEEE Computer Society , vol. 2 , 264--271. Chuang, Y.-Y., Curless, B., Salesin, D. H., and Szeliski, R. 2001. A bayesian approach to digital matting. In Proc. CVPR 2001, IEEE Computer Society, vol. 2, 264--271."},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/566654.566572"},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-006-8711-1"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2013.231"},{"key":"e_1_2_1_14_1","volume-title":"Proc. BMVC","author":"Faktor A.","year":"2014","unstructured":"Faktor , A. , and Irani , M . 2014. Video segmentation by non-local consensus voting . In Proc. BMVC 2014 . Faktor, A., and Irani, M. 2014. Video segmentation by non-local consensus voting. In Proc. BMVC 2014."},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/1073204.1073234"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2007.41"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/TVCG.2011.77"},{"key":"e_1_2_1_18_1","volume-title":"International Conference on Computer Vision Theory and Applications (VISAPP'09)","author":"Muja M.","unstructured":"Muja , M. , and Lowe , D. G . 2009. Fast approximate nearest neighbors with automatic algorithm configuration . In International Conference on Computer Vision Theory and Applications (VISAPP'09) . Muja, M., and Lowe, D. G. 2009. Fast approximate nearest neighbors with automatic algorithm configuration. In International Conference on Computer Vision Theory and Applications (VISAPP'09)."},{"key":"e_1_2_1_19_1","doi-asserted-by":"crossref","unstructured":"Osher S. J. and Fedkiw R. P. 2003. Level Set Methods and Dynamic Implicit Surfaces 1st ed. Springer-Verlag.  Osher S. J. and Fedkiw R. P. 2003. Level Set Methods and Dynamic Implicit Surfaces 1st ed. Springer-Verlag.","DOI":"10.1007\/b98879"},{"key":"e_1_2_1_20_1","volume-title":"Proc. ICCV, 779--786","author":"Price B.","unstructured":"Price , B. , Morse , B. , and Cohen , S . 2009. LIVEcut: Learning-based interactive video segmentation by evaluation of multiple propagated cues . In Proc. ICCV, 779--786 . Price, B., Morse, B., and Cohen, S. 2009. LIVEcut: Learning-based interactive video segmentation by evaluation of multiple propagated cues. In Proc. ICCV, 779--786."},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.55"},{"key":"e_1_2_1_22_1","volume-title":"Psychology","author":"Rubin E.","unstructured":"Rubin , E. 2001. Figure and ground . In Visual Perception, S. Yantis, Ed. Psychology Press , Philadelphia , 225--229. Rubin, E. 2001. Figure and ground. In Visual Perception, S. Yantis, Ed. Psychology Press, Philadelphia, 225--229."},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/1141911.1141920"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/237170.237263"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.1467-8659.2011.02038.x"},{"key":"e_1_2_1_26_1","volume-title":"Proc. BMVC.","author":"Tsai D.","unstructured":"Tsai , D. , Flagg , M. , and Rehg , J. M . 2010. Motion coherent tracking with multi-label MRF optimization . In Proc. BMVC. Tsai, D., Flagg, M., and Rehg, J. M. 2010. Motion coherent tracking with multi-label MRF optimization. In Proc. BMVC."},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/1073204.1073233"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cviu.2013.10.013"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2013.175"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/2366145.2366194"}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2816795.2818105","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2816795.2818105","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T05:48:19Z","timestamp":1750225699000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2816795.2818105"}},"subtitle":["non-successive mask transfer and interpolation for video cutout"],"short-title":[],"issued":{"date-parts":[[2015,11,2]]},"references-count":30,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2015,11,4]]}},"alternative-id":["10.1145\/2816795.2818105"],"URL":"https:\/\/doi.org\/10.1145\/2816795.2818105","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"value":"0730-0301","type":"print"},{"value":"1557-7368","type":"electronic"}],"subject":[],"published":{"date-parts":[[2015,11,2]]},"assertion":[{"value":"2015-11-02","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}