{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,18]],"date-time":"2026-07-18T02:35:32Z","timestamp":1784342132289,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":44,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,5,6]],"date-time":"2021-05-06T00:00:00Z","timestamp":1620259200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,5,6]]},"DOI":"10.1145\/3411764.3445721","type":"proceedings-article","created":{"date-parts":[[2021,5,8]],"date-time":"2021-05-08T03:50:37Z","timestamp":1620445837000},"page":"1-16","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":54,"title":["Automatic Generation of Two-Level Hierarchical Tutorials from Instructional Makeup Videos"],"prefix":"10.1145","author":[{"given":"Anh","family":"Truong","sequence":"first","affiliation":[{"name":"Stanford University, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Peggy","family":"Chi","sequence":"additional","affiliation":[{"name":"Google Research, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"David","family":"Salesin","sequence":"additional","affiliation":[{"name":"Google Research, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Irfan","family":"Essa","sequence":"additional","affiliation":[{"name":"Google Research Google, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Maneesh","family":"Agrawala","sequence":"additional","affiliation":[{"name":"Stanford University, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2021,5,7]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/882262.882352"},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.495"},{"key":"e_1_3_2_2_3_1","volume-title":"Bobbi Brown Makeup Manual: For Everyone from Beginner to Pro","author":"Brown B.","unstructured":"B. Brown . 2008. Bobbi Brown Makeup Manual: For Everyone from Beginner to Pro . Grand Central Publishing , New York, NY, USA . B. Brown. 2008. Bobbi Brown Makeup Manual: For Everyone from Beginner to Pro. Grand Central Publishing, New York, NY, USA."},{"key":"e_1_3_2_2_4_1","volume-title":"PairedCycleGAN: Asymmetric Style Transfer for Applying and Removing Makeup. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Computer Vision Foundation \/ IEEE","author":"Chang Huiwen","year":"2018","unstructured":"Huiwen Chang , Jingwan Lu , Fisher Yu , and Adam Finkelstein . 2018 . PairedCycleGAN: Asymmetric Style Transfer for Applying and Removing Makeup. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Computer Vision Foundation \/ IEEE , Salt Lake City, UT, USA, 40\u201348. Huiwen Chang, Jingwan Lu, Fisher Yu, and Adam Finkelstein. 2018. PairedCycleGAN: Asymmetric Style Transfer for Applying and Removing Makeup. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Computer Vision Foundation \/ IEEE, Salt Lake City, UT, USA, 40\u201348."},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/3290605.3300931"},{"key":"e_1_3_2_2_6_1","volume-title":"BeautyGlow: On-Demand Makeup Transfer Framework With Reversible Generative Network. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Computer Vision Foundation \/ IEEE","author":"Chen Hung-Jen","year":"2019","unstructured":"Hung-Jen Chen , Ka-Ming Hui , Szu-Yu Wang , Li-Wu Tsao , Hong-Han Shuai , and Wen-Huang Cheng . 2019 . BeautyGlow: On-Demand Makeup Transfer Framework With Reversible Generative Network. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Computer Vision Foundation \/ IEEE , Long Beach, CA, USA, 10042\u201310050. Hung-Jen Chen, Ka-Ming Hui, Szu-Yu Wang, Li-Wu Tsao, Hong-Han Shuai, and Wen-Huang Cheng. 2019. BeautyGlow: On-Demand Makeup Transfer Framework With Reversible Generative Network. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Computer Vision Foundation \/ IEEE, Long Beach, CA, USA, 10042\u201310050."},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/2380116.2380130"},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/2501988.2502052"},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/1964921.1964961"},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"crossref","unstructured":"Logan Fiorella and Richard\u00a0E Mayer. 2018. What works and doesn\u2019t work with instructional video.  Logan Fiorella and Richard\u00a0E Mayer. 2018. What works and doesn\u2019t work with instructional video.","DOI":"10.1016\/j.chb.2018.07.015"},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/3313831.3376437"},{"key":"e_1_3_2_2_12_1","unstructured":"Google. 2020. Cloud Natural Language documentation | Cloud Natural Language API. Google Cloud. https:\/\/cloud.google.com\/natural-language\/docs  Google. 2020. Cloud Natural Language documentation | Cloud Natural Language API. Google Cloud. https:\/\/cloud.google.com\/natural-language\/docs"},{"key":"e_1_3_2_2_13_1","unstructured":"Google. 2020. Cloud Speech-to-Text - Speech Recognition | Google Cloud. Google Cloud. https:\/\/cloud.google.com\/speech-to-text  Google. 2020. Cloud Speech-to-Text - Speech Recognition | Google Cloud. Google Cloud. https:\/\/cloud.google.com\/speech-to-text"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/1576246.1531372"},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/1866029.1866054"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/989863.989917"},{"key":"e_1_3_2_2_17_1","volume-title":"Unsupervised Visual-Linguistic Reference Resolution in Instructional Videos. In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Computer Vision Foundation \/ IEEE","author":"Huang D.","year":"2017","unstructured":"D. Huang , J.\u00a0 J. Lim , L. Fei-Fei , and J.\u00a0 C. Niebles . 2017 . Unsupervised Visual-Linguistic Reference Resolution in Instructional Videos. In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Computer Vision Foundation \/ IEEE , Honolulu, HI, USA, 1032\u20131041. D. Huang, J.\u00a0J. Lim, L. Fei-Fei, and J.\u00a0C. Niebles. 2017. Unsupervised Visual-Linguistic Reference Resolution in Instructional Videos. In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Computer Vision Foundation \/ IEEE, Honolulu, HI, USA, 1032\u20131041."},{"key":"e_1_3_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00623"},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.351"},{"key":"e_1_3_2_2_20_1","volume-title":"Makeup Masterclass: A Complete Course in Makeup for All Levels, Beginner to Advanced","author":"Jones R.","unstructured":"R. Jones . 2017. Robert Jones \u2019 Makeup Masterclass: A Complete Course in Makeup for All Levels, Beginner to Advanced . Fair Winds Press , Beverly, MA, USA . R. Jones. 2017. Robert Jones\u2019 Makeup Masterclass: A Complete Course in Makeup for All Levels, Beginner to Advanced. Fair Winds Press, Beverly, MA, USA."},{"key":"e_1_3_2_2_21_1","volume-title":"International conference on virtual systems and multimedia. VSMM","author":"Kameda Yoshinari","year":"1996","unstructured":"Yoshinari Kameda and Michihiko Minoh . 1996 . A human motion estimation method using 3-successive video frames . In International conference on virtual systems and multimedia. VSMM , Gifu, Japan, 135\u2013140. Yoshinari Kameda and Michihiko Minoh. 1996. A human motion estimation method using 3-successive video frames. In International conference on virtual systems and multimedia. VSMM, Gifu, Japan, 135\u2013140."},{"key":"e_1_3_2_2_22_1","unstructured":"Archana Kannan. 2020. Shopping for a beauty product? Try it on with Google. https:\/\/blog.google\/products\/shopping\/shopping-beauty-product-try-it-google\/  Archana Kannan. 2020. Shopping for a beauty product? Try it on with Google. https:\/\/blog.google\/products\/shopping\/shopping-beauty-product-try-it-google\/"},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/2556288.2556986"},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/3313831.3376845"},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1117\/12.333848"},{"key":"e_1_3_2_2_26_1","unstructured":"Jonathan Malmaud Jonathan Huang Vivek Rathod Nick Johnston Andrew Rabinovich and Kevin Murphy. 2015. What\u2019s Cookin\u2019? Interpreting Cooking Videos using Text Speech and Vision. CoRR abs\/1503.01558(2015) 1\u201310. arxiv:1503.01558http:\/\/arxiv.org\/abs\/1503.01558  Jonathan Malmaud Jonathan Huang Vivek Rathod Nick Johnston Andrew Rabinovich and Kevin Murphy. 2015. What\u2019s Cookin\u2019? Interpreting Cooking Videos using Text Speech and Vision. CoRR abs\/1503.01558(2015) 1\u201310. arxiv:1503.01558http:\/\/arxiv.org\/abs\/1503.01558"},{"key":"e_1_3_2_2_27_1","volume-title":"Animation: Does it facilitate learning. In AAAI spring symposium on smart graphics, Vol.\u00a05359","author":"Morrison Julie\u00a0Bauer","year":"2000","unstructured":"Julie\u00a0Bauer Morrison , Barbara Tversky , and Mireille Betrancourt . 2000 . Animation: Does it facilitate learning. In AAAI spring symposium on smart graphics, Vol.\u00a05359 . The AAAI Press , Menlo Park, CA, USA , 8. Julie\u00a0Bauer Morrison, Barbara Tversky, and Mireille Betrancourt. 2000. Animation: Does it facilitate learning. In AAAI spring symposium on smart graphics, Vol.\u00a05359. The AAAI Press, Menlo Park, CA, USA, 8."},{"key":"e_1_3_2_2_28_1","volume-title":"Graphics Interface","author":"Nawhal Megha","unstructured":"Megha Nawhal , Jacqueline\u00a0 B Lang , Greg Mori , and Parmit\u00a0 K Chilana . 2019. VideoWhiz: Non-Linear Interactive Overviews for Recipe Videos .. In Graphics Interface . Canadian Information Processing Society , Kingston, Ontario , Canada, 15\u20131. Megha Nawhal, Jacqueline\u00a0B Lang, Greg Mori, and Parmit\u00a0K Chilana. 2019. VideoWhiz: Non-Linear Interactive Overviews for Recipe Videos.. In Graphics Interface. Canadian Information Processing Society, Kingston, Ontario, Canada, 15\u20131."},{"key":"e_1_3_2_2_29_1","unstructured":"Celie O\u2019Neil-Hart. 2017. Self-directed learning from YouTube - Think with Google. https:\/\/www.thinkwithgoogle.com\/advertising-channels\/video\/self-directed-learning-youtube\/  Celie O\u2019Neil-Hart. 2017. Self-directed learning from YouTube - Think with Google. https:\/\/www.thinkwithgoogle.com\/advertising-channels\/video\/self-directed-learning-youtube\/"},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/2642918.2647400"},{"key":"e_1_3_2_2_31_1","unstructured":"Promise Phan. 2015. \u2019INSIDE OUT\u2019 Makeup Tutorial (Disgust Sadness Joy Anger & Fear). https:\/\/www.youtube.com\/watch?v=C1DXqkOCBt0  Promise Phan. 2015. \u2019INSIDE OUT\u2019 Makeup Tutorial (Disgust Sadness Joy Anger & Fear). https:\/\/www.youtube.com\/watch?v=C1DXqkOCBt0"},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/2047196.2047213"},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.509"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00756"},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/2984511.2984569"},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/3313831.3376759"},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/3313831.3376759"},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.180"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/2675133.2675219"},{"key":"e_1_3_2_2_40_1","unstructured":"Wikipedia contributors. 2020. F1 score. https:\/\/en.m.wikipedia.org\/wiki\/F1_score  Wikipedia contributors. 2020. F1 score. https:\/\/en.m.wikipedia.org\/wiki\/F1_score"},{"key":"e_1_3_2_2_41_1","unstructured":"Ann Yuan and Andrey Vakunov. 2020. Face and hand tracking in the browser with MediaPipe and TensorFlow.js. https:\/\/blog.tensorflow.org\/2020\/03\/face-and-hand-tracking-in-browser-with-mediapipe-and-tensorflowjs.html  Ann Yuan and Andrey Vakunov. 2020. Face and hand tracking in the browser with MediaPipe and TensorFlow.js. https:\/\/blog.tensorflow.org\/2020\/03\/face-and-hand-tracking-in-browser-with-mediapipe-and-tensorflowjs.html"},{"key":"e_1_3_2_2_42_1","volume-title":"Event structure in perception and conception.Psychological bulletin 127, 1","author":"Zacks M","year":"2001","unstructured":"Jeffrey\u00a0 M Zacks and Barbara Tversky . 2001. Event structure in perception and conception.Psychological bulletin 127, 1 ( 2001 ), 3. Jeffrey\u00a0M Zacks and Barbara Tversky. 2001. Event structure in perception and conception.Psychological bulletin 127, 1 (2001), 3."},{"key":"e_1_3_2_2_43_1","first-page":"29","article-title":"Perceiving, remembering, and communicating structure in events.Journal of experimental psychology","volume":"130","author":"Zacks M","year":"2001","unstructured":"Jeffrey\u00a0 M Zacks , Barbara Tversky , and Gowri Iyer . 2001 . Perceiving, remembering, and communicating structure in events.Journal of experimental psychology : General 130 , 1 (2001), 29 . Jeffrey\u00a0M Zacks, Barbara Tversky, and Gowri Iyer. 2001. Perceiving, remembering, and communicating structure in events.Journal of experimental psychology: General 130, 1 (2001), 29.","journal-title":"General"},{"key":"e_1_3_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00365"}],"event":{"name":"CHI '21: CHI Conference on Human Factors in Computing Systems","location":"Yokohama Japan","acronym":"CHI '21","sponsor":["SIGCHI ACM Special Interest Group on Computer-Human Interaction"]},"container-title":["Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3411764.3445721","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3411764.3445721","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T21:24:47Z","timestamp":1750195487000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3411764.3445721"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,5,6]]},"references-count":44,"alternative-id":["10.1145\/3411764.3445721","10.1145\/3411764"],"URL":"https:\/\/doi.org\/10.1145\/3411764.3445721","relation":{},"subject":[],"published":{"date-parts":[[2021,5,6]]},"assertion":[{"value":"2021-05-07","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}