{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,13]],"date-time":"2026-04-13T15:59:27Z","timestamp":1776095967761,"version":"3.50.1"},"reference-count":34,"publisher":"Association for Computing Machinery (ACM)","issue":"6","license":[{"start":{"date-parts":[[2015,11,2]],"date-time":"2015-11-02T00:00:00Z","timestamp":1446422400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Quanta","award":["6918829"],"award-info":[{"award-number":["6918829"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2015,11,4]]},"abstract":"<jats:p>\n            Blackboard-style lecture videos are popular, but learning using existing video player interfaces can be challenging. Viewers cannot consume the lecture material at their own pace, and the content is also difficult to search or skim. For these reasons, some people prefer lecture notes to videos. To address these limitations, we present\n            <jats:italic>Visual Transcripts<\/jats:italic>\n            , a readable representation of lecture videos that combines visual information with transcript text. To generate a Visual Transcript, we first segment the visual content of a lecture into discrete visual entities that correspond to equations, figures, or lines of text. Then, we analyze the temporal correspondence between the transcript and visuals to determine how sentences relate to visual entities. Finally, we arrange the text and visuals in a linear layout based on these relationships. We compare our result with a standard video player, and a state-of-the-art interface designed specifically for blackboard-style lecture videos. User evaluation suggests that users prefer our interface for learning and that our interface is effective in helping them browse or search through lecture videos.\n          <\/jats:p>","DOI":"10.1145\/2816795.2818123","type":"journal-article","created":{"date-parts":[[2015,10,27]],"date-time":"2015-10-27T12:36:39Z","timestamp":1445949399000},"page":"1-10","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":37,"title":["Visual transcripts"],"prefix":"10.1145","volume":"34","author":[{"given":"Hijung Valentina","family":"Shin","sequence":"first","affiliation":[{"name":"MIT CSAIL"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Floraine","family":"Berthouzoz","sequence":"additional","affiliation":[{"name":"Adobe Research"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wilmot","family":"Li","sequence":"additional","affiliation":[{"name":"Adobe Research"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fr\u00e9do","family":"Durand","sequence":"additional","affiliation":[{"name":"MIT CSAIL"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2015,11,2]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"crossref","unstructured":"Agnihotri L. Devara K. V. McGee T. and Dimitrova N. 2001. Summarization of video programs based on closed captions. In Photonics West 2001-Electronic Imaging International Society for Optics and Photonics 599--607.  Agnihotri L. Devara K. V. McGee T. and Dimitrova N. 2001. Summarization of video programs based on closed captions. In Photonics West 2001-Electronic Imaging International Society for Optics and Photonics 599--607.","DOI":"10.1117\/12.410973"},{"key":"e_1_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1778765.1778826"},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/332040.332428"},{"key":"e_1_2_2_4_1","volume-title":"Proc. of the EuroGraphics conf., State of the Art Report, Citeseer, 1--23","author":"Borgo R.","unstructured":"Borgo , R. , Chen , M. , Daubney , B. , Grundy , E. , Janicke , H. , Heidemann , G. , Hoferlin , B. , Hoferlin , M. , Weiskopf , D. , and Xie , X . 2011. A survey on video-based graphics and video visualization . In Proc. of the EuroGraphics conf., State of the Art Report, Citeseer, 1--23 . Borgo, R., Chen, M., Daubney, B., Grundy, E., Janicke, H., Heidemann, G., Hoferlin, B., Hoferlin, M., Weiskopf, D., and Xie, X. 2011. A survey on video-based graphics and video visualization. In Proc. of the EuroGraphics conf., State of the Art Report, Citeseer, 1--23."},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/2380116.2380130"},{"key":"e_1_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/2501988.2502052"},{"key":"e_1_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2007.906602"},{"key":"e_1_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2001.941193"},{"key":"e_1_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/641007.641120"},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1007\/11919629_58"},{"key":"e_1_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2003.812758"},{"key":"e_1_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/319463.319691"},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/2632111"},{"key":"e_1_2_2_14_1","volume-title":"-G","author":"Hwang W.-I.","year":"2006","unstructured":"Hwang , W.-I. , Lee , P.-J. , Chun , B.-K. , Ryu , D.-S. , and Cho , H . -G . 2006 . Cinema comics: Cartoon generation from video stream. In GRAPP , 299--304. Hwang, W.-I., Lee, P.-J., Chun, B.-K., Ryu, D.-S., and Cho, H.-G. 2006. Cinema comics: Cartoon generation from video stream. In GRAPP, 299--304."},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/2501988.2502038"},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/2642918.2647389"},{"key":"e_1_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/2556288.2556986"},{"key":"e_1_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1002\/spe.4380111102"},{"key":"e_1_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/237170.237260"},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1002\/(SICI)1097-4571(199506)46:5%3C340::AID-ASI5%3E3.0.CO;2-S"},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/332040.332425"},{"key":"e_1_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.350"},{"key":"e_1_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/2470654.2466147"},{"key":"e_1_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/302979.303108"},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2004.841694"},{"key":"e_1_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/2642918.2647400"},{"key":"e_1_2_2_27_1","volume-title":"ANSES: Summarisation of news video. In Image and Video Retrieval","author":"Pickering M. J.","year":"2003","unstructured":"Pickering , M. J. , Wong , L. , and R\u00fcger , S. M . 2003 . ANSES: Summarisation of news video. In Image and Video Retrieval . Springer , 425--434. Pickering, M. J., Wong, L., and R\u00fcger, S. M. 2003. ANSES: Summarisation of news video. In Image and Video Retrieval. Springer, 425--434."},{"key":"e_1_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/2501988.2501993"},{"key":"e_1_2_2_29_1","unstructured":"Shah D. 2014. \"MOOCs in 2014: Breaking down the numbers (edsurge news)\".  Shah D. 2014. \"MOOCs in 2014: Breaking down the numbers (edsurge news)\"."},{"key":"e_1_2_2_30_1","volume-title":"Electronic Imaging: Science & Technology","author":"Shahraray B.","year":"1995","unstructured":"Shahraray , B. , and Gibbon , D. C . 1995 . Automatic generation of pictorial transcripts of video programs. In IS&T\/SPIE's Symposium on Electronic Imaging: Science & Technology , International Society for Optics and Photonics , 512--518. Shahraray, B., and Gibbon, D. C. 1995. Automatic generation of pictorial transcripts of video programs. In IS&T\/SPIE's Symposium on Electronic Imaging: Science & Technology, International Society for Optics and Photonics, 512--518."},{"key":"e_1_2_2_31_1","volume-title":"Multimedia Signal Processing, 1997., IEEE First Workshop on, IEEE, 581--586","author":"Shahraray B.","unstructured":"Shahraray , B. , and Gibbon , D. C . 1997. Pictorial transcripts: Multimedia processing applied to digital library creation . In Multimedia Signal Processing, 1997., IEEE First Workshop on, IEEE, 581--586 . Shahraray, B., and Gibbon, D. C. 1997. Pictorial transcripts: Multimedia processing applied to digital library creation. In Multimedia Signal Processing, 1997., IEEE First Workshop on, IEEE, 581--586."},{"key":"e_1_2_2_32_1","volume-title":"Content-Based Access of Image and Video Database, 1998. Proceedings., 1998 IEEE International Workshop on, IEEE, 61--70","author":"Smith M. A.","unstructured":"Smith , M. A. , and Kanade , T . 1998. Video skimming and characterization through the combination of image and language understanding . In Content-Based Access of Image and Video Database, 1998. Proceedings., 1998 IEEE International Workshop on, IEEE, 61--70 . Smith, M. A., and Kanade, T. 1998. Video skimming and characterization through the combination of image and language understanding. In Content-Based Access of Image and Video Database, 1998. Proceedings., 1998 IEEE International Workshop on, IEEE, 61--70."},{"key":"e_1_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/1198302.1198305"},{"key":"e_1_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/319463.319654"}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2816795.2818123","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2816795.2818123","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T05:48:19Z","timestamp":1750225699000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2816795.2818123"}},"subtitle":["lecture notes from blackboard-style lecture videos"],"short-title":[],"issued":{"date-parts":[[2015,11,2]]},"references-count":34,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2015,11,4]]}},"alternative-id":["10.1145\/2816795.2818123"],"URL":"https:\/\/doi.org\/10.1145\/2816795.2818123","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"value":"0730-0301","type":"print"},{"value":"1557-7368","type":"electronic"}],"subject":[],"published":{"date-parts":[[2015,11,2]]},"assertion":[{"value":"2015-11-02","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}