{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,7,9]],"date-time":"2025-07-09T22:45:52Z","timestamp":1752101152166,"version":"3.41.0"},"reference-count":44,"publisher":"Association for Computing Machinery (ACM)","issue":"2s","license":[{"start":{"date-parts":[[2020,4,30]],"date-time":"2020-04-30T00:00:00Z","timestamp":1588204800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"National Key R8D Program of China","award":["2018YFB1003703"],"award-info":[{"award-number":["2018YFB1003703"]}]},{"DOI":"10.13039\/501100001809","name":"NSFC","doi-asserted-by":"crossref","award":["61936011, 61521002"],"award-info":[{"award-number":["61936011, 61521002"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Beijing Key Lab of Networked Multimedia"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2020,4,30]]},"abstract":"<jats:p>Scheduling viewers effectively among different Content Delivery Network (CDN) providers is challenging owing to the extreme diversity in the crowdsourced live streaming (CLS) scenarios. Abundant algorithms have been proposed in recent years, which, however, suffer from a critical limitation: Due to their inaccurate feature engineering or naive rules, they cannot optimally schedule viewers. To address this concern, we put forward LTS (Learn to Schedule), a novel scheduling algorithm that can adapt to the dynamics from both viewer traffics and CDN performance. In detail, we first propose LTS-RL, an approach that schedules CLS viewers based on deep reinforcement learning (DRL). Since LTS-RL is trained in an end-to-end way, it can automatically learn scheduling algorithms without any pre-programmed models or assumptions about the environment dynamics. At the same time, to practically deploy LTS-RL, we then use the decision tree and imitation learning to convert LTS-RL into a more light-weighted and interpretable model, which is denoted as Fast-LTS. After the extensive evaluation of the real data from a leading CLS platform in China, we demonstrate that our proposed model (both LTS-RL and Fast-LTS) can improve the average quality of experience (QoE) over state-of-the-art approaches by 8.71--15.63%. At the same time, we also demonstrate that Fast-LTS can faithfully convert the complicated LTS-RL with slight performance degradation (&lt; 2%), while significantly reducing the decision time (\u00d77--10).<\/jats:p>","DOI":"10.1145\/3397226","type":"journal-article","created":{"date-parts":[[2020,6,22]],"date-time":"2020-06-22T02:49:20Z","timestamp":1592794160000},"page":"1-22","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["A Practical Learning-based Approach for Viewer Scheduling in the Crowdsourced Live Streaming"],"prefix":"10.1145","volume":"16","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-4251-825X","authenticated-orcid":false,"given":"Rui-Xiao","family":"Zhang","sequence":"first","affiliation":[{"name":"BNRist, Department of Computer Science and Technology, Tsinghua University; Key Laboratory of Pervasive Computing, Ministry of Education, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ming","family":"Ma","sequence":"additional","affiliation":[{"name":"Kuaishou Technology Co., Ltd, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tianchi","family":"Huang","sequence":"additional","affiliation":[{"name":"BNRist, Department of Computer Science and Technology, Tsinghua University, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Haitian","family":"Pang","sequence":"additional","affiliation":[{"name":"BNRist, Department of Computer Science and Technology, Tsinghua University, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3563-3951","authenticated-orcid":false,"given":"Xin","family":"Yao","sequence":"additional","affiliation":[{"name":"Key Laboratory of Pervasive Computing, Ministry of Education, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chenglei","family":"Wu","sequence":"additional","affiliation":[{"name":"Key Laboratory of Pervasive Computing, Ministry of Education, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Lifeng","family":"Sun","sequence":"additional","affiliation":[{"name":"BNRist, Department of Computer Science and Technology, Tsinghua University; Key Laboratory of Pervasive Computing, Ministry of Education, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,6,21]]},"reference":[{"volume-title":"Proceedings of the IEEE International Conference on Computer Communications (INFOCOM\u201912)","author":"Kumar Vijay","key":"e_1_2_1_1_1","unstructured":"Vijay Kumar Adhikari et al. 2012. Unreeling netflix: Understanding and improving multi-cdn movie delivery . In Proceedings of the IEEE International Conference on Computer Communications (INFOCOM\u201912) . IEEE, 1620--1628. Vijay Kumar Adhikari et al. 2012. Unreeling netflix: Understanding and improving multi-cdn movie delivery. In Proceedings of the IEEE International Conference on Computer Communications (INFOCOM\u201912). IEEE, 1620--1628."},{"volume-title":"Proceedings of the 2012 IEEE Conference on Computer Communications Workshops (INFOCOM WKSHPS\u201912)","author":"Kumar Vijay","key":"e_1_2_1_2_1","unstructured":"Vijay Kumar Adhikari et al. 2012. A tale of three CDNs: An active measurement study of Hulu and its CDNs . In Proceedings of the 2012 IEEE Conference on Computer Communications Workshops (INFOCOM WKSHPS\u201912) . IEEE, 7--12. Vijay Kumar Adhikari et al. 2012. A tale of three CDNs: An active measurement study of Hulu and its CDNs. In Proceedings of the 2012 IEEE Conference on Computer Communications Workshops (INFOCOM WKSHPS\u201912). IEEE, 7--12."},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.110"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0004-3702(98)00034-4"},{"volume-title":"Classification and Regression Trees","author":"Breiman Leo","key":"e_1_2_1_5_1","unstructured":"Leo Breiman . 2017. Classification and Regression Trees . Routledge . Leo Breiman. 2017. Classification and Regression Trees. Routledge."},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/3341216.3342210"},{"key":"e_1_2_1_7_1","volume-title":"ACM SIGCOMM Computer Communication Review","volume":"41","author":"Florin","unstructured":"Florin Dobrian and et al. 2011. Understanding the impact of video quality on user engagement . In ACM SIGCOMM Computer Communication Review , Vol. 41 . ACM, 362--373. Florin Dobrian and et al. 2011. Understanding the impact of video quality on user engagement. In ACM SIGCOMM Computer Communication Review, Vol. 41. ACM, 362--373."},{"key":"e_1_2_1_8_1","volume-title":"Proceedings of the International Conference on Machine Learning. 1928--1937","author":"Mnih","year":"2016","unstructured":"Mnih et al. 2016 . Asynchronous methods for deep reinforcement learning . In Proceedings of the International Conference on Machine Learning. 1928--1937 . Mnih et al. 2016. Asynchronous methods for deep reinforcement learning. In Proceedings of the International Conference on Machine Learning. 1928--1937."},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2015.2460193"},{"volume-title":"2016 31st Youth Academic Annual Conference of Chinese Association of Automation (YAC\u201916)","author":"Fu Rui","key":"e_1_2_1_10_1","unstructured":"Rui Fu , Zuo Zhang , and L. Li . 2016. Using LSTM and GRU neural network methods for traffic flow prediction . In 2016 31st Youth Academic Annual Conference of Chinese Association of Automation (YAC\u201916) . IEEE, 324\u2013328. Rui Fu, Zuo Zhang, and L. Li. 2016. Using LSTM and GRU neural network methods for traffic flow prediction. In 2016 31st Youth Academic Annual Conference of Chinese Association of Automation (YAC\u201916). IEEE, 324\u2013328."},{"key":"e_1_2_1_11_1","volume-title":"On upper-confidence bound policies for non-stationary bandit problems. arXiv preprint arXiv:0805.3415","author":"Garivier Aur\u00e9lien","year":"2008","unstructured":"Aur\u00e9lien Garivier and Eric Moulines . 2008. On upper-confidence bound policies for non-stationary bandit problems. arXiv preprint arXiv:0805.3415 ( 2008 ). Aur\u00e9lien Garivier and Eric Moulines. 2008. On upper-confidence bound policies for non-stationary bandit problems. arXiv preprint arXiv:0805.3415 (2008)."},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/3243734.3243792"},{"key":"e_1_2_1_13_1","volume-title":"QARC: Video quality aware rate control for real-time video streaming based on deep reinforcement learning. arXiv preprint arXiv:1805.02482","author":"Huang Tianchi","year":"2018","unstructured":"Tianchi Huang , Rui-Xiao Zhang , Chao Zhou , and Lifeng Sun . 2018 . QARC: Video quality aware rate control for real-time video streaming based on deep reinforcement learning. arXiv preprint arXiv:1805.02482 (2018). Tianchi Huang, Rui-Xiao Zhang, Chao Zhou, and Lifeng Sun. 2018. QARC: Video quality aware rate control for real-time video streaming based on deep reinforcement learning. arXiv preprint arXiv:1805.02482 (2018)."},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/3343031.3351014"},{"volume-title":"Proceedings of the USENIX Symposium on Networked Systems Design and Implementation (NSDI\u201916)","author":"Junchen","key":"e_1_2_1_15_1","unstructured":"Junchen Jiang et al. 2016. CFA: A practical prediction system for video QoE optimization . In Proceedings of the USENIX Symposium on Networked Systems Design and Implementation (NSDI\u201916) . 137--150. Junchen Jiang et al. 2016. CFA: A practical prediction system for video QoE optimization. In Proceedings of the USENIX Symposium on Networked Systems Design and Implementation (NSDI\u201916). 137--150."},{"key":"e_1_2_1_16_1","volume-title":"Proceedings of the USENIX Symposium on Networked Systems Design and Implementation (NSDI\u201917)","volume":"1","author":"Junchen","unstructured":"Junchen Jiang et al. 2017. Pytheas: Enabling data-driven quality of experience optimization using group-based exploration-exploitation . In Proceedings of the USENIX Symposium on Networked Systems Design and Implementation (NSDI\u201917) , Vol. 1 . 3. Junchen Jiang et al. 2017. Pytheas: Enabling data-driven quality of experience optimization using group-based exploration-exploitation. In Proceedings of the USENIX Symposium on Networked Systems Design and Implementation (NSDI\u201917), Vol. 1. 3."},{"key":"e_1_2_1_17_1","volume-title":"Twitch Ended 2017 With 15 Million Daily Visitors, 27K Partnered Streamers. Retrieved February, 6","author":"Klein Jessica","year":"2018","unstructured":"Jessica Klein . 2018. Twitch Ended 2017 With 15 Million Daily Visitors, 27K Partnered Streamers. Retrieved February, 6 , 2018 from https:\/\/www.tubefilter.com\/2018\/02\/06\/twitch-2017-year-in-review\/. Jessica Klein. 2018. Twitch Ended 2017 With 15 Million Daily Visitors, 27K Partnered Streamers. Retrieved February, 6, 2018 from https:\/\/www.tubefilter.com\/2018\/02\/06\/twitch-2017-year-in-review\/."},{"key":"e_1_2_1_18_1","first-page":"371","article-title":"Optimizing cost and performance for content multihoming","volume":"42","author":"Liu Hongqiang Harry","year":"2012","unstructured":"Hongqiang Harry Liu , Ye Wang , Yang Richard Yang , Hao Wang , and Chen Tian . 2012 . Optimizing cost and performance for content multihoming . ACM Spec. Interest Group Data Commun. 42 , 4 (2012), 371 -- 382 . Hongqiang Harry Liu, Ye Wang, Yang Richard Yang, Hao Wang, and Chen Tian. 2012. Optimizing cost and performance for content multihoming. ACM Spec. Interest Group Data Commun. 42, 4 (2012), 371--382.","journal-title":"ACM Spec. Interest Group Data Commun."},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/2377677.2377752"},{"key":"e_1_2_1_20_1","volume-title":"Lundberg and Su-In Lee","author":"Scott","year":"2017","unstructured":"Scott M. Lundberg and Su-In Lee . 2017 . A unified approach to interpreting model predictions. In Advances in Neural Information Processing Systems . 4765--4774. Scott M. Lundberg and Su-In Lee. 2017. A unified approach to interpreting model predictions. In Advances in Neural Information Processing Systems. 4765--4774."},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/2797211"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1145\/3098822.3098843"},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/3343031.3350866"},{"key":"e_1_2_1_24_1","volume-title":"Explaining deep learning-based networked systems. arXiv preprint arXiv:1910.03835","author":"Meng Zili","year":"2019","unstructured":"Zili Meng , Minhu Wang , Mingwei Xu , Hongzi Mao , Jiasong Bai , and Hongxin Hu. 2019. Explaining deep learning-based networked systems. arXiv preprint arXiv:1910.03835 ( 2019 ). Zili Meng, Minhu Wang, Mingwei Xu, Hongzi Mao, Jiasong Bai, and Hongxin Hu. 2019. Explaining deep learning-based networked systems. arXiv preprint arXiv:1910.03835 (2019)."},{"key":"e_1_2_1_25_1","doi-asserted-by":"crossref","unstructured":"Volodymyr Mnih et al. 2015. Human-level control through deep reinforcement learning. Nature 518 7540 (2015) 529.  Volodymyr Mnih et al. 2015. Human-level control through deep reinforcement learning. Nature 518 7540 (2015) 529.","DOI":"10.1038\/nature14236"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/3240508.3240642"},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/2939672.2939778"},{"key":"e_1_2_1_28_1","volume-title":"Proceedings of the 14th International Conference on Artificial Intelligence and Statistics. 627--635","author":"Ross St\u00e9phane","year":"2011","unstructured":"St\u00e9phane Ross , Geoffrey Gordon , and Drew Bagnell . 2011 . A reduction of imitation learning and structured prediction to no-regret online learning . In Proceedings of the 14th International Conference on Artificial Intelligence and Statistics. 627--635 . St\u00e9phane Ross, Geoffrey Gordon, and Drew Bagnell. 2011. A reduction of imitation learning and structured prediction to no-regret online learning. In Proceedings of the 14th International Conference on Artificial Intelligence and Statistics. 627--635."},{"key":"e_1_2_1_29_1","volume-title":"Proceedings of the 31st International Conference on International Conference on Machine Learning","volume":"32","author":"Silver David","year":"2014","unstructured":"David Silver , Guy Lever , Nicolas Heess , Thomas Degris , Daan Wierstra , and Martin Riedmiller . 2014 . Deterministic policy gradient algorithms . In Proceedings of the 31st International Conference on International Conference on Machine Learning , Vol. 32 . I--387. David Silver, Guy Lever, Nicolas Heess, Thomas Degris, Daan Wierstra, and Martin Riedmiller. 2014. Deterministic policy gradient algorithms. In Proceedings of the 31st International Conference on International Conference on Machine Learning, Vol. 32. I--387."},{"key":"e_1_2_1_30_1","volume-title":"Proceedings of the IEEE International Conference on Computer Communications (INFOCOM\u201900)","volume":"1","author":"Mark","unstructured":"Mark Stemm and et al. 2000. A network measurement architecture for adaptive applications . In Proceedings of the IEEE International Conference on Computer Communications (INFOCOM\u201900) . Vol. 1 . IEEE, 285--294. Mark Stemm and et al. 2000. A network measurement architecture for adaptive applications. In Proceedings of the IEEE International Conference on Computer Communications (INFOCOM\u201900). Vol. 1. IEEE, 285--294."},{"key":"e_1_2_1_31_1","volume-title":"Barto","author":"Sutton Richard S.","year":"1998","unstructured":"Richard S. Sutton and Andrew G . Barto . 1998 . Reinforcement Learning : An Introduction. Vol. 1 . MIT Press , Cambridge, MA. Richard S. Sutton and Andrew G. Barto. 1998. Reinforcement Learning: An Introduction. Vol. 1. MIT Press, Cambridge, MA."},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/MWC.2017.1700244"},{"volume-title":"Proceedings of the 2011 31st International Conference on Distributed Computing Systems (ICDCS\u201911)","author":"Ruben","key":"e_1_2_1_33_1","unstructured":"Ruben Torres et al. 2011. Dissecting video server selection strategies in the youtube cdn . In Proceedings of the 2011 31st International Conference on Distributed Computing Systems (ICDCS\u201911) . IEEE, 248--257. Ruben Torres et al. 2011. Dissecting video server selection strategies in the youtube cdn. In Proceedings of the 2011 31st International Conference on Distributed Computing Systems (ICDCS\u201911). IEEE, 248--257."},{"key":"e_1_2_1_34_1","unstructured":"Andrew Trask et al. 2018. Neural arithmetic logic units. In Advances in Neural Information Processing Systems. 8046--8055.  Andrew Trask et al. 2018. Neural arithmetic logic units. In Advances in Neural Information Processing Systems. 8046--8055."},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICC.2014.6883800"},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/TWC.2017.2769644"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jpdc.2011.10.003"},{"key":"e_1_2_1_38_1","unstructured":"Zhiyuan Xu and etal 2018. Experience-driven networking: A deep reinforcement learning based approach. arXiv preprint arXiv:1801.05757 (2018).  Zhiyuan Xu and et al. 2018. Experience-driven networking: A deep reinforcement learning based approach. arXiv preprint arXiv:1801.05757 (2018)."},{"volume-title":"Proceedings of the 2017 ACM on Multimedia Conference. ACM, 73--81","author":"Bo","key":"e_1_2_1_39_1","unstructured":"Bo Yan and et al. 2017. LiveJack: Integrating CDNs and edge clouds for live content broadcasting . In Proceedings of the 2017 ACM on Multimedia Conference. ACM, 73--81 . Bo Yan and et al. 2017. LiveJack: Integrating CDNs and edge clouds for live content broadcasting. In Proceedings of the 2017 ACM on Multimedia Conference. ACM, 73--81."},{"key":"e_1_2_1_40_1","volume-title":"Proceedings of the 2018 USENIX Annual Technical Conference (USENIX ATC\u201918)","author":"Yan Francis Y.","year":"2018","unstructured":"Francis Y. Yan , Jestin Ma , Greg D. Hill , Deepti Raghavan , Riad S. Wahby , Philip Levis , and Keith Winstein . 2018 . Pantheon: The training ground for Internet congestion-control research . In Proceedings of the 2018 USENIX Annual Technical Conference (USENIX ATC\u201918) . 731--743. Francis Y. Yan, Jestin Ma, Greg D. Hill, Deepti Raghavan, Riad S. Wahby, Philip Levis, and Keith Winstein. 2018. Pantheon: The training ground for Internet congestion-control research. In Proceedings of the 2018 USENIX Annual Technical Conference (USENIX ATC\u201918). 731--743."},{"volume-title":"Proceedings of the 25th ACM International Conference on Multimedia. ACM, 1372--1380","author":"Zhu","key":"e_1_2_1_41_1","unstructured":"Zhu Yifei and et al. 2017. When cloud meets uncertain crowd: An auction approach for crowdsourced livecast transcoding . In Proceedings of the 25th ACM International Conference on Multimedia. ACM, 1372--1380 . Zhu Yifei and et al. 2017. When cloud meets uncertain crowd: An auction approach for crowdsourced livecast transcoding. In Proceedings of the 25th ACM International Conference on Multimedia. ACM, 1372--1380."},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/3304112.3325607"},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/3304112.3325607"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1145\/3232565.3232569"}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3397226","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3397226","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T21:31:37Z","timestamp":1750195897000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3397226"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,4,30]]},"references-count":44,"journal-issue":{"issue":"2s","published-print":{"date-parts":[[2020,4,30]]}},"alternative-id":["10.1145\/3397226"],"URL":"https:\/\/doi.org\/10.1145\/3397226","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"type":"print","value":"1551-6857"},{"type":"electronic","value":"1551-6865"}],"subject":[],"published":{"date-parts":[[2020,4,30]]},"assertion":[{"value":"2019-12-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2020-04-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2020-06-21","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}