{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,2]],"date-time":"2026-05-02T15:19:40Z","timestamp":1777735180232,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":43,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,10,19]],"date-time":"2020-10-19T00:00:00Z","timestamp":1603065600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"the Key Research and Development Project of Jiangsu Province","award":["Grant No. BE2019104"],"award-info":[{"award-number":["Grant No. BE2019104"]}]},{"name":"National Natural Science Foundation of China","award":["Grant No. 61672276"],"award-info":[{"award-number":["Grant No. 61672276"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,10,19]]},"DOI":"10.1145\/3340531.3411913","type":"proceedings-article","created":{"date-parts":[[2020,10,19]],"date-time":"2020-10-19T05:31:05Z","timestamp":1603085465000},"page":"1355-1364","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":18,"title":["Auxiliary-task Based Deep Reinforcement Learning for Participant Selection Problem in Mobile Crowdsourcing"],"prefix":"10.1145","author":[{"given":"Wei","family":"Shen","sequence":"first","affiliation":[{"name":"Baidu Inc., Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiaonan","family":"He","sequence":"additional","affiliation":[{"name":"Baidu Inc., Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chuheng","family":"Zhang","sequence":"additional","affiliation":[{"name":"Tsinghua University, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Qiang","family":"Ni","sequence":"additional","affiliation":[{"name":"Lancaster University, Lancaster, United Kingdom"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wanchun","family":"Dou","sequence":"additional","affiliation":[{"name":"Nanjing University, Nanjing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yan","family":"Wang","sequence":"additional","affiliation":[{"name":"Tencent Holdings Ltd., Guangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,10,19]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/3081333.3081345"},{"key":"e_1_3_2_2_2_1","volume-title":"AlphaStar: An Evolutionary Computation Perspective. arXiv preprint arXiv:1902.01724","author":"Arulkumaran Kai","year":"2019","unstructured":"Kai Arulkumaran , Antoine Cully , and Julian Togelius . 2019. AlphaStar: An Evolutionary Computation Perspective. arXiv preprint arXiv:1902.01724 ( 2019 ). Kai Arulkumaran, Antoine Cully, and Julian Togelius. 2019. AlphaStar: An Evolutionary Computation Perspective. arXiv preprint arXiv:1902.01724 (2019)."},{"key":"e_1_3_2_2_3_1","volume-title":"Jamie Ryan Kiros, and Geoffrey E Hinton","author":"Ba Jimmy Lei","year":"2016","unstructured":"Jimmy Lei Ba , Jamie Ryan Kiros, and Geoffrey E Hinton . 2016 . Layer normalization. arXiv preprint arXiv:1607.06450 (2016). Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton. 2016. Layer normalization. arXiv preprint arXiv:1607.06450 (2016)."},{"key":"e_1_3_2_2_4_1","volume-title":"Proceedings of the International Conference on Neural Information Processing Systems. 4352--4360","author":"Bonald Thomas","year":"2017","unstructured":"Thomas Bonald and Richard Combes . 2017 . A Minimax Optimal Algorithm for Crowdsourcing . In Proceedings of the International Conference on Neural Information Processing Systems. 4352--4360 . Thomas Bonald and Richard Combes. 2017. A Minimax Optimal Algorithm for Crowdsourcing. In Proceedings of the International Conference on Neural Information Processing Systems. 4352--4360."},{"key":"e_1_3_2_2_5_1","volume-title":"Openai gym. arXiv preprint arXiv:1606.01540","author":"Brockman Greg","year":"2016","unstructured":"Greg Brockman , Vicki Cheung , Ludwig Pettersson , Jonas Schneider , John Schulman , Jie Tang , and Wojciech Zaremba . 2016. Openai gym. arXiv preprint arXiv:1606.01540 ( 2016 ). Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016. Openai gym. arXiv preprint arXiv:1606.01540 (2016)."},{"key":"e_1_3_2_2_6_1","unstructured":"Tom B Brown Benjamin Mann Nick Ryder Melanie Subbiah Jared Kaplan Prafulla Dhariwal Arvind Neelakantan Pranav Shyam Girish Sastry Amanda Askell etal 2020. Language models are few-shot learners. arXiv preprint arXiv:2005.14165 (2020).  Tom B Brown Benjamin Mann Nick Ryder Melanie Subbiah Jared Kaplan Prafulla Dhariwal Arvind Neelakantan Pranav Shyam Girish Sastry Amanda Askell et al. 2020. Language models are few-shot learners. arXiv preprint arXiv:2005.14165 (2020)."},{"key":"e_1_3_2_2_7_1","volume-title":"The Value-Improvement Path: Towards Better Representations for Reinforcement Learning. arXiv preprint arXiv:2006.02243","author":"Dabney Will","year":"2020","unstructured":"Will Dabney , Andr\u00e9 Barreto , Mark Rowland , Robert Dadashi , John Quan , Marc G Bellemare , and David Silver . 2020. The Value-Improvement Path: Towards Better Representations for Reinforcement Learning. arXiv preprint arXiv:2006.02243 ( 2020 ). Will Dabney, Andr\u00e9 Barreto, Mark Rowland, Robert Dadashi, John Quan, Marc G Bellemare, and David Silver. 2020. The Value-Improvement Path: Towards Better Representations for Reinforcement Learning. arXiv preprint arXiv:2006.02243 (2020)."},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2016.7498417"},{"key":"e_1_3_2_2_9_1","volume-title":"Bilal Piot, Jean-bastien Grill, Florent Altch\u00e9, R\u00e9mi Munos, and Mohammad Gheshlaghi Azar.","author":"Guo Daniel","year":"2020","unstructured":"Daniel Guo , Bernardo Avila Pires , Bilal Piot, Jean-bastien Grill, Florent Altch\u00e9, R\u00e9mi Munos, and Mohammad Gheshlaghi Azar. 2020 . Bootstrap Latent-Predictive Representations for Multitask Reinforcement Learning . arXiv preprint arXiv:2004.14646 (2020). Daniel Guo, Bernardo Avila Pires, Bilal Piot, Jean-bastien Grill, Florent Altch\u00e9, R\u00e9mi Munos, and Mohammad Gheshlaghi Azar. 2020. Bootstrap Latent-Predictive Representations for Multitask Reinforcement Learning. arXiv preprint arXiv:2004.14646 (2020)."},{"key":"e_1_3_2_2_10_1","volume-title":"World models. arXiv preprint arXiv:1803.10122","author":"Ha David","year":"2018","unstructured":"David Ha and Jrgen Schmidhuber . 2018. World models. arXiv preprint arXiv:1803.10122 ( 2018 ). David Ha and Jrgen Schmidhuber. 2018. World models. arXiv preprint arXiv:1803.10122 (2018)."},{"key":"e_1_3_2_2_11_1","volume-title":"2019 a. Dream to control: Learning behaviors by latent imagination. arXiv preprint arXiv:1912.01603","author":"Hafner Danijar","year":"2019","unstructured":"Danijar Hafner , Timothy Lillicrap , Jimmy Ba , and Mohammad Norouzi . 2019 a. Dream to control: Learning behaviors by latent imagination. arXiv preprint arXiv:1912.01603 ( 2019 ). Danijar Hafner, Timothy Lillicrap, Jimmy Ba, and Mohammad Norouzi. 2019 a. Dream to control: Learning behaviors by latent imagination. arXiv preprint arXiv:1912.01603 (2019)."},{"key":"e_1_3_2_2_12_1","volume-title":"International Conference on Machine Learning. 2555--2565","author":"Hafner Danijar","year":"2019","unstructured":"Danijar Hafner , Timothy Lillicrap , Ian Fischer , Ruben Villegas , David Ha , Honglak Lee , and James Davidson . 2019 b. Learning latent dynamics for planning from pixels . In International Conference on Machine Learning. 2555--2565 . Danijar Hafner, Timothy Lillicrap, Ian Fischer, Ruben Villegas, David Ha, Honglak Lee, and James Davidson. 2019 b. Learning latent dynamics for planning from pixels. In International Conference on Machine Learning. 2555--2565."},{"key":"e_1_3_2_2_13_1","volume-title":"Batch normalization: Accelerating deep network training by reducing internal covariate shift. arXiv preprint arXiv:1502.03167","author":"Ioffe Sergey","year":"2015","unstructured":"Sergey Ioffe and Christian Szegedy . 2015. Batch normalization: Accelerating deep network training by reducing internal covariate shift. arXiv preprint arXiv:1502.03167 ( 2015 ). Sergey Ioffe and Christian Szegedy. 2015. Batch normalization: Accelerating deep network training by reducing internal covariate shift. arXiv preprint arXiv:1502.03167 (2015)."},{"key":"e_1_3_2_2_14_1","volume-title":"Tom Schaul, Joel Z Leibo, David Silver, and Koray Kavukcuoglu.","author":"Jaderberg Max","year":"2016","unstructured":"Max Jaderberg , Volodymyr Mnih , Wojciech Marian Czarnecki , Tom Schaul, Joel Z Leibo, David Silver, and Koray Kavukcuoglu. 2016 . Reinforcement learning with unsupervised auxiliary tasks. arXiv preprint arXiv:1611.05397 (2016). Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki, Tom Schaul, Joel Z Leibo, David Silver, and Koray Kavukcuoglu. 2016. Reinforcement learning with unsupervised auxiliary tasks. arXiv preprint arXiv:1611.05397 (2016)."},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3219819.3220033"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/3219819.3219993"},{"key":"e_1_3_2_2_17_1","first-page":"15","article-title":"A Survey of the Use of Crowdsourcing in Software Engineering","volume":"126","author":"Mao Ke","year":"2016","unstructured":"Ke Mao , Licia Capra , Mark Harman , and Yue Jia . 2016 . A Survey of the Use of Crowdsourcing in Software Engineering . Journal of Systems and Software , Vol. 126 , 22 (2016), 15 -- 27 . Ke Mao, Licia Capra, Mark Harman, and Yue Jia. 2016. A Survey of the Use of Crowdsourcing in Software Engineering. Journal of Systems and Software, Vol. 126, 22 (2016), 15--27.","journal-title":"Journal of Systems and Software"},{"key":"e_1_3_2_2_18_1","volume-title":"Proceedings of the 27th International Conference on Machine Learning. 807--814","author":"Nair Vinod","unstructured":"Vinod Nair and Geoffrey E. Hinton . 2010. Rectified Linear Units Improve Restricted Boltzmann Machines . In Proceedings of the 27th International Conference on Machine Learning. 807--814 . Vinod Nair and Geoffrey E. Hinton. 2010. Rectified Linear Units Improve Restricted Boltzmann Machines. In Proceedings of the 27th International Conference on Machine Learning. 807--814."},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v31i1.10708"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v30i1.9941"},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1093\/comjnl\/bxx015"},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/INFOCOM.2016.7524548"},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.24963\/ijcai.2019\/958"},{"key":"e_1_3_2_2_24_1","unstructured":"Alec Radford Jeffrey Wu Rewon Child David Luan Dario Amodei and Ilya Sutskever. 2019. Language models are unsupervised multitask learners. https:\/\/openai.com\/blog\/better-language-models.  Alec Radford Jeffrey Wu Rewon Child David Luan Dario Amodei and Ilya Sutskever. 2019. Language models are unsupervised multitask learners. https:\/\/openai.com\/blog\/better-language-models."},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/PERCOMW.2017.7917529"},{"key":"e_1_3_2_2_26_1","volume-title":"Proceedings of Advances in Neural Information Processing Systems. 2483--2493","author":"Santurkar Shibani","year":"2018","unstructured":"Shibani Santurkar , Dimitris Tsipras , Andrew Ilyas , and Aleksander Madry . 2018 . How does batch normalization help optimization? . In Proceedings of Advances in Neural Information Processing Systems. 2483--2493 . Shibani Santurkar, Dimitris Tsipras, Andrew Ilyas, and Aleksander Madry. 2018. How does batch normalization help optimization?. In Proceedings of Advances in Neural Information Processing Systems. 2483--2493."},{"key":"e_1_3_2_2_27_1","unstructured":"Todd W. Schneider. 2018a. Analyzing 1.1 Billion NYC Taxi and Uber Trips with a Vengeance. https:\/\/toddwschneider.com\/posts\/analyzing-1--1-billion-nyc-taxi-and-uber-trips-with-a-vengeance.  Todd W. Schneider. 2018a. Analyzing 1.1 Billion NYC Taxi and Uber Trips with a Vengeance. https:\/\/toddwschneider.com\/posts\/analyzing-1--1-billion-nyc-taxi-and-uber-trips-with-a-vengeance."},{"key":"e_1_3_2_2_28_1","unstructured":"Todd W. Schneider. 2018b. New York City Taxi and For-Hire Vehicle Data. https:\/\/github.com\/toddwschneider\/nyc-taxi-data.  Todd W. Schneider. 2018b. New York City Taxi and For-Hire Vehicle Data. https:\/\/github.com\/toddwschneider\/nyc-taxi-data."},{"key":"e_1_3_2_2_29_1","volume-title":"Proximal policy optimization algorithms. arXiv preprint arXiv:1707.06347","author":"Schulman John","year":"2017","unstructured":"John Schulman , Filip Wolski , Prafulla Dhariwal , Alec Radford , and Oleg Klimov . 2017. Proximal policy optimization algorithms. arXiv preprint arXiv:1707.06347 ( 2017 ). John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017. Proximal policy optimization algorithms. arXiv preprint arXiv:1707.06347 (2017)."},{"key":"e_1_3_2_2_30_1","volume-title":"Curl: Contrastive unsupervised representations for reinforcement learning. arXiv preprint arXiv:2004.04136","author":"Srinivas Aravind","year":"2020","unstructured":"Aravind Srinivas , Michael Laskin , and Pieter Abbeel . 2020 . Curl: Contrastive unsupervised representations for reinforcement learning. arXiv preprint arXiv:2004.04136 (2020). Aravind Srinivas, Michael Laskin, and Pieter Abbeel. 2020. Curl: Contrastive unsupervised representations for reinforcement learning. arXiv preprint arXiv:2004.04136 (2020)."},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/3292500.3330793"},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2016.7498228"},{"key":"e_1_3_2_2_33_1","volume-title":"Proceedings of the International Conference on Neural Information Processing Systems. 5998--6008","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani , Noam Shazeer , Niki Parmar , Jakob Uszkoreit , Llion Jones , Aidan N. Gomez , Lukasz Kaiser , and Illia Polosukhin . 2017 . Attention is All you Need . In Proceedings of the International Conference on Neural Information Processing Systems. 5998--6008 . Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017. Attention is All you Need. In Proceedings of the International Conference on Neural Information Processing Systems. 5998--6008."},{"key":"e_1_3_2_2_34_1","unstructured":"Vivek Veeriah Matteo Hessel Zhongwen Xu Janarthanan Rajendran Richard L Lewis Junhyuk Oh Hado P van Hasselt David Silver and Satinder Singh. 2019. Discovery of useful questions as auxiliary tasks. In Advances in Neural Information Processing Systems. 9310--9321.  Vivek Veeriah Matteo Hessel Zhongwen Xu Janarthanan Rajendran Richard L Lewis Junhyuk Oh Hado P van Hasselt David Silver and Satinder Singh. 2019. Discovery of useful questions as auxiliary tasks. In Advances in Neural Information Processing Systems. 9310--9321."},{"key":"e_1_3_2_2_35_1","volume-title":"Michelle Yeo, Alireza Makhzani, Heinrich K\u00fcttler, John Agapiou, Julian Schrittwieser, et al.","author":"Vinyals Oriol","year":"2017","unstructured":"Oriol Vinyals , Timo Ewalds , Sergey Bartunov , Petko Georgiev , Alexander Sasha Vezhnevets , Michelle Yeo, Alireza Makhzani, Heinrich K\u00fcttler, John Agapiou, Julian Schrittwieser, et al. 2017 . Starcraft ii: A new challenge for reinforcement learning. arXiv preprint arXiv:1708.04782 (2017). Oriol Vinyals, Timo Ewalds, Sergey Bartunov, Petko Georgiev, Alexander Sasha Vezhnevets, Michelle Yeo, Alireza Makhzani, Heinrich K\u00fcttler, John Agapiou, Julian Schrittwieser, et al. 2017. Starcraft ii: A new challenge for reinforcement learning. arXiv preprint arXiv:1708.04782 (2017)."},{"key":"e_1_3_2_2_36_1","unstructured":"Oriol Vinyals Meire Fortunato and Navdeep Jaitly. 2015. Pointer networks. In Proceeding of Advances in Neural Information Processing Systems. 2692--2700.  Oriol Vinyals Meire Fortunato and Navdeep Jaitly. 2015. Pointer networks. In Proceeding of Advances in Neural Information Processing Systems. 2692--2700."},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/3219819.3219900"},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1007\/bf00992696"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/3219819.3219824"},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v32i1.11836"},{"key":"e_1_3_2_2_41_1","volume-title":"Proceeding of International Conference on Learning Representations.","author":"Zambaldi Vinicius","year":"2019","unstructured":"Vinicius Zambaldi , David Raposo , Adam Santoro , Victor Bapst , Yujia Li , Igor Babuschkin , Karl Tuyls , David Reichert , Timothy Lillicrap , and Edward Lockhart . 2019 . Relational Deep Reinforcement Learning . In Proceeding of International Conference on Learning Representations. Vinicius Zambaldi, David Raposo, Adam Santoro, Victor Bapst, Yujia Li, Igor Babuschkin, Karl Tuyls, David Reichert, Timothy Lillicrap, and Edward Lockhart. 2019. Relational Deep Reinforcement Learning. In Proceeding of International Conference on Learning Representations."},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/2632048.2632059"},{"key":"e_1_3_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/3097983.3098138"}],"event":{"name":"CIKM '20: The 29th ACM International Conference on Information and Knowledge Management","location":"Virtual Event Ireland","acronym":"CIKM '20","sponsor":["SIGWEB ACM Special Interest Group on Hypertext, Hypermedia, and Web","SIGIR ACM Special Interest Group on Information Retrieval"]},"container-title":["Proceedings of the 29th ACM International Conference on Information &amp; Knowledge Management"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3340531.3411913","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3340531.3411913","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:01:22Z","timestamp":1750197682000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3340531.3411913"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,10,19]]},"references-count":43,"alternative-id":["10.1145\/3340531.3411913","10.1145\/3340531"],"URL":"https:\/\/doi.org\/10.1145\/3340531.3411913","relation":{},"subject":[],"published":{"date-parts":[[2020,10,19]]},"assertion":[{"value":"2020-10-19","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}