{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:24:16Z","timestamp":1750220656003,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":38,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,10,19]],"date-time":"2020-10-19T00:00:00Z","timestamp":1603065600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100006642","name":"U.S. Department of Education","doi-asserted-by":"publisher","award":["P200A150306"],"award-info":[{"award-number":["P200A150306"]}],"id":[{"id":"10.13039\/100006642","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100006435","name":"National Science Foundation","doi-asserted-by":"publisher","award":["IIS-1718310, IIS-1815866, IIS-1718310, CNS-1852498, CNS-1560229"],"award-info":[{"award-number":["IIS-1718310, IIS-1815866, IIS-1718310, CNS-1852498, CNS-1560229"]}],"id":[{"id":"10.13039\/100006435","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,10,19]]},"DOI":"10.1145\/3340531.3412018","type":"proceedings-article","created":{"date-parts":[[2020,10,19]],"date-time":"2020-10-19T06:32:45Z","timestamp":1603089165000},"page":"485-494","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Learning to Selectively Update State Neurons in Recurrent Networks"],"prefix":"10.1145","author":[{"given":"Thomas","family":"Hartvigsen","sequence":"first","affiliation":[{"name":"Worcester Polytechnic Institute, Worcester, MA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Cansu","family":"Sen","sequence":"additional","affiliation":[{"name":"Worcester Polytechnic Institute, Worcester, MA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiangnan","family":"Kong","sequence":"additional","affiliation":[{"name":"Worcester Polytechnic Institute, Worcester, MA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Elke","family":"Rundensteiner","sequence":"additional","affiliation":[{"name":"Worcester Polytechnic Institute, Worcester, MA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,10,19]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1103\/PhysRevE.64.061907"},{"key":"e_1_3_2_2_2_1","volume-title":"Conditional computation in neural networks for faster models. arXiv preprint arXiv:1511.06297","author":"Bengio Emmanuel","year":"2015","unstructured":"Emmanuel Bengio , Pierre-Luc Bacon , Joelle Pineau , and Doina Precup . 2015. Conditional computation in neural networks for faster models. arXiv preprint arXiv:1511.06297 ( 2015 ). Emmanuel Bengio, Pierre-Luc Bacon, Joelle Pineau, and Doina Precup. 2015. Conditional computation in neural networks for faster models. arXiv preprint arXiv:1511.06297 (2015)."},{"key":"e_1_3_2_2_3_1","volume-title":"Estimating or propagating gradients through stochastic neurons for conditional computation. arXiv preprint arXiv:1308.3432","author":"Bengio Yoshua","year":"2013","unstructured":"Yoshua Bengio , Nicholas L\u00e9onard , and Aaron Courville . 2013. Estimating or propagating gradients through stochastic neurons for conditional computation. arXiv preprint arXiv:1308.3432 ( 2013 ). Yoshua Bengio, Nicholas L\u00e9onard, and Aaron Courville. 2013. Estimating or propagating gradients through stochastic neurons for conditional computation. arXiv preprint arXiv:1308.3432 (2013)."},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/72.279181"},{"key":"e_1_3_2_2_5_1","volume-title":"International Conference on Learning Representations.","author":"Campos V'ictor","year":"2018","unstructured":"V'ictor Campos , Brendan Jou , Xavier Giro-i Nieto , Jordi Torres , and Shih-Fu Chang . 2018 . Skip RNN: Learning to Skip State Updates in Recurrent Neural Networks . In International Conference on Learning Representations. V'ictor Campos, Brendan Jou, Xavier Giro-i Nieto, Jordi Torres, and Shih-Fu Chang. 2018. Skip RNN: Learning to Skip State Updates in Recurrent Neural Networks. In International Conference on Learning Representations."},{"key":"e_1_3_2_2_6_1","volume-title":"Recurrent neural networks for multivariate time series with missing values. Scientific reports","author":"Che Zhengping","year":"2018","unstructured":"Zhengping Che , Sanjay Purushotham , Kyunghyun Cho , David Sontag , and Yan Liu . 2018. Recurrent neural networks for multivariate time series with missing values. Scientific reports , Vol. 8 , 1 ( 2018 ), 6085. Zhengping Che, Sanjay Purushotham, Kyunghyun Cho, David Sontag, and Yan Liu. 2018. Recurrent neural networks for multivariate time series with missing values. Scientific reports, Vol. 8, 1 (2018), 6085."},{"key":"e_1_3_2_2_7_1","volume-title":"A survey of model compression and acceleration for deep neural networks. arXiv preprint arXiv:1710.09282","author":"Cheng Yu","year":"2017","unstructured":"Yu Cheng , Duo Wang , Pan Zhou , and Tao Zhang . 2017. A survey of model compression and acceleration for deep neural networks. arXiv preprint arXiv:1710.09282 ( 2017 ). Yu Cheng, Duo Wang, Pan Zhou, and Tao Zhang. 2017. A survey of model compression and acceleration for deep neural networks. arXiv preprint arXiv:1710.09282 (2017)."},{"key":"e_1_3_2_2_8_1","volume-title":"Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio.","author":"Cho Kyunghyun","year":"2014","unstructured":"Kyunghyun Cho , Bart Van Merri\u00ebnboer , Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014 . Learning phrase representations using RNN encoder-decoder for statistical machine translation. In Empirical Methods in Natural Language Processing . Kyunghyun Cho, Bart Van Merri\u00ebnboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014. Learning phrase representations using RNN encoder-decoder for statistical machine translation. In Empirical Methods in Natural Language Processing."},{"key":"e_1_3_2_2_9_1","volume-title":"International Conference on Learning Representations.","author":"Chung Junyoung","year":"2017","unstructured":"Junyoung Chung , Sungjin Ahn , and Yoshua Bengio . 2017 . Hierarchical multiscale recurrent neural networks . In International Conference on Learning Representations. Junyoung Chung, Sungjin Ahn, and Yoshua Bengio. 2017. Hierarchical multiscale recurrent neural networks. In International Conference on Learning Representations."},{"key":"e_1_3_2_2_10_1","volume-title":"Finding structure in time. Cognitive science","author":"Elman Jeffrey L","year":"1990","unstructured":"Jeffrey L Elman . 1990. Finding structure in time. Cognitive science , Vol. 14 , 2 ( 1990 ), 179--211. Jeffrey L Elman. 1990. Finding structure in time. Cognitive science, Vol. 14, 2 (1990), 179--211."},{"key":"e_1_3_2_2_11_1","volume-title":"Frage: Frequency-agnostic word representation. In Advances in neural information processing systems. 1334--1345.","author":"Gong Chengyue","year":"2018","unstructured":"Chengyue Gong , Di He , Xu Tan , Tao Qin , Liwei Wang , and Tie-Yan Liu . 2018 . Frage: Frequency-agnostic word representation. In Advances in neural information processing systems. 1334--1345. Chengyue Gong, Di He, Xu Tan, Tao Qin, Liwei Wang, and Tie-Yan Liu. 2018. Frage: Frequency-agnostic word representation. In Advances in neural information processing systems. 1334--1345."},{"key":"e_1_3_2_2_12_1","unstructured":"Alex Graves. 2013. Generating sequences with recurrent neural networks. In arXiv preprint arXiv:1308.0850.  Alex Graves. 2013. Generating sequences with recurrent neural networks. In arXiv preprint arXiv:1308.0850."},{"key":"e_1_3_2_2_13_1","unstructured":"David Ha and J\u00fcrgen Schmidhuber. 2018. Recurrent world models facilitate policy evolution. In Advances in Neural Information Processing Systems. 2450--2462.  David Ha and J\u00fcrgen Schmidhuber. 2018. Recurrent world models facilitate policy evolution. In Advances in Neural Information Processing Systems. 2450--2462."},{"key":"e_1_3_2_2_14_1","unstructured":"Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. 2016. Deep residual learning for image recognition. In Computer vision and pattern recognition. 770--778.  Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. 2016. Deep residual learning for image recognition. In Computer vision and pattern recognition. 770--778."},{"key":"e_1_3_2_2_15_1","volume-title":"International Conference on Machine Learning.","author":"Henaff Mikael","year":"2016","unstructured":"Mikael Henaff , Arthur Szlam , and Yann LeCun . 2016 . Recurrent orthogonal networks and long-memory tasks . International Conference on Machine Learning. Mikael Henaff, Arthur Szlam, and Yann LeCun. 2016. Recurrent orthogonal networks and long-memory tasks. International Conference on Machine Learning."},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1162\/neco.1997.9.8.1735"},{"key":"e_1_3_2_2_17_1","volume-title":"International Conference on Learning Representations.","author":"Jernite Yacine","year":"2017","unstructured":"Yacine Jernite , Edouard Grave , Armand Joulin , and Tomas Mikolov . 2017 . Variable computation in recurrent neural networks . In International Conference on Learning Representations. Yacine Jernite, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017. Variable computation in recurrent neural networks. In International Conference on Learning Representations."},{"key":"e_1_3_2_2_18_1","unstructured":"Francc ois Kawala Ahlame Douzal-Chouakria Eric Gaussier and Eustache Dimert. 2013. Pr\u00e9dictions d'activit\u00e9 dans les r\u00e9seaux sociaux en ligne. In 4i\u00e8me conf\u00e9rence sur les mod\u00e8les et l'analyse des r\u00e9seaux: Approches math\u00e9matiques et informatiques. 16.  Francc ois Kawala Ahlame Douzal-Chouakria Eric Gaussier and Eustache Dimert. 2013. Pr\u00e9dictions d'activit\u00e9 dans les r\u00e9seaux sociaux en ligne. In 4i\u00e8me conf\u00e9rence sur les mod\u00e8les et l'analyse des r\u00e9seaux: Approches math\u00e9matiques et informatiques. 16."},{"key":"e_1_3_2_2_19_1","volume-title":"International Conference on Learning Representations.","author":"Kingma Diederik P","year":"2015","unstructured":"Diederik P Kingma and Jimmy Ba . 2015 . Adam: A method for stochastic optimization . In International Conference on Learning Representations. Diederik P Kingma and Jimmy Ba. 2015. Adam: A method for stochastic optimization. In International Conference on Learning Representations."},{"key":"e_1_3_2_2_20_1","volume-title":"International Conference on Machine Learning. 1863--1871","author":"Koutnik Jan","year":"2014","unstructured":"Jan Koutnik , Klaus Greff , Faustino Gomez , and Juergen Schmidhuber . 2014 . A Clockwork RNN . In International Conference on Machine Learning. 1863--1871 . Jan Koutnik, Klaus Greff, Faustino Gomez, and Juergen Schmidhuber. 2014. A Clockwork RNN. In International Conference on Machine Learning. 1863--1871."},{"key":"e_1_3_2_2_21_1","volume-title":"International Conference on Learning Representations.","author":"Krueger David","year":"2017","unstructured":"David Krueger , Tegan Maharaj , J\u00e1nos Kram\u00e1r , Mohammad Pezeshki , Nicolas Ballas , Nan Rosemary Ke , Anirudh Goyal , Yoshua Bengio , Aaron Courville , and Chris Pal . 2017 . Zoneout: Regularizing rnns by randomly preserving hidden activations . In International Conference on Learning Representations. David Krueger, Tegan Maharaj, J\u00e1nos Kram\u00e1r, Mohammad Pezeshki, Nicolas Ballas, Nan Rosemary Ke, Anirudh Goyal, Yoshua Bengio, Aaron Courville, and Chris Pal. 2017. Zoneout: Regularizing rnns by randomly preserving hidden activations. In International Conference on Learning Representations."},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.24963\/ijcai.2019\/705"},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v32i1.11630"},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v32i1.11307"},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.3115\/1075812.1075835"},{"key":"e_1_3_2_2_26_1","volume-title":"Mogrifier LSTM. In International Conference on Learning Representations.","author":"Melis G\u00e1bor","year":"2020","unstructured":"G\u00e1bor Melis , Tom\u00e1vs Kovc isk\u1ef3, and Phil Blunsom . 2020 . Mogrifier LSTM. In International Conference on Learning Representations. G\u00e1bor Melis, Tom\u00e1vs Kovc isk\u1ef3, and Phil Blunsom. 2020. Mogrifier LSTM. In International Conference on Learning Representations."},{"key":"e_1_3_2_2_27_1","unstructured":"Daniel Neil Michael Pfeiffer and Shih-Chii Liu. 2016. Phased lstm: Accelerating recurrent network training for long or event-based sequences. In Advances in Neural Information Processing Systems. 3882--3890.  Daniel Neil Michael Pfeiffer and Shih-Chii Liu. 2016. Phased lstm: Accelerating recurrent network training for long or event-based sequences. In Advances in Neural Information Processing Systems. 3882--3890."},{"key":"e_1_3_2_2_28_1","unstructured":"Adam Paszke Sam Gross Soumith Chintala Gregory Chanan Edward Yang Zachary DeVito Zeming Lin Alban Desmaison Luca Antiga and Adam Lerer. 2017. Automatic differentiation in PyTorch. In Advances in neural information processing systems.  Adam Paszke Sam Gross Soumith Chintala Gregory Chanan Edward Yang Zachary DeVito Zeming Lin Alban Desmaison Luca Antiga and Adam Lerer. 2017. Automatic differentiation in PyTorch. In Advances in neural information processing systems."},{"key":"e_1_3_2_2_29_1","volume-title":"Self-delimiting neural networks. arXiv preprint arXiv:1210.0118","author":"Schmidhuber J\u00fcrgen","year":"2012","unstructured":"J\u00fcrgen Schmidhuber . 2012. Self-delimiting neural networks. arXiv preprint arXiv:1210.0118 ( 2012 ). J\u00fcrgen Schmidhuber. 2012. Self-delimiting neural networks. arXiv preprint arXiv:1210.0118 (2012)."},{"key":"e_1_3_2_2_30_1","volume-title":"Outrageously large neural networks: The sparsely-gated mixture-of-experts layer. arXiv preprint arXiv:1701.06538","author":"Shazeer Noam","year":"2017","unstructured":"Noam Shazeer , Azalia Mirhoseini , Krzysztof Maziarz , Andy Davis , Quoc Le , Geoffrey Hinton , and Jeff Dean . 2017. Outrageously large neural networks: The sparsely-gated mixture-of-experts layer. arXiv preprint arXiv:1701.06538 ( 2017 ). Noam Shazeer, Azalia Mirhoseini, Krzysztof Maziarz, Andy Davis, Quoc Le, Geoffrey Hinton, and Jeff Dean. 2017. Outrageously large neural networks: The sparsely-gated mixture-of-experts layer. arXiv preprint arXiv:1701.06538 (2017)."},{"key":"e_1_3_2_2_31_1","volume-title":"International Conference on Learning Representations.","author":"Shen Yikang","year":"2019","unstructured":"Yikang Shen , Shawn Tan , Alessandro Sordoni , and Aaron Courville . 2019 . Ordered Neurons: Incorporating Tree Structures into Recurrent Neural Networks . In International Conference on Learning Representations. Yikang Shen, Shawn Tan, Alessandro Sordoni, and Aaron Courville. 2019. Ordered Neurons: Incorporating Tree Structures into Recurrent Neural Networks. In International Conference on Learning Representations."},{"key":"e_1_3_2_2_32_1","unstructured":"Rupesh Kumar Srivastava Klaus Greff and J\u00fcrgen Schmidhuber. 2015. Highway networks. In Advances in Neural Information Processing Systems.  Rupesh Kumar Srivastava Klaus Greff and J\u00fcrgen Schmidhuber. 2015. Highway networks. In Advances in Neural Information Processing Systems."},{"key":"e_1_3_2_2_33_1","volume-title":"International Conference on Machine Learning. 6555--6565","author":"Wang Dilin","year":"2019","unstructured":"Dilin Wang , Chengyue Gong , and Qiang Liu . 2019 . Improving Neural Language Modeling via Adversarial Training . In International Conference on Machine Learning. 6555--6565 . Dilin Wang, Chengyue Gong, and Qiang Liu. 2019. Improving Neural Language Modeling via Adversarial Training. In International Conference on Machine Learning. 6555--6565."},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"crossref","unstructured":"Yiren Wang and Fei Tian. 2016. Recurrent residual learning for sequence classification. In Empirical methods in natural language processing. 938--943.  Yiren Wang and Fei Tian. 2016. Recurrent residual learning for sequence classification. In Empirical methods in natural language processing. 938--943.","DOI":"10.18653\/v1\/D16-1093"},{"key":"e_1_3_2_2_35_1","volume-title":"Simple statistical gradient-following algorithms for connectionist reinforcement learning. Machine learning","author":"Williams Ronald J","year":"1992","unstructured":"Ronald J Williams . 1992. Simple statistical gradient-following algorithms for connectionist reinforcement learning. Machine learning , Vol. 8 , 3--4 ( 1992 ), 229--256. Ronald J Williams. 1992. Simple statistical gradient-following algorithms for connectionist reinforcement learning. Machine learning, Vol. 8, 3--4 (1992), 229--256."},{"key":"e_1_3_2_2_36_1","unstructured":"Yonghui Wu Mike Schuster Zhifeng Chen Quoc V Le Mohammad Norouzi Wolfgang Macherey Maxim Krikun Yuan Cao Qin Gao Klaus Macherey etal 2016. Google's neural machine translation system: Bridging the gap between human and machine translation. arXiv preprint arXiv:1609.08144 (2016).  Yonghui Wu Mike Schuster Zhifeng Chen Quoc V Le Mohammad Norouzi Wolfgang Macherey Maxim Krikun Yuan Cao Qin Gao Klaus Macherey et al. 2016. Google's neural machine translation system: Bridging the gap between human and machine translation. arXiv preprint arXiv:1609.08144 (2016)."},{"key":"e_1_3_2_2_37_1","volume-title":"International Conference on Machine Learning. 2048--2057","author":"Xu Kelvin","year":"2015","unstructured":"Kelvin Xu , Jimmy Ba , Ryan Kiros , Kyunghyun Cho , Aaron Courville , Ruslan Salakhudinov , Rich Zemel , and Yoshua Bengio . 2015 . Show, attend and tell: Neural image caption generation with visual attention . In International Conference on Machine Learning. 2048--2057 . Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio. 2015. Show, attend and tell: Neural image caption generation with visual attention. In International Conference on Machine Learning. 2048--2057."},{"key":"e_1_3_2_2_38_1","volume-title":"International Conference on Machine Learning","volume":"70","author":"Zilly Julian Georg","year":"2017","unstructured":"Julian Georg Zilly , Rupesh Kumar Srivastava , Jan Koutnik , and J\u00fcrgen Schmidhuber . 2017 . Recurrent highway networks . In International Conference on Machine Learning , Vol. 70 . 4189--4198. Julian Georg Zilly, Rupesh Kumar Srivastava, Jan Koutnik, and J\u00fcrgen Schmidhuber. 2017. Recurrent highway networks. In International Conference on Machine Learning, Vol. 70. 4189--4198."}],"event":{"name":"CIKM '20: The 29th ACM International Conference on Information and Knowledge Management","sponsor":["SIGWEB ACM Special Interest Group on Hypertext, Hypermedia, and Web","SIGIR ACM Special Interest Group on Information Retrieval"],"location":"Virtual Event Ireland","acronym":"CIKM '20"},"container-title":["Proceedings of the 29th ACM International Conference on Information &amp; Knowledge Management"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3340531.3412018","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3340531.3412018","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:02:29Z","timestamp":1750197749000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3340531.3412018"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,10,19]]},"references-count":38,"alternative-id":["10.1145\/3340531.3412018","10.1145\/3340531"],"URL":"https:\/\/doi.org\/10.1145\/3340531.3412018","relation":{},"subject":[],"published":{"date-parts":[[2020,10,19]]},"assertion":[{"value":"2020-10-19","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}