{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:36:23Z","timestamp":1750221383874,"version":"3.41.0"},"reference-count":38,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2018,1,23]],"date-time":"2018-01-23T00:00:00Z","timestamp":1516665600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100007156","name":"Hong Kong Innovation and Technology Fund","doi-asserted-by":"crossref","award":["ITS\/041\/16"],"award-info":[{"award-number":["ITS\/041\/16"]}],"id":[{"id":"10.13039\/501100007156","id-type":"DOI","asserted-by":"crossref"}]},{"name":"National Basic Program of China, 973 Program","award":["2015CB351706"],"award-info":[{"award-number":["2015CB351706"]}]},{"name":"Research Grants Council of theHong Kong Special Administrative Region","award":["GRF 14225616"],"award-info":[{"award-number":["GRF 14225616"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Knowl. Discov. Data"],"published-print":{"date-parts":[[2018,6,30]]},"abstract":"<jats:p>Bayesian Probabilistic Matrix Factorization (BPMF) is a powerful model in many dyadic data prediction problems, especially the applications of Recommender system. However, its poor scalability has limited its wide applications on massive data. Based on the conditional independence property of observed entries in BPMF model, we propose a novel distributed memo-free variational inference method for large-scale matrix factorization problems. Compared with the state-of-the-art methods, the proposed method is favored for several attractive properties. Specifically, it does not require tuning of learning rate carefully, shuffling the training set at each iteration, or storing massive redundant variables, and can introduce new agents into the computations on the fly. We conduct extensive experiments on both synthetic and real-world datasets. The experimental results show that our method can converge significantly faster with better prediction performance than alternative algorithms.<\/jats:p>","DOI":"10.1145\/3161886","type":"journal-article","created":{"date-parts":[[2018,1,23]],"date-time":"2018-01-23T14:53:36Z","timestamp":1516719216000},"page":"1-24","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Large-Scale Bayesian Probabilistic Matrix Factorization with Memo-Free Distributed Variational Inference"],"prefix":"10.1145","volume":"12","author":[{"given":"Guangyong","family":"Chen","sequence":"first","affiliation":[{"name":"The Chinese University of Hong Kong, Hong Kong"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fengyuan","family":"Zhu","sequence":"additional","affiliation":[{"name":"The Chinese University of Hong Kong, Hong Kong"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Pheng Ann","family":"Heng","sequence":"additional","affiliation":[{"name":"The Chinese University of Hong Kong, Hong Kong"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2018,1,23]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/2783258.2783373"},{"key":"e_1_2_1_2_1","volume-title":"Shachter","author":"Azevedo-Filho Adriano","year":"1994","unstructured":"Adriano Azevedo-Filho and Ross D . Shachter . 1994 . Laplace\u2019s method approximations for probabilistic inference in belief networks with continuous variables. In Proceedings of the 10th International Conference on Uncertainty in artificial intelligence. Morgan Kaufmann Publishers Inc ., 28--36. Adriano Azevedo-Filho and Ross D. Shachter. 1994. Laplace\u2019s method approximations for probabilistic inference in belief networks with continuous variables. In Proceedings of the 10th International Conference on Uncertainty in artificial intelligence. Morgan Kaufmann Publishers Inc., 28--36."},{"key":"e_1_2_1_3_1","volume-title":"Proceedings of KDD Cup and Workshop","volume":"2007","author":"Bennett James","year":"2007","unstructured":"James Bennett and Stan Lanning . 2007 . The netflix prize . In Proceedings of KDD Cup and Workshop , Vol. 2007 . 35. James Bennett and Stan Lanning. 2007. The netflix prize. In Proceedings of KDD Cup and Workshop, Vol. 2007. 35."},{"volume-title":"Pattern Recognition and Machine Learning","author":"Bishop Christopher M.","key":"e_1_2_1_4_1","unstructured":"Christopher M. Bishop . 2006. Pattern Recognition and Machine Learning , Vol. 1 . Springer . Christopher M. Bishop. 2006. Pattern Recognition and Machine Learning, Vol. 1. Springer."},{"volume-title":"Convex Optimization","author":"Boyd Stephen","key":"e_1_2_1_5_1","unstructured":"Stephen Boyd and Lieven Vandenberghe . 2004. Convex Optimization . Cambridge University Press . Stephen Boyd and Lieven Vandenberghe. 2004. Convex Optimization. Cambridge University Press."},{"key":"e_1_2_1_6_1","volume-title":"Jordan","author":"Broderick Tamara","year":"2013","unstructured":"Tamara Broderick , Nicholas Boyd , Andre Wibisono , Ashia C. Wilson , and Michael I . Jordan . 2013 . Streaming variational Bayes. In Proceedings of Advances in Neural Information Processing Systems . 1727--1735. Tamara Broderick, Nicholas Boyd, Andre Wibisono, Ashia C. Wilson, and Michael I. Jordan. 2013. Streaming variational Bayes. In Proceedings of Advances in Neural Information Processing Systems. 1727--1735."},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10208-009-9045-5"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1155\/2009\/785152"},{"key":"e_1_2_1_9_1","volume-title":"Proceedings of the 28th Conference on Learning Theory. 797--842","author":"Ge Rong","year":"2015","unstructured":"Rong Ge , Furong Huang , Chi Jin , and Yang Yuan . 2015 . Escaping from saddle points online stochastic gradient for tensor decomposition . In Proceedings of the 28th Conference on Learning Theory. 797--842 . Rong Ge, Furong Huang, Chi Jin, and Yang Yuan. 2015. Escaping from saddle points online stochastic gradient for tensor decomposition. In Proceedings of the 28th Conference on Learning Theory. 797--842."},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/2020408.2020426"},{"key":"e_1_2_1_11_1","first-page":"449","article-title":"Variational inference for Bayesian mixtures of factor analysers","volume":"12","author":"Ghahramani Zoubin","year":"1999","unstructured":"Zoubin Ghahramani , Matthew J. Beal , and others. 1999 . Variational inference for Bayesian mixtures of factor analysers . In Proceedings of Advances in Neural Information Processing Systems , Vol. 12. 449 -- 455 . Zoubin Ghahramani, Matthew J. Beal, and others. 1999. Variational inference for Bayesian mixtures of factor analysers. In Proceedings of Advances in Neural Information Processing Systems, Vol. 12. 449--455.","journal-title":"Proceedings of Advances in Neural Information Processing Systems"},{"key":"e_1_2_1_12_1","volume-title":"Proceedings of Advances in Neural Information Processing Systems LCCC Workshop.","author":"Hall Keith B.","year":"2010","unstructured":"Keith B. Hall , Scott Gilpin , and Gideon Mann . 2010 . MapReduce\/bigtable for distributed optimization . In Proceedings of Advances in Neural Information Processing Systems LCCC Workshop. Keith B. Hall, Scott Gilpin, and Gideon Mann. 2010. MapReduce\/bigtable for distributed optimization. In Proceedings of Advances in Neural Information Processing Systems LCCC Workshop."},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/2827872"},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/2751562"},{"key":"e_1_2_1_15_1","volume-title":"Hughes and Erik Sudderth","author":"Michael","year":"2013","unstructured":"Michael C. Hughes and Erik Sudderth . 2013 . Memoized online variational inference for Dirichlet process mixture models. In Proceedings of Advances in Neural Information Processing Systems . 1133--1141. Michael C. Hughes and Erik Sudderth. 2013. Memoized online variational inference for Dirichlet process mixture models. In Proceedings of Advances in Neural Information Processing Systems. 1133--1141."},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/1644873.1644874"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.5555\/264772"},{"key":"e_1_2_1_18_1","volume-title":"Proceedings of KDD Cup and Workshop","volume":"7","author":"Lim Yew Jin","year":"2007","unstructured":"Yew Jin Lim and Yee Whye Teh . 2007 . Variational Bayesian approach to movie rating prediction . In Proceedings of KDD Cup and Workshop , Vol. 7 . Citeseer, 15--21. Yew Jin Lim and Yee Whye Teh. 2007. Variational Bayesian approach to movie rating prediction. In Proceedings of KDD Cup and Workshop, Vol. 7. Citeseer, 15--21."},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.dss.2013.04.002"},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/1961209.1961212"},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2014.2353639"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.5555\/2789272.2831142"},{"key":"e_1_2_1_23_1","volume-title":"Mann","author":"Mcdonald Ryan","year":"2009","unstructured":"Ryan Mcdonald , Mehryar Mohri , Nathan Silberman , Dan Walker , and Gideon S . Mann . 2009 . Efficient large-scale distributed training of conditional maximum entropy models. In Proceedings of Advances in Neural Information Processing Systems . 1231--1239. Ryan Mcdonald, Mehryar Mohri, Nathan Silberman, Dan Walker, and Gideon S. Mann. 2009. Efficient large-scale distributed training of conditional maximum entropy models. In Proceedings of Advances in Neural Information Processing Systems. 1231--1239."},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/2724720"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.5555\/2567709.2502582"},{"key":"e_1_2_1_26_1","volume-title":"Hinton","author":"Neal Radford M.","year":"1998","unstructured":"Radford M. Neal and Geoffrey E . Hinton . 1998 . A view of the EM algorithm that justifies incremental, sparse, and other variants. In Learning in Graphical Models, Michael I. Jordan (Ed.). Springer , 355--368. Radford M. Neal and Geoffrey E. Hinton. 1998. A view of the EM algorithm that justifies incremental, sparse, and other variants. In Learning in Graphical Models, Michael I. Jordan (Ed.). Springer, 355--368."},{"key":"e_1_2_1_27_1","volume-title":"Proceedings of AAAI.","author":"Porteous Ian","year":"2010","unstructured":"Ian Porteous , Arthur U. Asuncion , and Max Welling . 2010 . Bayesian matrix factorization with side information and dirichlet process mixtures . In Proceedings of AAAI. Ian Porteous, Arthur U. Asuncion, and Max Welling. 2010. Bayesian matrix factorization with side information and dirichlet process mixtures. In Proceedings of AAAI."},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1137\/070697835"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/1390156.1390267"},{"key":"e_1_2_1_30_1","volume-title":"If you liked this, you are sure to love that. New York Times Magazine","author":"Thompson C.","year":"2008","unstructured":"C. Thompson . 2008. If you liked this, you are sure to love that. New York Times Magazine ( 2008 ). Retrieved from http:\/\/www.nytimes.com\/2008\/11\/23\/magazine\/23Netflix-t.html. C. Thompson. 2008. If you liked this, you are sure to love that. New York Times Magazine (2008). Retrieved from http:\/\/www.nytimes.com\/2008\/11\/23\/magazine\/23Netflix-t.html."},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/2629474"},{"volume-title":"Proceedings of the 28th International Conference on Machine Learning. 681--688","author":"Welling Max","key":"e_1_2_1_32_1","unstructured":"Max Welling and Yee W. Teh . 2011. Bayesian learning via stochastic gradient Langevin dynamics . In Proceedings of the 28th International Conference on Machine Learning. 681--688 . Max Welling and Yee W. Teh. 2011. Bayesian learning via stochastic gradient Langevin dynamics. In Proceedings of the 28th International Conference on Machine Learning. 681--688."},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/800135.804414"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/2663356"},{"volume-title":"Proceedings of the 32nd International Conference on Machine Learning. 457--465","author":"Zhang Yuchen","key":"e_1_2_1_35_1","unstructured":"Yuchen Zhang , Martin J. Wainwright , and Michael I. Jordan . 2015. Distributed estimation of generalized matrix rank: Efficient algorithms and lower bounds . In Proceedings of the 32nd International Conference on Machine Learning. 457--465 . Yuchen Zhang, Martin J. Wainwright, and Michael I. Jordan. 2015. Distributed estimation of generalized matrix rank: Efficient algorithms and lower bounds. In Proceedings of the 32nd International Conference on Machine Learning. 457--465."},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/2842629"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/2507157.2507164"},{"key":"e_1_2_1_38_1","volume-title":"Smola","author":"Zinkevich Martin","year":"2010","unstructured":"Martin Zinkevich , Markus Weimer , Lihong Li , and Alex J . Smola . 2010 . Parallelized stochastic gradient descent. In Proceedings of Advances in Neural Information Processing Systems . 2595--2603. Martin Zinkevich, Markus Weimer, Lihong Li, and Alex J. Smola. 2010. Parallelized stochastic gradient descent. In Proceedings of Advances in Neural Information Processing Systems. 2595--2603."}],"container-title":["ACM Transactions on Knowledge Discovery from Data"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3161886","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3161886","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T02:26:54Z","timestamp":1750213614000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3161886"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,1,23]]},"references-count":38,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2018,6,30]]}},"alternative-id":["10.1145\/3161886"],"URL":"https:\/\/doi.org\/10.1145\/3161886","relation":{},"ISSN":["1556-4681","1556-472X"],"issn-type":[{"type":"print","value":"1556-4681"},{"type":"electronic","value":"1556-472X"}],"subject":[],"published":{"date-parts":[[2018,1,23]]},"assertion":[{"value":"2016-04-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2017-11-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2018-01-23","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}