{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,27]],"date-time":"2026-03-27T15:58:04Z","timestamp":1774627084520,"version":"3.50.1"},"reference-count":51,"publisher":"Springer Science and Business Media LLC","issue":"2","license":[{"start":{"date-parts":[[2008,2,14]],"date-time":"2008-02-14T00:00:00Z","timestamp":1202947200000},"content-version":"tdm","delay-in-days":0,"URL":"http:\/\/www.springer.com\/tdm"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Auton Agent Multi-Agent Syst"],"published-print":{"date-parts":[[2008,10]]},"DOI":"10.1007\/s10458-007-9026-5","type":"journal-article","created":{"date-parts":[[2008,2,13]],"date-time":"2008-02-13T06:32:39Z","timestamp":1202884359000},"page":"190-250","source":"Crossref","is-referenced-by-count":96,"title":["Formal models and algorithms for decentralized decision making under uncertainty"],"prefix":"10.1007","volume":"17","author":[{"given":"Sven","family":"Seuken","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shlomo","family":"Zilberstein","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2008,2,14]]},"reference":[{"key":"9026_CR1","unstructured":"Amato, C., Bernstein, D., & Zilberstein, S. (2007). Optimizing memory-bounded controllers for decentralized POMDP. In Proceedings of the Twenty-Third Conference on Uncertainty in Artificial Intelligence (UAI). Vancouver, Canada, July 2007."},{"key":"9026_CR2","unstructured":"Amato, C., Bernstein, D. S., & Zilberstein, S. (2007). Solving POMDPs using quadratically constrained linear programs. In Proceedings of the Twentieth International Joint Conference on Artificial Intelligence (IJCAI) (pp. 2418\u20132424). Hyderabad, India, January 2007."},{"key":"9026_CR3","doi-asserted-by":"crossref","first-page":"174","DOI":"10.1016\/0022-247X(65)90154-X","volume":"10","author":"K.J. Astr\u00f6m","year":"1965","unstructured":"Astr\u00f6m K.J. (1965) Optimal control of Markov decision processes with incomplete state estimation. Journal of Mathematical Analysis and Applications 10: 174\u2013205","journal-title":"Journal of Mathematical Analysis and Applications"},{"key":"9026_CR4","doi-asserted-by":"crossref","unstructured":"Becker, R., Zilberstein, S., Lesser, V., & Goldman, C. V. (2004). Solving transition independent decentralized Markov decision processes. Journal of Artificial Intelligence Research (JAIR), 22, 423\u2013455.","DOI":"10.1613\/jair.1497"},{"key":"9026_CR5","unstructured":"Bernstein, D. S. (2005). Complexity analysis and optimal algorithms for decentralized decision making. PhD thesis, Department of Computer Science, University of Massachusetts Amherst, Amherst, MA."},{"issue":"4","key":"9026_CR6","doi-asserted-by":"crossref","first-page":"819","DOI":"10.1287\/moor.27.4.819.297","volume":"27","author":"D.S. Bernstein","year":"2002","unstructured":"Bernstein D.S., Givan R., Immerman N., Zilberstein S. (2002) The complexity of decentralized control of Markov decision processes. Mathematics of Operations Research 27(4): 819\u2013840","journal-title":"Mathematics of Operations Research"},{"key":"9026_CR7","unstructured":"Bernstein, D. S., Hansen, E. A., & Zilberstein, S. (2005). Bounded policy iteration for decentralized POMDPs. In Proceedings of the Nineteenth International Joint Conference on Artificial Intelligence (IJCAI) (pp. 1287\u20131292). Edinburgh, Scotland, July 2005."},{"key":"9026_CR8","unstructured":"Bernstein, D. S., Zilberstein, S., & Immerman, N. (2000). The complexity of decentralized control of Markov decision processes. In Proceedings of the Sixteenth Conference on Uncertainty in Artificial Intelligence (UAI) (pp. 32\u201337). Stanford, California, June 2000."},{"key":"9026_CR9","unstructured":"Beynier, A., & Mouaddib, A. (2006). An iterative algorithm for solving constrained decentralized Markov decision processes. In Proceedings of the Twenty-First National Conference on Artificial Intelligence (AAAI) (pp. 1089\u20131094). Boston, MA, July 2006."},{"key":"9026_CR10","unstructured":"Boutilier, C. (1996). Planning, learning and coordination in multiagent decision processes. In Proceedings of the Conference on Theoretical Aspects of Rationality and Knowledge (pp. 195\u2013210). De Zeeuwse Stromen, The Netherlands."},{"key":"9026_CR11","unstructured":"Cogill, R., Rotkowitz, M., van Roy, B., & Lall, S. (2004). An approximate dynamic programming approach to decentralized control of stochastic systems. In Proceedings of the Allerton Conference on Communication, Control, and Computing (pp. 1040\u20131049). Urbana-Champaign, IL, 2004."},{"issue":"6","key":"9026_CR12","doi-asserted-by":"crossref","first-page":"850","DOI":"10.1287\/opre.51.6.850.24925","volume":"51","author":"D.P. Farias de","year":"2003","unstructured":"de Farias D.P., van Roy B. (2003) The linear programming approach to approximate dynamic programming. Operations Research 51(6): 850\u2013865","journal-title":"Operations Research"},{"key":"9026_CR13","unstructured":"Doshi, P., & Gmytrasiewicz, P. J. (2005). A particle filtering based approach to approximating interactive POMDPs. In Proceedings of the Twentieth National Conference on Artificial Intelligence (AAAI) (pp. 969\u2013974). Pittsburg, Pennsylvania, July 2005."},{"key":"9026_CR14","doi-asserted-by":"crossref","unstructured":"Doshi, P., & Gmytrasiewicz, P. J. (2005). Approximating state estimation in multiagent settings using particle filters. In Proceedings of the Fourth International Joint Conference on Autonomous Agents and Multi-Agent Systems (AAMAS) (pp. 320\u2013327). Utrecht, Netherlands, July 2005.","DOI":"10.1145\/1082473.1082522"},{"key":"9026_CR15","unstructured":"Emery-Montemerlo, R., Gordon, G., Schneider, J., & Thrun, S. (2004). Approximate solutions for partially observable stochastic games with common payoffs. In Proceedings of the Third International Joint Conference on Autonomous Agents and Multi-Agent Systems (AAMAS) (pp. 136\u2013143). New York, NY, July 2004."},{"key":"9026_CR16","doi-asserted-by":"crossref","first-page":"49","DOI":"10.1613\/jair.1579","volume":"24","author":"P.J. Gmytrasiewicz","year":"2005","unstructured":"Gmytrasiewicz P.J., Doshi P. (2005) A framework for sequential planning in multiagent settings. Journal of Artificial Intelligence Research JAIR) 24: 49\u201379","journal-title":"Journal of Artificial Intelligence Research JAIR)"},{"issue":"1","key":"9026_CR17","doi-asserted-by":"crossref","first-page":"47","DOI":"10.1007\/s10458-006-0008-9","volume":"15","author":"C.V. Goldman","year":"2007","unstructured":"Goldman C.V., Allen M., Zilberstein S. (2007) Learning to communicate in a decentralized environment. Autonomous Agents and Multi-Agent Systems 15(1): 47\u201390","journal-title":"Autonomous Agents and Multi-Agent Systems"},{"key":"9026_CR18","doi-asserted-by":"crossref","unstructured":"Goldman, C. V., & Zilberstein, S. (2003). Optimizing information exchange in cooperative multi agent systems. In Proceedings of the Second International Joint Conference on Autonomous Agents and Multi-Agent Systems (AAMAS) (pp. 137\u2013144). Melbourne, Australia, July 2003.","DOI":"10.1145\/860575.860598"},{"key":"9026_CR19","doi-asserted-by":"crossref","first-page":"143","DOI":"10.1613\/jair.1427","volume":"22","author":"C.V. Goldman","year":"2004","unstructured":"Goldman C.V., Zilberstein S. (2004) Decentralized control of cooperative systems: Categorization and complexity analysis. Journal of Artificial Intelligence Research (JAIR) 22: 143\u2013174","journal-title":"Journal of Artificial Intelligence Research (JAIR)"},{"key":"9026_CR20","unstructured":"Goldman, C. V., & Zilberstein, S. (2004). Goal-oriented DEC-MDPs with direct communication. Technical Report 04-44, Department of Computer Science, University of Massachusetts Amherst."},{"key":"9026_CR21","unstructured":"Hansen, E. A. (1998). Finite-memory control of partially observable systems. PhD thesis, Department of Computer Science, University of Massachuetts Amherst."},{"key":"9026_CR22","unstructured":"Hansen, E. A., Bernstein, D. S., & Zilberstein, S. (2004). Dynamic programming for partially observable stochastic games. In Proceedings of the Nineteenth National Conference on Artificial Intelligence (AAAI) (pp. 709\u2013715). San Jose, California, July 2004."},{"issue":"2","key":"9026_CR23","doi-asserted-by":"crossref","first-page":"99","DOI":"10.1016\/S0004-3702(98)00023-X","volume":"101","author":"L.P. Kaelbling","year":"1998","unstructured":"Kaelbling L.P., Littmann M.L., Cassandra A.R. (1998) Planning and acting in partially observable stochastic domains. Artificial Intelligence 101(2): 99\u2013134","journal-title":"Artificial Intelligence"},{"key":"9026_CR24","doi-asserted-by":"crossref","first-page":"1231","DOI":"10.2307\/2951500","volume":"1","author":"E. Kalai","year":"1993","unstructured":"Kalai E., Lehrer E. (1993) Rational learning leads to nash equilibrium. Econometrica 1: 1231\u20131240","journal-title":"Econometrica"},{"issue":"1\u20132","key":"9026_CR25","doi-asserted-by":"crossref","first-page":"5","DOI":"10.1016\/S0004-3702(02)00378-8","volume":"147","author":"O. Madani","year":"2003","unstructured":"Madani O., Hanks S., Condon A. (2003) On the undecidability of probabilistic planning and related stochastic optimization problems. Artificial Intelligence 147(1\u20132): 5\u201334","journal-title":"Artificial Intelligence"},{"key":"9026_CR26","unstructured":"Nair, R., Pynadath, D., Yokoo, M., Tambe, M., & Marsella, S. (2003). Taming decentralized POMDPs: Towards efficient policy computation for multiagent settings. In Proceedings of the Eighteenth International Joint Conference on Artificial Intelligence (IJCAI) (pp. 705\u2013711). Acapulco, Mexico, August 2003."},{"key":"9026_CR27","unstructured":"Nair, R., Varakantham, P., Tambe, M., & Yokoo, M. (2005). Networked distributed POMDPs: A synthesis of distributed constraint optimization and POMDPs. In Proceedings of the Twentieth National Conference on Artificial Intelligence (AAAI) (pp. 133\u2013139). Pittsburgh, Pennsylvania, July 2005."},{"key":"9026_CR28","volume-title":"Algorithmic Game Theory","year":"2007","unstructured":"Nisan, N., Roughgarden, T., Tardos, E., Vazirani, V.V. (eds) (2007) Algorithmic Game Theory. Cambridge University Press., New York, NY"},{"key":"9026_CR29","doi-asserted-by":"crossref","unstructured":"Ooi, J. M., & Wornell, G. W. (1996). Decentralized control of a multiple access broadcast channel: Performance bounds. In Proceedings of the Thirty-Fifth Conference on Decision and Control (pp. 293\u2013298). Kobe, Japan, December 1996.","DOI":"10.1109\/CDC.1996.574318"},{"key":"9026_CR30","volume-title":"A course in game theory","author":"M.J. Osborne","year":"1994","unstructured":"Osborne M.J., Rubinstein A. (1994) A course in game theory. The MIT Press, Cambridge, MA"},{"key":"9026_CR31","volume-title":"Computational complexity","author":"C.H. Papadimitriou","year":"1994","unstructured":"Papadimitriou C.H. (1994) Computational complexity. Addison Wesley, Reading, MA"},{"issue":"4","key":"9026_CR32","doi-asserted-by":"crossref","first-page":"639","DOI":"10.1137\/0324038","volume":"24","author":"C.H. Papadimitriou","year":"1986","unstructured":"Papadimitriou C.H., Tsitsiklis J.N. (1986) Intractable problems in control theory. SIAM Journal on Control and Optimization 24(4): 639\u2013654","journal-title":"SIAM Journal on Control and Optimization"},{"issue":"3","key":"9026_CR33","doi-asserted-by":"crossref","first-page":"441","DOI":"10.1287\/moor.12.3.441","volume":"12","author":"C.H. Papadimitriou","year":"1987","unstructured":"Papadimitriou C.H., Tsitsiklis J.N. (1987) The complexity of Markov decision processes. Mathematics of Operations Research 12(3): 441\u2013450","journal-title":"Mathematics of Operations Research"},{"key":"9026_CR34","unstructured":"Parkes, D. C. (2008). Computational mechanism design. In Lecture notes of Tutorials at 10th Conf. on Theoretical Aspectsof Rationality and Knowledge (TARK-05). Institute of Mathematical Sciences, University of Singapore (to appear)."},{"key":"9026_CR35","unstructured":"Peshkin, L., Kim, K.-E., Meuleau, N., & Kaelbling, L. P. (2000). Learning to cooperate via policy search. In Proceedings of the Sixteenth Conference on Uncertainty in Artificial Intelligence (UAI) (pp. 489\u2013496). Stanford, CA, July 2000."},{"key":"9026_CR36","unstructured":"Petrik, M., & Zilberstein, S. (2007). Anytime coordination using separable bilinear programs. In Proceedings of the Twenty-Second Conference on Artificial Intelligence (AAAI) (pp. 750\u2013755). Vancouver, Canada, July 2007."},{"key":"9026_CR37","unstructured":"Poupart, P., & Boutilier, C. (2003). Bounded finite state controllers. In Advances in Neural Information Processing Systems 16 (NIPS). Vancouver, Canada, December 2003."},{"key":"9026_CR38","doi-asserted-by":"crossref","DOI":"10.1002\/9780470316887","volume-title":"Markov decision processes: discrete stochastic dynamic programming","author":"M.L. Puterman","year":"1994","unstructured":"Puterman M.L. (1994) Markov decision processes: discrete stochastic dynamic programming. Wiley, New York, NY"},{"key":"9026_CR39","doi-asserted-by":"crossref","first-page":"389","DOI":"10.1613\/jair.1024","volume":"16","author":"D.V. Pynadath","year":"2002","unstructured":"Pynadath D.V., Tambe M. (2002) The communicative multiagent team decision problem: Analyzing teamwork theories and models. Journal of Artificial Intelligence Research (JAIR) 16: 389\u2013423","journal-title":"Journal of Artificial Intelligence Research (JAIR)"},{"key":"9026_CR40","doi-asserted-by":"crossref","unstructured":"Rabinovich, Z., Goldman, C. V., & Rosenschein, J. S. (2003). The complexity of multiagent systems: The price of silence. In Proceedings of the Second International Joint Conference on Autonomous Agents and Multi-Agent Systems (AAMAS) (pp. 1102\u20131103). Melbourne, Australia, 2003.","DOI":"10.1145\/860575.860816"},{"key":"9026_CR41","volume-title":"Artificial intelligence: A modern approach","author":"S. Russell","year":"2003","unstructured":"Russell S., Norvig P. (2003) Artificial intelligence: A modern approach. Prentice Hall, Upper Saddle River, NJ"},{"key":"9026_CR42","unstructured":"Seuken, S., & Zilberstein, S. (2007). Memory-bounded dynamic programming for DEC-POMDPs. In Proceedings of the Twentieth International Joint Conference on Artificial Intelligence (IJCAI) (pp. 2009\u20132015). Hyderabad, India, January 2007."},{"key":"9026_CR43","unstructured":"Seuken, S., & Zilberstein, S. (2007). Improved memory-bounded dynamic programming for decentralized POMDPs. In Proceedings of the 23rd Conference on Uncertainty in Artificial Intelligence (UAI). Vancouver, Canada, July 2007."},{"key":"9026_CR44","doi-asserted-by":"crossref","unstructured":"Shen, J., Becker, R., & Lesser, V. (2006). Agent interaction in distributed MDPs and its implications on complexity. In Proceedings of the Fifth International Joint Conference on Autonomous Agents and Multi-Agent Systems (AAMAS). Hakodate, Japan, May 2006.","DOI":"10.1145\/1160633.1160730"},{"issue":"4","key":"9026_CR45","doi-asserted-by":"crossref","first-page":"1347","DOI":"10.1137\/S009753979628292X","volume":"28","author":"I. Suzuki","year":"1999","unstructured":"Suzuki I., Yamashita M. (1999) Distributed anonymous mobile robots: Formation of geometric patterns. SIAM Journal on Computing 28(4): 1347\u20131363","journal-title":"SIAM Journal on Computing"},{"key":"9026_CR46","doi-asserted-by":"crossref","unstructured":"Szer, D., & Charpillet, F. (2005). An optimal best-first search algorithm for solving infinite horizon DEC-POMDPs. In Proceedings of the Sixteenth European Conference on Machine Learning (ECML) (pp. 389\u2013399). Porto, Portugal, October 2005.","DOI":"10.1007\/11564096_38"},{"key":"9026_CR47","unstructured":"Szer, D., & Charpillet, F. (2006). Point-based dynamic programming for DEC-POMDPs. In Proceedings of the Twenty-First National Conference on Artificial Intelligence (AAAI) (pp. 1233\u20131238). Boston, MA, July 2006."},{"key":"9026_CR48","unstructured":"Szer, D., Charpillet, F., & Zilberstein, S. (2005). MAA*: A heuristic search algorithm for solving decentralized POMDPs. In Proceedings of the Twenty-First Conference on Uncertainty in Artificial Intelligence (UAI) (pp. 576\u2013583). Edinburgh, Scotland, July 2005."},{"key":"9026_CR49","doi-asserted-by":"crossref","first-page":"83","DOI":"10.1613\/jair.433","volume":"7","author":"M. Tambe","year":"1997","unstructured":"Tambe M. (1997) Towards flexible teamwork. Journal of Artificial Intelligence Research (JAIR) 7: 83\u2013124","journal-title":"Journal of Artificial Intelligence Research (JAIR)"},{"key":"9026_CR50","doi-asserted-by":"crossref","unstructured":"Washington, R., Golden, K., Bresina, J., Smith, D. E., Anderson, C., & Smith, T. (1999). Autonomous rovers for Mars exploration. In Proceedings of the IEEE Aerospace Conference (pp. 237\u2013251). Snowmass, CO, March 1999.","DOI":"10.1109\/AERO.1999.794236"},{"key":"9026_CR51","doi-asserted-by":"crossref","unstructured":"Zilberstein, S., Washington, R.,Bernstein, D. S., & Mouaddib, A.-I. (2002). Decision-theoretic control of planetary rovers. In Beetz, M., Guibas, L., Hertzberg, J., Ghallab, M., & Pollack, M.E. (Eds.), Advances in Plan-Based Control of Robotic Agents, Lecture Notes in Computer Science (Vol. 2466, pp. 270\u2013289). Berlin, Germany: Springer.","DOI":"10.1007\/3-540-37724-7_16"}],"container-title":["Autonomous Agents and Multi-Agent Systems"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s10458-007-9026-5.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s10458-007-9026-5\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s10458-007-9026-5","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2019,5,29]],"date-time":"2019-05-29T17:28:23Z","timestamp":1559150903000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s10458-007-9026-5"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2008,2,14]]},"references-count":51,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2008,10]]}},"alternative-id":["9026"],"URL":"https:\/\/doi.org\/10.1007\/s10458-007-9026-5","relation":{},"ISSN":["1387-2532","1573-7454"],"issn-type":[{"value":"1387-2532","type":"print"},{"value":"1573-7454","type":"electronic"}],"subject":[],"published":{"date-parts":[[2008,2,14]]}}}