{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,21]],"date-time":"2026-05-21T02:20:27Z","timestamp":1779330027194,"version":"3.51.4"},"reference-count":28,"publisher":"SAGE Publications","issue":"6","license":[{"start":{"date-parts":[[2011,9,26]],"date-time":"2011-09-26T00:00:00Z","timestamp":1316995200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/journals.sagepub.com\/page\/policies\/text-and-data-mining-license"}],"content-domain":{"domain":["journals.sagepub.com"],"crossmark-restriction":true},"short-container-title":["SIMULATION"],"published-print":{"date-parts":[[2012,6]]},"abstract":"<jats:p>The direct simulation Monte Carlo (DSMC) is a computational method for fluid mechanics simulation in the regime of rarefied gas flow. It is a numerical solution of the Boltzmann equation based on an individual particle basis. Accurate simulations typically require particle numbers in the range of hundreds of thousands to millions. Such large simulations require an inordinate amount of time for processing using serial computing on central processing units (CPUs). In this paper we investigate data-parallel techniques on graphics processing units (GPUs) to execute very large scale DSMC simulations. We have designed and implemented Bird\u2019s method on a three-dimensional simulation domain that includes complex geometry interactions. We also have tested and verified the statistical and theoretical accuracy of our implementation. Our results show substantial performance improvements (nearly two orders of magnitude) over Bird\u2019s serial implementation without loss of accuracy.<\/jats:p>","DOI":"10.1177\/0037549711418787","type":"journal-article","created":{"date-parts":[[2011,9,27]],"date-time":"2011-09-27T00:37:49Z","timestamp":1317083869000},"page":"680-693","update-policy":"https:\/\/doi.org\/10.1177\/sage-journals-update-policy","source":"Crossref","is-referenced-by-count":13,"title":["Graphics processing unit based direct simulation Monte Carlo"],"prefix":"10.1177","volume":"88","author":[{"given":"Denis","family":"Gladkov","sequence":"first","affiliation":[{"name":"Department of Mechanical Engineering, College of Engineering and Applied Science, Milwaukee, WI, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jos\u00e9-Juan","family":"Tapia","sequence":"additional","affiliation":[{"name":"Department of Mechanical Engineering, College of Engineering and Applied Science, Milwaukee, WI, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Samuel","family":"Alberts","sequence":"additional","affiliation":[{"name":"Department of Mechanical Engineering, College of Engineering and Applied Science, Milwaukee, WI, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Roshan M","family":"D\u2019Souza","sequence":"additional","affiliation":[{"name":"Department of Mechanical Engineering, College of Engineering and Applied Science, Milwaukee, WI, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"179","published-online":{"date-parts":[[2011,9,26]]},"reference":[{"key":"bibr1-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1093\/oso\/9780198561958.001.0001"},{"key":"bibr2-0037549711418787","author":"Ivanov MS","year":"2007","journal-title":"51st IUVSTA Workshop on Modern Problems and Capability of Vacuum Gas Dynamics"},{"key":"bibr3-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1146\/annurev.fluid.30.1.403"},{"key":"bibr4-0037549711418787","first-page":"485","volume-title":"Parallel DSMC strategies for 3d computations. Proceedings of Parallel CFD","author":"Ivanov M","year":"1996"},{"key":"bibr5-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1006\/jcph.1996.0141"},{"key":"bibr6-0037549711418787","volume-title":"Proceedings of the AIAA Aerospace Sciences Meeting","author":"Gao D"},{"key":"bibr7-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1002\/nme.1232"},{"key":"bibr8-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1016\/S0045-7825(98)00302-8"},{"key":"bibr9-0037549711418787","first-page":"496","volume":"61","author":"Ota M","year":"1995","journal-title":"JSME Int J Ser B"},{"key":"bibr10-0037549711418787","unstructured":"Robinson CD. Particle Simulation on Parallel Computers with Dynamic Load Balancing. PhD thesis, Imperial College of Science, Technology and Medicine, 1998."},{"key":"bibr11-0037549711418787","volume-title":"Proceedings of the AIAA Aerospace Sciences Meeting","author":"Gao D"},{"key":"bibr12-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1111\/j.1467-8659.2007.01012.x"},{"key":"bibr13-0037549711418787","author":"NVIDIA Corp","year":"2008","journal-title":"NVIDIA CUDA Programming Guide 2.0"},{"key":"bibr14-0037549711418787","volume-title":"Programming Massively Parallel Processors: A Hand-on Approach","author":"Kirk DB","year":"2010"},{"key":"bibr15-0037549711418787","volume-title":"Parallel Programming in C with OpenMP and MPI","author":"Quinn M","year":"2004"},{"key":"bibr16-0037549711418787","volume-title":"Proceedings of the 27th International Symposium on Rarefied Gas Dynamics","author":"Smith MR"},{"key":"bibr17-0037549711418787","author":"Frezzotti A","year":"2010","journal-title":"Solving theB5oltzmann Equation on GPU"},{"key":"bibr18-0037549711418787","volume-title":"Proceedings of the 47th AIAA Aerospace Sciences Meeting Including The New Horizons Forum and Aerospace Exposition","author":"Thibault JC"},{"key":"bibr19-0037549711418787","first-page":"59","author":"Gladkov D","year":"2010","journal-title":"Proceedings of GCMS 2010"},{"key":"bibr20-0037549711418787","author":"NVIDIA Corp","year":"2010","journal-title":"NVIDIAs Next Generation CUDA Compute Architecture: Fermi"},{"key":"bibr21-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1145\/272991.272995"},{"key":"bibr22-0037549711418787","first-page":"805","author":"Howes L","year":"2008","journal-title":"GPU Gems 3"},{"key":"bibr23-0037549711418787","volume-title":"Parallel Mersenne Twister","author":"Podlozhnyuk V","year":"2007"},{"key":"bibr24-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-59657-5_3"},{"key":"bibr25-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1080\/10867651.2001.10487535"},{"key":"bibr26-0037549711418787","doi-asserted-by":"publisher","DOI":"10.1080\/10867651.1997.10487468"},{"key":"bibr27-0037549711418787","volume-title":"CUDA particles","author":"Green S","year":"2008"},{"key":"bibr28-0037549711418787","volume-title":"Proceedings 23rd IEEE International Parallel and Distributed Processing Symposium","author":"Satish N"}],"container-title":["SIMULATION"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/0037549711418787","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/0037549711418787","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,5,1]],"date-time":"2026-05-01T11:23:23Z","timestamp":1777634603000},"score":1,"resource":{"primary":{"URL":"https:\/\/journals.sagepub.com\/doi\/10.1177\/0037549711418787"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2011,9,26]]},"references-count":28,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2012,6]]}},"alternative-id":["10.1177\/0037549711418787"],"URL":"https:\/\/doi.org\/10.1177\/0037549711418787","relation":{},"ISSN":["0037-5497","1741-3133"],"issn-type":[{"value":"0037-5497","type":"print"},{"value":"1741-3133","type":"electronic"}],"subject":[],"published":{"date-parts":[[2011,9,26]]}}}