{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,5,2]],"date-time":"2025-05-02T22:43:49Z","timestamp":1746225829786},"reference-count":27,"publisher":"Springer Science and Business Media LLC","issue":"12","license":[{"start":{"date-parts":[[2017,5,8]],"date-time":"2017-05-08T00:00:00Z","timestamp":1494201600000},"content-version":"unspecified","delay-in-days":0,"URL":"http:\/\/www.springer.com\/tdm"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Sci. China Inf. Sci."],"published-print":{"date-parts":[[2017,12]]},"DOI":"10.1007\/s11432-015-0989-3","type":"journal-article","created":{"date-parts":[[2017,5,15]],"date-time":"2017-05-15T08:14:24Z","timestamp":1494836064000},"update-policy":"http:\/\/dx.doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Characterizing and optimizing Java-based HPC applications on Intel many-core architecture","\u57fa\u4e8eIntel\u4f17\u6838\u67b6\u6784\u7684Java\u9ad8\u6027\u80fd\u8ba1\u7b97\u7814\u7a76\u4e0e\u4f18\u5316"],"prefix":"10.1007","volume":"60","author":[{"given":"Yang","family":"Yu","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tianyang","family":"Lei","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Haibo","family":"Chen","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Binyu","family":"Zang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2017,5,8]]},"reference":[{"key":"989_CR1","volume-title":"Intel\u00ae Xeon PhiTM Coprocessor-the Architecture","author":"G Chrysos","year":"2014","unstructured":"Chrysos G. Intel\u00ae Xeon PhiTM Coprocessor-the Architecture. Intel Whitepaper, 2014"},{"key":"989_CR2","doi-asserted-by":"crossref","first-page":"532","DOI":"10.1016\/j.jpdc.2009.02.006","volume":"69","author":"A Shafi","year":"2009","unstructured":"Shafi A, Carpenter B, Baker M. Nested parallelism for multi-core HPC systems using Java. J Parall Distrib Comput, 2009, 69: 532\u2013545","journal-title":"J Parall Distrib Comput"},{"key":"989_CR3","first-page":"19","volume":"10","author":"J E Moreira","year":"2002","unstructured":"Moreira J E, Midkiff S P, Gupta M, et al. NINJA: Java for high performance numerical computing. Sci Program, 2002, 10: 19\u201333","journal-title":"Sci Program"},{"key":"989_CR4","volume-title":"Current state of Java for HPC","author":"B Amedro","year":"2008","unstructured":"Amedro B, Bodnartchouk V, Caromel D, et al. Current state of Java for HPC. Technical Report RT-0353. INRIA, 2008"},{"key":"989_CR5","doi-asserted-by":"crossref","first-page":"243","DOI":"10.1007\/s10686-011-9241-6","volume":"31","author":"W O\u2019Mullane","year":"2011","unstructured":"O\u2019Mullane W, Luri X, Parsons P, et al. Using Java for distributed computing in the Gaia satellite data processing. Exp Astron, 2011, 31: 243\u2013258","journal-title":"Exp Astron"},{"key":"989_CR6","doi-asserted-by":"crossref","first-page":"30","DOI":"10.1145\/1596655.1596661","volume-title":"Proceedings of the 7th International Conference on Principles and Practice of Programming in Java","author":"G L Taboada","year":"2009","unstructured":"Taboada G L, Touri\u02dcno J, Doallo R. Java for high performance computing: assessment of current research and practice. In: Proceedings of the 7th International Conference on Principles and Practice of Programming in Java. New York: ACM, 2009. 30\u201339"},{"key":"989_CR7","doi-asserted-by":"crossref","first-page":"18","DOI":"10.1109\/5992.908997","volume":"3","author":"R F Boisvert","year":"2001","unstructured":"Boisvert R F, Moreira J, Philippsen M, et al. Java and numerical computing. Comput Sci Eng, 2001, 3: 18\u201324","journal-title":"Comput Sci Eng"},{"key":"989_CR8","volume-title":"Intel R 64 and IA-32 Architectures Software Developer\u2019s Manual","author":"P Guide","year":"2010","unstructured":"Guide P. Intel R 64 and IA-32 Architectures Software Developer\u2019s Manual. 2010"},{"key":"989_CR9","first-page":"207","volume-title":"Proceedings of the 5th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming","author":"R D Blumofe","year":"1995","unstructured":"Blumofe R D, Joerg C F, Kuszmaul B C, et al. Cilk: an efficient multithreaded runtime system. In: Proceedings of the 5th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming. New York: ACM, 1995. 207\u2013216"},{"key":"989_CR10","volume-title":"The Java Virtual Machine Specification","author":"T Lindholm","year":"2014","unstructured":"Lindholm T, Yellin F, Bracha G, et al. The Java Virtual Machine Specification. 8th ed. Redwood City: Pearson Education, 2014"},{"key":"989_CR11","unstructured":"Intel. Intel \u00ae Xeon PhiTM Coprocessor Instruction Set Architecture Reference Manual. 2012"},{"key":"989_CR12","doi-asserted-by":"crossref","first-page":"8","DOI":"10.1145\/582034.582042","volume-title":"Proceedings of the 2001 ACM\/IEEE Conference on Supercomputing","author":"L A Smith","year":"2001","unstructured":"Smith L A, Bull J M, Obdrizalek J. A parallel Java grande benchmark suite. In: Proceedings of the 2001 ACM\/IEEE Conference on Supercomputing. New York: ACM, 2001. 8"},{"key":"989_CR13","first-page":"13","volume-title":"Proceedings of IEEE 13th International Symposium on High Performance Computer Architecture","author":"C Ranger","year":"2007","unstructured":"Ranger C, Raghuraman R, Penmetsa A, et al. Evaluating MapReduce for multi-core and multiprocessor systems. In: Proceedings of IEEE 13th International Symposium on High Performance Computer Architecture. Washington, DC: IEEE, 2007. 13\u201324"},{"key":"989_CR14","first-page":"55","volume":"11","author":"Z Fang","year":"2015","unstructured":"Fang Z, Mehta S, Yew P C, et al. Measuring microarchitectural details of multi-and many-core memory systems through microbenchmarking. ACM Trans Architect Code Optim, 2015, 11: 55","journal-title":"ACM Trans Architect Code Optim"},{"key":"989_CR15","unstructured":"Intel. Intel\u00ae Xeon PhiTM Coprocessor System Software Developers Guide. 2013"},{"key":"989_CR16","doi-asserted-by":"crossref","first-page":"73","DOI":"10.1145\/2597652.2597660","volume-title":"Proceedings of the 28th ACM International Conference on Supercomputing","author":"S Mehta","year":"2014","unstructured":"Mehta S, Fang Z, Zhai A, et al. Multi-stage coordinated prefetching for present-day processors. In: Proceedings of the 28th ACM International Conference on Supercomputing. New York: ACM, 2014. 73\u201382"},{"key":"989_CR17","first-page":"1575","volume-title":"Proceedings of IEEE 27th International Parallel and Distributed Processing Symposium Workshops & PhD Forum (IPDPSW), Cambridge","author":"R Krishnaiyer","year":"2013","unstructured":"Krishnaiyer R, Kultursay E, Chawla P, et al. Compiler-based data prefetching and streaming non-temporal store generation for the Intel\u00ae Xeon PhiTM coprocessor. In: Proceedings of IEEE 27th International Parallel and Distributed Processing Symposium Workshops & PhD Forum (IPDPSW), Cambridge, 2013. 1575\u20131586"},{"key":"989_CR18","first-page":"193","volume-title":"Proceedings of the Joint European Conferences on Theory and Practice of Software and the 17th International Conference on Compiler Construction. Berlin\/Heidelberg: Springer-Verlag","author":"T Wurthinger","year":"2008","unstructured":"Wurthinger T, Wimmer C, Mossenbock H. Visualization of program dependence graphs. In: Proceedings of the Joint European Conferences on Theory and Practice of Software and the 17th International Conference on Compiler Construction. Berlin\/Heidelberg: Springer-Verlag, 2008. 193\u2013196"},{"key":"989_CR19","first-page":"409","volume-title":"Proceedings of the 39th Annual IEEE\/ACM International Symposium on Microarchitecture","author":"J Tuck","year":"2006","unstructured":"Tuck J, Ceze L, Torrellas J. Scalable cache miss handling for high memory-level parallelism. In: Proceedings of the 39th Annual IEEE\/ACM International Symposium on Microarchitecture. Washington, DC: IEEE, 2006. 409\u2013422"},{"key":"989_CR20","unstructured":"Fang J, Varbanescu A L, Sips H, et al. An empirical study of Intel Xeon Phi. arXiv:1310.5842"},{"key":"989_CR21","first-page":"736","volume-title":"Proceedings of the 42nd International Conference on Parallel Processing, Lyon","author":"A Ramachandran","year":"2013","unstructured":"Ramachandran A, Vienne J, van der Wijngaart R, et al. Performance evaluation of NAS parallel benchmarks on Intel Xeon Phi. In: Proceedings of the 42nd International Conference on Parallel Processing, Lyon, 2013. 736\u2013743"},{"key":"989_CR22","first-page":"126","volume-title":"Proceedings of IEEE 27th International Symposium on Parallel & Distributed Processing (IPDPS), Boston","author":"A Heinecke","year":"2013","unstructured":"Heinecke A, Vaidyanathan K, Smelyanskiy M, et al. Design and implementation of the linpack benchmark for single and multi-node systems based on Intel R Xeon PhiTM coprocessor. In: Proceedings of IEEE 27th International Symposium on Parallel & Distributed Processing (IPDPS), Boston, 2013. 126\u2013137"},{"key":"989_CR23","doi-asserted-by":"crossref","first-page":"591","DOI":"10.1145\/2654822.2541954","volume":"42","author":"S Eyerman","year":"2014","unstructured":"Eyerman S, Eeckhout L. The benefit of SMT in the multi-core era: flexibility towards degrees of thread-level parallelism. ACM SIGARCH Comput Architect News, 2014, 42: 591\u2013606","journal-title":"ACM SIGARCH Comput Architect News"},{"key":"989_CR24","doi-asserted-by":"crossref","first-page":"1521","DOI":"10.1109\/TC.2010.232","volume":"60","author":"K Y Chen","year":"2011","unstructured":"Chen K Y, Chang J M, Hou T W. Multithreading in Java: performance and scalability on multicore systems. IEEE Trans Comput, 2011, 60: 1521\u20131534","journal-title":"IEEE Trans Comput"},{"key":"989_CR25","first-page":"229","volume-title":"Proceedings of the 18th International Conference on Architectural Support for Programming Languages and Operating Systems","author":"L Gidra","year":"2013","unstructured":"Gidra L, Thomas G, Sopena J, et al. A study of the scalability of stop-the-world garbage collectors on multicores. In: Proceedings of the 18th International Conference on Architectural Support for Programming Languages and Operating Systems. New York: ACM, 2013. 229\u2013240"},{"key":"989_CR26","first-page":"887","volume-title":"Proceedings of the 15th International Euro-Par Conference on Parallel Processing","author":"Y H Yan","year":"2009","unstructured":"Yan Y H, Grossman M, Sarkar V. JCUDA: a programmer-friendly interface for accelerating Java programs with CUDA. In: Proceedings of the 15th International Euro-Par Conference on Parallel Processing. Berlin\/Heidelberg: Springer-Verlag, 2009. 887\u2013899"},{"key":"989_CR27","first-page":"1398","volume-title":"Proceedings of the 27th International Conference on Advanced Information Networking and Applications Workshops","author":"J Docampo","year":"2013","unstructured":"Docampo J, Ramos S, Taboada G L, et al. Evaluation of Java for general purpose GPU computing. In: Proceedings of the 27th International Conference on Advanced Information Networking and Applications Workshops. Washington, DC: IEEE, 2013. 1398\u20131404"}],"container-title":["Science China Information Sciences"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11432-015-0989-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s11432-015-0989-3\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11432-015-0989-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,8,23]],"date-time":"2023-08-23T15:24:27Z","timestamp":1692804267000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s11432-015-0989-3"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2017,5,8]]},"references-count":27,"journal-issue":{"issue":"12","published-print":{"date-parts":[[2017,12]]}},"alternative-id":["989"],"URL":"https:\/\/doi.org\/10.1007\/s11432-015-0989-3","relation":{},"ISSN":["1674-733X","1869-1919"],"issn-type":[{"value":"1674-733X","type":"print"},{"value":"1869-1919","type":"electronic"}],"subject":[],"published":{"date-parts":[[2017,5,8]]},"article-number":"122106"}}