{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2022,4,2]],"date-time":"2022-04-02T11:42:46Z","timestamp":1648899766956},"reference-count":28,"publisher":"Springer Science and Business Media LLC","issue":"3","license":[{"start":{"date-parts":[[2014,7,18]],"date-time":"2014-07-18T00:00:00Z","timestamp":1405641600000},"content-version":"tdm","delay-in-days":0,"URL":"http:\/\/www.springer.com\/tdm"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J Sign Process Syst"],"published-print":{"date-parts":[[2015,9]]},"DOI":"10.1007\/s11265-014-0916-x","type":"journal-article","created":{"date-parts":[[2014,7,17]],"date-time":"2014-07-17T01:03:27Z","timestamp":1405559007000},"page":"245-259","update-policy":"http:\/\/dx.doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":2,"title":["An Adaptive Heterogeneous Runtime Framework for Irregular Applications"],"prefix":"10.1007","volume":"80","author":[{"given":"Chih-Chen","family":"Kao","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wei-Chung","family":"Hsu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2014,7,18]]},"reference":[{"key":"916_CR1","unstructured":"Nvidia, C. (2011). Nvidia Cuda Programming Guide."},{"key":"916_CR2","unstructured":"AMD, A. (2008). Ati Close to the Metal (ctm) Guide."},{"key":"916_CR3","first-page":"l1","volume":"1","author":"A Munshi","year":"2009","unstructured":"Munshi, A., & et al. (2009). The Opencl Specification. Khronos OpenCL Working Group, 1, l1\u201315.","journal-title":"Khronos OpenCL Working Group"},{"issue":"4","key":"916_CR4","doi-asserted-by":"crossref","first-page":"3","DOI":"10.1145\/1594835.1504181","volume":"44","author":"M Kulkarni","year":"2009","unstructured":"Kulkarni, M., Burtscher, M., Inkulu, R., Pingali, K., Cas\u010baval, C. (2009). How much parallelism is there in irregular applications? In ACM sigplan notices (Vol. 44 pp. 3\u201314). ACM.","journal-title":"ACM Sigplan Notices"},{"issue":"11","key":"916_CR5","doi-asserted-by":"crossref","first-page":"32","DOI":"10.1109\/MC.2005.379","volume":"38","author":"R Kumar","year":"2005","unstructured":"Kumar, R., Tullsen, D., Jouppi, N., Ranganathan, P. (2005). Heterogeneous chip multiprocessors. Computer, 38 (11), 32\u201338.","journal-title":"Computer"},{"key":"916_CR6","unstructured":"Pingali, K., Kulkarni, M., Nguyen, D., Burtscher, M., Mendez-Lojo, M., Prountzos, D., Sui, X., Zhong, Z. Amorphous Data-Parallelism in Irregular Algorithms."},{"key":"916_CR7","doi-asserted-by":"crossref","unstructured":"Cook, R. L., Porter, T., Carpenter, L. (1984). Distributed ray tracing. In Proceedings of the 11th annual conference on computer graphics and interactive techniques, Ser. SIGGRAPH \u201984, (pp. 137\u2013145). New York, NY, USA: ACM. [Online]. Available. doi: 10.1145\/800031.808590 .","DOI":"10.1145\/800031.808590"},{"key":"916_CR8","doi-asserted-by":"crossref","unstructured":"Purcell, T. J., Buck, I., Mark, W. R., Hanrahan, P. (2002). Ray tracing on programmable graphics hardware, In Proceedings of the 29th annual conference on computer graphics and interactive techniques, series. SIGGRAPH \u201902 (pp. 703\u2013712). New York: ACM. [Online]. Available. doi: 10.1145\/566570.566640 .","DOI":"10.1145\/566570.566640"},{"key":"916_CR9","unstructured":"Pharr, M., & Humphreys, G. (2010). Physically based rendering: from theory to implementation. Morgan Kaufmann."},{"key":"916_CR10","doi-asserted-by":"crossref","unstructured":"Burtscher, M., Nasre, R., Pingali, K. (2012). A quantitative study of irregular programs on Gpus. In 2012 IEEE international symposium on IEEE workload characterization (IISWC) (pp. 141\u2013151).","DOI":"10.1109\/IISWC.2012.6402918"},{"issue":"1","key":"916_CR11","doi-asserted-by":"crossref","first-page":"369","DOI":"10.1145\/1961295.1950408","volume":"39","author":"EZ Zhang","year":"2011","unstructured":"Zhang, E.Z., Jiang, Y., Guo, Z., Tian, K., Shen, X. (2011). On-the-fly elimination of dynamic irregularities for Gpu computing. ACM SIGARCH Computer Architecture News, 39(1), 369\u2013380. ACM.","journal-title":"ACM SIGARCH Computer Architecture News"},{"key":"916_CR12","doi-asserted-by":"crossref","unstructured":"Monteiro, P., & Monteiro, M. P. (2010). A pattern language for parallelizing irregular algorithms. In Proceedings of the 2010 workshop on parallel programming patterns, series. ParaPLoP \u201910 (pp. 13:1\u201313:14). New York: ACM. [Online]. Available. doi: 10.1145\/1953611.1953624 .","DOI":"10.1145\/1953611.1953624"},{"key":"916_CR13","doi-asserted-by":"crossref","unstructured":"Nasre, R., Burtscher, M., Pingali, K. (2013). Data-driven versus topology-driven irregular computations on Gpus. In 2013 IEEE 27th international symposium on IEEE parallel and distributed processing(IPDPS) (pp. 463\u2013474).","DOI":"10.1109\/IPDPS.2013.28"},{"key":"916_CR14","doi-asserted-by":"crossref","unstructured":"Gummaraju, J., Morichetti, L., Houston, M., Sander, B., Gaster, B. R., Zheng, B. (2010). Twin peaks: a software platform for heterogeneous computing on general-purpose and graphics processors. In Proceedings of the 19th international conference on parallel architectures and compilation techniques, series. PACT \u201910 (pp. 205\u2013216). New York: ACM. [Online]. Available: doi 10.1145\/1854273.1854302 .","DOI":"10.1145\/1854273.1854302"},{"key":"916_CR15","doi-asserted-by":"crossref","unstructured":"Lattner, C., & Adve, V. (2004). Llvm: a compilation framework for lifelong program analysis transformation. In International symposium on code generation and optimization, 2004, CGO 2004 (pp. 75\u201386).","DOI":"10.1109\/CGO.2004.1281665"},{"key":"916_CR16","doi-asserted-by":"crossref","unstructured":"Stratton, J. A., Stone, S. S., Wen-mei, W.H. (2008). Mcuda: an efficient implementation of cuda kernels for multi-core Cpus. In Languages and compilers for parallel computing (pp. 16\u201330). Springer.","DOI":"10.1007\/978-3-540-89740-8_2"},{"key":"916_CR17","doi-asserted-by":"crossref","unstructured":"Wang, P. H., Collins, J.D., Chinya, G. N., Jiang, H., Tian, X., Girkar, M., Yang, N. Y., Lueh, G.-Y., Wang, H. (2007). Exochi: architecture and programming environment for a heterogeneous multi-core multithreaded system. In Proceedings of the 2007 ACM SIGPLAN conference on programming language design and implementation. series. PLDI \u201907 (pp. 156\u2013166). New York: ACM. [Online]. Available. doi: 10.1145\/1250734.1250753 .","DOI":"10.1145\/1250734.1250753"},{"key":"916_CR18","doi-asserted-by":"crossref","unstructured":"Linderman, M. D., Collins, J. D., Wang, H., Meng, T. H. (2008). Merge: a programming model for heterogeneous multi-core systems. In Proceedings of the 13th international conference on architectural support for programming languages and operating systems series. ASPLOS XIII (pp. 287\u2013296). New York: ACM. [Online]. Available. doi: 10.1145\/1346281.1346318 .","DOI":"10.1145\/1346281.1346318"},{"key":"916_CR19","doi-asserted-by":"crossref","unstructured":"Luk, C.-K., Hong, S., Kim, H. (2009). Qilin: exploiting parallelism on heterogeneous multiprocessors with adaptive mapping. In Proceedings of the 42nd annual IEEE\/ACM international symposium on microarchitecture, series. MICRO 42 (pp. 45\u201355). New York: ACM. [Online]. Available. doi: 10.1145\/1669112.1669121 .","DOI":"10.1145\/1669112.1669121"},{"key":"916_CR20","doi-asserted-by":"crossref","unstructured":"Lee, C., Ro, W.W., Gaudiot, J.-L. (2012). Cooperative heterogeneous computing for parallel processing on CPU\/GPU hybrids. In 2012 16th workshop on IEEE interaction between compilers and computer architectures (INTERACT) (pp. 33\u201340).","DOI":"10.1109\/INTERACT.2012.6339624"},{"key":"916_CR21","doi-asserted-by":"crossref","unstructured":"Burtscher, M., & Rabeti, H. (2013). A scalable heterogeneous parallelization framework for iterative local searches. In 2013 IEEE 27th international symposium on parallel distributed processing (IPDPS) (pp. 1289\u20131298).","DOI":"10.1109\/IPDPS.2013.27"},{"key":"916_CR22","unstructured":"Komatsu, K., Sato, K., Arai, Y., Koyama, K., Takizawa, H., Kobayashi, H. (June 2010). Evaluating performance and portability of opencl programs. In The 5th international workshop on automatic performance tuning."},{"key":"916_CR23","doi-asserted-by":"crossref","unstructured":"Aila, T., & Laine, S. (2009). Understanding the efficiency of ray traversal on Gpus. In Proceedings of the conference on high performance graphics 2009 (pp. 145\u2013149). ACM.","DOI":"10.1145\/1572769.1572792"},{"key":"916_CR24","doi-asserted-by":"crossref","unstructured":"Benthin, C., Wald, I., Scherbaum, M., Friedrich, H. (2006). Ray tracing on the cell processor. In IEEE symposium on interactive ray tracing 2006 (pp. 15\u201323). IEEE.","DOI":"10.1109\/RT.2006.280210"},{"key":"916_CR25","doi-asserted-by":"crossref","unstructured":"Billeter, M., Olsson, O, Assarsson, U. (2009). Efficient stream compaction on wide simd many-core architectures. In Proceedings of the conference on high performance graphics 2009 (pp. 159\u2013166). ACM.","DOI":"10.1145\/1572769.1572795"},{"key":"916_CR26","doi-asserted-by":"crossref","unstructured":"Wald, I., & Havran, V. (2006). On building fast kd-trees for ray tracing, and on doing that in o(n Log n). In IEEE symposium on interactive ray tracing 2006 (pp. 61\u201369).","DOI":"10.1109\/RT.2006.280216"},{"key":"916_CR27","doi-asserted-by":"crossref","unstructured":"Soupikov, A., Shevtsov, M., Kapustin, A. (2008). Improving Kd-tree quality at a reasonable construction cost. In IEEE Symposium on interactive ray tracing, 2008. RT 2008 (pp. 67\u201372).","DOI":"10.1109\/RT.2008.4634623"},{"key":"916_CR28","doi-asserted-by":"crossref","unstructured":"Dammertz, H., Hanika, J., Keller, A. (2008). Shallow bounding volume hierarchies for fast simd ray tracing of incoherent rays. In Proceedings of the 19th eurographics conference on rendering, series. EGSR\u201908, (pp. 1225\u20131233). Aire-la-Ville, Switzerland: Eurographics Association. [Online]. Available. doi: 10.1111\/j.1467-8659.2008.01261.x .","DOI":"10.1111\/j.1467-8659.2008.01261.x"}],"container-title":["Journal of Signal Processing Systems"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11265-014-0916-x.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s11265-014-0916-x\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11265-014-0916-x","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2019,8,12]],"date-time":"2019-08-12T18:34:33Z","timestamp":1565634873000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s11265-014-0916-x"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2014,7,18]]},"references-count":28,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2015,9]]}},"alternative-id":["916"],"URL":"https:\/\/doi.org\/10.1007\/s11265-014-0916-x","relation":{},"ISSN":["1939-8018","1939-8115"],"issn-type":[{"value":"1939-8018","type":"print"},{"value":"1939-8115","type":"electronic"}],"subject":[],"published":{"date-parts":[[2014,7,18]]}}}