{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:12:54Z","timestamp":1750219974898,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":33,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,21]],"date-time":"2022-10-21T00:00:00Z","timestamp":1666310400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,21]]},"DOI":"10.1145\/3569966.3570075","type":"proceedings-article","created":{"date-parts":[[2022,12,20]],"date-time":"2022-12-20T22:24:41Z","timestamp":1671575081000},"page":"383-387","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Heterogeneous Computing and Applications in Deep Learning: A Survey"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-6511-4736","authenticated-orcid":false,"given":"Qiong","family":"Wu","sequence":"first","affiliation":[{"name":"Beijing Institute of Computer Technology and Application, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1209-4174","authenticated-orcid":false,"given":"Yuefeng","family":"Shen","sequence":"additional","affiliation":[{"name":"Beijing Institute of Computer Technology and Application, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5629-199X","authenticated-orcid":false,"given":"Mingqing","family":"Zhang","sequence":"additional","affiliation":[{"name":"Beijing Institute of Computer Technology and Application, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,12,20]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"ImageNet: A large-scale hierarchical image database[C]\/\/2009 IEEE conference on computer vision and pattern recognition","author":"Deng J","year":"2009","unstructured":"Deng J , Dong W , Socher R , ImageNet: A large-scale hierarchical image database[C]\/\/2009 IEEE conference on computer vision and pattern recognition . IEEE , 2009 : 248-255. Deng J, Dong W, Socher R, ImageNet: A large-scale hierarchical image database[C]\/\/2009 IEEE conference on computer vision and pattern recognition. IEEE, 2009: 248-255."},{"key":"e_1_3_2_1_2_1","volume-title":"Reducing the dimensionality of data with neural networks[J]. science","author":"Hinton G E","year":"2006","unstructured":"Hinton G E , Salakhutdinov R R . Reducing the dimensionality of data with neural networks[J]. science , 2006 , 313(5786): 504-507. Hinton G E, Salakhutdinov R R. Reducing the dimensionality of data with neural networks[J]. science, 2006, 313(5786): 504-507."},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.19328\/j.cnki.2096-8655.2021.04.001"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.11896\/jsjkx.200600045"},{"key":"e_1_3_2_1_5_1","volume-title":"The open standard for parallel programming of heterogeneous systems","author":"KHRONOS Group","year":"2022","unstructured":"KHRONOS Group . The open standard for parallel programming of heterogeneous systems . 2022 . http:\/\/www.khronos.org\/opencl KHRONOS Group. The open standard for parallel programming of heterogeneous systems. 2022. http:\/\/www.khronos.org\/opencl"},{"key":"e_1_3_2_1_6_1","unstructured":"NVIDIA CUDA toolkit. 2022. https:\/\/developer.nvidia.com\/cuda-downloads  NVIDIA CUDA toolkit. 2022. https:\/\/developer.nvidia.com\/cuda-downloads"},{"key":"e_1_3_2_1_7_1","volume-title":"Directives for accelerators","author":"OpenAcc","year":"2022","unstructured":"OpenAcc : Directives for accelerators . 2022 . http:\/\/www.openacc.org\/ OpenAcc: Directives for accelerators. 2022. http:\/\/www.openacc.org\/"},{"key":"e_1_3_2_1_8_1","volume-title":"CUDA-lite: Reducing GPU programming complexity[C]\/\/International Workshop on Languages and Compilers for Parallel Computing","author":"Ueng S Z","year":"2008","unstructured":"Ueng S Z , Lathara M , Baghsorkhi S S , CUDA-lite: Reducing GPU programming complexity[C]\/\/International Workshop on Languages and Compilers for Parallel Computing . Springer , Berlin, Heidelberg , 2008 : 1-15. Ueng S Z, Lathara M, Baghsorkhi S S, CUDA-lite: Reducing GPU programming complexity[C]\/\/International Workshop on Languages and Compilers for Parallel Computing. Springer, Berlin, Heidelberg, 2008: 1-15."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2010.62"},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/1186562.1015800"},{"key":"e_1_3_2_1_12_1","volume-title":"accelerated massive parallelism with Microsoft Visual C++[M]","author":"Gregory K","year":"2012","unstructured":"Gregory K , Miller A. C++ AMP : accelerated massive parallelism with Microsoft Visual C++[M] . Microsoft Press , 2012 . Gregory K, Miller A. C++ AMP: accelerated massive parallelism with Microsoft Visual C++[M]. Microsoft Press, 2012."},{"key":"e_1_3_2_1_13_1","volume-title":"a programming model for heterogeneous multi-core systems[J]. ACM SIGOPS operating systems review","author":"Linderman M D","year":"2008","unstructured":"Linderman M D , Collins J D , Wang H , Merge : a programming model for heterogeneous multi-core systems[J]. ACM SIGOPS operating systems review , 2008 , 42(2): 287-296. Linderman M D, Collins J D, Wang H, Merge: a programming model for heterogeneous multi-core systems[J]. ACM SIGOPS operating systems review, 2008, 42(2): 287-296."},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"crossref","unstructured":"Auerbach J Bacon D F Cheng P Lime: a java-compatible and synthesizable language for heterogeneous architectures[C]\/\/Proceedings of the ACM international conference on Object oriented programming systems languages and applications. 2010: 89-108.  Auerbach J Bacon D F Cheng P Lime: a java-compatible and synthesizable language for heterogeneous architectures[C]\/\/Proceedings of the ACM international conference on Object oriented programming systems languages and applications. 2010: 89-108.","DOI":"10.1145\/1932682.1869469"},{"volume-title":"Compiling python to a hybrid execution environment[C]\/\/Proceedings of the 3rd Workshop on General-Purpose Computation on Graphics Processing Units. 2010: 19-30","author":"Garg R","key":"e_1_3_2_1_15_1","unstructured":"Garg R , Amaral J N . Compiling python to a hybrid execution environment[C]\/\/Proceedings of the 3rd Workshop on General-Purpose Computation on Graphics Processing Units. 2010: 19-30 . Garg R, Amaral J N. Compiling python to a hybrid execution environment[C]\/\/Proceedings of the 3rd Workshop on General-Purpose Computation on Graphics Processing Units. 2010: 19-30."},{"volume-title":"Copperhead: compiling an embedded data parallel language[C]\/\/Proceedings of the 16th ACM symposium on Principles and practice of parallel programming. 2011: 47-56","author":"Catanzaro B","key":"e_1_3_2_1_16_1","unstructured":"Catanzaro B , Garland M , Keutzer K. Copperhead: compiling an embedded data parallel language[C]\/\/Proceedings of the 16th ACM symposium on Principles and practice of parallel programming. 2011: 47-56 . Catanzaro B, Garland M, Keutzer K. Copperhead: compiling an embedded data parallel language[C]\/\/Proceedings of the 16th ACM symposium on Principles and practice of parallel programming. 2011: 47-56."},{"key":"e_1_3_2_1_17_1","volume-title":"The landscape of parallel computing research: A view from berkeley[J]","author":"Asanovic K","year":"2006","unstructured":"Asanovic K , Bodik R , Catanzaro B C , The landscape of parallel computing research: A view from berkeley[J] . 2006 . Asanovic K, Bodik R, Catanzaro B C, The landscape of parallel computing research: A view from berkeley[J]. 2006."},{"key":"e_1_3_2_1_18_1","first-page":"1","article-title":"A hybrid list-based task scheduling scheme for heterogeneous computing[J]","volume":"2021","author":"Sulaiman M","unstructured":"Sulaiman M , Halim Z , Waqas M , A hybrid list-based task scheduling scheme for heterogeneous computing[J] . The Journal of Supercomputing , 2021 : 1 - 37 . Sulaiman M, Halim Z, Waqas M, A hybrid list-based task scheduling scheme for heterogeneous computing[J]. The Journal of Supercomputing, 2021: 1-37.","journal-title":"The Journal of Supercomputing"},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.5987"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2020.3041829"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.anucene.2019.106988"},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.5923"},{"key":"e_1_3_2_1_23_1","first-page":"5446","volume-title":"KSII Transactions on Internet and Information Systems (TIIS)","author":"Nidaw B Y","year":"2019","unstructured":"Nidaw B Y , Oh M H , Kim Y W . Appropriate Synchronization Time Allocation for Distributed Heterogeneous Parallel Computing Systems[J] . KSII Transactions on Internet and Information Systems (TIIS) , 2019 , 13(11): 5446 - 5463 . Nidaw B Y, Oh M H, Kim Y W. Appropriate Synchronization Time Allocation for Distributed Heterogeneous Parallel Computing Systems[J]. KSII Transactions on Internet and Information Systems (TIIS), 2019, 13(11): 5446-5463."},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.peva.2010.08.020"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.7544\/issn1000-1239.2021.20210166"},{"volume-title":"Large-scale deep unsupervised learning using graphics processors[C]\/\/Proceedings of the 26th annual international conference on machine learning. 2009: 873-880","author":"Raina R","key":"e_1_3_2_1_26_1","unstructured":"Raina R , Madhavan A , Ng A Y . Large-scale deep unsupervised learning using graphics processors[C]\/\/Proceedings of the 26th annual international conference on machine learning. 2009: 873-880 . Raina R, Madhavan A, Ng A Y. Large-scale deep unsupervised learning using graphics processors[C]\/\/Proceedings of the 26th annual international conference on machine learning. 2009: 873-880."},{"key":"e_1_3_2_1_27_1","volume-title":"Multi-GPU training of convnets[J]. arXiv preprint arXiv:1312.5853","author":"Yadan O","year":"2013","unstructured":"Yadan O , Adams K , Taigman Y , Multi-GPU training of convnets[J]. arXiv preprint arXiv:1312.5853 , 2013 . Yadan O, Adams K, Taigman Y, Multi-GPU training of convnets[J]. arXiv preprint arXiv:1312.5853, 2013."},{"key":"e_1_3_2_1_28_1","first-page":"50","article-title":"Opencl accelerated caffe for convolutional neural networks[C]\/\/2016 IEEE international parallel and distributed processing symposium workshops (IPDPSW)","volume":"2016","author":"Bottleson J","unstructured":"Bottleson J , Kim S Y , Andrews J , cl Caffe : Opencl accelerated caffe for convolutional neural networks[C]\/\/2016 IEEE international parallel and distributed processing symposium workshops (IPDPSW) . IEEE , 2016 : 50 - 57 . Bottleson J, Kim S Y, Andrews J, clCaffe: Opencl accelerated caffe for convolutional neural networks[C]\/\/2016 IEEE international parallel and distributed processing symposium workshops (IPDPSW). IEEE, 2016: 50-57.","journal-title":"IEEE"},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11227-017-1994-x"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"crossref","unstructured":"Suda N Chandra V Dasika G Throughput-optimized OpenCL-based FPGA accelerator for large-scale convolutional neural networks[C]\/\/Proceedings of the 2016 ACM\/SIGDA International Symposium on Field-Programmable Gate Arrays. 2016: 16-25.  Suda N Chandra V Dasika G Throughput-optimized OpenCL-based FPGA accelerator for large-scale convolutional neural networks[C]\/\/Proceedings of the 2016 ACM\/SIGDA International Symposium on Field-Programmable Gate Arrays. 2016: 16-25.","DOI":"10.1145\/2847263.2847276"},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/2654822.2541967"},{"key":"e_1_3_2_1_32_1","volume-title":"Restricted Boltzmann machines and deep belief networks on sunway cluster[C]\/\/2016 IEEE 18th International Conference on High Performance Computing and Communications","author":"Song K","year":"2016","unstructured":"Song K , Liu Y , Wang R , Restricted Boltzmann machines and deep belief networks on sunway cluster[C]\/\/2016 IEEE 18th International Conference on High Performance Computing and Communications ; IEEE 14th International Conference on Smart City; IEEE 2nd International Conference on Data Science and Systems (HPCC\/SmartCity\/DSS). IEEE , 2016 : 245-252. Song K, Liu Y, Wang R, Restricted Boltzmann machines and deep belief networks on sunway cluster[C]\/\/2016 IEEE 18th International Conference on High Performance Computing and Communications; IEEE 14th International Conference on Smart City; IEEE 2nd International Conference on Data Science and Systems (HPCC\/SmartCity\/DSS). IEEE, 2016: 245-252."},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/3149485.3149511"},{"key":"e_1_3_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.13328\/j.cnki.jos.004608"}],"event":{"name":"CSSE 2022: 2022 5th International Conference on Computer Science and Software Engineering","acronym":"CSSE 2022","location":"Guilin China"},"container-title":["Proceedings of the 5th International Conference on Computer Science and Software Engineering"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3569966.3570075","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3569966.3570075","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:49:20Z","timestamp":1750182560000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3569966.3570075"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,21]]},"references-count":33,"alternative-id":["10.1145\/3569966.3570075","10.1145\/3569966"],"URL":"https:\/\/doi.org\/10.1145\/3569966.3570075","relation":{},"subject":[],"published":{"date-parts":[[2022,10,21]]},"assertion":[{"value":"2022-12-20","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}