{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,18]],"date-time":"2026-01-18T01:31:37Z","timestamp":1768699897351,"version":"3.49.0"},"reference-count":165,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2018,7,16]],"date-time":"2018-07-16T00:00:00Z","timestamp":1531699200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000001","name":"National Science Foundation","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Comput. Surv."],"published-print":{"date-parts":[[2019,5,31]]},"abstract":"<jats:p>The gap is widening between the processor clock speed of end-system architectures and network throughput capabilities. It is now physically possible to provide single-flow throughput of speeds up to 100\u00a0Gbps, and 400\u00a0Gbps will soon be possible. Most current research into high-speed data networking focuses on managing expanding network capabilities within datacenter Local Area Networks (LANs) or efficiently multiplexing millions of relatively small flows through a Wide Area Network (WAN). However, datacenter hyper-convergence places high-throughput networking workloads on general-purpose hardware, and distributed High-Performance Computing (HPC) applications require time-sensitive, high-throughput end-to-end flows (also referred to as \u201celephant flows\u201d) to occur over WANs. For these applications, the bottleneck is often the end-system and not the intervening network. Since the problem of the end-system bottleneck was uncovered, many techniques have been developed which address this mismatch with varying degrees of effectiveness. In this survey, we describe the most promising techniques, beginning with network architectures and NIC design, continuing with operating and end-system architectures, and concluding with clean-slate protocol design.<\/jats:p>","DOI":"10.1145\/3184899","type":"journal-article","created":{"date-parts":[[2018,7,16]],"date-time":"2018-07-16T13:25:21Z","timestamp":1531747521000},"page":"1-36","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":11,"title":["A Survey of End-System Optimizations for High-Speed Networks"],"prefix":"10.1145","volume":"51","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-2214-7447","authenticated-orcid":false,"given":"Nathan","family":"Hanford","sequence":"first","affiliation":[{"name":"University of California, Davis, CA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Vishal","family":"Ahuja","sequence":"additional","affiliation":[{"name":"University of California, Davis, CA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Matthew K.","family":"Farrens","sequence":"additional","affiliation":[{"name":"University of California, Davis, CA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Brian","family":"Tierney","sequence":"additional","affiliation":[{"name":"ESnet, Berkeley, CA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Dipak","family":"Ghosal","sequence":"additional","affiliation":[{"name":"University of California, Davis, CA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2018,7,16]]},"reference":[{"key":"e_1_2_1_1_1","volume-title":"Proceedings of the International Conference on Supercomputing. ACM","unstructured":"2011. Platform IO DMA transaction acceleration . In Proceedings of the International Conference on Supercomputing. ACM , New York. 415 2011. Platform IO DMA transaction acceleration. In Proceedings of the International Conference on Supercomputing. ACM, New York. 415"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2012.18"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1996130.1996140"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/2396556.2396564"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/CCGrid.2012.54"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2010.23"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/2903150.2911718"},{"key":"e_1_2_1_8_1","volume-title":"Supplement to InfiniBand Architecture Specification","author":"InfiniBand Trade Association","unstructured":"InfiniBand Trade Association . 2010. Supplement to InfiniBand Architecture Specification Volume 1 Release 1.2.1 Annex A16: RDMA over Converged Ethernet. Technical Report Release 1.2.1. InfiniBand Trade Association , Beaverton, OR. Retrieved from http:\/\/www.infinibandta.org\/content\/pages.php?pg&equal;technology_download. InfiniBand Trade Association. 2010. Supplement to InfiniBand Architecture Specification Volume 1 Release 1.2.1 Annex A16: RDMA over Converged Ethernet. Technical Report Release 1.2.1. InfiniBand Trade Association, Beaverton, OR. Retrieved from http:\/\/www.infinibandta.org\/content\/pages.php?pg&equal;technology_download."},{"key":"e_1_2_1_9_1","volume-title":"Proceedings of the 5th International Workshop on Protocols for Fast Long-Distance Networks (PFLDnet)","volume":"7","author":"Baiocchi Andrea","year":"2007","unstructured":"Andrea Baiocchi , Angelo P Castellani , and Francesco Vacirca . 2007 . YeAH-TCP: Yet another highspeed TCP . In Proceedings of the 5th International Workshop on Protocols for Fast Long-Distance Networks (PFLDnet) , Vol. 7 . Information Sciences Institute, University of Southern California Viterbi School of Engineering, Marina Del Rey, Los Angeles, CA, 37--42. Andrea Baiocchi, Angelo P Castellani, and Francesco Vacirca. 2007. YeAH-TCP: Yet another highspeed TCP. In Proceedings of the 5th International Workshop on Protocols for Fast Long-Distance Networks (PFLDnet), Vol. 7. Information Sciences Institute, University of Southern California Viterbi School of Engineering, Marina Del Rey, Los Angeles, CA, 37--42."},{"key":"e_1_2_1_10_1","volume-title":"Proceedings of the RAIT Workshop 04. IEEE Conference on Cluster Computing","author":"Balaji Pavan","year":"2004","unstructured":"Pavan Balaji . 2004 . Sockets vs RDMA interface over 10-gigabit networks: An in-depth analysis of the memory traffic bottleneck . In Proceedings of the RAIT Workshop 04. IEEE Conference on Cluster Computing , San Diego, CA. Pavan Balaji. 2004. Sockets vs RDMA interface over 10-gigabit networks: An in-depth analysis of the memory traffic bottleneck. In Proceedings of the RAIT Workshop 04. IEEE Conference on Cluster Computing, San Diego, CA."},{"key":"e_1_2_1_11_1","volume-title":"Proceedings of the IEEE 20th International Parallel Distributed Processing Symposium.","author":"Balaji P.","unstructured":"P. Balaji , S. Bhagvat , H. W. Jin , and D. K. Panda . 2006. Asynchronous zero-copy communication for synchronous sockets in the sockets direct protocol (SDP) over InfiniBand . In Proceedings of the IEEE 20th International Parallel Distributed Processing Symposium. P. Balaji, S. Bhagvat, H. W. Jin, and D. K. Panda. 2006. Asynchronous zero-copy communication for synchronous sockets in the sockets direct protocol (SDP) over InfiniBand. In Proceedings of the IEEE 20th International Parallel Distributed Processing Symposium."},{"key":"e_1_2_1_12_1","volume-title":"Proceedings of the IEEE International Conference on Cluster Computing. 1--10","author":"Balaji P.","unstructured":"P. Balaji , W. Feng , Q. Gao , R. Noronha , W. Yu , and D. K. Panda . 2005a. Head-to-TOE evaluation of high-performance sockets over protocol offload engines . In Proceedings of the IEEE International Conference on Cluster Computing. 1--10 . P. Balaji, W. Feng, Q. Gao, R. Noronha, W. Yu, and D. K. Panda. 2005a. Head-to-TOE evaluation of high-performance sockets over protocol offload engines. In Proceedings of the IEEE International Conference on Cluster Computing. 1--10."},{"key":"e_1_2_1_13_1","volume-title":"Proceedings of the IEEE International Conference on Cluster Computing","author":"Balaji P.","year":"2005","unstructured":"P. Balaji , H. W. Jin , K. Vaidyanathan , and D. K. Panda . 2005b. Supporting iWARP compatibility and features for regular network adapters . In Proceedings of the IEEE International Conference on Cluster Computing ( 2005 ). 1--10. P. Balaji, H. W. Jin, K. Vaidyanathan, and D. K. Panda. 2005b. Supporting iWARP compatibility and features for regular network adapters. In Proceedings of the IEEE International Conference on Cluster Computing (2005). 1--10."},{"key":"e_1_2_1_14_1","volume-title":"Proceedings of the 20th International Parallel and Distributed Processing Symposium (IPDPS\u201906)","author":"Banerjee A.","unstructured":"A. Banerjee , Wu chun Feng , B. Mukherjee , and D. Ghosal . 2006. RAPID: An end-system aware protocol for intelligent data transfer over lambda grids . In Proceedings of the 20th International Parallel and Distributed Processing Symposium (IPDPS\u201906) . A. Banerjee, Wu chun Feng, B. Mukherjee, and D. Ghosal. 2006. RAPID: An end-system aware protocol for intelligent data transfer over lambda grids. In Proceedings of the 20th International Parallel and Distributed Processing Symposium (IPDPS\u201906)."},{"key":"e_1_2_1_15_1","volume-title":"Sandia National Laboratories","author":"Barrett Brian W.","year":"2017","unstructured":"Brian W. Barrett , Ron Brightwell , Ryan E. Grant , Scott Hemmert , Kevin Pedretti , Kyle Wheeler , Keith Underwood , Rolf Riesen , Arthur B. Maccabe , and Trammell Hudson . 2017 . The portals 4.1 network programming interface . Sandia National Laboratories , November 2012, Technical Report SAND 2017-3825 (April 2017). Brian W. Barrett, Ron Brightwell, Ryan E. Grant, Scott Hemmert, Kevin Pedretti, Kyle Wheeler, Keith Underwood, Rolf Riesen, Arthur B. Maccabe, and Trammell Hudson. 2017. The portals 4.1 network programming interface. Sandia National Laboratories, November 2012, Technical Report SAND2017-3825 (April 2017)."},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/1400713.1400718"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/1629575.1629579"},{"key":"e_1_2_1_18_1","volume-title":"Proceedings of the USENIX 10th Conference on Operating Systems Design and Implementation (OSDI\u201912)","author":"Belay Adam","year":"2012","unstructured":"Adam Belay , Andrea Bittau , Ali Mashtizadeh , David Terei , David Mazi\u00e8res , and Christos Kozyrakis . 2012 . Dune: Safe user-level access to privileged CPU features . In Proceedings of the USENIX 10th Conference on Operating Systems Design and Implementation (OSDI\u201912) . USENIX Association, Berkeley, CA, 335--348. Adam Belay, Andrea Bittau, Ali Mashtizadeh, David Terei, David Mazi\u00e8res, and Christos Kozyrakis. 2012. Dune: Safe user-level access to privileged CPU features. In Proceedings of the USENIX 10th Conference on Operating Systems Design and Implementation (OSDI\u201912). USENIX Association, Berkeley, CA, 335--348."},{"key":"e_1_2_1_19_1","volume-title":"Proceedings of the USENIX 11th Symposium on Operating Systems Design and Implementation (OSDI 14)","author":"Belay Adam","year":"2014","unstructured":"Adam Belay , George Prekas , Ana Klimovic , Samuel Grossman , Christos Kozyrakis , and Edouard Bugnion . 2014 . IX: A protected dataplane operating system for high throughput and low latency . In Proceedings of the USENIX 11th Symposium on Operating Systems Design and Implementation (OSDI 14) . USENIX Association, 49--65. Adam Belay, George Prekas, Ana Klimovic, Samuel Grossman, Christos Kozyrakis, and Edouard Bugnion. 2014. IX: A protected dataplane operating system for high throughput and low latency. In Proceedings of the USENIX 11th Symposium on Operating Systems Design and Implementation (OSDI 14). USENIX Association, 49--65."},{"key":"e_1_2_1_20_1","unstructured":"Christian Benvenuti. 2005. Understanding Linux Network Internals - Guided Tour to Networking on Linux. O\u2019Reilly. http:\/\/www.oreilly.de\/catalog\/understandlni\/index.html.   Christian Benvenuti. 2005. Understanding Linux Network Internals - Guided Tour to Networking on Linux. O\u2019Reilly. http:\/\/www.oreilly.de\/catalog\/understandlni\/index.html."},{"key":"e_1_2_1_21_1","volume-title":"Proceedings of the IEEE 2016 International Conference on Systems, Man, and Cybernetics (SMC). 004681--004686","author":"Beserra D.","unstructured":"D. Beserra , E. D. Moreno , P. T. Endo , and J. Barreto . 2016. Performance evaluation of a lightweight virtualization solution for HPC I\/O scenarios . In Proceedings of the IEEE 2016 International Conference on Systems, Man, and Cybernetics (SMC). 004681--004686 . D. Beserra, E. D. Moreno, P. T. Endo, and J. Barreto. 2016. Performance evaluation of a lightweight virtualization solution for HPC I\/O scenarios. In Proceedings of the IEEE 2016 International Conference on Systems, Man, and Cybernetics (SMC). 004681--004686."},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/1168918.1168897"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/2658260.2658272"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/40.342015"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1364\/OFC.2015.Th3J.3"},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/1217935.1217961"},{"key":"e_1_2_1_28_1","volume-title":"Proceedings of the IEEE\/ACM 3rd International Symposium on Cluster Computing and the Grid (CCGrid\u201903)","author":"Buntinas D.","unstructured":"D. Buntinas , D. K. Panda , and R. Brightwell . 2003. Application-bypass broadcast in MPICH over GM . In Proceedings of the IEEE\/ACM 3rd International Symposium on Cluster Computing and the Grid (CCGrid\u201903) . 2--9. D. Buntinas, D. K. Panda, and R. Brightwell. 2003. Application-bypass broadcast in MPICH over GM. In Proceedings of the IEEE\/ACM 3rd International Symposium on Cluster Computing and the Grid (CCGrid\u201903). 2--9."},{"key":"e_1_2_1_29_1","doi-asserted-by":"crossref","first-page":"4","DOI":"10.1109\/MNET.2013.6574662","article-title":"Survey on converged data center networks with DCB and FCoE: Standards and protocols","volume":"27","author":"Cai Yueping","year":"2013","unstructured":"Yueping Cai , Yao Yan , Zhenghao Zhang , and Yuanyuan Yang . 2013 . Survey on converged data center networks with DCB and FCoE: Standards and protocols . IEEE Network 27 , 4 (July 2013), 27--32. Yueping Cai, Yao Yan, Zhenghao Zhang, and Yuanyuan Yang. 2013. Survey on converged data center networks with DCB and FCoE: Standards and protocols. IEEE Network 27, 4 (July 2013), 27--32.","journal-title":"IEEE Network"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/3012426.3022184"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/35.917506"},{"key":"e_1_2_1_32_1","first-page":"1124","article-title":"AIO-TFRC: A light-weight rate control scheme for streaming over wireless. In Proceedings of the 2005 IEEE International Conference on Wireless Networks","volume":"2","author":"Chen Minghua","year":"2005","unstructured":"Minghua Chen and A. Zakhor . 2005 . AIO-TFRC: A light-weight rate control scheme for streaming over wireless. In Proceedings of the 2005 IEEE International Conference on Wireless Networks , Communications and Mobile Computing , Vol. 2. 1124 -- 1129 . Minghua Chen and A. Zakhor. 2005. AIO-TFRC: A light-weight rate control scheme for streaming over wireless. In Proceedings of the 2005 IEEE International Conference on Wireless Networks, Communications and Mobile Computing, Vol. 2. 1124--1129.","journal-title":"Communications and Mobile Computing"},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11277-015-3157-9"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/2534645.2534651"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/52324.52336"},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/1005062.1005069"},{"key":"e_1_2_1_37_1","volume-title":"International Conference on Parallel and Distributed Computing Systems (PDCS\u201905)","author":"Dalessandro Dennis","year":"2005","unstructured":"Dennis Dalessandro , Ananth Devulapalli , and Pete Wyckoff . 2005 . Design and implementation of the iWarp protocol in software . In International Conference on Parallel and Distributed Computing Systems (PDCS\u201905) . 471--476. Dennis Dalessandro, Ananth Devulapalli, and Pete Wyckoff. 2005. Design and implementation of the iWarp protocol in software. In International Conference on Parallel and Distributed Computing Systems (PDCS\u201905). 471--476."},{"key":"e_1_2_1_38_1","volume-title":"Proceedings of the IEEE 20th International Parallel Distributed Processing Symposium.","author":"Dalessandro D.","unstructured":"D. Dalessandro , A. Devulapalli , and P. Wyckoff . 2006a. iWarp protocol kernel space software implementation . In Proceedings of the IEEE 20th International Parallel Distributed Processing Symposium. D. Dalessandro, A. Devulapalli, and P. Wyckoff. 2006a. iWarp protocol kernel space software implementation. In Proceedings of the IEEE 20th International Parallel Distributed Processing Symposium."},{"key":"e_1_2_1_39_1","volume-title":"Proceedings of the IEEE International Conference on Cluster Computing","author":"Dalessandro D.","year":"2006","unstructured":"D. Dalessandro , P. Wyckoff , and G. Montry . 2006b. Initial performance evaluation of the NetEffect 10 gigabit iWARP adapter . In Proceedings of the IEEE International Conference on Cluster Computing ( 2006 ). 1--7. D. Dalessandro, P. Wyckoff, and G. Montry. 2006b. Initial performance evaluation of the NetEffect 10 gigabit iWARP adapter. In Proceedings of the IEEE International Conference on Cluster Computing (2006). 1--7."},{"key":"e_1_2_1_40_1","unstructured":"Patrick Darling and Jeff Bruner (Eds.). 2010. Intel Developer Forum (IDF) San Francisco 2010. Intel Corporation.  Patrick Darling and Jeff Bruner (Eds.). 2010. Intel Developer Forum (IDF) San Francisco 2010. Intel Corporation."},{"key":"e_1_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2015.15"},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLUSTER.2015.55"},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2009.37"},{"key":"e_1_2_1_44_1","volume-title":"Linux Weekly News. Eklektix","author":"Edge Jake","unstructured":"Jake Edge . 2010. Receive flow steering . In Linux Weekly News. Eklektix , Inc . Jake Edge. 2010. Receive flow steering. In Linux Weekly News. Eklektix, Inc."},{"key":"e_1_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.5555\/1247415.1247436"},{"key":"e_1_2_1_46_1","volume-title":"Proceedings of the International Symposium on Performance Evaluation of Computer and Telecommunication Systems (SPECTS","author":"Emmerich P.","year":"2015","unstructured":"P. Emmerich , D. Raumer , A. Beifu , L. Erlacher , F. Wohlfart , T. M. Runge , S. Gallenmller , and G. Carle . 2015. Optimizing latency and CPU load in packet processing systems . In Proceedings of the International Symposium on Performance Evaluation of Computer and Telecommunication Systems (SPECTS 2015 ). 1--8. P. Emmerich, D. Raumer, A. Beifu, L. Erlacher, F. Wohlfart, T. M. Runge, S. Gallenmller, and G. Carle. 2015. Optimizing latency and CPU load in packet processing systems. In Proceedings of the International Symposium on Performance Evaluation of Computer and Telecommunication Systems (SPECTS 2015). 1--8."},{"key":"e_1_2_1_47_1","volume-title":"Proceedings of the IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS\u201915)","author":"Felter W.","unstructured":"W. Felter , A. Ferreira , R. Rajamony , and J. Rubio . 2015. An updated performance comparison of virtual machines and Linux containers . In Proceedings of the IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS\u201915) . 171--172. W. Felter, A. Ferreira, R. Rajamony, and J. Rubio. 2015. An updated performance comparison of virtual machines and Linux containers. In Proceedings of the IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS\u201915). 171--172."},{"key":"e_1_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.5555\/792761.793233"},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.17487\/RFC3649"},{"key":"e_1_2_1_50_1","doi-asserted-by":"crossref","unstructured":"A. Ford C. Raiciu M. Handley and O. Bonaventure. 2013. TCP Extensions for Multipath Operation with Multiple Addresses. RFC 6824 (Experimental). (Jan. 2013). http:\/\/www.ietf.org\/rfc\/rfc6824.txt.  A. Ford C. Raiciu M. Handley and O. Bonaventure. 2013. TCP Extensions for Multipath Operation with Multiple Addresses. RFC 6824 (Experimental). (Jan. 2013). http:\/\/www.ietf.org\/rfc\/rfc6824.txt.","DOI":"10.17487\/rfc6824"},{"key":"e_1_2_1_51_1","volume-title":"Proceedings of the USENIX Annual Technical Conference","author":"Freimuth Douglas","year":"2005","unstructured":"Douglas Freimuth , Elbert C. Hu , Jason D. LaVoie , Ronald Mraz , Erich M. Nahum , Prashant Pradhan , and John M. Tracey . 2005. Server network scalability and TCP offload . In Proceedings of the USENIX Annual Technical Conference , April 10-15, 2005 , Anaheim, CA. 209--222. http:\/\/www.usenix.org\/events\/usenix05\/tech\/general\/freimuth.html. Douglas Freimuth, Elbert C. Hu, Jason D. LaVoie, Ronald Mraz, Erich M. Nahum, Prashant Pradhan, and John M. Tracey. 2005. Server network scalability and TCP offload. In Proceedings of the USENIX Annual Technical Conference, April 10-15, 2005, Anaheim, CA. 209--222. http:\/\/www.usenix.org\/events\/usenix05\/tech\/general\/freimuth.html."},{"key":"e_1_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.1145\/2491463"},{"key":"e_1_2_1_53_1","volume-title":"Data Traffic Monitoring and Analysis: From Measurement, Classification, and Anomaly Detection to Quality of Experience","author":"Garc\u00eda-Dorado Jos\u00e9 Luis","unstructured":"Jos\u00e9 Luis Garc\u00eda-Dorado , Felipe Mata , Javier Ramos , Pedro M. Santiago del R\u00edo , Victor Moreno , and Javier Aracil . 2013. Data Traffic Monitoring and Analysis: From Measurement, Classification, and Anomaly Detection to Quality of Experience . Springer , Berlin , High-Performance Network Traffic Processing Systems Using Commodity Hardware, 3--27. Jos\u00e9 Luis Garc\u00eda-Dorado, Felipe Mata, Javier Ramos, Pedro M. Santiago del R\u00edo, Victor Moreno, and Javier Aracil. 2013. Data Traffic Monitoring and Analysis: From Measurement, Classification, and Anomaly Detection to Quality of Experience. Springer, Berlin, High-Performance Network Traffic Processing Systems Using Commodity Hardware, 3--27."},{"key":"e_1_2_1_54_1","volume-title":"Proceedings of the USENIX Annual Technical Conference (USENIX ATC 14)","author":"Gaud Fabien","year":"2014","unstructured":"Fabien Gaud , Baptiste Lepers , Jeremie Decouchant , Justin Funston , Alexandra Fedorova , and Vivien Quema . 2014 . Large pages may be harmful on NUMA systems . In Proceedings of the USENIX Annual Technical Conference (USENIX ATC 14) . USENIX Association, Philadelphia, PA, 231--242. Fabien Gaud, Baptiste Lepers, Jeremie Decouchant, Justin Funston, Alexandra Fedorova, and Vivien Quema. 2014. Large pages may be harmful on NUMA systems. In Proceedings of the USENIX Annual Technical Conference (USENIX ATC 14). USENIX Association, Philadelphia, PA, 231--242."},{"key":"e_1_2_1_55_1","volume-title":"Proceedings of the IEEE International Parallel and Distributed Processing Symposium (IPDPS\u201916)","author":"Gerofi B.","unstructured":"B. Gerofi , M. Takagi , A. Hori , G. Nakamura , T. Shirasawa , and Y. Ishikawa . 2016. On the scalability, performance isolation and device driver transparency of the IHK\/McKernel hybrid lightweight kernel . In Proceedings of the IEEE International Parallel and Distributed Processing Symposium (IPDPS\u201916) . 1041--1050. B. Gerofi, M. Takagi, A. Hori, G. Nakamura, T. Shirasawa, and Y. Ishikawa. 2016. On the scalability, performance isolation and device driver transparency of the IHK\/McKernel hybrid lightweight kernel. In Proceedings of the IEEE International Parallel and Distributed Processing Symposium (IPDPS\u201916). 1041--1050."},{"key":"e_1_2_1_56_1","doi-asserted-by":"publisher","DOI":"10.1145\/2063166.2071893"},{"key":"e_1_2_1_57_1","volume-title":"Network Processors: Architecture, Programming, and Implementation. Morgan Kaufmann","author":"Giladi Ran","year":"2008","unstructured":"Ran Giladi . 2008 . Network Processors: Architecture, Programming, and Implementation. Morgan Kaufmann , Burlington, MA . Ran Giladi. 2008. Network Processors: Architecture, Programming, and Implementation. Morgan Kaufmann, Burlington, MA."},{"key":"e_1_2_1_58_1","volume-title":"Proceedings of the IEEE International Symposium on Performance Analysis of Systems Software (ISPASS\u201910)","author":"Grant R. E.","unstructured":"R. E. Grant , P. Balaji , and A. Afsahi . 2010. A study of hardware assisted IP over InfiniBand and its impact on enterprise data center performance . In Proceedings of the IEEE International Symposium on Performance Analysis of Systems Software (ISPASS\u201910) . 144--153. R. E. Grant, P. Balaji, and A. Afsahi. 2010. A study of hardware assisted IP over InfiniBand and its impact on enterprise data center performance. In Proceedings of the IEEE International Symposium on Performance Analysis of Systems Software (ISPASS\u201910). 144--153."},{"key":"e_1_2_1_59_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPP-W.2008.25"},{"key":"e_1_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2011.66"},{"key":"e_1_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2015.19"},{"key":"e_1_2_1_62_1","doi-asserted-by":"publisher","DOI":"10.1023\/B:GRID.0000037553.18581.3b"},{"key":"e_1_2_1_63_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.comnet.2006.11.009"},{"key":"e_1_2_1_64_1","volume-title":"Proceedings of the 9th International Conference on Network Protocols. 180--188","author":"Guo Liang","unstructured":"Liang Guo and I. Matta . 2001. The war between mice and elephants . In Proceedings of the 9th International Conference on Network Protocols. 180--188 . Liang Guo and I. Matta. 2001. The war between mice and elephants. In Proceedings of the 9th International Conference on Network Protocols. 180--188."},{"key":"e_1_2_1_65_1","doi-asserted-by":"publisher","DOI":"10.1145\/1400097.1400105"},{"key":"e_1_2_1_66_1","volume-title":"USENIX 10th Symposium on Operating Systems Design and Implementation (OSDI\u201912)","author":"Han Sangjin","year":"2012","unstructured":"Sangjin Han , Scott Marshall , Byung-Gon Chun , and Sylvia Ratnasamy . 2012 . MegaPipe: A new programming interface for scalable network I\/O . In USENIX 10th Symposium on Operating Systems Design and Implementation (OSDI\u201912) . 135--148. Sangjin Han, Scott Marshall, Byung-Gon Chun, and Sylvia Ratnasamy. 2012. MegaPipe: A new programming interface for scalable network I\/O. In USENIX 10th Symposium on Operating Systems Design and Implementation (OSDI\u201912). 135--148."},{"key":"e_1_2_1_67_1","doi-asserted-by":"publisher","DOI":"10.5555\/2688394.2688396"},{"key":"e_1_2_1_68_1","volume-title":"Proceedings of the IEEE 2002 International Conference on Cluster Computing. 317--324","author":"He E.","unstructured":"E. He , J. Leigh , O. Yu , and T. A. DeFanti . 2002. Reliable blast UDP : Predictable high performance bulk data transfer . In Proceedings of the IEEE 2002 International Conference on Cluster Computing. 317--324 . E. He, J. Leigh, O. Yu, and T. A. DeFanti. 2002. Reliable blast UDP : Predictable high performance bulk data transfer. In Proceedings of the IEEE 2002 International Conference on Cluster Computing. 317--324."},{"key":"e_1_2_1_69_1","volume-title":"RSockets. In Proceedings of the 2012 OpenFabrics Alliance Workshop. OpenFabrics Alliance.","author":"Hefty Sean","year":"2012","unstructured":"Sean Hefty . 2012 . RSockets. In Proceedings of the 2012 OpenFabrics Alliance Workshop. OpenFabrics Alliance. Sean Hefty. 2012. RSockets. In Proceedings of the 2012 OpenFabrics Alliance Workshop. OpenFabrics Alliance."},{"key":"e_1_2_1_70_1","volume-title":"Patterson","author":"Hennessy John L.","year":"2012","unstructured":"John L. Hennessy and David A . Patterson . 2012 . Computer Architecture : A Quantitative Approach (5. ed.). Morgan Kaufmann . John L. Hennessy and David A. Patterson. 2012. Computer Architecture: A Quantitative Approach (5. ed.). Morgan Kaufmann."},{"key":"e_1_2_1_71_1","doi-asserted-by":"publisher","DOI":"10.1145\/1362622.1362692"},{"key":"e_1_2_1_72_1","doi-asserted-by":"publisher","DOI":"10.1145\/2602204.2602212"},{"key":"e_1_2_1_73_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2015.2453960"},{"key":"e_1_2_1_75_1","doi-asserted-by":"publisher","DOI":"10.1145\/1080695.1069976"},{"key":"e_1_2_1_76_1","volume-title":"IEEE standard for local and metropolitan area networks--media access control (MAC) bridges and virtual bridged local area networks--amendment 17: Priority-based flow control","author":"IEEE.","year":"2011","unstructured":"IEEE. 2011. IEEE standard for local and metropolitan area networks--media access control (MAC) bridges and virtual bridged local area networks--amendment 17: Priority-based flow control . IEEE Std 802.1Qbb- 2011 (Amendment to IEEE Std 802.1Q-2011 as amended by IEEE Std 802.1Qbe-2011 and IEEE Std 802.1Qbc-2011) (Sept 2011), 1--40. IEEE. 2011. IEEE standard for local and metropolitan area networks--media access control (MAC) bridges and virtual bridged local area networks--amendment 17: Priority-based flow control. IEEE Std 802.1Qbb-2011 (Amendment to IEEE Std 802.1Q-2011 as amended by IEEE Std 802.1Qbe-2011 and IEEE Std 802.1Qbc-2011) (Sept 2011), 1--40."},{"key":"e_1_2_1_77_1","first-page":"3","article-title":"IEEE standard for ethernet - section 1","volume":"802","author":"IEEE.","year":"2012","unstructured":"IEEE. 2012 . IEEE standard for ethernet - section 1 . IEEE Std 802 . 3 - 2012 (Revision to IEEE Std 802.3-2008) (Dec 2012), 1--0. IEEE. 2012. IEEE standard for ethernet - section 1. IEEE Std 802.3-2012 (Revision to IEEE Std 802.3-2008) (Dec 2012), 1--0.","journal-title":"IEEE Std"},{"key":"e_1_2_1_81_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2009.19"},{"key":"e_1_2_1_82_1","volume-title":"Proceedings of the USENIX 11th Symposium on Networked Systems Design and Implementation (NSDI 14)","author":"Jeong EunYoung","year":"2014","unstructured":"EunYoung Jeong , Shinae Wood , Muhammad Jamshed , Haewon Jeong , Sunghwan Ihm , Dongsu Han , and KyoungSoo Park . 2014 . mTCP: A highly scalable user-level TCP stack for multicore systems . In Proceedings of the USENIX 11th Symposium on Networked Systems Design and Implementation (NSDI 14) . USENIX Association, Seattle, WA, 489--502. https:\/\/www.usenix.org\/conference\/nsdi14\/technical-sessions\/presentation\/jeong. EunYoung Jeong, Shinae Wood, Muhammad Jamshed, Haewon Jeong, Sunghwan Ihm, Dongsu Han, and KyoungSoo Park. 2014. mTCP: A highly scalable user-level TCP stack for multicore systems. In Proceedings of the USENIX 11th Symposium on Networked Systems Design and Implementation (NSDI 14). USENIX Association, Seattle, WA, 489--502. https:\/\/www.usenix.org\/conference\/nsdi14\/technical-sessions\/presentation\/jeong."},{"key":"e_1_2_1_83_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPP.2011.37"},{"key":"e_1_2_1_84_1","doi-asserted-by":"publisher","DOI":"10.1145\/2391229.2391238"},{"key":"e_1_2_1_85_1","doi-asserted-by":"publisher","DOI":"10.1145\/964725.633035"},{"key":"e_1_2_1_86_1","volume-title":"Weissman","author":"Katz Daniel S.","year":"2012","unstructured":"Daniel S. Katz , Shantenu Jha , Manish Parashar , Omer F. Rana , and Jon B . Weissman . 2012 . Survey and analysis of production distributed computing infrastructures. CoRR abs\/1208.2649 (2012). http:\/\/arxiv.org\/abs\/1208.2649 Daniel S. Katz, Shantenu Jha, Manish Parashar, Omer F. Rana, and Jon B. Weissman. 2012. Survey and analysis of production distributed computing infrastructures. CoRR abs\/1208.2649 (2012). http:\/\/arxiv.org\/abs\/1208.2649"},{"key":"e_1_2_1_87_1","doi-asserted-by":"publisher","DOI":"10.1109\/NCCA.2014.19"},{"key":"e_1_2_1_88_1","volume-title":"Proceedings of the IEEE 1st Conference on Network Softwarization (NetSoft). 1--8.","author":"Kawashima R.","unstructured":"R. Kawashima , S. Muramatsu , H. Nakayama , T. Hayashi , and H. Matsuo . 2015. SCLP: Segment-oriented connection-less protocol for high-performance software tunneling in datacenter networks . In Proceedings of the IEEE 1st Conference on Network Softwarization (NetSoft). 1--8. R. Kawashima, S. Muramatsu, H. Nakayama, T. Hayashi, and H. Matsuo. 2015. SCLP: Segment-oriented connection-less protocol for high-performance software tunneling in datacenter networks. In Proceedings of the IEEE 1st Conference on Network Softwarization (NetSoft). 1--8."},{"key":"e_1_2_1_89_1","doi-asserted-by":"publisher","DOI":"10.1109\/CloudNet.2012.6483669"},{"key":"e_1_2_1_90_1","doi-asserted-by":"publisher","DOI":"10.1145\/956981.956989"},{"key":"e_1_2_1_91_1","volume-title":"Proceedings of the 24th Annual Joint Conference of the IEEE Computer and Communications Societies (INFOCOM\u201905)","author":"King R.","year":"2005","unstructured":"R. King , Richard G. Baraniuk , and Rudolf H. Riedi . 2005. TCP-Africa: An adaptive and fair rapid increase rule for scalable TCP . In Proceedings of the 24th Annual Joint Conference of the IEEE Computer and Communications Societies (INFOCOM\u201905) , 13-17 March 2005 , Miami, FL. 1838--1848. R. King, Richard G. Baraniuk, and Rudolf H. Riedi. 2005. TCP-Africa: An adaptive and fair rapid increase rule for scalable TCP. In Proceedings of the 24th Annual Joint Conference of the IEEE Computer and Communications Societies (INFOCOM\u201905), 13-17 March 2005, Miami, FL. 1838--1848."},{"key":"e_1_2_1_92_1","doi-asserted-by":"publisher","DOI":"10.1109\/HPCC.2012.113"},{"key":"e_1_2_1_93_1","doi-asserted-by":"publisher","DOI":"10.1145\/2534695.2534699"},{"key":"e_1_2_1_94_1","doi-asserted-by":"publisher","DOI":"10.1109\/JPROC.2014.2371999"},{"key":"e_1_2_1_95_1","doi-asserted-by":"publisher","DOI":"10.5555\/1331699.1331716"},{"key":"e_1_2_1_96_1","volume-title":"Proceedings of the IEEE 15th International Symposium on High Performance Computer Architecture (HPCA). IEEE, 341--352","author":"Kumar A.","unstructured":"A. Kumar , R. Huggahalli , and S. Makineni . 2009. Characterization of direct cache access on multi-core systems and 10GbE . In Proceedings of the IEEE 15th International Symposium on High Performance Computer Architecture (HPCA). IEEE, 341--352 . A. Kumar, R. Huggahalli, and S. Makineni. 2009. Characterization of direct cache access on multi-core systems and 10GbE. In Proceedings of the IEEE 15th International Symposium on High Performance Computer Architecture (HPCA). IEEE, 341--352."},{"key":"e_1_2_1_97_1","volume-title":"Ross","author":"Kurose James F.","year":"2001","unstructured":"James F. Kurose and Keith W . Ross . 2001 . Computer Networking : A Top-Down Approach Featuring the Internet. Addison-Wesley-Longman . James F. Kurose and Keith W. Ross. 2001. Computer Networking: A Top-Down Approach Featuring the Internet. Addison-Wesley-Longman."},{"key":"e_1_2_1_98_1","volume-title":"Proceedings of the IEEE International Symposium on Parallel Distributed Processing (IPDPS\u201910)","author":"Lange J.","unstructured":"J. Lange , K. Pedretti , T. Hudson , P. Dinda , Z. Cui , L. Xia , P. Bridges , A. Gocke , S. Jaconette , M. Levenhagen , and R. Brightwell . 2010. Palacios and kitten: New high performance operating systems for scalable virtualized and native supercomputing . In Proceedings of the IEEE International Symposium on Parallel Distributed Processing (IPDPS\u201910) . 1--12. J. Lange, K. Pedretti, T. Hudson, P. Dinda, Z. Cui, L. Xia, P. Bridges, A. Gocke, S. Jaconette, M. Levenhagen, and R. Brightwell. 2010. Palacios and kitten: New high performance operating systems for scalable virtualized and native supercomputing. In Proceedings of the IEEE International Symposium on Parallel Distributed Processing (IPDPS\u201910). 1--12."},{"key":"e_1_2_1_99_1","volume-title":"Proceedings of the 2nd International Workshop on Protocols for Fast Long-Distance Networks (PFLDnet)","volume":"2","author":"Leith Douglas","year":"2004","unstructured":"Douglas Leith and Robert Shorten . 2004 . H-TCP: TCP for high-speed and long-distance networks . In Proceedings of the 2nd International Workshop on Protocols for Fast Long-Distance Networks (PFLDnet) , Vol. 2 . Argonne National Laboratory, Argonne, IL. Douglas Leith and Robert Shorten. 2004. H-TCP: TCP for high-speed and long-distance networks. In Proceedings of the 2nd International Workshop on Protocols for Fast Long-Distance Networks (PFLDnet), Vol. 2. Argonne National Laboratory, Argonne, IL."},{"key":"e_1_2_1_100_1","volume-title":"Proceedings of the USENIX 7th Symposium on Operating Systems Design and Implementation (OSDI06)","author":"Le\u00f3n E. A.","unstructured":"E. A. Le\u00f3n and A. B. Maccabe . 2006. Reducing memory bandwidth for chip-multiprocessors using cache injection . In Proceedings of the USENIX 7th Symposium on Operating Systems Design and Implementation (OSDI06) . Poster Session, Seattle, WA. E. A. Le\u00f3n and A. B. Maccabe. 2006. Reducing memory bandwidth for chip-multiprocessors using cache injection. In Proceedings of the USENIX 7th Symposium on Operating Systems Design and Implementation (OSDI06). Poster Session, Seattle, WA."},{"key":"e_1_2_1_101_1","volume-title":"Proceedings of the USENIX Annual Technical Conference (USENIX ATC 15)","author":"Lepers Baptiste","year":"2015","unstructured":"Baptiste Lepers , Vivien Quema , and Alexandra Fedorova . 2015 . Thread and memory placement on NUMA systems: Asymmetry matters . In Proceedings of the USENIX Annual Technical Conference (USENIX ATC 15) . USENIX Association, Santa Clara, CA, 277--289. https:\/\/www.usenix.org\/conference\/atc15\/technical-session\/presentation\/lepers. Baptiste Lepers, Vivien Quema, and Alexandra Fedorova. 2015. Thread and memory placement on NUMA systems: Asymmetry matters. In Proceedings of the USENIX Annual Technical Conference (USENIX ATC 15). USENIX Association, Santa Clara, CA, 277--289. https:\/\/www.usenix.org\/conference\/atc15\/technical-session\/presentation\/lepers."},{"key":"e_1_2_1_102_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2009.16"},{"key":"e_1_2_1_103_1","doi-asserted-by":"publisher","DOI":"10.1145\/1882486.1882503"},{"key":"e_1_2_1_104_1","doi-asserted-by":"publisher","DOI":"10.1145\/1872007.1872039"},{"key":"e_1_2_1_105_1","volume-title":"Proceedings of the IEEE 17th International Symposium on High Performance Computer Architecture (HPCA","author":"Liao Guangdeng","year":"2011","unstructured":"Guangdeng Liao , Xia Zhu , and L. Bnuyan . 2011. A new server I\/O architecture for high speed networks . In Proceedings of the IEEE 17th International Symposium on High Performance Computer Architecture (HPCA 2011 ). 255--265. Guangdeng Liao, Xia Zhu, and L. Bnuyan. 2011. A new server I\/O architecture for high speed networks. In Proceedings of the IEEE 17th International Symposium on High Performance Computer Architecture (HPCA 2011). 255--265."},{"key":"e_1_2_1_106_1","unstructured":"Robert Love. 2007. Linux System Programming: System and Library Calls Every Programmer Needs to Know. O\u2019Reilly. http:\/\/www.oreilly.com\/catalog\/9780596009588\/index.html  Robert Love. 2007. Linux System Programming: System and Library Calls Every Programmer Needs to Know. O\u2019Reilly. http:\/\/www.oreilly.com\/catalog\/9780596009588\/index.html"},{"key":"e_1_2_1_107_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPP.2013.78"},{"key":"e_1_2_1_108_1","doi-asserted-by":"crossref","unstructured":"Xukang Lu Qishi Wu Nageswara S. V. Rao and Zongmin Wang. 2009. Performance-adaptive prediction-based transport control over dedicated links. In Quality of Service in Heterogeneous Networks Novella Bartolini Sotiris Nikoletseas Prasun Sinha Valeria Cardellini and Anirban Mahanti (Eds.). Lecture Notes of the Institute for Computer Sciences Social Informatics and Telecommunications Engineering Vol. 22. Springer Berlin 265--279.  Xukang Lu Qishi Wu Nageswara S. V. Rao and Zongmin Wang. 2009. Performance-adaptive prediction-based transport control over dedicated links. In Quality of Service in Heterogeneous Networks Novella Bartolini Sotiris Nikoletseas Prasun Sinha Valeria Cardellini and Anirban Mahanti (Eds.). Lecture Notes of the Institute for Computer Sciences Social Informatics and Telecommunications Engineering Vol. 22. Springer Berlin 265--279.","DOI":"10.1007\/978-3-642-10625-5_17"},{"key":"e_1_2_1_109_1","doi-asserted-by":"publisher","DOI":"10.1145\/357783.331677"},{"key":"e_1_2_1_110_1","doi-asserted-by":"publisher","DOI":"10.1145\/2619239.2626311"},{"key":"e_1_2_1_111_1","volume-title":"Tsunami: A High-Speed Rate-Controlled Protocol for File Transfer.","author":"Meiss Mark R.","year":"2002","unstructured":"Mark R. Meiss . 2002 . Tsunami: A High-Speed Rate-Controlled Protocol for File Transfer. (2002). http:\/\/steinbeck.ucs.indiana.edu\/mmeiss\/papers\/tsunami.pdf. Mark R. Meiss. 2002. Tsunami: A High-Speed Rate-Controlled Protocol for File Transfer. (2002). http:\/\/steinbeck.ucs.indiana.edu\/mmeiss\/papers\/tsunami.pdf."},{"key":"e_1_2_1_112_1","unstructured":"C. B. Melzer J. Rosen R. O\u2019Gorman P. A. Wood M. C. Drummond and D. Hiller. 1999. IP checksum offload. (April 27 1999). http:\/\/www.google.com\/patents\/US5898713 US Patent 5 898 713.  C. B. Melzer J. Rosen R. O\u2019Gorman P. A. Wood M. C. Drummond and D. Hiller. 1999. IP checksum offload. (April 27 1999). http:\/\/www.google.com\/patents\/US5898713 US Patent 5 898 713."},{"key":"e_1_2_1_113_1","volume-title":"Proceedings of the IEEE 17th Symposium on Reliable Distributed Systems. IEEE, 341--346","author":"Milenkovic A.","unstructured":"A. Milenkovic and V. Milutinovic . 1998. Cache injection on bus based multiprocessors . In Proceedings of the IEEE 17th Symposium on Reliable Distributed Systems. IEEE, 341--346 . A. Milenkovic and V. Milutinovic. 1998. Cache injection on bus based multiprocessors. In Proceedings of the IEEE 17th Symposium on Reliable Distributed Systems. IEEE, 341--346."},{"key":"e_1_2_1_114_1","volume-title":"Proceedings of the IEEE International Conference on Communications (ICC","author":"Millnert V.","year":"2017","unstructured":"V. Millnert , J. Eker , and E. Bini . 2017. Dynamic control of NFV forwarding graphs with end-to-end deadline constraints . In Proceedings of the IEEE International Conference on Communications (ICC 2017 ). 1--7. V. Millnert, J. Eker, and E. Bini. 2017. Dynamic control of NFV forwarding graphs with end-to-end deadline constraints. In Proceedings of the IEEE International Conference on Communications (ICC 2017). 1--7."},{"key":"e_1_2_1_115_1","volume-title":"Proceedings of the 5th International Symposium on Modeling, Analysis, and Simulation of Computer and Telecommunication Systems (MASCOTS\u201997)","author":"Milutinovic V.","unstructured":"V. Milutinovic , A. Milenkovic , and G. Sheaffer . 1997. The cache injection\/cofetch architecture: Initial performance evaluation . In Proceedings of the 5th International Symposium on Modeling, Analysis, and Simulation of Computer and Telecommunication Systems (MASCOTS\u201997) . IEEE, 63--64. V. Milutinovic, A. Milenkovic, and G. Sheaffer. 1997. The cache injection\/cofetch architecture: Initial performance evaluation. In Proceedings of the 5th International Symposium on Modeling, Analysis, and Simulation of Computer and Telecommunication Systems (MASCOTS\u201997). IEEE, 63--64."},{"key":"e_1_2_1_116_1","doi-asserted-by":"publisher","DOI":"10.1145\/2856125"},{"key":"e_1_2_1_117_1","volume-title":"9th Workshop on Hot Topics in Operating Systems (HotOS IX).","author":"Mogul J. C.","year":"2003","unstructured":"J. C. Mogul . 2003 . TCP offload is a dumb idea whose time has come . In 9th Workshop on Hot Topics in Operating Systems (HotOS IX). J. C. Mogul. 2003. TCP offload is a dumb idea whose time has come. In 9th Workshop on Hot Topics in Operating Systems (HotOS IX)."},{"key":"e_1_2_1_118_1","doi-asserted-by":"publisher","DOI":"10.1145\/263326.263335"},{"key":"e_1_2_1_119_1","doi-asserted-by":"publisher","DOI":"10.5555\/1647539.1647545"},{"key":"e_1_2_1_120_1","doi-asserted-by":"publisher","DOI":"10.1145\/2209249.2209264"},{"key":"e_1_2_1_121_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11390-013-1352-2"},{"key":"e_1_2_1_122_1","doi-asserted-by":"publisher","DOI":"10.1109\/PDP.2009.29"},{"key":"e_1_2_1_123_1","doi-asserted-by":"publisher","DOI":"10.1145\/2578901"},{"key":"e_1_2_1_124_1","doi-asserted-by":"publisher","DOI":"10.1145\/1542275.1542308"},{"key":"e_1_2_1_125_1","doi-asserted-by":"publisher","DOI":"10.1145\/2168836.2168870"},{"key":"e_1_2_1_126_1","doi-asserted-by":"publisher","DOI":"10.1145\/2812806"},{"key":"e_1_2_1_127_1","doi-asserted-by":"publisher","DOI":"10.1109\/40.988689"},{"key":"e_1_2_1_128_1","doi-asserted-by":"publisher","DOI":"10.5555\/2789770.2789779"},{"key":"e_1_2_1_129_1","volume-title":"Proceedings of the 11th USENIX Symposium on Networked Systems Design and Implementation (NSDI\u201914)","author":"Radhakrishnan Sivasankar","year":"2014","unstructured":"Sivasankar Radhakrishnan , Yilong Geng , Vimalkumar Jeyakumar , Abdul Kabbani , George Porter , and Amin Vahdat . 2014 . SENIC: Scalable NIC for end-host rate limiting . In Proceedings of the 11th USENIX Symposium on Networked Systems Design and Implementation (NSDI\u201914) . 475--488. https:\/\/www.usenix.org\/conference\/nsdi14\/technical-sessions\/presentation\/radhakrishnan. Sivasankar Radhakrishnan, Yilong Geng, Vimalkumar Jeyakumar, Abdul Kabbani, George Porter, and Amin Vahdat. 2014. SENIC: Scalable NIC for end-host rate limiting. In Proceedings of the 11th USENIX Symposium on Networked Systems Design and Implementation (NSDI\u201914). 475--488. https:\/\/www.usenix.org\/conference\/nsdi14\/technical-sessions\/presentation\/radhakrishnan."},{"key":"e_1_2_1_130_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCT.2013.27"},{"key":"e_1_2_1_131_1","volume-title":"Proceedings of the 5th International Symposium on Electronic System Design (ISED","author":"Rajesh R.","year":"2014","unstructured":"R. Rajesh , K. B. Ramia , and M. Kulkarni . 2014. Integration of LwIP stack over Intel(R) DPDK for high throughput packet delivery to applications . In Proceedings of the 5th International Symposium on Electronic System Design (ISED 2014 ). 130--134. R. Rajesh, K. B. Ramia, and M. Kulkarni. 2014. Integration of LwIP stack over Intel(R) DPDK for high throughput packet delivery to applications. In Proceedings of the 5th International Symposium on Electronic System Design (ISED 2014). 130--134."},{"key":"e_1_2_1_133_1","doi-asserted-by":"publisher","DOI":"10.1007\/BF03219967"},{"key":"e_1_2_1_134_1","doi-asserted-by":"publisher","DOI":"10.1145\/2038916.2038941"},{"key":"e_1_2_1_135_1","volume-title":"Proceedings of the 34th International Conference of the Chilean Computer Science Society (SCCC","author":"Rivera D.","year":"2015","unstructured":"D. Rivera , S. Blasco , J. Bustos-Jimenez , and J. Simmonds . 2015. Spin lock killed the performance star . In Proceedings of the 34th International Conference of the Chilean Computer Science Society (SCCC 2015 ). 1--6. D. Rivera, S. Blasco, J. Bustos-Jimenez, and J. Simmonds. 2015. Spin lock killed the performance star. In Proceedings of the 34th International Conference of the Chilean Computer Science Society (SCCC 2015). 1--6."},{"key":"e_1_2_1_136_1","doi-asserted-by":"publisher","DOI":"10.1145\/2090147.2103536"},{"key":"e_1_2_1_137_1","doi-asserted-by":"publisher","DOI":"10.1109\/MM.2012.12"},{"key":"e_1_2_1_138_1","doi-asserted-by":"publisher","DOI":"10.1145\/1851276.1851283"},{"key":"e_1_2_1_139_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCNT.2010.49"},{"key":"e_1_2_1_140_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.aeue.2006.04.007"},{"key":"e_1_2_1_141_1","volume-title":"Proceedings of the IEEE INFOCOM","author":"Sanadhya S.","year":"2011","unstructured":"S. Sanadhya and R. Sivakumar . 2011. Adaptive flow control for TCP on mobile phones . In Proceedings of the IEEE INFOCOM 2011 . 2912--2920. S. Sanadhya and R. Sivakumar. 2011. Adaptive flow control for TCP on mobile phones. In Proceedings of the IEEE INFOCOM 2011. 2912--2920."},{"key":"e_1_2_1_142_1","volume-title":"IP: When does hardware support help?","author":"Sarkar P.","year":"2003","unstructured":"P. Sarkar , S. Uttamchandani , and K. Voruganti . 2003 . Storage over IP: When does hardware support help? (2003). P. Sarkar, S. Uttamchandani, and K. Voruganti. 2003. Storage over IP: When does hardware support help? (2003)."},{"key":"e_1_2_1_143_1","doi-asserted-by":"publisher","DOI":"10.1145\/1362622.1362672"},{"key":"e_1_2_1_144_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPP.2013.73"},{"key":"e_1_2_1_145_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2006.18"},{"key":"e_1_2_1_146_1","volume-title":"USENIX Annual Technical Conference","author":"Shalev Leah","year":"2010","unstructured":"Leah Shalev , Julian Satran , Eran Borovik , and Muli Ben-Yehuda . 2010 . IsoStack - highly efficient network processing on dedicated cores . In USENIX Annual Technical Conference , Boston, MA , June 23-25, 2010. https:\/\/www.usenix.org\/conference\/usenix-atc-10\/isostack%E2%80%94highly-efficient-network-processing-dedicated-cores. Leah Shalev, Julian Satran, Eran Borovik, and Muli Ben-Yehuda. 2010. IsoStack - highly efficient network processing on dedicated cores. In USENIX Annual Technical Conference, Boston, MA, June 23-25, 2010. https:\/\/www.usenix.org\/conference\/usenix-atc-10\/isostack%E2%80%94highly-efficient-network-processing-dedicated-cores."},{"key":"e_1_2_1_147_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2015.13"},{"key":"e_1_2_1_148_1","doi-asserted-by":"publisher","DOI":"10.5555\/2490483.2490484"},{"key":"e_1_2_1_149_1","doi-asserted-by":"publisher","DOI":"10.1145\/944747.944750"},{"key":"e_1_2_1_150_1","doi-asserted-by":"publisher","DOI":"10.1145\/2366231.2337196"},{"key":"e_1_2_1_151_1","volume-title":"Proceedings of the 9th USENIX Conference on Operating Systems Design and Implementation (OSDI\u201910)","author":"Soares Livio","year":"2010","unstructured":"Livio Soares and Michael Stumm . 2010 . FlexSC: Flexible system call scheduling with exception-less system calls . In Proceedings of the 9th USENIX Conference on Operating Systems Design and Implementation (OSDI\u201910) . USENIX Association, Berkeley, CA, 1--8. Livio Soares and Michael Stumm. 2010. FlexSC: Flexible system call scheduling with exception-less system calls. In Proceedings of the 9th USENIX Conference on Operating Systems Design and Implementation (OSDI\u201910). USENIX Association, Berkeley, CA, 1--8."},{"key":"e_1_2_1_152_1","unstructured":"G. Adam Stanislav. 2016. Chapter 7. Sockets. Retrieved from https:\/\/www.freebsd.org\/doc\/en\/books\/developers-handbook\/sockets.html.  G. Adam Stanislav. 2016. Chapter 7. Sockets. Retrieved from https:\/\/www.freebsd.org\/doc\/en\/books\/developers-handbook\/sockets.html."},{"key":"e_1_2_1_153_1","doi-asserted-by":"publisher","DOI":"10.17487\/RFC2001"},{"key":"e_1_2_1_154_1","doi-asserted-by":"publisher","DOI":"10.1109\/HPCC.2012.75"},{"key":"e_1_2_1_155_1","volume-title":"IEEE International Conference on Cluster Computing and Workshops (CLUSTER\u201909)","author":"Subramoni H.","unstructured":"H. Subramoni , Ping Lai , Miao Luo , and D. K. Panda . 2009. RDMA over ethernet: A preliminary study . In IEEE International Conference on Cluster Computing and Workshops (CLUSTER\u201909) . 1--9. H. Subramoni, Ping Lai, Miao Luo, and D. K. Panda. 2009. RDMA over ethernet: A preliminary study. In IEEE International Conference on Cluster Computing and Workshops (CLUSTER\u201909). 1--9."},{"key":"e_1_2_1_156_1","doi-asserted-by":"publisher","DOI":"10.1145\/1095408.1095421"},{"key":"e_1_2_1_157_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11432-014-5255-9"},{"key":"e_1_2_1_158_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISDA.2008.102"},{"key":"e_1_2_1_159_1","doi-asserted-by":"publisher","DOI":"10.1109\/eScience.2012.6404462"},{"key":"e_1_2_1_160_1","unstructured":"Willem de Brujin Tom Herbert. 2010. Scaling in the Linux Networking Stack (2.6.35 ed.). Google. https:\/\/www.kernel.org\/doc\/Documentation\/networking\/scaling.txt Linux Kernel Documentation.  Willem de Brujin Tom Herbert. 2010. Scaling in the Linux Networking Stack (2.6.35 ed.). Google. https:\/\/www.kernel.org\/doc\/Documentation\/networking\/scaling.txt Linux Kernel Documentation."},{"key":"e_1_2_1_161_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2011.15"},{"key":"e_1_2_1_162_1","doi-asserted-by":"publisher","DOI":"10.1504\/IJHPCN.2004.008896"},{"key":"e_1_2_1_163_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNET.2006.886335"},{"key":"e_1_2_1_164_1","doi-asserted-by":"publisher","DOI":"10.1145\/2612262.2612263"},{"key":"e_1_2_1_165_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2011.195"},{"key":"e_1_2_1_166_1","volume-title":"23rd Annual Joint Conference of the IEEE Computer and Communications Societies (INFOCOM","volume":"4","author":"Xu Lisong","year":"2004","unstructured":"Lisong Xu , K. Harfoush , and Injong Rhee . 2004 . Binary increase congestion control (BIC) for fast long-distance networks . In 23rd Annual Joint Conference of the IEEE Computer and Communications Societies (INFOCOM 2004), Vol. 4 . 2514--2524. Lisong Xu, K. Harfoush, and Injong Rhee. 2004. Binary increase congestion control (BIC) for fast long-distance networks. In 23rd Annual Joint Conference of the IEEE Computer and Communications Societies (INFOCOM 2004), Vol. 4. 2514--2524."},{"key":"e_1_2_1_167_1","volume-title":"International Conference for High Performance Computing, Networking, Storage and Analysis (SC","author":"Yoshino T.","year":"2008","unstructured":"T. Yoshino , Y. Sugawara , K. Inagami , J. Tamatsukuri , M. Inaba , and K. Hiraki . 2008. Performance optimization of TCP\/IP over 10 gigabit ethernet by precise instrumentation . In International Conference for High Performance Computing, Networking, Storage and Analysis (SC 2008 ). 1--12. T. Yoshino, Y. Sugawara, K. Inagami, J. Tamatsukuri, M. Inaba, and K. Hiraki. 2008. Performance optimization of TCP\/IP over 10 gigabit ethernet by precise instrumentation. In International Conference for High Performance Computing, Networking, Storage and Analysis (SC 2008). 1--12."},{"key":"e_1_2_1_168_1","doi-asserted-by":"publisher","DOI":"10.1145\/2228360.2228410"},{"key":"e_1_2_1_169_1","volume-title":"USENIX 11th Symposium on Operating Systems Design and Implementation (OSDI 14)","author":"Zellweger Gerd","year":"2014","unstructured":"Gerd Zellweger , Simon Gerber , Kornilios Kourtis , and Timothy Roscoe . 2014 . Decoupling cores, kernels, and operating systems . In USENIX 11th Symposium on Operating Systems Design and Implementation (OSDI 14) . USENIX Association, Broomfield, CO, 17--31. https:\/\/www.usenix.org\/conference\/osdi14\/technical-sessions\/presentation\/zellweger. Gerd Zellweger, Simon Gerber, Kornilios Kourtis, and Timothy Roscoe. 2014. Decoupling cores, kernels, and operating systems. In USENIX 11th Symposium on Operating Systems Design and Implementation (OSDI 14). USENIX Association, Broomfield, CO, 17--31. https:\/\/www.usenix.org\/conference\/osdi14\/technical-sessions\/presentation\/zellweger."},{"key":"e_1_2_1_170_1","volume-title":"Proceedings of the 1st Workshop on Provisioning and Transport for Hybrid Networks (PATHNets 2004","author":"Zheng Xuan","year":"2004","unstructured":"Xuan Zheng , Anant Padmanath Mudambi , and Malathi Veeraraghavan . 2004 . FRTP: Fixed rate transport protocol\u2014A modified version of SABUL for end-to-end circuits . In Proceedings of the 1st Workshop on Provisioning and Transport for Hybrid Networks (PATHNets 2004 ). Broadnets, San Jose, CA, 11. http:\/\/broadnets.org\/ 2004\/workshop-papers\/Pathnets\/01_FRTP-XuanZheng.pdf. Xuan Zheng, Anant Padmanath Mudambi, and Malathi Veeraraghavan. 2004. FRTP: Fixed rate transport protocol\u2014A modified version of SABUL for end-to-end circuits. In Proceedings of the 1st Workshop on Provisioning and Transport for Hybrid Networks (PATHNets 2004). Broadnets, San Jose, CA, 11. http:\/\/broadnets.org\/2004\/workshop-papers\/Pathnets\/01_FRTP-XuanZheng.pdf."},{"key":"e_1_2_1_171_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11390-016-1614-x"}],"container-title":["ACM Computing Surveys"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3184899","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3184899","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3184899","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T02:11:25Z","timestamp":1750212685000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3184899"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,7,16]]},"references-count":165,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2019,5,31]]}},"alternative-id":["10.1145\/3184899"],"URL":"https:\/\/doi.org\/10.1145\/3184899","relation":{},"ISSN":["0360-0300","1557-7341"],"issn-type":[{"value":"0360-0300","type":"print"},{"value":"1557-7341","type":"electronic"}],"subject":[],"published":{"date-parts":[[2018,7,16]]},"assertion":[{"value":"2016-12-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2018-01-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2018-07-16","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}