{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,8]],"date-time":"2026-01-08T05:13:42Z","timestamp":1767849222932,"version":"3.49.0"},"publisher-location":"New York, NY, USA","reference-count":69,"publisher":"ACM","license":[{"start":{"date-parts":[[2019,10,12]],"date-time":"2019-10-12T00:00:00Z","timestamp":1570838400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000001","name":"National Science Foundation","doi-asserted-by":"publisher","award":["CNS-1705047"],"award-info":[{"award-number":["CNS-1705047"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2019,10,12]]},"DOI":"10.1145\/3352460.3358278","type":"proceedings-article","created":{"date-parts":[[2019,10,11]],"date-time":"2019-10-11T11:16:45Z","timestamp":1570792605000},"page":"699-711","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":22,"title":["NetDIMM"],"prefix":"10.1145","author":[{"given":"Mohammad","family":"Alian","sequence":"first","affiliation":[{"name":"University of Illinois Urbana-Champaign"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Nam Sung","family":"Kim","sequence":"additional","affiliation":[{"name":"University of Illinois Urbana-Champaign"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2019,10,12]]},"reference":[{"key":"e_1_3_2_1_1_1","unstructured":"[n. d.]. A prelude to nonvolatile DIMM technology. Future of NVDIMM-P. https:\/\/gigglehd.com\/gg\/hard\/1893698 Accessed: 06\/30\/2019."},{"key":"e_1_3_2_1_2_1","unstructured":"[n. d.]. Add 4GB DMA32 zone. https:\/\/lwn.net\/Articles\/152337\/. Accessed: 2018-11-20."},{"key":"e_1_3_2_1_3_1","unstructured":"[n. d.]. Arista 7130 Connect Series Ultra-low latency switches. https:\/\/www.arista.com\/en\/products\/7130-series Accessed: 03\/30\/2019."},{"key":"e_1_3_2_1_4_1","volume-title":"d.]. ConnectX-2 EN with RDMA over Ethernet (RoCE)","unstructured":"[n. d.]. ConnectX-2 EN with RDMA over Ethernet (RoCE). http:\/\/www.mellanox.com\/related-docs\/prod_software\/ConnectX-2_RDMA_RoCE.pdf. Accessed: 2018-11-28."},{"key":"e_1_3_2_1_5_1","unstructured":"[n. d.]. Data Plane Development Kit. http:\/\/dpdk.org\/."},{"key":"e_1_3_2_1_6_1","unstructured":"[n. d.]. Diablo conjures up hell of a DIMM: 128GB NAND pretend-RAM summoned. https:\/\/www.theregister.co.uk\/2016\/07\/22\/diablos_devilishly_clever_nandbased_pretend_dram_dimms_now_shipping\/ Accessed: 06\/30\/2019."},{"key":"e_1_3_2_1_7_1","unstructured":"[n. d.]. High memory. https:\/\/en.wikipedia.org\/wiki\/High_memory"},{"key":"e_1_3_2_1_8_1","unstructured":"[n. d.]. Intel announces Optane DC Persistent Memory DIMMs. https:\/\/www.techspot.com\/news\/79483-intel-announces-optane-dc-persistent-memory-dimms.html Accessed: 06\/30\/2019."},{"key":"e_1_3_2_1_9_1","unstructured":"[n. d.]. Intel Data Direct I\/O Technology (Intel DDIO): A Primer. https:\/\/www.intel.com\/content\/www\/us\/en\/io\/data-direct-i-o-technology-brief.html"},{"key":"e_1_3_2_1_10_1","unstructured":"[n. d.]. Iperf: The ultimate speed test tool for TCP UDP and SCTP. https:\/\/iperf.fr\/"},{"key":"e_1_3_2_1_11_1","unstructured":"[n. d.]. Large Send Offload. https:\/\/en.wikipedia.org\/wiki\/Large_send_offload. Accessed: 2018-11-28."},{"key":"e_1_3_2_1_12_1","unstructured":"[n. d.]. Low-latency Ethernet device polling. https:\/\/lwn.net\/Articles\/551284\/. Accessed: 2018-11-28."},{"key":"e_1_3_2_1_13_1","unstructured":"[n. d.]. Micron\u00e2\u0102&Zacute;s NVDIMM Delivers Persistent Memory. https:\/\/www.electronicdesign.com\/industrial-automation\/micron-s-nvdimm-delivers-persistent-memory Accessed: 06\/30\/2019."},{"key":"e_1_3_2_1_14_1","unstructured":"[n. d.]. Open Fast Path. https:\/\/openfastpath.org\/. Accessed: 2018-11-28."},{"key":"e_1_3_2_1_15_1","unstructured":"[n. d.]. Overcoming System Memory Challenges with Persistent Memory and NVDIMM-P. https:\/\/www.jedec.org\/sites\/default\/files\/Bill_Gervasi.pdf. Accessed: 2018-12-5."},{"key":"e_1_3_2_1_16_1","unstructured":"[n. d.]. Single- and Multichannel Memory Modes. https:\/\/www.intel.com\/content\/www\/us\/en\/support\/articles\/000005657\/boards-and-kits.html Accessed: 06\/30\/2019."},{"key":"e_1_3_2_1_17_1","unstructured":"[n. d.]. TCP frame. https:\/\/en.wikipedia.org\/wiki\/Transmission_Control_Protocol. Accessed: 2018-03-25."},{"key":"e_1_3_2_1_18_1","unstructured":"[n. d.]. Wall Street's Quest To Process Data At The Speed Of Light. https:\/\/www.informationweek.com\/wall-streets-quest-to-process-data-at-the-speed-of-light\/d\/d-id\/1054287. Accessed: 2018-12-3."},{"key":"e_1_3_2_1_19_1","volume-title":"Application-Transparent Near-Memory Processing Architecture with Memory Channel Network. In The 51st Annual IEEE\/ACM International Symposium on Microarchitecture. IEEE.","author":"Alian Mohammad","year":"2018","unstructured":"Mohammad Alian, Seung Won Min, Hadi Asgharimoghaddam, Ashutosh Dhar, Dong Kai Wang, Thomas Roewer, Adam McPadden, Oliver O'Halloran, Deming Chen, Jinjun Xiong, Daehoon Kim, Wen-mei Hwu, and Nam Sung Kim. 2018. Application-Transparent Near-Memory Processing Architecture with Memory Channel Network. In The 51st Annual IEEE\/ACM International Symposium on Microarchitecture. IEEE."},{"key":"e_1_3_2_1_20_1","volume-title":"Simulating PCI-Express Interconnect for Future System Exploration. In 2018 IEEE International Symposium on Workload Characterization (IISWC). IEEE, 168--178","author":"Alian Mohammad","year":"2018","unstructured":"Mohammad Alian, Krishna Parasuram Srinivasan, and Nam Sung Kim. 2018. Simulating PCI-Express Interconnect for Future System Exploration. In 2018 IEEE International Symposium on Workload Characterization (IISWC). IEEE, 168--178."},{"key":"e_1_3_2_1_21_1","volume-title":"Data center tcp (dctcp). ACM SIGCOMM computer communication review 41, 4","author":"Alizadeh Mohammad","year":"2011","unstructured":"Mohammad Alizadeh, Albert Greenberg, David A Maltz, Jitendra Padhye, Parveen Patel, Balaji Prabhakar, Sudipta Sengupta, and Murari Sridharan. 2011. Data center tcp (dctcp). ACM SIGCOMM computer communication review 41, 4 (2011), 63--74."},{"key":"e_1_3_2_1_22_1","volume-title":"Proceedings of the 9th USENIX conference on Networked Systems Design and Implementation. USENIX Association, 19--19","author":"Alizadeh Mohammad","year":"2012","unstructured":"Mohammad Alizadeh, Abdul Kabbani, Tom Edsall, Balaji Prabhakar, Amin Vahdat, and Masato Yasuda. 2012. Less is more: trading a little bandwidth for ultra-low latency in the data center. In Proceedings of the 9th USENIX conference on Networked Systems Design and Implementation. USENIX Association, 19--19."},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.5555\/2043535.2043537"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/1879141.1879175"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/2024716.2024718"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/1168918.1168897"},{"key":"e_1_3_2_1_27_1","unstructured":"Eduard Br\u00f6se. [n. d.]. ZeroCopy: Techniques Benefits and Pitfalls. ([n. d.])."},{"key":"e_1_3_2_1_28_1","volume-title":"PCI express system architecture","author":"Budruk Ravi","unstructured":"Ravi Budruk, Don Anderson, and Tom Shanley. 2004. PCI express system architecture. Addison-Wesley Professional."},{"key":"e_1_3_2_1_29_1","unstructured":"Willem de Bruijn and Eric Dumazet. [n. d.]. sendmsg copy avoidance with MSG_ZEROCOPY. ([n. d.])."},{"key":"e_1_3_2_1_30_1","volume-title":"NDA: Near-DRAM Acceleration Architecture Leveraging Commodity DRAM Devices and Standard Memory Modules. In HPCA.","author":"Farmahini-Farahani A.","year":"2015","unstructured":"A. Farmahini-Farahani, J. Ahn, K. Morrow, and N. S. Kim. 2015. NDA: Near-DRAM Acceleration Architecture Leveraging Commodity DRAM Devices and Standard Memory Modules. In HPCA."},{"key":"e_1_3_2_1_31_1","volume-title":"High Performance Interconnects, 2005. Proceedings. 13th Symposium on. IEEE, 58--63","author":"Balaji Pavan","year":"2005","unstructured":"Wu-chun Feng, Pavan Balaji, Chris Baron, Laxmi N Bhuyan, and Dhabaleswar K Panda. 2005. Performance characterization of a 10-Gigabit Ethernet TOE. In High Performance Interconnects, 2005. Proceedings. 13th Symposium on. IEEE, 58--63."},{"key":"e_1_3_2_1_32_1","volume-title":"14th {USENIX} Symposium on Networked Systems Design and Implementation ({NSDI} 17). 315--328.","author":"Firestone Daniel","unstructured":"Daniel Firestone. 2017. {VFP}: A Virtual Switch Platform for Host {SDN} in the Public Cloud. In 14th {USENIX} Symposium on Networked Systems Design and Implementation ({NSDI} 17). 315--328."},{"key":"e_1_3_2_1_33_1","unstructured":"Daniel Firestone Andrew Putnam Sambhrama Mundkur Derek Chiou Alireza Dabagh Mike Andrewartha Hari Angepat Vivek Bhanu Adrian Caulfield Eric Chung et al. 2018. Azure accelerated networking: SmartNICs in the public cloud. In 15th {USENIX} Symposium on Networked Systems Design and Implementation ({NSDI} 18). 51--66."},{"key":"e_1_3_2_1_34_1","volume-title":"USENIX Annual Technical Conference. 333--346","author":"Flajslik Mario","year":"2013","unstructured":"Mario Flajslik and Mendel Rosenblum. 2013. Network Interface Design for Low Latency Request-Response Protocols.. In USENIX Annual Technical Conference. 333--346."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL.2014.6927487"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/3098822.3098825"},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISPASS.2014.6844484"},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISCA.2005.23"},{"key":"e_1_3_2_1_39_1","unstructured":"Intel. 2018. Intel Ethernet Controller X710\/ XXV710\/XL710 Datasheet."},{"key":"e_1_3_2_1_40_1","unstructured":"Intel. 2018. Intel Memory Latency Checker v3.5."},{"key":"e_1_3_2_1_41_1","volume-title":"ACM SIGCOMM computer communication review","author":"Jacobson Van","unstructured":"Van Jacobson. 1988. Congestion avoidance and control. In ACM SIGCOMM computer communication review, Vol. 18. ACM, 314--329."},{"key":"e_1_3_2_1_42_1","unstructured":"James Hongyi Zeng. [n. d.]. Data Sharing on traffic pattern inside Facebook datacenter network. https:\/\/research.fb.com\/data-sharing-on-traffic-pattern-inside-facebooks-datacenter-network\/ Accessed: 03\/30\/2019."},{"key":"e_1_3_2_1_43_1","first-page":"493","article-title":"An Efficient Architecture for a TCP Offload Engine Based on Hardware\/Software Co-design","volume":"27","author":"Jang Hankook","year":"2011","unstructured":"Hankook Jang, Sang-Hwa Chung, Dong Kyue Kim, and Yun-Sung Lee. 2011. An Efficient Architecture for a TCP Offload Engine Based on Hardware\/Software Co-design. J. Inf. Sci. Eng. 27, 2 (2011), 493--509.","journal-title":"J. Inf. Sci. Eng."},{"key":"e_1_3_2_1_44_1","volume-title":"Proceedings of the 11th USENIX Conference on Networked Systems Design and Implementation (NSDI'14)","author":"Jeong Eun Young","year":"2014","unstructured":"Eun Young Jeong, Shinae Woo, Muhammad Jamshed, Haewon Jeong, Sunghwan Ihm, Dongsu Han, and KyoungSoo Park. 2014. mTCP: A Highly Scalable User-level TCP Stack for Multicore Systems. In Proceedings of the 11th USENIX Conference on Networked Systems Design and Implementation (NSDI'14). USENIX Association, Berkeley, CA, USA, 489--502."},{"key":"e_1_3_2_1_45_1","volume-title":"2016 USENIX Annual Technical Conference. 437","author":"Michael Kaminsky Anuj Kalia","year":"2016","unstructured":"Anuj Kalia Michael Kaminsky and David G Andersen. 2016. Design guidelines for high performance RDMA systems. In 2016 USENIX Annual Technical Conference. 437."},{"key":"e_1_3_2_1_46_1","volume-title":"Thomas Anderson, and Arvind Krishnamurthy.","author":"Kaufmann Antoine","year":"2016","unstructured":"Antoine Kaufmann, SImon Peter, Naveen Kr Sharma, Thomas Anderson, and Arvind Krishnamurthy. 2016. High performance packet processing with flexnic. In ACM SIGARCH Computer Architecture News, Vol. 44. ACM, 67--81."},{"key":"e_1_3_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1145\/605432.605423"},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2005.185"},{"key":"e_1_3_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1145\/1592568.1592599"},{"key":"e_1_3_2_1_50_1","volume-title":"International Conference on Supercomputing (ICS) Workshop on Characterizing Applications for Heterogeneous Exascale Systems (CACHES).","author":"Larsen Steen","year":"2011","unstructured":"Steen Larsen and Ben Lee. 2011. Platform IO DMA transaction acceleration. In International Conference on Supercomputing (ICS) Workshop on Characterizing Applications for Heterogeneous Exascale Systems (CACHES)."},{"key":"e_1_3_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10766-009-0109-6"},{"key":"e_1_3_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.5555\/2014698.2014863"},{"key":"e_1_3_2_1_53_1","doi-asserted-by":"publisher","DOI":"10.1145\/2485922.2485926"},{"key":"e_1_3_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1147\/JRD.2015.2429031"},{"key":"e_1_3_2_1_55_1","unstructured":"Micron. [n. d.]. 3D XPoint\u2122 Technology. https:\/\/www.micron.com\/products\/advanced-solutions\/3d-xpoint-technology."},{"key":"e_1_3_2_1_56_1","unstructured":"Micron. [n. d.]. Micron DDR4 SDRAM Datasheet. https:\/\/www.micron.com\/-\/media\/client\/global\/documents\/products\/data-sheet\/dram\/ddr4\/8gb_ddr4_sdram.pdf"},{"key":"e_1_3_2_1_57_1","doi-asserted-by":"publisher","DOI":"10.1109\/40.342013"},{"key":"e_1_3_2_1_58_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISPASS.2017.7975287"},{"key":"e_1_3_2_1_59_1","doi-asserted-by":"publisher","DOI":"10.1145\/3230543.3230560"},{"key":"e_1_3_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1145\/2785956.2787472"},{"key":"e_1_3_2_1_61_1","volume-title":"Mowry","author":"Seshadri Vivek","year":"2013","unstructured":"Vivek Seshadri, Yoongu Kim, Chris Fallin, Donghyuk Lee, Rachata Ausavarungnirun, Gennady Pekhimenko, Yixin Luo, Onur Mutlu, Phillip B. Gibbons, Michael A. Kozuch, and Todd C. Mowry. 2013. RowClone: Fast and Energy-Efficient In-DRAM Bulk Data Copy and Initialization. In MICRO."},{"key":"e_1_3_2_1_62_1","volume-title":"Design, and Implementation in Linux","author":"Seth Sameer","unstructured":"Sameer Seth and M Ajaykumar Venkatesulu. 2009. TCP\/IP Architecture, Design, and Implementation in Linux. Vol. 68. John Wiley & Sons."},{"key":"e_1_3_2_1_63_1","first-page":"256","article-title":"Performance review of zero copy techniques","volume":"6","author":"Song Jia","year":"2012","unstructured":"Jia Song and Jim Alves-Foss. 2012. Performance review of zero copy techniques. International Journal of Computer Science and Security (IJCSS) 6, 4 (2012), 256.","journal-title":"International Journal of Computer Science and Security (IJCSS)"},{"key":"e_1_3_2_1_64_1","doi-asserted-by":"publisher","DOI":"10.1145\/3123939.3124535"},{"key":"e_1_3_2_1_65_1","volume-title":"2006 International Conference on Field Programmable Logic and Applications. IEEE, 1--4.","author":"Tanabe N","year":"2006","unstructured":"N Tanabe, H Nakajyo, H Amano, M Yoshimi, A Kitamura, and T Miyashiro. 2006. DIMMnet-2: A reconfigurable board connected into a memory slot. In 2006 International Conference on Field Programmable Logic and Applications. IEEE, 1--4."},{"key":"e_1_3_2_1_66_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLUSTR.2000.888988"},{"key":"e_1_3_2_1_67_1","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA.2010.5416638"},{"key":"e_1_3_2_1_68_1","volume-title":"ResQ: Enabling SLOs in Network Function Virtualization. In 15th USENIX Symposium on Networked Systems Design and Implementation (NSDI 18)","author":"Tootoonchian Amin","year":"2018","unstructured":"Amin Tootoonchian, Aurojit Panda, Chang Lan, Melvin Walls, Katerina Argyraki, Sylvia Ratnasamy, and Scott Shenker. 2018. ResQ: Enabling SLOs in Network Function Virtualization. In 15th USENIX Symposium on Networked Systems Design and Implementation (NSDI 18). USENIX Association."},{"key":"e_1_3_2_1_69_1","volume-title":"Parallel and Distributed Systems, 2005. Proceedings. 11th International Conference on","volume":"1","author":"Wang Wen-Fong","year":"2005","unstructured":"Wen-Fong Wang, Jun-Yau Wang, and Jin-Jie Li. 2005. Study on enhanced strategies for TCP\/IP offload engines. In Parallel and Distributed Systems, 2005. Proceedings. 11th International Conference on, Vol. 1. IEEE, 398--404."}],"event":{"name":"MICRO '52: The 52nd Annual IEEE\/ACM International Symposium on Microarchitecture","location":"Columbus OH USA","acronym":"MICRO '52","sponsor":["SIGMICRO ACM Special Interest Group on Microarchitectural Research and Processing","IEEE CS"]},"container-title":["Proceedings of the 52nd Annual IEEE\/ACM International Symposium on Microarchitecture"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3352460.3358278","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3352460.3358278","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3352460.3358278","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,7,29]],"date-time":"2025-07-29T22:24:03Z","timestamp":1753827843000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3352460.3358278"}},"subtitle":["Low-Latency Near-Memory Network Interface Architecture"],"short-title":[],"issued":{"date-parts":[[2019,10,12]]},"references-count":69,"alternative-id":["10.1145\/3352460.3358278","10.1145\/3352460"],"URL":"https:\/\/doi.org\/10.1145\/3352460.3358278","relation":{},"subject":[],"published":{"date-parts":[[2019,10,12]]},"assertion":[{"value":"2019-10-12","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}