{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,1]],"date-time":"2026-06-01T23:22:14Z","timestamp":1780356134468,"version":"3.54.1"},"reference-count":76,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2022,5,26]],"date-time":"2022-05-26T00:00:00Z","timestamp":1653523200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"The European High-Performance Computing Joint Undertaking (JU) and the German Ministry of Education and Research","award":["955811"],"award-info":[{"award-number":["955811"]}]},{"name":"HPDS Corp."}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Meas. Anal. Comput. Syst."],"published-print":{"date-parts":[[2022,5,26]]},"abstract":"<jats:p>All-flash storage (AFS) systems have become an essential infrastructure component to support enterprise applications, where sub-millisecond latency and very high throughput are required. Nevertheless, the price per capacity ofsolid-state drives (SSDs) is relatively high, which has encouraged system architects to adoptdata reduction techniques, mainlydeduplication andcompression, in enterprise storage solutions. To provide higher reliability and performance, SSDs are typically grouped usingredundant array of independent disk (RAID) configurations. Data reduction on top of RAID arrays, however, adds I\/O overheads and also complicates the I\/O patterns redirected to the underlying backend SSDs, which invalidates the best-practice configurations used in AFS. Unfortunately, existing works on the performance of data reduction do not consider its interaction and I\/O overheads with other enterprise storage components including SSD arrays and RAID controllers. In this paper, using a real setup with enterprise-grade components and based on the open-source data reduction module RedHat VDO, we reveal novel observations on the performance gap between the state-of-the-art and the optimal all-flash storage stack with integrated data reduction. We therefore explore the I\/O patterns at the storage entry point and compare them with those at the disk subsystem. Our analysis shows a significant amount of I\/O overheads for guaranteeing consistency and avoiding data loss through data journaling, frequent small-sized metadata updates, and duplicate content verification. We accompany these observations with cross-layer optimizations to enhance the performance of AFS, which range from deriving new optimal hardware RAID configurations up to introducing changes to the enterprise storage stack. By analyzing the characteristics of I\/O types and their overheads, we propose three techniques: (a) application-aware lazy persistence, (b) a fast, read-only I\/O cache for duplicate verification, and (c) disaggregation of block maps and data by offloading block maps to a very fast persistent memory device. By consolidating all proposed optimizations and implementing them in an enterprise AFS, we show 1.3\u00d7 to 12.5\u00d7 speedup over the baseline AFS with 90% data reduction, and from 7.8\u00d7 up to 57\u00d7 performance\/cost improvement over an optimized AFS (with no data reduction) running applications ranging from 100% read-only to 100% write-only accesses.<\/jats:p>","DOI":"10.1145\/3530896","type":"journal-article","created":{"date-parts":[[2022,6,6]],"date-time":"2022-06-06T17:16:18Z","timestamp":1654535778000},"page":"1-27","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":4,"title":["An Enterprise-Grade Open-Source Data Reduction Architecture for All-Flash Storage Systems"],"prefix":"10.1145","volume":"6","author":[{"given":"Mohammadamin","family":"Ajdari","sequence":"first","affiliation":[{"name":"Sharif University of Technology, Tehran, Iran"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Patrick","family":"Raaf","sequence":"additional","affiliation":[{"name":"Johannes Gutenberg University, Mainz, Germany"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Mostafa","family":"Kishani","sequence":"additional","affiliation":[{"name":"Sharif University of Technology, Tehran, Iran"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Reza","family":"Salkhordeh","sequence":"additional","affiliation":[{"name":"Johannes Gutenberg University, Mainz, Germany"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Hossein","family":"Asadi","sequence":"additional","affiliation":[{"name":"Sharif University of Technology, Tehran, Iran"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Andr\u00e9","family":"Brinkmann","sequence":"additional","affiliation":[{"name":"Johannes Gutenberg University, Mainz, Germany"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,6,6]]},"reference":[{"key":"e_1_2_1_1_1","volume-title":"Proceedings of the International Workshop on OpenCL, IWOCL 2013 & 2014","author":"Abdelfattah Mohamed S.","year":"2013","unstructured":"014)]% AbdelfattahHS14, Mohamed S. Abdelfattah , Andrei Hagiescu , and Deshanand P. Singh . 2014. Gzip on a chip: high performance lossless data compression on FPGAs using OpenCL . In Proceedings of the International Workshop on OpenCL, IWOCL 2013 & 2014 , May 13 --14 , 2013 , Georgia Tech, Atlanta, GA, USA \/ Bristol, UK, May 12--13, 2014 . ACM, 4:1--4:9. https:\/\/doi.org\/10.1145\/2664666.2664670 10.1145\/2664666.2664670 014)]% AbdelfattahHS14, Mohamed S. Abdelfattah, Andrei Hagiescu, and Deshanand P. Singh. 2014. Gzip on a chip: high performance lossless data compression on FPGAs using OpenCL. In Proceedings of the International Workshop on OpenCL, IWOCL 2013 & 2014, May 13--14, 2013, Georgia Tech, Atlanta, GA, USA \/ Bristol, UK, May 12--13, 2014 . ACM, 4:1--4:9. https:\/\/doi.org\/10.1145\/2664666.2664670"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/3179412"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2021.3066308"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3352460.3358303"},{"key":"e_1_2_1_5_1","volume-title":"CIDR: A Cost-Effective In-Line Data Reduction System for Terabit-Per-Second Scale SSD Arrays. In 25th IEEE International Symposium on High Performance Computer Architecture, HPCA 2019","author":"Ajdari Mohammadamin","year":"2019","unstructured":"019b)]% AjdariPKKK19, Mohammadamin Ajdari , Pyeongsu Park , Joonsung Kim , Dongup Kwon , and Jangwoo Kim . 2019 b . CIDR: A Cost-Effective In-Line Data Reduction System for Terabit-Per-Second Scale SSD Arrays. In 25th IEEE International Symposium on High Performance Computer Architecture, HPCA 2019 , Washington, DC, USA, February 16--20 , 2019. IEEE, 28--41. https:\/\/doi.org\/10.1109\/HPCA.2019.00025 10.1109\/HPCA.2019.00025 019b)]% AjdariPKKK19, Mohammadamin Ajdari, Pyeongsu Park, Joonsung Kim, Dongup Kwon, and Jangwoo Kim. 2019 b. CIDR: A Cost-Effective In-Line Data Reduction System for Terabit-Per-Second Scale SSD Arrays. In 25th IEEE International Symposium on High Performance Computer Architecture, HPCA 2019, Washington, DC, USA, February 16--20, 2019. IEEE, 28--41. https:\/\/doi.org\/10.1109\/HPCA.2019.00025"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/LCA.2017.2753258"},{"key":"e_1_2_1_7_1","volume-title":"Proceedings of the 10th USENIX conference on File and Storage Technologies, FAST 2012","author":"Bhatotia Pramod","year":"2012","unstructured":"012)]% BhatotiaRV12, Pramod Bhatotia , Rodrigo Rodrigues , and Akshat Verma . 2012 . Shredder: GPU-accelerated incremental storage and computation . In Proceedings of the 10th USENIX conference on File and Storage Technologies, FAST 2012 , San Jose, CA, USA, February 14--17 , 2012. USENIX Association, 14. https:\/\/www.usenix.org\/conference\/fast12\/shredder-gpu-accelerated-incremental-storage-and-computation 012)]% BhatotiaRV12, Pramod Bhatotia, Rodrigo Rodrigues, and Akshat Verma. 2012. Shredder: GPU-accelerated incremental storage and computation. In Proceedings of the 10th USENIX conference on File and Storage Technologies, FAST 2012, San Jose, CA, USA, February 14--17, 2012. USENIX Association, 14. https:\/\/www.usenix.org\/conference\/fast12\/shredder-gpu-accelerated-incremental-storage-and-computation"},{"key":"e_1_2_1_8_1","volume-title":"https:\/\/www.broadcom.com\/support\/knowledgebase\/1211161498420\/performance-tuning-on-the-mr-sas-2108-lsi-sas-2208-sas-3108-base Retrieved","author":"Broadcom Inc. 2022. Knowledge Base.","year":"2022","unstructured":"022)]% broadcomWBCache, Broadcom Inc. 2022. Knowledge Base. https:\/\/www.broadcom.com\/support\/knowledgebase\/1211161498420\/performance-tuning-on-the-mr-sas-2108-lsi-sas-2208-sas-3108-base Retrieved January 20, 2022 from 022)]% broadcomWBCache, Broadcom Inc. 2022. Knowledge Base. https:\/\/www.broadcom.com\/support\/knowledgebase\/1211161498420\/performance-tuning-on-the-mr-sas-2108-lsi-sas-2208-sas-3108-base Retrieved January 20, 2022 from"},{"key":"e_1_2_1_9_1","volume-title":"et almbox","author":"Bucy John S","year":"2003","unstructured":"003)]% bucy2003disksim, John S Bucy , Gregory R Ganger , et almbox . 2003 . The DiskSim simulation environment version 3.0 reference manual .School of Computer Science, Carnegie Mellon University . 003)]% bucy2003disksim, John S Bucy, Gregory R Ganger, et almbox. 2003. The DiskSim simulation environment version 3.0 reference manual .School of Computer Science, Carnegie Mellon University."},{"key":"e_1_2_1_10_1","volume-title":"9th USENIX Conference on File and Storage Technologies","author":"Chen Feng","year":"2011","unstructured":"011)]% chenLZ11, Feng Chen , Tian Luo , and Xiaodong Zhang . 2011 . CAFTL: A Content-Aware Flash Translation Layer Enhancing the Lifespan of Flash Memory based Solid State Drives . In 9th USENIX Conference on File and Storage Technologies , San Jose, CA, USA, February 15--17. 77--90. 011)]% chenLZ11, Feng Chen, Tian Luo, and Xiaodong Zhang. 2011. CAFTL: A Content-Aware Flash Translation Layer Enhancing the Lifespan of Flash Memory based Solid State Drives. In 9th USENIX Conference on File and Storage Technologies, San Jose, CA, USA, February 15--17. 77--90."},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/12.543706"},{"key":"e_1_2_1_12_1","unstructured":"017)]% HP_Reduce Chris M. Evans. Jan 2017. HPE 3PAR Adaptive Data reduction: A competitive comparison of array-based data reduction . https:\/\/www.hpe.com\/h20195\/v2\/getpdf.aspx\/4AA6--6256ENW.pdf .  017)]% HP_Reduce Chris M. Evans. Jan 2017. HPE 3PAR Adaptive Data reduction: A competitive comparison of array-based data reduction . https:\/\/www.hpe.com\/h20195\/v2\/getpdf.aspx\/4AA6--6256ENW.pdf ."},{"key":"e_1_2_1_13_1","volume-title":"Mixing Deduplication and Compression on Active Data Sets. In 2011 Data Compression Conference (DCC 2011","author":"Constantinescu Cornel","year":"2011","unstructured":"011)]% ConstantinescuGC11, Cornel Constantinescu , Joseph S. Glider , and David D. Chambliss . 2011 . Mixing Deduplication and Compression on Active Data Sets. In 2011 Data Compression Conference (DCC 2011 ), 29--31 March 2011 , Snowbird, UT, USA. IEEE Computer Society, 393--402. https:\/\/doi.org\/10.1109\/DCC. 2011.46 10.1109\/DCC.2011.46 011)]% ConstantinescuGC11, Cornel Constantinescu, Joseph S. Glider, and David D. Chambliss. 2011. Mixing Deduplication and Compression on Active Data Sets. In 2011 Data Compression Conference (DCC 2011), 29--31 March 2011, Snowbird, UT, USA. IEEE Computer Society, 393--402. https:\/\/doi.org\/10.1109\/DCC.2011.46"},{"key":"e_1_2_1_14_1","volume-title":"ChunkStash: Speeding Up Inline Storage Deduplication Using Flash Memory. In USENIX Annual Technical Conference (ATC)","author":"Debnath Biplob K.","year":"2010","unstructured":"010)]% DebnathS010, Biplob K. Debnath , Sudipta Sengupta , and Jin Li . 2010 . ChunkStash: Speeding Up Inline Storage Deduplication Using Flash Memory. In USENIX Annual Technical Conference (ATC) , Boston, MA, USA, June 23--25 . https:\/\/www.usenix.org\/conference\/usenix-atc-10\/chunkstash-speeding-inline-storage-deduplication-using-flash-memory 010)]% DebnathS010, Biplob K. Debnath, Sudipta Sengupta, and Jin Li. 2010. ChunkStash: Speeding Up Inline Storage Deduplication Using Flash Memory. In USENIX Annual Technical Conference (ATC), Boston, MA, USA, June 23--25 . https:\/\/www.usenix.org\/conference\/usenix-atc-10\/chunkstash-speeding-inline-storage-deduplication-using-flash-memory"},{"key":"e_1_2_1_15_1","volume-title":"R-Dedup: Content Aware Redundancy Management for SSD-Based RAID Systems. In 43rd International Conference on Parallel Processing (ICPP)","author":"Du Yimo","year":"2014","unstructured":"014)]% DuZX14, Yimo Du , Youtao Zhang , and Nong Xiao . 2014 . R-Dedup: Content Aware Redundancy Management for SSD-Based RAID Systems. In 43rd International Conference on Parallel Processing (ICPP) , Minneapolis, MN, USA, September 9--12. 111--120. https:\/\/doi.org\/10.1109\/ICPP. 2014.20 10.1109\/ICPP.2014.20 014)]% DuZX14, Yimo Du, Youtao Zhang, and Nong Xiao. 2014. R-Dedup: Content Aware Redundancy Management for SSD-Based RAID Systems. In 43rd International Conference on Parallel Processing (ICPP), Minneapolis, MN, USA, September 9--12. 111--120. https:\/\/doi.org\/10.1109\/ICPP.2014.20"},{"key":"e_1_2_1_16_1","volume-title":"Primary Data Deduplication - Large Scale Study and System Design. In 2012 USENIX Annual Technical Conference","author":"El-Shimi Ahmed","year":"2012","unstructured":"imi Ahmed El-Shimi , Ran Kalach , Ankit Kumar , Adi Ottean , Jin Li , and Sudipta Sengupta . 2012 . Primary Data Deduplication - Large Scale Study and System Design. In 2012 USENIX Annual Technical Conference , Boston, MA, USA, June 13--15 , 2012. USENIX Association, 285--296. https:\/\/www.usenix.org\/conference\/atc12\/technical-sessions\/presentation\/el-shimi imi et almbox.(2012)]% El-ShimiKKO0S12, Ahmed El-Shimi, Ran Kalach, Ankit Kumar, Adi Ottean, Jin Li, and Sudipta Sengupta. 2012. Primary Data Deduplication - Large Scale Study and System Design. In 2012 USENIX Annual Technical Conference, Boston, MA, USA, June 13--15, 2012. USENIX Association, 285--296. https:\/\/www.usenix.org\/conference\/atc12\/technical-sessions\/presentation\/el-shimi"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/FCCM.2015.46"},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2018.2808496"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/IISWC.2016.7581280"},{"key":"e_1_2_1_20_1","volume-title":"USENIX Annual Technical Conference (ATC)","author":"Guo Fanglu","year":"2011","unstructured":", Fanglu Guo and Petros Efstathopoulos . 2011 . Building a High-performance Deduplication System . In USENIX Annual Technical Conference (ATC) , Portland, OR, USA, June 15--17 . https:\/\/www.usenix.org\/conference\/usenixatc11\/building-high-performance-deduplication-system , Fanglu Guo and Petros Efstathopoulos. 2011. Building a High-performance Deduplication System. In USENIX Annual Technical Conference (ATC), Portland, OR, USA, June 15--17 . https:\/\/www.usenix.org\/conference\/usenixatc11\/building-high-performance-deduplication-system"},{"key":"e_1_2_1_21_1","volume-title":"an ultra fast hash algorithm for C# \/ .NET . https:\/\/blog.teamleadnet.com\/2012\/08\/murmurhash3-ultra-fast-hash-algorithm.html Retrieved","author":"Horvath Adam","year":"2022","unstructured":"Adam Horvath . 2021. MurMurHash3 , an ultra fast hash algorithm for C# \/ .NET . https:\/\/blog.teamleadnet.com\/2012\/08\/murmurhash3-ultra-fast-hash-algorithm.html Retrieved January 25, 2022 from , Adam Horvath. 2021. MurMurHash3, an ultra fast hash algorithm for C# \/ .NET . https:\/\/blog.teamleadnet.com\/2012\/08\/murmurhash3-ultra-fast-hash-algorithm.html Retrieved January 25, 2022 from"},{"key":"e_1_2_1_22_1","volume-title":"Flash Memory Summit, 2018","author":"Imershein Louis","year":"2018","unstructured":", Louis Imershein . 2018 . Open Source Data Reduction for High Performance Flash Storage . In Flash Memory Summit, 2018 , Santa Clara, CA, USA. https:\/\/www.flashmemorysummit.com\/English\/Collaterals\/Proceedings\/ 2018\/20180808_SOFT-202--1_Imershein.pdf , Louis Imershein. 2018. Open Source Data Reduction for High Performance Flash Storage. In Flash Memory Summit, 2018, Santa Clara, CA, USA. https:\/\/www.flashmemorysummit.com\/English\/Collaterals\/Proceedings\/2018\/20180808_SOFT-202--1_Imershein.pdf"},{"key":"e_1_2_1_23_1","volume-title":"2022 a","author":"Intel ISAL","year":"2022","unstructured":"022a)]% intel ISAL , Intel . 2022 a . Intel(R) Intelligent Storage Acceleration Library . https:\/\/github.com\/01org\/isa-l Retrieved January 20, 2022 from 022a)]% intelISAL, Intel. 2022 a. Intel(R) Intelligent Storage Acceleration Library . https:\/\/github.com\/01org\/isa-l Retrieved January 20, 2022 from"},{"key":"e_1_2_1_24_1","volume-title":"2022 b. Open Cache Acceleration Software . https:\/\/open-cas.github.io\/ Retrieved","year":"2022","unstructured":"022b)]% opencasdoc, Intel. 2022 b. Open Cache Acceleration Software . https:\/\/open-cas.github.io\/ Retrieved January 20, 2022 from 022b)]% opencasdoc, Intel. 2022 b. Open Cache Acceleration Software . https:\/\/open-cas.github.io\/ Retrieved January 20, 2022 from"},{"key":"e_1_2_1_25_1","volume-title":"2022 c. Open CAS Linux . https:\/\/https:\/\/github.com\/Open-CAS\/open-cas-linux\/ Retrieved","year":"2022","unstructured":"022c)]% opencasgit, Intel. 2022 c. Open CAS Linux . https:\/\/https:\/\/github.com\/Open-CAS\/open-cas-linux\/ Retrieved January 20, 2022 from 022c)]% opencasgit, Intel. 2022 c. Open CAS Linux . https:\/\/https:\/\/github.com\/Open-CAS\/open-cas-linux\/ Retrieved January 20, 2022 from"},{"key":"e_1_2_1_26_1","volume-title":"Raj Jain and Shawn Routhier","year":"1986","unstructured":", Raj Jain and Shawn Routhier . 1986 . Packet trains--measurements and a new model for computer network traffic. IEEE journal on selected areas in Communications , Vol. 4 , 6 (1986), 986--995. , Raj Jain and Shawn Routhier. 1986. Packet trains--measurements and a new model for computer network traffic. IEEE journal on selected areas in Communications , Vol. 4, 6 (1986), 986--995."},{"key":"e_1_2_1_27_1","volume-title":"FusionRAID: Achieving Consistent Low Latency for Commodity SSD Arrays. In 19th USENIX Conference on File and Storage Technologies, FAST February 23--25","author":"Jiang Tianyang","year":"2021","unstructured":"021)]% JiangZHMWLZ21, Tianyang Jiang , Guangyan Zhang , Zican Huang , Xiaosong Ma , Junyu Wei , Zhiyue Li , and Weimin Zheng . 2021 . FusionRAID: Achieving Consistent Low Latency for Commodity SSD Arrays. In 19th USENIX Conference on File and Storage Technologies, FAST February 23--25 , 2021 . USENIX Association, 355--370. https:\/\/www.usenix.org\/conference\/fast21\/presentation\/jiang 021)]% JiangZHMWLZ21, Tianyang Jiang, Guangyan Zhang, Zican Huang, Xiaosong Ma, Junyu Wei, Zhiyue Li, and Weimin Zheng. 2021. FusionRAID: Achieving Consistent Low Latency for Commodity SSD Arrays. In 19th USENIX Conference on File and Storage Technologies, FAST February 23--25, 2021 . USENIX Association, 355--370. https:\/\/www.usenix.org\/conference\/fast21\/presentation\/jiang"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/2741948.2741952"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/MSST.2016.7897082"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/SIMUL.2009.17"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2019.2962691"},{"key":"e_1_2_1_32_1","volume-title":"What SMART Hard Disk Errors Actually Tell Us . https:\/\/www.backblaze.com\/blog\/what-smart-stats-indicate-hard-drive-failures Retrieved","author":"Klein Andy","year":"2022","unstructured":", Andy Klein . 2022. What SMART Hard Disk Errors Actually Tell Us . https:\/\/www.backblaze.com\/blog\/what-smart-stats-indicate-hard-drive-failures Retrieved January 20, 2022 from , Andy Klein. 2022. What SMART Hard Disk Errors Actually Tell Us . https:\/\/www.backblaze.com\/blog\/what-smart-stats-indicate-hard-drive-failures Retrieved January 20, 2022 from"},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/1837915.1837921"},{"key":"e_1_2_1_34_1","volume-title":"Quantitative system performance: computer system analysis using queueing network models","author":"Lazowska Edward D","unstructured":"984)]% lazowska1984quantitative, Edward D Lazowska , John Zahorjan , G Scott Graham , and Kenneth C Sevcik . 1984. Quantitative system performance: computer system analysis using queueing network models . Prentice-Hall, Inc. 984)]% lazowska1984quantitative, Edward D Lazowska, John Zahorjan, G Scott Graham, and Kenneth C Sevcik. 1984. Quantitative system performance: computer system analysis using queueing network models .Prentice-Hall, Inc."},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/FCCM.2015.43"},{"key":"e_1_2_1_36_1","volume-title":"SOSP ACM SIGOPS 28th Symposium on Operating Systems Principles, Virtual Event \/ Koblenz, Germany, October 26--29 . 263--279","author":"Li Huaicheng","unstructured":"021)]% LiPSLGG21, Huaicheng Li , Martin L. Putra , Ronald Shi , Xing Lin , Gregory R. Ganger , and Haryadi S. Gunawi . 2021. lODA: A Host\/Device Co-Design for Strong Predictability Contract on Modern Flash Storage . In SOSP ACM SIGOPS 28th Symposium on Operating Systems Principles, Virtual Event \/ Koblenz, Germany, October 26--29 . 263--279 . https:\/\/doi.org\/10.1145\/3477132.3483573 10.1145\/3477132.3483573 021)]% LiPSLGG21, Huaicheng Li, Martin L. Putra, Ronald Shi, Xing Lin, Gregory R. Ganger, and Haryadi S. Gunawi. 2021. lODA: A Host\/Device Co-Design for Strong Predictability Contract on Modern Flash Storage. In SOSP ACM SIGOPS 28th Symposium on Operating Systems Principles, Virtual Event \/ Koblenz, Germany, October 26--29 . 263--279. https:\/\/doi.org\/10.1145\/3477132.3483573"},{"key":"e_1_2_1_37_1","volume-title":"14th USENIX Conference on File and Storage Technologies (FAST 16)","author":"Li Wenji","year":"2016","unstructured":"016)]% li2016cachededup, Wenji Li , Gregory Jean-Baptise , Juan Riveros , Giri Narasimhan , Tony Zhang , and Ming Zhao . 2016 . CacheDedup: In-line deduplication for flash caching . In 14th USENIX Conference on File and Storage Technologies (FAST 16) . 301--314. 016)]% li2016cachededup, Wenji Li, Gregory Jean-Baptise, Juan Riveros, Giri Narasimhan, Tony Zhang, and Ming Zhao. 2016. CacheDedup: In-line deduplication for flash caching. In 14th USENIX Conference on File and Storage Technologies (FAST 16). 301--314."},{"key":"e_1_2_1_38_1","volume-title":"Inline Deduplication Using Sampling and Locality. In 7th USENIX Conference on File and Storage Technologies, February 24--27, 2009, San Francisco, CA, USA. Proceedings. 111--123","author":"Lillibridge Mark","year":"2009","unstructured":"009)]% LillibridgeEBDTC09, Mark Lillibridge , Kave Eshghi , Deepavali Bhagwat , Vinay Deolalikar , Greg Trezis , and Peter Camble . 2009 . Sparse Indexing: Large Scale , Inline Deduplication Using Sampling and Locality. In 7th USENIX Conference on File and Storage Technologies, February 24--27, 2009, San Francisco, CA, USA. Proceedings. 111--123 . http:\/\/www.usenix.org\/events\/fast09\/tech\/full_papers\/lillibridge\/lillibridge.pdf 009)]% LillibridgeEBDTC09, Mark Lillibridge, Kave Eshghi, Deepavali Bhagwat, Vinay Deolalikar, Greg Trezis, and Peter Camble. 2009. Sparse Indexing: Large Scale, Inline Deduplication Using Sampling and Locality. In 7th USENIX Conference on File and Storage Technologies, February 24--27, 2009, San Francisco, CA, USA. Proceedings. 111--123. http:\/\/www.usenix.org\/events\/fast09\/tech\/full_papers\/lillibridge\/lillibridge.pdf"},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11390-018-1808-5"},{"key":"e_1_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/3307650.3322206"},{"key":"e_1_2_1_41_1","volume-title":"Reference-Counter Aware Deduplication in Erasure-Coded Distributed Storage System. In IEEE International Conference on Networking, Architecture and Storage (NAS)","author":"Liu Tong","year":"2018","unstructured":"018b)]% LiuHAW18, Tong Liu , Xubin He , Shakeel Alibhai , and Chentao Wu . 2018 b. Reference-Counter Aware Deduplication in Erasure-Coded Distributed Storage System. In IEEE International Conference on Networking, Architecture and Storage (NAS) , Chongqing, China, October 11--14 . 1--10. https:\/\/doi.org\/10.1109\/NAS. 2018.8515697 10.1109\/NAS.2018.8515697 018b)]% LiuHAW18, Tong Liu, Xubin He, Shakeel Alibhai, and Chentao Wu. 2018b. Reference-Counter Aware Deduplication in Erasure-Coded Distributed Storage System. In IEEE International Conference on Networking, Architecture and Storage (NAS), Chongqing, China, October 11--14 . 1--10. https:\/\/doi.org\/10.1109\/NAS.2018.8515697"},{"key":"e_1_2_1_42_1","volume-title":"LSI MegaRAID Advanced Software Evaluation Guide V3.0. https:\/\/docs.broadcom.com\/doc\/12350183 Retrieved","author":"LSI Corporation","year":"2022","unstructured":"011)]% LSI11, LSI Corporation . 2011. LSI MegaRAID Advanced Software Evaluation Guide V3.0. https:\/\/docs.broadcom.com\/doc\/12350183 Retrieved January 20, 2022 from 011)]% LSI11, LSI Corporation. 2011. LSI MegaRAID Advanced Software Evaluation Guide V3.0. https:\/\/docs.broadcom.com\/doc\/12350183 Retrieved January 20, 2022 from"},{"key":"e_1_2_1_43_1","volume-title":"Modeling workload IO size mixes with Oracle's Vdbench tool . https:\/\/blog.purestorage.com\/purely-technical\/modeling-io-size-mixes-with-vdbench\/ Retrieved","author":"Lydiksen Lou","year":"2022","unstructured":", Lou Lydiksen . 2015. Modeling workload IO size mixes with Oracle's Vdbench tool . https:\/\/blog.purestorage.com\/purely-technical\/modeling-io-size-mixes-with-vdbench\/ Retrieved January 20, 2022 from , Lou Lydiksen. 2015. Modeling workload IO size mixes with Oracle's Vdbench tool . https:\/\/blog.purestorage.com\/purely-technical\/modeling-io-size-mixes-with-vdbench\/ Retrieved January 20, 2022 from"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2014.84"},{"key":"e_1_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/MSST.2010.5496992"},{"key":"e_1_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/PCCC.2018.8710792"},{"key":"e_1_2_1_47_1","unstructured":"019)]% Oracle19 Oracle. 2019. Architectural Overview of the Oracle ZFS Storage Appliance . https:\/\/www.oracle.com\/technetwork\/server-storage\/sun-unified-storage\/documentation\/o14-001-architecture-overview-zfsa-2099942.pdf Retrieved January 20 2022 from  019)]% Oracle19 Oracle. 2019. Architectural Overview of the Oracle ZFS Storage Appliance . https:\/\/www.oracle.com\/technetwork\/server-storage\/sun-unified-storage\/documentation\/o14-001-architecture-overview-zfsa-2099942.pdf Retrieved January 20 2022 from"},{"key":"e_1_2_1_48_1","volume-title":"1st International Qemu Users' Forum. Citeseer, 29--30","author":"Patel Avadh","year":"2011","unstructured":"011)]% patel2011marss, Avadh Patel , Furat Afram , and Kanad Ghose . 2011 . Marss-x86: A qemu-based micro-architectural and systems simulator for x86 multicore processors. In 1st International Qemu Users' Forum. Citeseer, 29--30 . 011)]% patel2011marss, Avadh Patel, Furat Afram, and Kanad Ghose. 2011. Marss-x86: A qemu-based micro-architectural and systems simulator for x86 multicore processors. In 1st International Qemu Users' Forum. Citeseer, 29--30."},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1145\/50202.50214"},{"key":"e_1_2_1_50_1","volume-title":"dm-vdo . https:\/\/github.com\/dm-vdo\/ Retrieved","year":"2022","unstructured":"021)]% redHat21git, RedHat. 2021. dm-vdo . https:\/\/github.com\/dm-vdo\/ Retrieved January 20, 2022 from 021)]% redHat21git, RedHat. 2021. dm-vdo . https:\/\/github.com\/dm-vdo\/ Retrieved January 20, 2022 from"},{"key":"e_1_2_1_51_1","unstructured":"021)]% redhat21doc RedHat Inc. 2021. RedHat Enterprise Linux 8 Deduplicating and Compressing Storage: Using VDO to Optimize Storage Capacity in RHEL 8 . https:\/\/access.redhat.com\/documentation\/en-us\/red_hat_enterprise_linux\/8\/pdf\/deduplicating_and_compressing_storage\/red_hat_enterprise_linux-8-deduplicating_and_compressing_storage-en-us.pdf Retrieved January 20 2022 from  021)]% redhat21doc RedHat Inc. 2021. RedHat Enterprise Linux 8 Deduplicating and Compressing Storage: Using VDO to Optimize Storage Capacity in RHEL 8 . https:\/\/access.redhat.com\/documentation\/en-us\/red_hat_enterprise_linux\/8\/pdf\/deduplicating_and_compressing_storage\/red_hat_enterprise_linux-8-deduplicating_and_compressing_storage-en-us.pdf Retrieved January 20 2022 from"},{"key":"e_1_2_1_52_1","first-page":"34","article-title":"omplete, Mendel Rosenblum, Stephen A Herrod, Emmett Witchel, and Anoop Gupta. 1995. Complete computer system simulation: The SimOS approach","volume":"3","year":"1995","unstructured":"995)]% rosenblum 1995 c omplete, Mendel Rosenblum, Stephen A Herrod, Emmett Witchel, and Anoop Gupta. 1995. Complete computer system simulation: The SimOS approach . IEEE Parallel & Distributed Technology: Systems & Applications , Vol. 3 , 4 (1995), 34 -- 43 . 995)]% rosenblum1995complete, Mendel Rosenblum, Stephen A Herrod, Emmett Witchel, and Anoop Gupta. 1995. Complete computer system simulation: The SimOS approach. IEEE Parallel & Distributed Technology: Systems & Applications , Vol. 3, 4 (1995), 34--43.","journal-title":"IEEE Parallel & Distributed Technology: Systems & Applications"},{"key":"e_1_2_1_53_1","volume-title":"Modeling the Fault Tolerance Consequences of Deduplication. In 30th IEEE Symposium on Reliable Distributed Systems (SRDS)","author":"Rozier Eric","year":"2011","unstructured":"011)]% RozierSZMUY11, Eric Rozier , William H. Sanders , Pin Zhou , NagaPramod Mandagere , Sandeep Uttamchandani , and Mark L. Yakushev . 2011 . Modeling the Fault Tolerance Consequences of Deduplication. In 30th IEEE Symposium on Reliable Distributed Systems (SRDS) , Madrid, Spain, October 4--7 . 75--84. https:\/\/doi.org\/10.1109\/SRDS. 2011 .18 10.1109\/SRDS.2011.18 011)]% RozierSZMUY11, Eric Rozier, William H. Sanders, Pin Zhou, NagaPramod Mandagere, Sandeep Uttamchandani, and Mark L. Yakushev. 2011. Modeling the Fault Tolerance Consequences of Deduplication. In 30th IEEE Symposium on Reliable Distributed Systems (SRDS), Madrid, Spain, October 4--7 . 75--84. https:\/\/doi.org\/10.1109\/SRDS.2011.18"},{"key":"e_1_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2018.2883745"},{"key":"e_1_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2019.2906597"},{"key":"e_1_2_1_56_1","volume-title":"Communicating With Your SSD: Understanding SMART Attributes . https:\/\/www.samsung.com Retrieved","author":"Samsung SMART","year":"2022","unstructured":"022)]% Samsung SMART , Samsung . 2022. Communicating With Your SSD: Understanding SMART Attributes . https:\/\/www.samsung.com Retrieved January 20, 2022 from 022)]% SamsungSMART, Samsung. 2022. Communicating With Your SSD: Understanding SMART Attributes . https:\/\/www.samsung.com Retrieved January 20, 2022 from"},{"key":"e_1_2_1_57_1","volume-title":"https:\/\/opendedup.org\/odd\/overview\/ Retrieved","author":"Silverberg Sam","year":"2022","unstructured":", Sam Silverberg . 2022. OpenDedup Overview . https:\/\/opendedup.org\/odd\/overview\/ Retrieved January 20, 2022 from , Sam Silverberg. 2022. OpenDedup Overview . https:\/\/opendedup.org\/odd\/overview\/ Retrieved January 20, 2022 from"},{"key":"e_1_2_1_58_1","volume-title":"Proceedings of the 10th USENIX conference on File and Storage Technologies, FAST 2012","author":"Srinivasan Kiran","year":"2012","unstructured":"012)]% SrinivasanBGV12, Kiran Srinivasan , Timothy Bisson , Garth R. Goodson , and Kaladhar Voruganti . 2012 . iDedup: latency-aware, inline data deduplication for primary storage . In Proceedings of the 10th USENIX conference on File and Storage Technologies, FAST 2012 , San Jose, CA, USA, February 14--17 , 2012. USENIX Association, 24. https:\/\/www.usenix.org\/conference\/fast12\/idedup-latency-aware-inline-data-deduplication-primary-storage 012)]% SrinivasanBGV12, Kiran Srinivasan, Timothy Bisson, Garth R. Goodson, and Kaladhar Voruganti. 2012. iDedup: latency-aware, inline data deduplication for primary storage. In Proceedings of the 10th USENIX conference on File and Storage Technologies, FAST 2012, San Jose, CA, USA, February 14--17, 2012. USENIX Association, 24. https:\/\/www.usenix.org\/conference\/fast12\/idedup-latency-aware-inline-data-deduplication-primary-storage"},{"key":"e_1_2_1_59_1","volume-title":"Lossless Data Compresion on FPGAs. In IEEE 19th Annual International Symposium on Field-Programmable Custom Computing Machines, FCCM 2011","author":"Sukhwani Bharat","year":"2011","unstructured":"011)]% SukhwaniABA11, Bharat Sukhwani , B\u00fc lent Abali , Bernard Brezzo , and Sameh W. Asaad . 2011. High-Throughput , Lossless Data Compresion on FPGAs. In IEEE 19th Annual International Symposium on Field-Programmable Custom Computing Machines, FCCM 2011 , Salt Lake City, Utah, USA, 1- -3 May 2011 . IEEE Computer Society, 113--116. https:\/\/doi.org\/10.1109\/FCCM.2011.56 10.1109\/FCCM.2011.56 011)]% SukhwaniABA11, Bharat Sukhwani, B\u00fc lent Abali, Bernard Brezzo, and Sameh W. Asaad. 2011. High-Throughput, Lossless Data Compresion on FPGAs. In IEEE 19th Annual International Symposium on Field-Programmable Custom Computing Machines, FCCM 2011, Salt Lake City, Utah, USA, 1--3 May 2011. IEEE Computer Society, 113--116. https:\/\/doi.org\/10.1109\/FCCM.2011.56"},{"key":"e_1_2_1_60_1","volume-title":"2014 Ottawa Linux Symposium (OLS) .","author":"Tarasov Vasily","year":"2014","unstructured":"014)]% TarasovJKMPSTZ14, Vasily Tarasov , Deepak Jain , Geoff Kuenning , Sonam Mandal , Karthikeyani Palanisami , Philip Shilane , Sagar Trehan , and Erez Zadok . 2014 . Dmdedup: Device mapper target for data deduplication . In 2014 Ottawa Linux Symposium (OLS) . 014)]% TarasovJKMPSTZ14, Vasily Tarasov, Deepak Jain, Geoff Kuenning, Sonam Mandal, Karthikeyani Palanisami, Philip Shilane, Sagar Trehan, and Erez Zadok. 2014. Dmdedup: Device mapper target for data deduplication. In 2014 Ottawa Linux Symposium (OLS) ."},{"key":"e_1_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1145\/2745844.2745856"},{"key":"e_1_2_1_62_1","volume-title":"Quick Generation of SSD Performance Models Using Machine Learning","author":"Tarihi Mojtaba","year":"2022","unstructured":"022)]% tetc\/Tarihi2022, Mojtaba Tarihi , Soheil Azadvar , Arash Tavakkol , Hossein Asadi , and Hamid Sarbazi-Azad . 2022. Quick Generation of SSD Performance Models Using Machine Learning . IEEE Transactions on Emerging Topics in Computing ( 2022 ). https:\/\/doi.org\/10.1109\/TETC.2021.3116197 10.1109\/TETC.2021.3116197 022)]% tetc\/Tarihi2022, Mojtaba Tarihi, Soheil Azadvar, Arash Tavakkol, Hossein Asadi, and Hamid Sarbazi-Azad. 2022. Quick Generation of SSD Performance Models Using Machine Learning. IEEE Transactions on Emerging Topics in Computing (2022). https:\/\/doi.org\/10.1109\/TETC.2021.3116197"},{"key":"e_1_2_1_63_1","volume-title":"The kernel development community. 2022 a. BTRFS Main Page . https:\/\/btrfs.wiki.kernel.org\/index.php\/Main_Page Retrieved","year":"2022","unstructured":"022a)]% kernel22 , The kernel development community. 2022 a. BTRFS Main Page . https:\/\/btrfs.wiki.kernel.org\/index.php\/Main_Page Retrieved January 20, 2022 from 022a)]% kernel22, The kernel development community. 2022 a. BTRFS Main Page . https:\/\/btrfs.wiki.kernel.org\/index.php\/Main_Page Retrieved January 20, 2022 from"},{"key":"e_1_2_1_64_1","volume-title":"The kernel development community. 2022 b. dm-linear . https:\/\/www.kernel.org\/doc\/html\/latest\/admin-guide\/device-mapper\/linear.html Retrieved","year":"2022","unstructured":"022b)]% kernelDML21 , The kernel development community. 2022 b. dm-linear . https:\/\/www.kernel.org\/doc\/html\/latest\/admin-guide\/device-mapper\/linear.html Retrieved January 20, 2022 from 022b)]% kernelDML21, The kernel development community. 2022 b. dm-linear . https:\/\/www.kernel.org\/doc\/html\/latest\/admin-guide\/device-mapper\/linear.html Retrieved January 20, 2022 from"},{"key":"e_1_2_1_65_1","volume-title":"IEEETransactions on Computers","volume":"67","author":"Wang Chundong","year":"2017","unstructured":"017)]% NVDedup, Chundong Wang , Qingsong Wei , Jun Yang , Cheng Chen , Yechao Yang , and Mingdi Xue . 2017 . NV-Dedup: High-performance inline deduplication for non-volatile memory . In IEEETransactions on Computers , Vol. 67 . IEEE, 658--671. 017)]% NVDedup, Chundong Wang, Qingsong Wei, Jun Yang, Cheng Chen, Yechao Yang, and Mingdi Xue. 2017. NV-Dedup: High-performance inline deduplication for non-volatile memory. In IEEETransactions on Computers, Vol. 67. IEEE, 658--671."},{"key":"e_1_2_1_66_1","volume-title":"Austere Flash Caching with Deduplication and Compression. In 2020 USENIX Annual Technical Conference, USENIX ATC 2020","author":"Wang Qiuping","year":"2020","unstructured":"020)]% wangLXKDL20, Qiuping Wang , Jinhong Li , Wen Xia , Erik Kruus , Biplob Debnath , and Patrick P. C. Lee . 2020 . Austere Flash Caching with Deduplication and Compression. In 2020 USENIX Annual Technical Conference, USENIX ATC 2020 , July 15 --17 , 2020 . USENIX Association, 713--726. https:\/\/www.usenix.org\/conference\/atc20\/presentation\/wang-qiuping 020)]% wangLXKDL20, Qiuping Wang, Jinhong Li, Wen Xia, Erik Kruus, Biplob Debnath, and Patrick P. C. Lee. 2020. Austere Flash Caching with Deduplication and Compression. In 2020 USENIX Annual Technical Conference, USENIX ATC 2020, July 15--17, 2020 . USENIX Association, 713--726. https:\/\/www.usenix.org\/conference\/atc20\/presentation\/wang-qiuping"},{"key":"e_1_2_1_67_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2013.6544846"},{"key":"e_1_2_1_68_1","volume-title":"HPDedup: A Hybrid Prioritized Data Deduplication Mechanism for Primary Storage in the Cloud. CoRR","author":"Wu Huijun","year":"2017","unstructured":"017)]% WuWFSZL17, Huijun Wu , Chen Wang , Yinjin Fu , Sherif Sakr , Liming Zhu , and Kai Lu. 2017. HPDedup: A Hybrid Prioritized Data Deduplication Mechanism for Primary Storage in the Cloud. CoRR , Vol. abs\/ 1702 .08153 ( 2017 ). showeprint[arXiv]1702.08153 http:\/\/arxiv.org\/abs\/1702.08153 017)]% WuWFSZL17, Huijun Wu, Chen Wang, Yinjin Fu, Sherif Sakr, Liming Zhu, and Kai Lu. 2017. HPDedup: A Hybrid Prioritized Data Deduplication Mechanism for Primary Storage in the Cloud. CoRR , Vol. abs\/1702.08153 (2017). showeprint[arXiv]1702.08153 http:\/\/arxiv.org\/abs\/1702.08153"},{"key":"e_1_2_1_69_1","volume-title":"IEEE Transactions on Computers, 2021","author":"Suzhen Wu HR","year":"2021","unstructured":"021)]% dedup HR , Suzhen Wu , Chunfeng Du , Weiwei Zhang , Bo Mao , and Hong Jiang . 2021 . DedupHR: Exploiting Content Locality to Alleviate Read\/Write Interference in Deduplication-based Flash Storage . In IEEE Transactions on Computers, 2021 . IEEE Computer Society. https:\/\/doi.org\/10.1109\/TC. 2021.3084116 10.1109\/TC.2021.3084116 021)]% dedupHR, Suzhen Wu, Chunfeng Du, Weiwei Zhang, Bo Mao, and Hong Jiang. 2021. DedupHR: Exploiting Content Locality to Alleviate Read\/Write Interference in Deduplication-based Flash Storage. In IEEE Transactions on Computers, 2021 . IEEE Computer Society. https:\/\/doi.org\/10.1109\/TC.2021.3084116"},{"key":"e_1_2_1_70_1","volume-title":"2011 USENIX Annual Technical Conference","author":"Xia Wen","year":"2011","unstructured":"011)]% xiaJFH11, Wen Xia , Hong Jiang , Dan Feng , and Yu Hua . 2011 . SiLo: A Similarity-Locality based Near-Exact Deduplication Scheme with Low RAM Overhead and High Throughput . In 2011 USENIX Annual Technical Conference , Portland, OR, USA, June 15--17 , 2011. USENIX Association. https:\/\/www.usenix.org\/conference\/usenixatc11\/silo-similarity-locality-based-near-exact-deduplication-scheme-low-ram 011)]% xiaJFH11, Wen Xia, Hong Jiang, Dan Feng, and Yu Hua. 2011. SiLo: A Similarity-Locality based Near-Exact Deduplication Scheme with Low RAM Overhead and High Throughput. In 2011 USENIX Annual Technical Conference, Portland, OR, USA, June 15--17, 2011. USENIX Association. https:\/\/www.usenix.org\/conference\/usenixatc11\/silo-similarity-locality-based-near-exact-deduplication-scheme-low-ram"},{"key":"e_1_2_1_71_1","volume-title":"P-Dedupe: Exploiting Parallelism in Data Deduplication System. In 7th IEEE International Conference on Networking, Architecture, and Storage (NAS)","author":"Xia Wen","year":"2012","unstructured":"012)]% xia2012p, Wen Xia , Hong Jiang , Dan Feng , Lei Tian , Min Fu , and Zhongtao Wang . 2012 . P-Dedupe: Exploiting Parallelism in Data Deduplication System. In 7th IEEE International Conference on Networking, Architecture, and Storage (NAS) , Xiamen, China, June 28--30. 338--347. https:\/\/doi.org\/10.1109\/NAS. 2012.46 10.1109\/NAS.2012.46 012)]% xia2012p, Wen Xia, Hong Jiang, Dan Feng, Lei Tian, Min Fu, and Zhongtao Wang. 2012. P-Dedupe: Exploiting Parallelism in Data Deduplication System. In 7th IEEE International Conference on Networking, Architecture, and Storage (NAS), Xiamen, China, June 28--30. 338--347. https:\/\/doi.org\/10.1109\/NAS.2012.46"},{"key":"e_1_2_1_72_1","doi-asserted-by":"publisher","DOI":"10.1109\/MSST.2019.00009"},{"key":"e_1_2_1_73_1","volume-title":"DupHunter: Flexible High-Performance Deduplication for Docker Registries. In USENIX Annual Technical Conference, USENIX ATC, July 15--17 . 769--783","author":"Zhao Nannan","year":"2020","unstructured":"020)]% ZhaoAACTSRAB20, Nannan Zhao , Hadeel Albahar , Subil Abraham , Keren Chen , Vasily Tarasov , Dimitrios Skourtis , Lukas Rupprecht , Ali Anwar , and Ali Raza Butt . 2020 . DupHunter: Flexible High-Performance Deduplication for Docker Registries. In USENIX Annual Technical Conference, USENIX ATC, July 15--17 . 769--783 . https:\/\/www.usenix.org\/conference\/atc20\/presentation\/zhao 020)]% ZhaoAACTSRAB20, Nannan Zhao, Hadeel Albahar, Subil Abraham, Keren Chen, Vasily Tarasov, Dimitrios Skourtis, Lukas Rupprecht, Ali Anwar, and Ali Raza Butt. 2020. DupHunter: Flexible High-Performance Deduplication for Docker Registries. In USENIX Annual Technical Conference, USENIX ATC, July 15--17 . 769--783. https:\/\/www.usenix.org\/conference\/atc20\/presentation\/zhao"},{"key":"e_1_2_1_74_1","volume-title":"Avoiding the Disk Bottleneck in the Data Domain Deduplication File System. In 6th USENIX Conference on File and Storage Technologies (FAST)","author":"Zhu Benjamin","unstructured":"008)]% ZhuLP08, Benjamin Zhu , Kai Li , and R. Hugo Patterson . 2008 . Avoiding the Disk Bottleneck in the Data Domain Deduplication File System. In 6th USENIX Conference on File and Storage Technologies (FAST) , San Jose, CA, USA, February 26--29. 269--282. http:\/\/www.usenix.org\/events\/fast08\/tech\/zhu.html 008)]% ZhuLP08, Benjamin Zhu, Kai Li, and R. Hugo Patterson. 2008. Avoiding the Disk Bottleneck in the Data Domain Deduplication File System. In 6th USENIX Conference on File and Storage Technologies (FAST), San Jose, CA, USA, February 26--29. 269--282. http:\/\/www.usenix.org\/events\/fast08\/tech\/zhu.html"},{"key":"e_1_2_1_75_1","volume-title":"19th USENIX Conference on File and Storage Technologies, FAST February 23--25 . USENIX Association, 171--185","author":"Zou Xiangyu","year":"2021","unstructured":"021)]% ZouYSX0W21, Xiangyu Zou , Jingsong Yuan , Philip Shilane , Wen Xia , Haijun Zhang , and Xuan Wang . 2021 . The Dilemma between Deduplication and Locality: Can Both be Achieved? . In 19th USENIX Conference on File and Storage Technologies, FAST February 23--25 . USENIX Association, 171--185 . https:\/\/www.usenix.org\/conference\/fast21\/presentation\/zou 021)]% ZouYSX0W21, Xiangyu Zou, Jingsong Yuan, Philip Shilane, Wen Xia, Haijun Zhang, and Xuan Wang. 2021. The Dilemma between Deduplication and Locality: Can Both be Achieved?. In 19th USENIX Conference on File and Storage Technologies, FAST February 23--25 . USENIX Association, 171--185. https:\/\/www.usenix.org\/conference\/fast21\/presentation\/zou"},{"key":"e_1_2_1_76_1","doi-asserted-by":"publisher","DOI":"10.1109\/MICRO.2018.00043"}],"container-title":["Proceedings of the ACM on Measurement and Analysis of Computing Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3530896","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3530896","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T18:09:26Z","timestamp":1750183766000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3530896"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,5,26]]},"references-count":76,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2022,5,26]]}},"alternative-id":["10.1145\/3530896"],"URL":"https:\/\/doi.org\/10.1145\/3530896","relation":{},"ISSN":["2476-1249"],"issn-type":[{"value":"2476-1249","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,5,26]]},"assertion":[{"value":"2022-06-06","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}