{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:16:47Z","timestamp":1750220207435,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":21,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,30]],"date-time":"2022-10-30T00:00:00Z","timestamp":1667088000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,30]]},"DOI":"10.1145\/3508352.3549435","type":"proceedings-article","created":{"date-parts":[[2022,12,22]],"date-time":"2022-12-22T12:10:54Z","timestamp":1671711054000},"page":"1-9","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Seprox"],"prefix":"10.1145","author":[{"given":"Aradhana Mohan","family":"Parvathy","sequence":"first","affiliation":[{"name":"Purdue University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Sarada","family":"Krithivasan","sequence":"additional","affiliation":[{"name":"Purdue University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Sanchari","family":"Sen","sequence":"additional","affiliation":[{"name":"IBM T. J. Watson Research Center"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Anand","family":"Raghunathan","sequence":"additional","affiliation":[{"name":"Purdue University"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,12,22]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_2_1_2_1","first-page":"4510","volume-title":"2018 IEEE Conference on Computer Vision and Pattern Recognition, CVPR","author":"Mark","year":"2018","unstructured":"Mark Sandler et al. 2018. MobileNetV2: Inverted Residuals and Linear Bottlenecks. In 2018 IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2018. pp. 4510--4520."},{"key":"e_1_3_2_1_3_1","first-page":"1","volume-title":"MASR: A Modular Accelerator for Sparse RNNs. In 2019 28th International Conference on Parallel Architectures and Compilation Techniques (PACT).","author":"Gupta Udit","year":"2019","unstructured":"Udit Gupta, Brandon Reagen, Lillian Pentecost, Marco Donato, Thierry Tambe, Alexander M. Rush, Gu-Yeon Wei, and David Brooks. 2019. MASR: A Modular Accelerator for Sparse RNNs. In 2019 28th International Conference on Parallel Architectures and Compilation Techniques (PACT). pp. 1--14."},{"key":"e_1_3_2_1_4_1","volume-title":"Trained Quantization and Huffman Coding. In 4th International Conference on Learning Representations, ICLR","author":"Han Song","year":"2016","unstructured":"Song Han, Huizi Mao, and William J. Dally. 2016. Deep Compression: Compressing Deep Neural Network with Pruning, Trained Quantization and Huffman Coding. In 4th International Conference on Learning Representations, ICLR 2016,."},{"key":"e_1_3_2_1_5_1","first-page":"770","volume-title":"Deep Residual Learning for Image Recognition. In 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"He Kaiming","year":"2016","unstructured":"Kaiming He, X. Zhang, Shaoqing Ren, and Jian Sun. 2016. Deep Residual Learning for Image Recognition. In 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). pp. 770--778."},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.5555\/3495724.3497438"},{"key":"e_1_3_2_1_7_1","unstructured":"Byung Soo Ko. 2020. ImageNet Classification Leaderboard. https:\/\/kobiso.github.io\/Computer-Vision-Leaderboard\/imagenet.html"},{"key":"e_1_3_2_1_9_1","first-page":"1097","volume-title":"Proceedings of the 25th International Conference on Neural Information Processing Systems -","volume":"1","author":"Krizhevsky Alex","unstructured":"Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. 2012. ImageNet Classification with Deep Convolutional Neural Networks. In Proceedings of the 25th International Conference on Neural Information Processing Systems - Volume 1 (NIPS'12). pp. 1097--1105."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.5555\/3524938.3525605"},{"key":"e_1_3_2_1_11_1","first-page":"27","volume-title":"2017 ACM\/IEEE 44th Annual International Symposium on Computer Architecture (ISCA).","author":"Angshuman","unstructured":"Angshuman Parashar et al. 2017. SCNN: An accelerator for compressed-sparse convolutional neural networks. In 2017 ACM\/IEEE 44th Annual International Symposium on Computer Architecture (ISCA). pp. 27--40."},{"key":"e_1_3_2_1_12_1","first-page":"8024","article-title":"PyTorch: An Imperative Style, High-Performance Deep Learning Library","volume":"32","author":"Adam Paszke","year":"2019","unstructured":"Adam Paszke et al. 2019. PyTorch: An Imperative Style, High-Performance Deep Learning Library. In Advances in Neural Information Processing Systems 32. pp. 8024--8035.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_13_1","unstructured":"Jeff Pool Abhishek Sawarkar and Jay Rodge. 2021. Accelerating Inference with Sparsity Using the NVIDIA Ampere Architecture and NVIDIA TensorRT. https:\/\/developer.nvidia.com\/blog\/accelerating-inference-with-sparsity-using-ampere-and-tensorrt\/"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCAD.2004.826558"},{"volume-title":"Iterative methods for sparse linear systems","author":"Saad Y.","key":"e_1_3_2_1_15_1","unstructured":"Y. Saad. 2003. Iterative methods for sparse linear systems (second ed.). SIAM."},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","unstructured":"Ananda Samajdar Yuhao Zhu Paul Whatmough Matthew Mattina and Tushar Krishna. 2018. SCALE-Sim: Systolic CNN Accelerator Simulator. 10.48550\/ARXIV.1811.02883","DOI":"10.48550\/ARXIV.1811.02883"},{"key":"e_1_3_2_1_17_1","first-page":"1","volume-title":"Efficacy of Pruning in Ultra-Low Precision DNNs. In 2021 IEEE\/ACM International Symposium on Low Power Electronics and Design (ISLPED).","author":"Sen Sanchari","year":"2021","unstructured":"Sanchari Sen, Swagath Venkataramani, and Anand Raghunathan. 2021. Efficacy of Pruning in Ultra-Low Precision DNNs. In 2021 IEEE\/ACM International Symposium on Low Power Electronics and Design (ISLPED). pp. 1--6."},{"key":"e_1_3_2_1_18_1","first-page":"5308","article-title":"Robust Quantization: One Model to Rule Them All","volume":"33","author":"Shkolnik Moran","year":"2020","unstructured":"Moran Shkolnik, Brian Chmiel, Ron Banner, Gil Shomron, Yury Nahshan, Alex Bronstein, and Uri Weiser. 2020. Robust Quantization: One Model to Rule Them All. In Advances in Neural Information Processing Systems, Vol. 33. pp. 5308--5317.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_19_1","first-page":"1796","article-title":"Ultra-Low Precision 4-bit Training of Deep Neural Networks","volume":"33","author":"Xiao Sun","year":"2020","unstructured":"Xiao Sun et al. 2020. Ultra-Low Precision 4-bit Training of Deep Neural Networks. In Advances in Neural Information Processing Systems, Vol. 33. pp. 1796--1807.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_20_1","first-page":"382","volume-title":"International Colloquium on Automata, Languages and Programming (ICALP).","author":"van Leeuwen Jan","unstructured":"Jan van Leeuwen. 1976. On the Construction of Huffman Trees. In International Colloquium on Automata, Languages and Programming (ICALP). pp. 382--410."},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2019.2910073"},{"key":"e_1_3_2_1_22_1","volume-title":"6th International Conference on Learning Representations, ICLR","author":"Zhu Michael","year":"2018","unstructured":"Michael Zhu and Suyog Gupta. 2018. To Prune, or Not to Prune: Exploring the Efficacy of Pruning for Model Compression. In 6th International Conference on Learning Representations, ICLR 2018,."}],"event":{"name":"ICCAD '22: IEEE\/ACM International Conference on Computer-Aided Design","sponsor":["SIGDA ACM Special Interest Group on Design Automation","IEEE-EDS Electronic Devices Society","IEEE CAS","IEEE CEDA"],"location":"San Diego California","acronym":"ICCAD '22"},"container-title":["Proceedings of the 41st IEEE\/ACM International Conference on Computer-Aided Design"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3508352.3549435","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3508352.3549435","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:02:57Z","timestamp":1750186977000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3508352.3549435"}},"subtitle":["Sequence-Based Approximations for Compressing Ultra-Low Precision Deep Neural Networks"],"short-title":[],"issued":{"date-parts":[[2022,10,30]]},"references-count":21,"alternative-id":["10.1145\/3508352.3549435","10.1145\/3508352"],"URL":"https:\/\/doi.org\/10.1145\/3508352.3549435","relation":{},"subject":[],"published":{"date-parts":[[2022,10,30]]},"assertion":[{"value":"2022-12-22","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}