{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,19]],"date-time":"2026-06-19T16:00:42Z","timestamp":1781884842658,"version":"3.54.5"},"reference-count":71,"publisher":"Association for Computing Machinery (ACM)","issue":"5","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Des. Autom. Electron. Syst."],"published-print":{"date-parts":[[2026,9,30]]},"abstract":"<jats:p>\n                    Low-rank adaptation (LoRA) is a predominant parameter-efficient finetuning method for adapting large language models (LLMs) to downstream tasks. Meanwhile, Compute-in-Memory (CIM) architectures demonstrate superior energy efficiency due to their array-level parallel in-memory computing designs. In this article, we propose deploying the LoRA-finetuned LLMs on the hybrid CIM architecture (i.e., pretrained weights onto energy-efficient Resistive Random-Access Memory (RRAM) and LoRA branches onto noise-free Static Random-Access Memory (SRAM)), reducing the energy cost to about 3% compared with the Nvidia A100 GPU. However, the inherent noise of RRAM on the saved weights leads to performance degradation, simultaneously. To address this issue, we design a novel Hardware-aware Low-rank Adaptation (HaLoRA) method. The key insight is to train a LoRA branch that is robust toward such noise and then deploy it on noise-free SRAM, while the extra cost is negligible since the parameters of LoRAs are much fewer than pretrained weights (e.g., 0.15% for LLaMA-3.2 1B model). To improve the robustness towards the noise, we theoretically analyze the gap between the optimization trajectories of the LoRA branch under both ideal and noisy conditions and further design an extra loss to minimize the upper bound of this gap. Therefore, we can enjoy both energy efficiency and accuracy during inference. Experiments finetuning the Qwen and LLaMA series demonstrate the effectiveness of HaLoRA across multiple reasoning tasks, achieving up to\n                    <jats:bold>22.7<\/jats:bold>\n                    improvement in average score while maintaining robustness at various noise types and noise levels.\n                  <\/jats:p>","DOI":"10.1145\/3801559","type":"journal-article","created":{"date-parts":[[2026,3,9]],"date-time":"2026-03-09T21:23:51Z","timestamp":1773091431000},"page":"1-23","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":4,"title":["Hardware-aware Low-Rank Adaptation for Large Language Models Based on Hybrid Compute-in-Memory Architecture"],"prefix":"10.1145","volume":"31","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-3664-3513","authenticated-orcid":false,"given":"Taiqiang","family":"Wu","sequence":"first","affiliation":[{"name":"EEE, The University of Hong Kong","place":["Hong Kong, Hong Kong"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-2026-1449","authenticated-orcid":false,"given":"Chenchen","family":"Ding","sequence":"additional","affiliation":[{"name":"EEE, The University of Hong Kong","place":["Hong Kong, Hong Kong"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0008-0427-7935","authenticated-orcid":false,"given":"Wenyong","family":"Zhou","sequence":"additional","affiliation":[{"name":"EEE, The University of Hong Kong","place":["Hong Kong, Hong Kong"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3826-9844","authenticated-orcid":false,"given":"Yuxin","family":"Cheng","sequence":"additional","affiliation":[{"name":"EEE, The University of Hong Kong","place":["Hong Kong, Hong Kong"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-4604-9028","authenticated-orcid":false,"given":"Xincheng","family":"Feng","sequence":"additional","affiliation":[{"name":"EEE, The University of Hong Kong","place":["Hong Kong, Hong Kong"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1281-0484","authenticated-orcid":false,"given":"Shuqi","family":"Wang","sequence":"additional","affiliation":[{"name":"EEE, The University of Hong Kong","place":["Hong Kong, Hong Kong"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5857-7393","authenticated-orcid":false,"given":"Wendong","family":"Xu","sequence":"additional","affiliation":[{"name":"EEE, The University of Hong Kong","place":["Hong Kong, Hong Kong"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0009-9585-7456","authenticated-orcid":false,"given":"Chufan","family":"Shi","sequence":"additional","affiliation":[{"name":"Tsinghua University","place":["Beijing, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7968-9469","authenticated-orcid":false,"given":"Zhengwu","family":"Liu","sequence":"additional","affiliation":[{"name":"The University of Hong Kong","place":["Hong Kong, Hong Kong"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3026-0108","authenticated-orcid":false,"given":"Ngai","family":"Wong","sequence":"additional","affiliation":[{"name":"The University of Hong Kong","place":["Hong Kong, Hong Kong"]}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,4,11]]},"reference":[{"key":"e_1_3_1_2_2","article-title":"Gpt-4 technical report","author":"Achiam Josh","year":"2023","unstructured":"Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et\u00a0al. 2023. Gpt-4 technical report. arXiv preprint arXiv:2303.08774 (2023).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_3_2","article-title":"The llama 3 herd of models","author":"Dubey Abhimanyu","year":"2024","unstructured":"Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Amy Yang, Angela Fan, et\u00a0al. 2024. The llama 3 herd of models. arXiv preprint arXiv:2407.21783 (2024).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_4_2","article-title":"Qwen technical report","author":"Bai Jinze","year":"2023","unstructured":"Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et\u00a0al. 2023. Qwen technical report. arXiv preprint arXiv:2309.16609 (2023).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_5_2","doi-asserted-by":"crossref","unstructured":"Taiqiang Wu Jiahao Wang Zhe Zhao and Ngai Wong. 2024. Mixture-of-subspaces in low-rank adaptation. (2024) 7880\u20137899.","DOI":"10.18653\/v1\/2024.emnlp-main.450"},{"key":"e_1_3_1_6_2","doi-asserted-by":"publisher","DOI":"10.21203\/rs.3.rs-4240043\/v1"},{"key":"e_1_3_1_7_2","first-page":"5737","volume-title":"Proceedings of the 31st International Conference on Computational Linguistics","author":"Wu Taiqiang","year":"2025","unstructured":"Taiqiang Wu, Chaofan Tao, Jiahao Wang, Runming Yang, Zhe Zhao, and Ngai Wong. 2025. Rethinking kullback-leibler divergence in knowledge distillation for large language models. In Proceedings of the 31st International Conference on Computational Linguistics. 5737\u20135755."},{"key":"e_1_3_1_8_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPEC58863.2023.10363447"},{"key":"e_1_3_1_9_2","first-page":"2790","volume-title":"International Conference on Machine Learning","author":"Houlsby Neil","year":"2019","unstructured":"Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin De Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly. 2019. Parameter-efficient transfer learning for NLP. In International Conference on Machine Learning. PMLR, 2790\u20132799."},{"key":"e_1_3_1_10_2","article-title":"Prefix-tuning: Optimizing continuous prompts for generation","author":"Li Xiang Lisa","year":"2021","unstructured":"Xiang Lisa Li and Percy Liang. 2021. Prefix-tuning: Optimizing continuous prompts for generation. arXiv preprint arXiv:2101.00190 (2021).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_11_2","article-title":"Bitfit: Simple parameter-efficient fine-tuning for transformer-based masked language-models","author":"Zaken Elad Ben","year":"2021","unstructured":"Elad Ben Zaken, Shauli Ravfogel, and Yoav Goldberg. 2021. Bitfit: Simple parameter-efficient fine-tuning for transformer-based masked language-models. arXiv preprint arXiv:2106.10199 (2021).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_12_2","volume-title":"International Conference on Learning Representations","author":"Hu Edward J.","year":"2022","unstructured":"Edward J. Hu, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, Weizhu Chen, et\u00a0al. 2022. LoRA: Low-rank adaptation of large language models. In International Conference on Learning Representations."},{"key":"e_1_3_1_13_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41586-020-1942-4"},{"key":"e_1_3_1_14_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC42613.2021.9365766"},{"key":"e_1_3_1_15_2","doi-asserted-by":"publisher","DOI":"10.1126\/science.abj9979"},{"key":"e_1_3_1_16_2","doi-asserted-by":"publisher","DOI":"10.1109\/TCSI.2021.3064189"},{"key":"e_1_3_1_17_2","doi-asserted-by":"publisher","DOI":"10.1109\/TCAD.2022.3197516"},{"key":"e_1_3_1_18_2","doi-asserted-by":"publisher","DOI":"10.1126\/science.adf5538"},{"key":"e_1_3_1_19_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41586-025-08639-2"},{"key":"e_1_3_1_20_2","doi-asserted-by":"publisher","DOI":"10.1109\/TCAD.2025.3531255"},{"key":"e_1_3_1_21_2","doi-asserted-by":"publisher","DOI":"10.1109\/TVLSI.2023.3282046"},{"key":"e_1_3_1_22_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41928-018-0092-2"},{"key":"e_1_3_1_23_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41565-020-0655-z"},{"key":"e_1_3_1_24_2","unstructured":"Qingru Zhang Minshuo Chen Alexander Bukharin Nikos Karampatziakis Pengcheng He Yu Cheng Weizhu Chen and Tuo Zhao. 2023. AdaLoRA: Adaptive Budget Allocation for Parameter-Efficient Fine-Tuning. (2023). arxiv:cs.CL\/2303.10512https:\/\/arxiv.org\/abs\/2303.10512"},{"key":"e_1_3_1_25_2","article-title":"Dylora: Parameter efficient tuning of pre-trained models using dynamic search-free low-rank adaptation","author":"Valipour Mojtaba","year":"2022","unstructured":"Mojtaba Valipour, Mehdi Rezagholizadeh, Ivan Kobyzev, and Ali Ghodsi. 2022. Dylora: Parameter efficient tuning of pre-trained models using dynamic search-free low-rank adaptation. arXiv preprint arXiv:2210.07558 (2022).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_26_2","article-title":"Autolora: Automatically tuning matrix ranks in low-rank adaptation based on meta learning","author":"Zhang Ruiyi","year":"2024","unstructured":"Ruiyi Zhang, Rushi Qiang, Sai Ashish Somayajula, and Pengtao Xie. 2024. Autolora: Automatically tuning matrix ranks in low-rank adaptation based on meta learning. arXiv preprint arXiv:2403.09113 (2024).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_27_2","article-title":"Dora: Enhancing parameter-efficient fine-tuning with dynamic rank distribution","author":"Mao Yulong","year":"2024","unstructured":"Yulong Mao, Kaiyu Huang, Changhao Guan, Ganglin Bao, Fengran Mo, and Jinan Xu. 2024. Dora: Enhancing parameter-efficient fine-tuning with dynamic rank distribution. arXiv preprint arXiv:2405.17357 (2024).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_28_2","first-page":"17783","volume-title":"Proceedings of the 41st International Conference on Machine Learning (Proceedings of Machine Learning Research)","volume":"235","author":"Hayou Soufiane","year":"2024","unstructured":"Soufiane Hayou, Nikhil Ghosh, and Bin Yu. 2024. LoRA+: Efficient low rank adaptation of large models. In Proceedings of the 41st International Conference on Machine Learning (Proceedings of Machine Learning Research). Ruslan Salakhutdinov, Zico Kolter, Katherine Heller, Adrian Weller, Nuria Oliver, Jonathan Scarlett, and Felix Berkenkamp (Eds.), Vol. 235. PMLR, 17783\u201317806. Retrieved from https:\/\/proceedings.mlr.press\/v235\/hayou24a.html"},{"key":"e_1_3_1_29_2","doi-asserted-by":"publisher","DOI":"10.52202\/079017-3846"},{"key":"e_1_3_1_30_2","article-title":"Milora: Harnessing minor singular components for parameter-efficient llm finetuning","author":"Wang Hanqing","year":"2024","unstructured":"Hanqing Wang, Yixia Li, Shuo Wang, Guanhua Chen, and Yun Chen. 2024. Milora: Harnessing minor singular components for parameter-efficient llm finetuning. arXiv preprint arXiv:2406.09044 (2024).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_31_2","first-page":"54905","article-title":"Lora-ga: Low-rank adaptation with gradient approximation","volume":"37","author":"Wang Shaowen","year":"2024","unstructured":"Shaowen Wang, Linxi Yu, and Jian Li. 2024. Lora-ga: Low-rank adaptation with gradient approximation. Advances in Neural Information Processing Systems 37 (2024), 54905\u201354931.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_1_32_2","doi-asserted-by":"publisher","DOI":"10.1126\/science.adf5538"},{"key":"e_1_3_1_33_2","first-page":"1","volume-title":"2024 IEEE Symposium on VLSI Technology and Circuits (VLSI Technology and Circuits)","author":"Prabhu Kartik","year":"2024","unstructured":"Kartik Prabhu, Robert M. Radway, Y. Jeffrey, Kai Bartolone, Massimo Giordano, Fabian Peddinghaus, Yonatan Urman, Win-San Khwa, Yu-Der Chih, Meng-Fan Chang, et\u00a0al. 2024. MINOTAUR: An edge transformer inference and training accelerator with 12 MBytes on-chip resistive RAM and fine-grained spatiotemporal power gating. In 2024 IEEE Symposium on VLSI Technology and Circuits (VLSI Technology and Circuits). IEEE, 1\u20132."},{"key":"e_1_3_1_34_2","doi-asserted-by":"publisher","DOI":"10.1109\/TVLSI.2023.3299509"},{"key":"e_1_3_1_35_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC42614.2022.9731679"},{"key":"e_1_3_1_36_2","doi-asserted-by":"publisher","DOI":"10.1109\/JSSC.2022.3140753"},{"key":"e_1_3_1_37_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC49657.2024.10454500"},{"key":"e_1_3_1_38_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC42615.2023.10067544"},{"key":"e_1_3_1_39_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC42614.2022.9731679"},{"key":"e_1_3_1_40_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC49657.2024.10454468"},{"key":"e_1_3_1_41_2","doi-asserted-by":"publisher","DOI":"10.1109\/IEDM13553.2020.9371982"},{"key":"e_1_3_1_42_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41586-022-04992-8"},{"issue":"1","key":"e_1_3_1_43_2","doi-asserted-by":"crossref","first-page":"2123","DOI":"10.1038\/s41467-025-57183-0","article-title":"A full-stack memristor-based computation-in-memory system with software-hardware co-development","volume":"16","author":"Yu Ruihua","year":"2025","unstructured":"Ruihua Yu, Ze Wang, Qi Liu, Bin Gao, Zhenqi Hao, Tao Guo, Sanchuan Ding, Junyang Zhang, Qi Qin, Dong Wu, et\u00a0al. 2025. A full-stack memristor-based computation-in-memory system with software-hardware co-development. Nature Communications 16, 1 (2025), 2123.","journal-title":"Nature Communications"},{"key":"e_1_3_1_44_2","doi-asserted-by":"publisher","DOI":"10.1109\/DAC18072.2020.9218605"},{"key":"e_1_3_1_45_2","doi-asserted-by":"publisher","DOI":"10.1109\/DAC18074.2021.9586160"},{"key":"e_1_3_1_46_2","doi-asserted-by":"publisher","DOI":"10.1109\/DAC18074.2021.9586115"},{"key":"e_1_3_1_47_2","doi-asserted-by":"publisher","DOI":"10.1109\/DAC56929.2023.10247835"},{"key":"e_1_3_1_48_2","doi-asserted-by":"publisher","DOI":"10.1109\/DAC56929.2023.10247934"},{"key":"e_1_3_1_49_2","article-title":"Attention is all you need","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017. Attention is all you need. Advances in Neural Information Processing Systems 30 (2017), 5998\u20136008.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_1_50_2","article-title":"Weight-inherited distillation for task-agnostic bert compression","author":"Wu Taiqiang","year":"2023","unstructured":"Taiqiang Wu, Cheng Hou, Shanshan Lao, Jiayi Li, Ngai Wong, Zhe Zhao, and Yujiu Yang. 2023. Weight-inherited distillation for task-agnostic bert compression. arXiv preprint arXiv:2305.09098 (2023).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_51_2","first-page":"1","volume-title":"2020 IEEE\/ACM International Conference On Computer Aided Design (ICCAD)","author":"Yang Xiaoxuan","year":"2020","unstructured":"Xiaoxuan Yang, Bonan Yan, Hai Li, and Yiran Chen. 2020. ReTransformer: ReRAM-based processing-in-memory architecture for transformer acceleration. In 2020 IEEE\/ACM International Conference On Computer Aided Design (ICCAD). 1\u20139."},{"key":"e_1_3_1_52_2","doi-asserted-by":"publisher","DOI":"10.1109\/TCSI.2021.3136355"},{"key":"e_1_3_1_53_2","doi-asserted-by":"publisher","DOI":"10.1109\/IEDM19573.2019.8993491"},{"key":"e_1_3_1_54_2","first-page":"277","volume-title":"Proceedings of the 59th ACM\/IEEE Design Automation Conference","author":"Yan Zheyu","year":"2022","unstructured":"Zheyu Yan, Xiaobo Sharon Hu, and Yiyu Shi. 2022. Swim: Selective write-verify for computing-in-memory neural accelerators. In Proceedings of the 59th ACM\/IEEE Design Automation Conference. 277\u2013282."},{"key":"e_1_3_1_55_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41928-024-01213-0"},{"key":"e_1_3_1_56_2","doi-asserted-by":"publisher","DOI":"10.1002\/aisy.202200338"},{"issue":"1","key":"e_1_3_1_57_2","doi-asserted-by":"crossref","first-page":"5897","DOI":"10.1038\/s41467-025-61025-4","article-title":"A near-threshold memristive computing-in-memory engine for edge intelligence","volume":"16","author":"Wang Linfang","year":"2025","unstructured":"Linfang Wang, Weizeng Li, Zhidao Zhou, Junjie An, Wang Ye, Zhi Li, Hanghang Gao, Hongyang Hu, Jing Liu, Xiaoming Chen, et\u00a0al. 2025. A near-threshold memristive computing-in-memory engine for edge intelligence. Nature Communications 16, 1 (2025), 5897.","journal-title":"Nature Communications"},{"key":"e_1_3_1_58_2","first-page":"87","article-title":"Awq: Activation-aware weight quantization for on-device llm compression and acceleration","volume":"6","author":"Lin Ji","year":"2024","unstructured":"Ji Lin, Jiaming Tang, Haotian Tang, Shang Yang, Wei-Ming Chen, Wei-Chen Wang, Guangxuan Xiao, Xingyu Dang, Chuang Gan, and Song Han. 2024. Awq: Activation-aware weight quantization for on-device llm compression and acceleration. Proceedings of Machine Learning and Systems 6 (2024), 87\u2013100.","journal-title":"Proceedings of Machine Learning and Systems"},{"key":"e_1_3_1_59_2","doi-asserted-by":"publisher","DOI":"10.1145\/3649329.3658498"},{"key":"e_1_3_1_60_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11432-023-3785-8"},{"key":"e_1_3_1_61_2","article-title":"Llm-adapters: An adapter family for parameter-efficient fine-tuning of large language models","author":"Hu Zhiqiang","year":"2023","unstructured":"Zhiqiang Hu, Lei Wang, Yihuai Lan, Wanyu Xu, Ee-Peng Lim, Lidong Bing, Xing Xu, Soujanya Poria, and Roy Ka-Wei Lee. 2023. Llm-adapters: An adapter family for parameter-efficient fine-tuning of large language models. arXiv preprint arXiv:2304.01933 (2023).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_62_2","doi-asserted-by":"crossref","first-page":"929","DOI":"10.1109\/IJCNN.2016.7727298","volume-title":"2016 International Joint Conference on Neural Networks (IJCNN)","author":"Agarwal Sapan","year":"2016","unstructured":"Sapan Agarwal, Steven J. Plimpton, David R. Hughart, Alexander H. Hsia, Isaac Richter, Jonathan A. Cox, Conrad D. James, and Matthew J. Marinella. 2016. Resistive memory device requirements for a neural algorithm accelerator. In 2016 International Joint Conference on Neural Networks (IJCNN). IEEE, 929\u2013938."},{"key":"e_1_3_1_63_2","article-title":"Noisytune: A little noise can help you finetune pretrained language models better","author":"Wu Chuhan","year":"2022","unstructured":"Chuhan Wu, Fangzhao Wu, Tao Qi, Yongfeng Huang, and Xing Xie. 2022. Noisytune: A little noise can help you finetune pretrained language models better. arXiv preprint arXiv:2202.12024 (2022).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_64_2","first-page":"1","volume-title":"2024 IEEE Latin American Electron Devices Conference (LAEDC)","author":"Sawal Vedant","year":"2024","unstructured":"Vedant Sawal and Hiu Yung Wong. 2024. Stuck-at faults in ReRAM neuromorphic circuit array and their correction through machine learning. In 2024 IEEE Latin American Electron Devices Conference (LAEDC). IEEE, 1\u20134."},{"issue":"1","key":"e_1_3_1_65_2","first-page":"102","article-title":"Stuck-at fault tolerance in RRAM computing systems","volume":"8","author":"Xia Lixue","year":"2017","unstructured":"Lixue Xia, Wenqin Huangfu, Tianqi Tang, Xiling Yin, Krishnendu Chakrabarty, Yuan Xie, Yu Wang, and Huazhong Yang. 2017. Stuck-at fault tolerance in RRAM computing systems. IEEE Journal on Emerging and Selected Topics in Circuits and Systems 8, 1 (2017), 102\u2013115.","journal-title":"IEEE Journal on Emerging and Selected Topics in Circuits and Systems"},{"key":"e_1_3_1_66_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC42614.2022.9731645"},{"key":"e_1_3_1_67_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC42613.2021.9365766"},{"key":"e_1_3_1_68_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41586-023-05759-5"},{"key":"e_1_3_1_69_2","first-page":"1","volume-title":"Proceedings of the 56th Annual Design Automation Conference 2019","author":"Zhang Wenqiang","year":"2019","unstructured":"Wenqiang Zhang, Xiaochen Peng, Huaqiang Wu, Bin Gao, Hu He, Youhui Zhang, Shimeng Yu, and He Qian. 2019. Design guidelines of RRAM based neural-processing-unit: A joint device-circuit-algorithm analysis. In Proceedings of the 56th Annual Design Automation Conference 2019. 1\u20136."},{"key":"e_1_3_1_70_2","article-title":"Loftq: Lora-fine-tuning-aware quantization for large language models","author":"Li Yixiao","year":"2023","unstructured":"Yixiao Li, Yifan Yu, Chen Liang, Pengcheng He, Nikos Karampatziakis, Weizhu Chen, and Tuo Zhao. 2023. Loftq: Lora-fine-tuning-aware quantization for large language models. arXiv preprint arXiv:2310.08659 (2023).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_71_2","article-title":"Mathqa: Towards interpretable math word problem solving with operation-based formalisms","author":"Amini Aida","year":"2019","unstructured":"Aida Amini, Saadia Gabriel, Peter Lin, Rik Koncel-Kedziorski, Yejin Choi, and Hannaneh Hajishirzi. 2019. Mathqa: Towards interpretable math word problem solving with operation-based formalisms. arXiv preprint arXiv:1905.13319 (2019).","journal-title":"arXiv preprint"},{"key":"e_1_3_1_72_2","article-title":"A survey on large language models for code generation","author":"Jiang Juyong","year":"2024","unstructured":"Juyong Jiang, Fan Wang, Jiasi Shen, Sungju Kim, and Sunghun Kim. 2024. A survey on large language models for code generation. arXiv preprint arXiv:2406.00515 (2024).","journal-title":"arXiv preprint"}],"container-title":["ACM Transactions on Design Automation of Electronic Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3801559","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,11]],"date-time":"2026-04-11T11:27:48Z","timestamp":1775906868000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3801559"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,4,11]]},"references-count":71,"journal-issue":{"issue":"5","published-print":{"date-parts":[[2026,9,30]]}},"alternative-id":["10.1145\/3801559"],"URL":"https:\/\/doi.org\/10.1145\/3801559","relation":{},"ISSN":["1084-4309","1557-7309"],"issn-type":[{"value":"1084-4309","type":"print"},{"value":"1557-7309","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,4,11]]},"assertion":[{"value":"2025-11-10","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-02-10","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-04-11","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}