{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,7]],"date-time":"2025-10-07T21:10:16Z","timestamp":1759871416363,"version":"build-2065373602"},"publisher-location":"New York, NY, USA","reference-count":35,"publisher":"ACM","funder":[{"DOI":"10.13039\/501100005799","name":"National Chiao Tung University","doi-asserted-by":"publisher","id":[{"id":"10.13039\/501100005799","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,7,14]]},"DOI":"10.1145\/3712256.3726482","type":"proceedings-article","created":{"date-parts":[[2025,7,8]],"date-time":"2025-07-08T12:28:18Z","timestamp":1751977698000},"page":"1226-1234","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Integrating Neural Architecture Search and Rematerialization for Efficient On-Device Learning"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0005-3943-8071","authenticated-orcid":false,"given":"Chih-Ling","family":"Chen","sequence":"first","affiliation":[{"name":"National Yang Ming Chiao Tung University, Hsinchu, Taiwan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0000-6931-6538","authenticated-orcid":false,"given":"Kai-Chiang","family":"Wu","sequence":"additional","affiliation":[{"name":"National Yang Ming Chiao Tung University, Hsinchu, Taiwan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4663-9099","authenticated-orcid":false,"given":"Ning-Chi","family":"Huang","sequence":"additional","affiliation":[{"name":"National Yang Ming Chiao Tung University, Hsinchu, Taiwan"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,7,13]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"IoT-enabled healthcare transformation leveraging deep learning for advanced patient monitoring and diagnosis. Multimedia Tools and Applications","author":"Alharbe Nawaf","year":"2024","unstructured":"Nawaf Alharbe and Manal Almalki. 2024. IoT-enabled healthcare transformation leveraging deep learning for advanced patient monitoring and diagnosis. Multimedia Tools and Applications (2024), 1\u201314."},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.naacl-long.345"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2022.acl-short.1"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.imu.2021.100550"},{"volume-title":"Neural Networks: Tricks of the Trade","author":"Bottou L\u00e9on","key":"e_1_3_2_1_5_1","unstructured":"L\u00e9on Bottou. 2012. Stochastic gradient descent tricks. In Neural Networks: Tricks of the Trade: Second Edition. Springer, 421\u2013436."},{"key":"e_1_3_2_1_6_1","volume-title":"International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=rydeCEhs-","author":"Brock Andrew","year":"2018","unstructured":"Andrew Brock, Theo Lim, J.M. Ritchie, and Nick Weston. 2018. SMASH: One-Shot Model Architecture Search through HyperNetworks. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=rydeCEhs-"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.5555\/3495724.3495883"},{"key":"e_1_3_2_1_8_1","volume-title":"Proceedings of the 34th International Conference on Neural Information Processing Systems","author":"Cai Han","year":"2020","unstructured":"Han Cai, Chuang Gan, Ligeng Zhu Massachusetts Institute of Technology, and Song Han Massachusetts Institute of Technology. 2020. TinyTL: reduce memory, not parameters for efficient on-device learning. In Proceedings of the 34th International Conference on Neural Information Processing Systems (Vancouver, BC, Canada) (NIPS '20). Curran Associates Inc., Red Hook, NY, USA, Article 947, 13 pages."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV57701.2024.00247"},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00008"},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01205"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/N19-1423"},{"key":"e_1_3_2_1_14_1","volume-title":"International conference on machine learning. PMLR, 647\u2013655","author":"Donahue Jeff","year":"2014","unstructured":"Jeff Donahue, Yangqing Jia, Oriol Vinyals, Judy Hoffman, Ning Zhang, Eric Tzeng, and Trevor Darrell. 2014. Decaf: A deep convolutional activation feature for generic visual recognition. In International conference on machine learning. PMLR, 647\u2013655."},{"key":"e_1_3_2_1_15_1","unstructured":"Raspberry Pi Foundation. 2019. Raspberry Pi 4. Available: https:\/\/www.raspberrypi.org\/products\/raspberry-pi-4-model-b\/."},{"volume-title":"International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=vYeQQ29Tbvx","author":"Frankle Jonathan","key":"e_1_3_2_1_16_1","unstructured":"Jonathan Frankle, David J. Schwab, and Ari S. Morcos. 2021. Training BatchNorm and Only BatchNorm: On the Expressive Power of Random Features in {CNN}s. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=vYeQQ29Tbvx"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/3498361.3539765"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP40776.2020.9053036"},{"key":"e_1_3_2_1_19_1","volume-title":"Proceedings, Part XVI 16","author":"Guo Zichao","year":"2020","unstructured":"Zichao Guo, Xiangyu Zhang, Haoyuan Mu, Wen Heng, Zechun Liu, Yichen Wei, and Jian Sun. 2020. Single path one-shot neural architecture search with uniform sampling. In Computer Vision-ECCV 2020: 16th European Conference, Glasgow, UK, August 23\u201328, 2020, Proceedings, Part XVI 16. Springer, 544\u2013560."},{"key":"e_1_3_2_1_20_1","volume-title":"LoRA: Low-Rank Adaptation of Large Language Models. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=nZeVKeeFYf9","author":"Hu Edward J","year":"2022","unstructured":"Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022. LoRA: Low-Rank Adaptation of Large Language Models. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=nZeVKeeFYf9"},{"key":"e_1_3_2_1_21_1","first-page":"497","article-title":"Checkmate: Breaking the memory wall with optimal tensor rematerialization","volume":"2","author":"Jain Paras","year":"2020","unstructured":"Paras Jain, Ajay Jain, Aniruddha Nrusimha, Amir Gholami, Pieter Abbeel, Joseph Gonzalez, Kurt Keutzer, and Ion Stoica. 2020. Checkmate: Breaking the memory wall with optimal tensor rematerialization. Proceedings of Machine Learning and Systems 2 (2020), 497\u2013511.","journal-title":"Proceedings of Machine Learning and Systems"},{"key":"e_1_3_2_1_22_1","first-page":"29248","article-title":"Back razor: Memory-efficient transfer learning by self-sparsified backpropagation","volume":"35","author":"Jiang Ziyu","year":"2022","unstructured":"Ziyu Jiang, Xuxi Chen, Xueqin Huang, Xianzhi Du, Denny Zhou, and Zhangyang Wang. 2022. Back razor: Memory-efficient transfer learning by self-sparsified backpropagation. Advances in Neural Information Processing Systems 35 (2022), 29248\u201329261.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_23_1","volume-title":"Kingma and Jimmy Ba","author":"Diederik","year":"2015","unstructured":"Diederik P. Kingma and Jimmy Ba. 2015. Adam: A Method for Stochastic Optimization. In 3rd International Conference on Learning Representations, ICLR 2015, San Diego, CA, USA, May 7\u20139, 2015, Conference Track Proceedings, Yoshua Bengio and Yann LeCun (Eds.). http:\/\/arxiv.org\/abs\/1412.6980"},{"key":"e_1_3_2_1_24_1","unstructured":"Alex Krizhevsky Geoffrey Hinton et al. 2009. Learning multiple layers of features from tiny images. (2009)."},{"key":"e_1_3_2_1_25_1","volume-title":"International Conference on Learning Representations. https:\/\/arxiv.org\/abs\/1810","author":"Mudrakarta Pramod Kaushik","year":"2019","unstructured":"Pramod Kaushik Mudrakarta, Mark Sandler, Andrey Zhmoginov, and Andrew Howard. 2019. K for the price of 1. Parameter efficient multi-task and transfer learning. In International Conference on Learning Representations. https:\/\/arxiv.org\/abs\/1810.10703"},{"key":"e_1_3_2_1_26_1","volume-title":"International Conference on Machine Learning. PMLR, 26363\u201326381","author":"Novikov Georgii Sergeevich","year":"2023","unstructured":"Georgii Sergeevich Novikov, Daniel Bershatsky, Julia Gusak, Alex Shonenkov, Denis Valerievich Dimitrov, and Ivan Oseledets. 2023. Few-bit backward: Quantized gradients of activation functions for memory footprint reduction. In International Conference on Machine Learning. PMLR, 26363\u201326381."},{"key":"e_1_3_2_1_27_1","volume-title":"Mesa: A memory-saving training framework for transformers. arXiv preprint arXiv:2111.11124","author":"Pan Zizheng","year":"2021","unstructured":"Zizheng Pan, Peng Chen, Haoyu He, Jing Liu, Jianfei Cai, and Bohan Zhuang. 2021. Mesa: A memory-saving training framework for transformers. arXiv preprint arXiv:2111.11124 (2021)."},{"key":"e_1_3_2_1_28_1","volume-title":"ZeroFL: Efficient On-Device Training for Federated Learning with Local Sparsity. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=2sDQwC_hmnM","author":"Qiu Xinchi","year":"2022","unstructured":"Xinchi Qiu, Javier Fernandez-Marques, Pedro PB Gusmao, Yan Gao, Titouan Parcollet, and Nicholas Donald Lane. 2022. ZeroFL: Efficient On-Device Training for Federated Learning with Local Sparsity. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=2sDQwC_hmnM"},{"key":"e_1_3_2_1_29_1","volume-title":"2019 5th International Conference on Advances in Electrical Engineering (ICAEE). IEEE, 280\u2013285","author":"Monir Rabby Md Khurram","year":"2019","unstructured":"Md Khurram Monir Rabby, Muhammad Mobaidul Islam, and Salman Monowar Imon. 2019. A review of IoT application in a smart traffic management system. In 2019 5th International Conference on Advances in Electrical Engineering (ICAEE). IEEE, 280\u2013285."},{"key":"e_1_3_2_1_30_1","unstructured":"Alec Radford Jeff Wu Rewon Child David Luan Dario Amodei and Ilya Sutskever. 2019. Language Models are Unsupervised Multitask Learners. (2019)."},{"key":"e_1_3_2_1_31_1","volume-title":"Low-memory neural network training: A technical report. arXiv preprint arXiv:1904.10631","author":"Sohoni Nimit S","year":"2019","unstructured":"Nimit S Sohoni, Christopher R Aberger, Megan Leszczynski, Jian Zhang, and Christopher R\u00e9. 2019. Low-memory neural network training: A technical report. arXiv preprint arXiv:1904.10631 (2019)."},{"key":"e_1_3_2_1_32_1","volume-title":"International conference on machine learning. PMLR, 10347\u201310357","author":"Touvron Hugo","year":"2021","unstructured":"Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Herv\u00e9 J\u00e9gou. 2021. Training data-efficient image transformers & distillation through attention. In International conference on machine learning. PMLR, 10347\u201310357."},{"key":"e_1_3_2_1_33_1","unstructured":"Catherine Wah Steve Branson Peter Welinder Pietro Perona and Serge Belongie. 2011. The caltech-ucsd birds-200-2011 dataset. (2011)."},{"key":"e_1_3_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA57654.2024.00071"},{"key":"e_1_3_2_1_35_1","volume-title":"International Conference on Machine Learning. PMLR, 42018\u201342045","author":"Zhao Xunyi","year":"2023","unstructured":"Xunyi Zhao, Th\u00e9otime Le Hellard, Lionel Eyraud-Dubois, Julia Gusak, and Olivier Beaumont. 2023. Rockmate: an efficient, fast, automatic and generic tool for re-materialization in pytorch. In International Conference on Machine Learning. PMLR, 42018\u201342045."}],"event":{"name":"GECCO '25: Genetic and Evolutionary Computation Conference","sponsor":["SIGEVO ACM Special Interest Group on Genetic and Evolutionary Computation"],"location":"NH Malaga Hotel Malaga Spain","acronym":"GECCO '25"},"container-title":["Proceedings of the Genetic and Evolutionary Computation Conference"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3712256.3726482","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,7]],"date-time":"2025-10-07T20:43:12Z","timestamp":1759869792000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3712256.3726482"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,7,13]]},"references-count":35,"alternative-id":["10.1145\/3712256.3726482","10.1145\/3712256"],"URL":"https:\/\/doi.org\/10.1145\/3712256.3726482","relation":{},"subject":[],"published":{"date-parts":[[2025,7,13]]},"assertion":[{"value":"2025-07-13","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}