{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,18]],"date-time":"2026-07-18T22:31:53Z","timestamp":1784413913903,"version":"3.55.0"},"reference-count":32,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2023,6,13]],"date-time":"2023-06-13T00:00:00Z","timestamp":1686614400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100012166","name":"National Key Research and Development Program of China","doi-asserted-by":"publisher","award":["2022YFE0200500"],"award-info":[{"award-number":["2022YFE0200500"]}],"id":[{"id":"10.13039\/501100012166","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Theme-based project TRS","award":["T41-603\/20R"],"award-info":[{"award-number":["T41-603\/20R"]}]},{"DOI":"10.13039\/501100001809","name":"National Science Foundation of China","doi-asserted-by":"crossref","award":["No. U22B2060"],"award-info":[{"award-number":["No. U22B2060"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100021171","name":"Guangdong Basic and Applied Basic Research Foundation","doi-asserted-by":"crossref","award":["2019B151530001"],"award-info":[{"award-number":["2019B151530001"]}],"id":[{"id":"10.13039\/501100021171","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Hong Kong ITC ITF grants","award":["MHX\/078\/21 and PRP\/004\/22FX"],"award-info":[{"award-number":["MHX\/078\/21 and PRP\/004\/22FX"]}]},{"name":"the Hong Kong RGC GRF Project","award":["16209519"],"award-info":[{"award-number":["16209519"]}]},{"name":"CRF Project","award":["C6030-18G, C2004-21GF"],"award-info":[{"award-number":["C6030-18G, C2004-21GF"]}]},{"name":"Shanghai Municipal Science and Technology Major Project","award":["2021SHZDZX0102"],"award-info":[{"award-number":["2021SHZDZX0102"]}]},{"name":"SJTU Global Strategic Partnership Fund"},{"name":"AOE Project","award":["AoE\/E-603\/18"],"award-info":[{"award-number":["AoE\/E-603\/18"]}]},{"name":"RIF Project","award":["R6020-19"],"award-info":[{"award-number":["R6020-19"]}]},{"name":"Microsoft Research Asia Collaborative Research Grant"},{"name":"HKUST Global Strategic Partnership Fund"},{"name":"China NSFC","award":["No. 61729201"],"award-info":[{"award-number":["No. 61729201"]}]},{"name":"HKUST-Webank joint research lab grant"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Proc. ACM Manag. Data"],"published-print":{"date-parts":[[2023,6,13]]},"abstract":"<jats:p>The acquisition of accurate rainfall distribution in space is an important task in hydrological analysis and natural disaster pre-warning. However, it is impossible to install rain gauges on every corner. Spatial interpolation is a common way to infer rainfall distribution based on available raingauge data. However, the existing works rely on some unrealistic pre-settings to capture spatial correlations, which limits their performance in real scenarios. To tackle this issue, we propose the SSIN, which is a novel data-driven self-supervised learning framework for rainfall spatial interpolation by mining latent spatial patterns from historical observation data. Inspired by the Cloze task and BERT, we fully consider the characteristics of spatial interpolation and design the SpaFormer model based on the Transformer architecture as the core of SSIN. Our main idea is: by constructing rich self-supervision signals via random masking, SpaFormer can learn informative embeddings for raw data and then adaptively model spatial correlations based on rainfall spatial context. Extensive experiments on two real-world raingauge datasets show that our method outperforms the state-of-the-art solutions. In addition, we take traffic spatial interpolation as another use case to further explore the performance of our method, and SpaFormer achieves the best performance on one large real-world traffic dataset, which further confirms the effectiveness and generality of our method.<\/jats:p>","DOI":"10.1145\/3589321","type":"journal-article","created":{"date-parts":[[2023,6,20]],"date-time":"2023-06-20T20:26:45Z","timestamp":1687292805000},"page":"1-21","source":"Crossref","is-referenced-by-count":3,"title":["SSIN: Self-Supervised Learning for Rainfall Spatial Interpolation"],"prefix":"10.1145","volume":"1","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-2051-635X","authenticated-orcid":false,"given":"Jia","family":"Li","sequence":"first","affiliation":[{"name":"The Hong Kong University of Science and Technology, Hong Kong SAR, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8364-3674","authenticated-orcid":false,"given":"Yanyan","family":"Shen","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University, Shanghai, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8257-5806","authenticated-orcid":false,"given":"Lei","family":"Chen","sequence":"additional","affiliation":[{"name":"The Hong Kong University of Science and Technology, Hong Kong SAR, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6693-3151","authenticated-orcid":false,"given":"Charles Wang Wai","family":"Ng","sequence":"additional","affiliation":[{"name":"The Hong Kong University of Science and Technology, Hong Kong SAR, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,6,20]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v34i04.5716"},{"key":"e_1_2_2_2_1","volume-title":"Data2vec: A general framework for self-supervised learning in speech, vision and language. arXiv preprint arXiv:2202.03555","author":"Baevski Alexei","year":"2022","unstructured":"Alexei Baevski, Wei-Ning Hsu, Qiantong Xu, Arun Babu, Jiatao Gu, and Michael Auli. 2022. Data2vec: A general framework for self-supervised learning in speech, vision and language. arXiv preprint arXiv:2202.03555 (2022)."},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10333-012-0319-1"},{"key":"e_1_2_2_4_1","volume-title":"13th USENIX Symposium on Operating Systems Design and Implementation (OSDI 18)","author":"Chen Tianqi","year":"2018","unstructured":"Tianqi Chen, Thierry Moreau, Ziheng Jiang, Lianmin Zheng, Eddie Yan, Haichen Shen, Meghan Cowan, Leyuan Wang, Yuwei Hu, Luis Ceze, et al. 2018. $$TVM$$: An automated $$End-to-End$$ optimizing compiler for deep learning. In 13th USENIX Symposium on Operating Systems Design and Implementation (OSDI 18). 578--594."},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1007\/978--1--4614--8265--9_437"},{"key":"e_1_2_2_6_1","volume-title":"BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In NAACL-HLT. 4171--4186.","author":"Devlin Jacob","year":"2019","unstructured":"Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019. BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In NAACL-HLT. 4171--4186."},{"key":"e_1_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0022-1694(98)00155-3"},{"key":"e_1_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10346-017-0904-x"},{"key":"e_1_2_2_9_1","doi-asserted-by":"crossref","unstructured":"Huifeng Guo Ruiming Tang Yunming Ye Zhenguo Li and Xiuqiang He. 2017. DeepFM: A Factorization-Machine based Neural Network for CTR Prediction. In IJCAI. 1725--1731.","DOI":"10.24963\/ijcai.2017\/239"},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1080\/02693799508902045"},{"key":"e_1_2_2_11_1","doi-asserted-by":"crossref","unstructured":"Sharon A Jewell and Nicolas Gaussiat. 2015. An assessment of kriging-based rain-gauge--radar merging techniques. Q. J. R. Meteorol. Soc. (2015).","DOI":"10.1002\/qj.2522"},{"key":"e_1_2_2_12_1","unstructured":"Guolin Ke Di He and Tie-Yan Liu. 2021. Rethinking Positional Encoding in Language Pre-training. In ICLR."},{"key":"e_1_2_2_13_1","volume-title":"Kingma and Jimmy Ba","author":"Diederik","year":"2015","unstructured":"Diederik P. Kingma and Jimmy Ba. 2015. Adam: A Method for Stochastic Optimization. In ICLR."},{"key":"e_1_2_2_14_1","unstructured":"Chunyuan Li Jianwei Yang Pengchuan Zhang Mei Gao Bin Xiao Xiyang Dai Lu Yuan and Jianfeng Gao. 2022. Efficient Self-supervised Vision Transformers for Representation Learning. In ICLR."},{"key":"e_1_2_2_15_1","unstructured":"Yaguang Li Rose Yu Cyrus Shahabi and Yan Liu. 2018. Diffusion Convolutional Recurrent Neural Network: Data-Driven Traffic Forecasting. In ICLR."},{"key":"e_1_2_2_16_1","volume-title":"Roberta: A robustly optimized bert pretraining approach. arXiv preprint arXiv:1907.11692","author":"Liu Yinhan","year":"2019","unstructured":"Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019. Roberta: A robustly optimized bert pretraining approach. arXiv preprint arXiv:1907.11692 (2019)."},{"key":"e_1_2_2_17_1","volume-title":"Agronomie, Soci\u00e9t\u00e9 et Environnement","author":"Ly Sarann","year":"2013","unstructured":"Sarann Ly, Catherine Charles, and Aurore Degr\u00e9. 2013. Different methods for spatial interpolation of rainfall data for operational hydrology and hydrological modeling at watershed scale: a review. Biotechnologie, Agronomie, Soci\u00e9t\u00e9 et Environnement (2013)."},{"key":"e_1_2_2_18_1","doi-asserted-by":"crossref","unstructured":"J Eamonn Nash and Jonh V Sutcliffe. 1970. River flow forecasting through conceptual models part I-A discussion of principles. J. Hydrol. (1970).","DOI":"10.1016\/0022-1694(70)90255-6"},{"key":"e_1_2_2_19_1","doi-asserted-by":"crossref","unstructured":"Peter Shaw Jakob Uszkoreit and Ashish Vaswani. 2018. Self-Attention with Relative Position Representations. In NAACL-HLT. 464--468.","DOI":"10.18653\/v1\/N18-2074"},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1002\/qj.2188"},{"key":"e_1_2_2_21_1","volume-title":"De Bilt","author":"Sluiter Raymond","year":"2009","unstructured":"Raymond Sluiter. 2009. Interpolation methods for climate data: literature review. KNMI, De Bilt (2009)."},{"key":"e_1_2_2_22_1","doi-asserted-by":"crossref","unstructured":"Weiping Song Chence Shi Zhiping Xiao Zhijian Duan Yewen Xu Ming Zhang and Jian Tang. 2019. AutoInt: Automatic Feature Interaction Learning via Self-Attentive Neural Networks. In CIKM. 1161--1170.","DOI":"10.1145\/3357384.3357925"},{"key":"e_1_2_2_23_1","doi-asserted-by":"crossref","unstructured":"Xiu Tang Sai Wu Mingli Song Shanshan Ying Feifei Li and Gang Chen. 2022. PreQR: Pre-training Representation for SQL Understanding. In SIGMOD. 204--216.","DOI":"10.1145\/3514221.3517878"},{"key":"e_1_2_2_24_1","doi-asserted-by":"crossref","unstructured":"Immanuel Trummer. 2022. DB-BERT: a Database Tuning Tool that \"Reads the Manual\". In SIGMOD. 190--203.","DOI":"10.1145\/3514221.3517843"},{"key":"e_1_2_2_25_1","volume-title":"Attention is all you need. NeurIPS","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, \u0141ukasz Kaiser, and Illia Polosukhin. 2017a. Attention is all you need. NeurIPS (2017), 5998--6008."},{"key":"e_1_2_2_26_1","unstructured":"Ashish Vaswani Noam Shazeer Niki Parmar Jakob Uszkoreit Llion Jones Aidan N. Gomez Lukasz Kaiser and Illia Polosukhin. 2017b. Attention is All you Need. In NeurIPS. 5998--6008."},{"key":"e_1_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1007\/978--3--662-03098--1_11"},{"key":"e_1_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1007\/978--3--662-03098--1_28"},{"key":"e_1_2_2_29_1","unstructured":"Max Welling and Thomas N Kipf. 2017. Semi-supervised classification with graph convolutional networks. In ICLR."},{"key":"e_1_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.3390\/w12041165"},{"key":"e_1_2_2_31_1","unstructured":"Yuankai Wu Dingyi Zhuang Aur\u00e9lie Labbe and Lijun Sun. 2021. Inductive Graph Neural Networks for Spatiotemporal Kriging. In AAAI."},{"key":"e_1_2_2_32_1","volume-title":"Yutao Zhu, Sirui Wang, Fuzheng Zhang, Zhongyuan Wang, and Ji-Rong Wen.","author":"Zhou Kun","year":"2020","unstructured":"Kun Zhou, Hui Wang, Wayne Xin Zhao, Yutao Zhu, Sirui Wang, Fuzheng Zhang, Zhongyuan Wang, and Ji-Rong Wen. 2020. S3-rec: Self-supervised learning for sequential recommendation with mutual information maximization. In CIKM. 1893--1902."}],"container-title":["Proceedings of the ACM on Management of Data"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3589321","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3589321","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T16:46:14Z","timestamp":1750178774000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3589321"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,6,13]]},"references-count":32,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2023,6,13]]}},"alternative-id":["10.1145\/3589321"],"URL":"https:\/\/doi.org\/10.1145\/3589321","relation":{},"ISSN":["2836-6573"],"issn-type":[{"value":"2836-6573","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,6,13]]}}}