{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,1]],"date-time":"2026-05-01T17:11:34Z","timestamp":1777655494719,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":50,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"National Key Research and Development Program of China","award":["2020AAA0108600"],"award-info":[{"award-number":["2020AAA0108600"]}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61901435, 62131003, 62021001, 62032006"],"award-info":[{"award-number":["61901435, 62131003, 62021001, 62032006"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3547771","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:43:01Z","timestamp":1665416581000},"page":"4867-4876","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":7,"title":["RPPformer-Flow: Relative Position Guided Point Transformer for Scene Flow Estimation"],"prefix":"10.1145","author":[{"given":"Hanlin","family":"Li","sequence":"first","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Guanting","family":"Dong","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yueyi","family":"Zhang","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiaoyan","family":"Sun","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhiwei","family":"Xiong","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision. 13126--13136","author":"Baur Stefan Andreas","year":"2021","unstructured":"Stefan Andreas Baur , David Josef Emmerichs , Frank Moosmann , Peter Pinggera , Bj\u00f6rn Ommer , and Andreas Geiger . 2021 . SLIM: Self-supervised LiDAR scene flow and motion segmentation . In Proceedings of the IEEE\/CVF International Conference on Computer Vision. 13126--13136 . Stefan Andreas Baur, David Josef Emmerichs, Frank Moosmann, Peter Pinggera, Bj\u00f6rn Ommer, and Andreas Geiger. 2021. SLIM: Self-supervised LiDAR scene flow and motion segmentation. In Proceedings of the IEEE\/CVF International Conference on Computer Vision. 13126--13136."},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58452-8_13"},{"key":"e_1_3_2_2_3_1","volume-title":"Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies","volume":"1","author":"Devlin Jacob","year":"2019","unstructured":"Jacob Devlin , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . 2019 . Bert: Pre-training of deep bidirectional transformers for language understanding . In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies Volume 1 (Long and Short Papers), Jill Burstein, Christy Doran, and Thamar Solorio (Eds.). Association for Computational Linguistics, 4171--4186. https:\/\/doi.org\/10. 18653\/v1\/n19--1423 Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019. Bert: Pre-training of deep bidirectional transformers for language understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies Volume 1 (Long and Short Papers), Jill Burstein, Christy Doran, and Thamar Solorio (Eds.). Association for Computational Linguistics, 4171--4186. https:\/\/doi.org\/10.18653\/v1\/n19--1423"},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3474085.3475195"},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01244"},{"key":"e_1_3_2_2_6_1","volume-title":"9th International Conference on Learning Representations. OpenReview.net.","author":"Dosovitskiy Alexey","year":"2021","unstructured":"Alexey Dosovitskiy , Lucas Beyer , Alexander Kolesnikov , Dirk Weissenborn , Xiaohua Zhai , Thomas Unterthiner , Mostafa Dehghani , Matthias Minderer , Georg Heigold , Sylvain Gelly , 2021 . An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale . In 9th International Conference on Learning Representations. OpenReview.net. Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al. 2021. An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. In 9th International Conference on Learning Representations. OpenReview.net."},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.316"},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/83.623193"},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00564"},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00337"},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1007\/s41095-021-0229-5"},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/3474085.3475285"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-021-01551-y"},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3394171.3413671"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.482"},{"key":"e_1_3_2_2_17_1","volume-title":"Permutohedral Lattice CNNs. In 3rd International Conference on Learning RepresentationsWorkshop Track Proceedings, Yoshua Bengio and Yann LeCun (Eds.).","author":"Kiefel Martin","year":"2015","unstructured":"Martin Kiefel , Varun Jampani , and Peter V Gehler . 2015 . Permutohedral Lattice CNNs. In 3rd International Conference on Learning RepresentationsWorkshop Track Proceedings, Yoshua Bengio and Yann LeCun (Eds.). Martin Kiefel, Varun Jampani, and Peter V Gehler. 2015. Permutohedral Lattice CNNs. In 3rd International Conference on Learning RepresentationsWorkshop Track Proceedings, Yoshua Bengio and Yann LeCun (Eds.)."},{"key":"e_1_3_2_2_18_1","volume-title":"3rd International Conference on Learning Representations.","author":"Kingma Diederik P","year":"2015","unstructured":"Diederik P Kingma and Jimmy Ba . 2015 . Adam: A method for stochastic optimization . In 3rd International Conference on Learning Representations. Diederik P Kingma and Jimmy Ba. 2015. Adam: A method for stochastic optimization. In 3rd International Conference on Learning Representations."},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00410"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/3474085.3475409"},{"key":"e_1_3_2_2_21_1","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 364--373","author":"Li Ruibo","year":"2021","unstructured":"Ruibo Li , Guosheng Lin , Tong He , Fayao Liu , and Chunhua Shen . 2021 . HCRFFlow: Scene flow from point clouds with continuous high-order CRFs and position-aware flow embedding . In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 364--373 . Ruibo Li, Guosheng Lin, Tong He, Fayao Liu, and Chunhua Shen. 2021. HCRFFlow: Scene flow from point clouds with continuous high-order CRFs and position-aware flow embedding. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 364--373."},{"key":"e_1_3_2_2_22_1","first-page":"7838","article-title":"Neural scene flow prior","volume":"34","author":"Li Xueqian","year":"2021","unstructured":"Xueqian Li , Jhony Kaesemodel Pontes , and Simon Lucey . 2021 . Neural scene flow prior . Advances in Neural Information Processing Systems 34 (2021), 7838 -- 7851 . Xueqian Li, Jhony Kaesemodel Pontes, and Simon Lucey. 2021. Neural scene flow prior. Advances in Neural Information Processing Systems 34 (2021), 7838--7851.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_2_23_1","first-page":"820","article-title":"Pointcnn: Convolution on x-transformed points","volume":"31","author":"Li Yangyan","year":"2018","unstructured":"Yangyan Li , Rui Bu , Mingchao Sun ,WeiWu, Xinhan Di , and Baoquan Chen . 2018 . Pointcnn: Convolution on x-transformed points . Advances in Neural Information Processing Systems 31 (2018), 820 -- 830 . Yangyan Li, Rui Bu, Mingchao Sun,WeiWu, Xinhan Di, and Baoquan Chen. 2018. Pointcnn: Convolution on x-transformed points. Advances in Neural Information Processing Systems 31 (2018), 820--830.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00062"},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00986"},{"key":"e_1_3_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01228-1_29"},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00315"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.438"},{"key":"e_1_3_2_2_29_1","first-page":"427","article-title":"Joint 3d estimation of vehicles and scene flow. ISPRS Annals of the Photogrammetry","volume":"2","author":"Menze Moritz","year":"2015","unstructured":"Moritz Menze , Christian Heipke , and Andreas Geiger . 2015 . Joint 3d estimation of vehicles and scene flow. ISPRS Annals of the Photogrammetry , Remote Sensing and Spatial Information Sciences 2 (2015), 427 . Moritz Menze, Christian Heipke, and Andreas Geiger. 2015. Joint 3d estimation of vehicles and scene flow. ISPRS Annals of the Photogrammetry, Remote Sensing and Spatial Information Sciences 2 (2015), 427.","journal-title":"Remote Sensing and Spatial Information Sciences"},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.isprsjprs.2017.09.013"},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00290"},{"key":"e_1_3_2_2_32_1","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 11177--11185","author":"Mittal Himangi","year":"2020","unstructured":"Himangi Mittal , Brian Okorn , and David Held . 2020 . Just go with the flow: Selfsupervised scene flow estimation . In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 11177--11185 . Himangi Mittal, Brian Okorn, and David Held. 2020. Just go with the flow: Selfsupervised scene flow estimation. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 11177--11185."},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00738"},{"key":"e_1_3_2_2_34_1","volume-title":"International Conference on Machine Learning. PMLR, 4055--4064","author":"Parmar Niki","year":"2018","unstructured":"Niki Parmar , Ashish Vaswani , Jakob Uszkoreit , Lukasz Kaiser , Noam Shazeer , Alexander Ku , and Dustin Tran . 2018 . Image transformer . In International Conference on Machine Learning. PMLR, 4055--4064 . Niki Parmar, Ashish Vaswani, Jakob Uszkoreit, Lukasz Kaiser, Noam Shazeer, Alexander Ku, and Dustin Tran. 2018. Image transformer. In International Conference on Machine Learning. PMLR, 4055--4064."},{"key":"e_1_3_2_2_35_1","volume-title":"Pytorch: An imperative style, high-performance deep learning library. Advances in Neural Information Processing Systems 32","author":"Paszke Adam","year":"2019","unstructured":"Adam Paszke , Sam Gross , Francisco Massa , Adam Lerer , James Bradbury , Gregory Chanan , Trevor Killeen , Zeming Lin , Natalia Gimelshein , Luca Antiga , 2019 . Pytorch: An imperative style, high-performance deep learning library. Advances in Neural Information Processing Systems 32 (2019). Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, et al. 2019. Pytorch: An imperative style, high-performance deep learning library. Advances in Neural Information Processing Systems 32 (2019)."},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58604-1_32"},{"key":"e_1_3_2_2_37_1","volume-title":"Pointnet: Deep hierarchical feature learning on point sets in a metric space. Advances in Neural Information Processing Systems 30","author":"Qi Charles Ruizhongtai","year":"2017","unstructured":"Charles Ruizhongtai Qi , Li Yi , Hao Su , and Leonidas J Guibas . 2017 . Pointnet: Deep hierarchical feature learning on point sets in a metric space. Advances in Neural Information Processing Systems 30 (2017). Charles Ruizhongtai Qi, Li Yi, Hao Su, and Leonidas J Guibas. 2017. Pointnet: Deep hierarchical feature learning on point sets in a metric space. Advances in Neural Information Processing Systems 30 (2017)."},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58536-5_24"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00827"},{"key":"e_1_3_2_2_40_1","volume-title":"Attention is all you need. Advances in Neural Information Processing Systems 30","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani , Noam Shazeer , Niki Parmar , Jakob Uszkoreit , Llion Jones , Aidan N Gomez , Lukasz Kaiser , and Illia Polosukhin . 2017. Attention is all you need. Advances in Neural Information Processing Systems 30 ( 2017 ). Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017. Attention is all you need. Advances in Neural Information Processing Systems 30 (2017)."},{"key":"e_1_3_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.1999.790293"},{"key":"e_1_3_2_2_42_1","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 6954--6963","author":"Rao Yongming","year":"2021","unstructured":"YiWei, ZiyiWang, Yongming Rao , Jiwen Lu , and Jie Zhou . 2021 . PV-RAFT: Point-Voxel Correlation Fields for Scene Flow Estimation of Point Clouds . In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 6954--6963 . YiWei, ZiyiWang, Yongming Rao, Jiwen Lu, and Jie Zhou. 2021. PV-RAFT: Point-Voxel Correlation Fields for Scene Flow Estimation of Point Clouds. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 6954--6963."},{"key":"e_1_3_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00985"},{"key":"e_1_3_2_2_44_1","volume-title":"European Conference on Computer Vision. Springer, 88--107","author":"YuanWang Zhi","year":"2020","unstructured":"WenxuanWu, Zhi YuanWang , Zhuwen Li , Wei Liu , and Li Fuxin . 2020 . Pointpwcnet: Cost volume on point clouds for (self-) supervised scene flow estimation . In European Conference on Computer Vision. Springer, 88--107 . WenxuanWu, Zhi YuanWang, Zhuwen Li,Wei Liu, and Li Fuxin. 2020. Pointpwcnet: Cost volume on point clouds for (self-) supervised scene flow estimation. In European Conference on Computer Vision. Springer, 88--107."},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01261-8_1"},{"key":"e_1_3_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00545"},{"key":"e_1_3_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.1145\/3474085.3475272"},{"key":"e_1_3_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01595"},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00681"},{"key":"e_1_3_2_2_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00472"}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","location":"Lisboa Portugal","acronym":"MM '22","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547771","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3547771","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:30:41Z","timestamp":1750188641000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547771"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":50,"alternative-id":["10.1145\/3503161.3547771","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3547771","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}