{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,17]],"date-time":"2026-07-17T06:10:30Z","timestamp":1784268630475,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":82,"publisher":"ACM","funder":[{"name":"National Nature Science Foundation of China","award":["62402406"],"award-info":[{"award-number":["62402406"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,12,15]]},"DOI":"10.1145\/3757377.3763872","type":"proceedings-article","created":{"date-parts":[[2025,12,8]],"date-time":"2025-12-08T16:30:41Z","timestamp":1765211441000},"page":"1-12","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["OmniPart: Part-Aware 3D Generation with Semantic Decoupling and Structural Cohesion"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0008-0141-1545","authenticated-orcid":false,"given":"Yunhan","family":"Yang","sequence":"first","affiliation":[{"name":"University of Hong Kong, Hong Kong, Hong Kong"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-2671-1308","authenticated-orcid":false,"given":"Yufan","family":"Zhou","sequence":"additional","affiliation":[{"name":"Harbin Institute of Technology, Shandong, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6164-8343","authenticated-orcid":false,"given":"Yuan-Chen","family":"Guo","sequence":"additional","affiliation":[{"name":"VAST, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2945-552X","authenticated-orcid":false,"given":"Zi-Xin","family":"Zou","sequence":"additional","affiliation":[{"name":"VAST, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5322-2884","authenticated-orcid":false,"given":"Yukun","family":"Huang","sequence":"additional","affiliation":[{"name":"University of Hong Kong, Hong Kong, Hong Kong"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8293-3223","authenticated-orcid":false,"given":"Ying-Tian","family":"Liu","sequence":"additional","affiliation":[{"name":"Tsinghua University, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5690-367X","authenticated-orcid":false,"given":"Hao","family":"Xu","sequence":"additional","affiliation":[{"name":"State Key Laboratory of CAD&amp;CG, Zhejiang University, Zhejiang, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9774-4687","authenticated-orcid":false,"given":"Ding","family":"Liang","sequence":"additional","affiliation":[{"name":"VAST, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0416-4374","authenticated-orcid":false,"given":"Yan-Pei","family":"Cao","sequence":"additional","affiliation":[{"name":"VAST, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1831-9952","authenticated-orcid":false,"given":"Xihui","family":"Liu","sequence":"additional","affiliation":[{"name":"University of Hong Kong, Hong Kong, Hong Kong"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2025,12,14]]},"reference":[{"key":"e_1_3_3_2_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.01392"},{"key":"e_1_3_3_2_3_1","volume-title":"ECCV","author":"Chatterjee Agneet","year":"2024","unstructured":"Agneet Chatterjee, Gabriela Ben\u00a0Melech Stan, Estelle Aflalo, Sayak Paul, Dhruba Ghosh, Tejas Gokhale, Ludwig Schmidt, Hannaneh Hajishirzi, Vasudev Lal, Chitta Baral, et\u00a0al. 2024. Getting it right: Improving spatial consistency in text-to-image models. In ECCV."},{"key":"e_1_3_3_2_4_1","doi-asserted-by":"crossref","unstructured":"Minghao Chen Roman Shapovalov Iro Laina Tom Monnier Jianyuan Wang David Novotny and Andrea Vedaldi. 2024b. PartGen: Part-level 3D Generation and Reconstruction with Multi-View Diffusion Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2412.18608 (2024).","DOI":"10.1109\/CVPR52734.2025.00552"},{"key":"e_1_3_3_2_5_1","doi-asserted-by":"publisher","DOI":"10.52202\/079017-3080"},{"key":"e_1_3_3_2_6_1","volume-title":"ICLR","author":"Chen Yiwen","year":"2025","unstructured":"Yiwen Chen, Tong He, Di Huang, Weicai Ye, Sijin Chen, Jiaxiang Tang, Xin Chen, Zhongang Cai, Lei Yang, Gang Yu, et\u00a0al. 2025. Meshanything: Artist-created mesh generation with autoregressive transformers. In ICLR."},{"key":"e_1_3_3_2_7_1","volume-title":"ECCV","author":"Chen Yongwei","year":"2024","unstructured":"Yongwei Chen, Tengfei Wang, Tong Wu, Xingang Pan, Kui Jia, and Ziwei Liu. 2024d. Comboverse: Compositional 3d assets creation using spatially-aware diffusion guidance. In ECCV."},{"key":"e_1_3_3_2_8_1","unstructured":"Yiwen Chen Yikai Wang Yihao Luo Zhengyi Wang Zilong Chen Jun Zhu Chi Zhang and Guosheng Lin. 2024c. Meshanything v2: Artist-created mesh generation with adjacent mesh tokenization. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.02555 (2024)."},{"key":"e_1_3_3_2_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.00215"},{"key":"e_1_3_3_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.693"},{"key":"e_1_3_3_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.01263"},{"key":"e_1_3_3_2_12_1","unstructured":"Ken Deng Yunhan Yang Jingxiang Sun Xihui Liu Yebin Liu Ding Liang and Yan-Pei Cao. 2025. GeoSAM2: Unleashing the Power of SAM2 for 3D Part Segmentation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2508.14036 (2025)."},{"key":"e_1_3_3_2_13_1","unstructured":"Woojung Han Yeonkyung Lee Chanyoung Kim Kwanghyun Park and Seong\u00a0Jae Hwang. 2025. Spatial Transport Optimization by Repositioning Attention Map for Training-Free Text-to-Image Synthesis. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2503.22168 (2025)."},{"key":"e_1_3_3_2_14_1","unstructured":"Zekun Hao David\u00a0W Romero Tsung-Yi Lin and Ming-Yu Liu. 2024. Meshtron: High-Fidelity Artist-Like 3D Mesh Generation at Scale. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2412.09548 (2024)."},{"key":"e_1_3_3_2_15_1","doi-asserted-by":"crossref","unstructured":"Kaiyi Huang Chengqi Duan Kaiyue Sun Enze Xie Zhenguo Li and Xihui Liu. 2025a. T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-Image Generation. IEEE Transactions on Pattern Analysis and Machine Intelligence (2025).","DOI":"10.1109\/TPAMI.2025.3531907"},{"key":"e_1_3_3_2_16_1","volume-title":"NeurIPS","author":"Huang Kaiyi","year":"2023","unstructured":"Kaiyi Huang, Kaiyue Sun, Enze Xie, Zhenguo Li, and Xihui Liu. 2023. T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation. In NeurIPS."},{"key":"e_1_3_3_2_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52734.2025.02202"},{"key":"e_1_3_3_2_18_1","unstructured":"Zehuan Huang Yuan-Chen Guo Haoran Wang Ran Yi Lizhuang Ma Yan-Pei Cao and Lu Sheng. 2024a. Mv-adapter: Multi-view consistent image generation made easy. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2412.03632 (2024)."},{"key":"e_1_3_3_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.00934"},{"key":"e_1_3_3_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/3550469.3555394"},{"key":"e_1_3_3_2_21_1","unstructured":"Peidong Jia Chenxuan Li Yuhui Yuan Zeyu Liu Yichao Shen Bohan Chen Xingru Chen Yinglin Zheng Dong Chen Ji Li et\u00a0al. 2023. COLE: A Hierarchical Generation Framework for Multi-Layered and Editable Graphic Design. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2311.16974 (2023)."},{"key":"e_1_3_3_2_22_1","doi-asserted-by":"crossref","unstructured":"Bernhard Kerbl Georgios Kopanas Thomas Leimk\u00fchler and George Drettakis. 2023. 3d gaussian splatting for real-time radiance field rendering. ACM Trans. Graph. 42 4 (2023) 139\u20131.","DOI":"10.1145\/3592433"},{"key":"e_1_3_3_2_23_1","volume-title":"ECCV","author":"Kim Hyunjin","year":"2024","unstructured":"Hyunjin Kim and Minhyuk Sung. 2024. PartSTAD: 2D-to-3D Part Segmentation Task Adaptation. In ECCV."},{"key":"e_1_3_3_2_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.00371"},{"key":"e_1_3_3_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.01328"},{"key":"e_1_3_3_2_26_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-73235-5_7"},{"key":"e_1_3_3_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01069"},{"key":"e_1_3_3_2_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.01216"},{"key":"e_1_3_3_2_29_1","unstructured":"Songlin Li Despoina Paschalidou and Leonidas Guibas. 2024b. PASTA: Controllable Part-Aware Shape Generation with Autoregressive Transformers. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2407.13677 (2024)."},{"key":"e_1_3_3_2_30_1","unstructured":"Weiyu Li Jiarui Liu Rui Chen Yixun Liang Xuelin Chen Ping Tan and Xiaoxiao Long. 2024a. CraftsMan: High-fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2405.14979 (2024)."},{"key":"e_1_3_3_2_31_1","volume-title":"NeurIPS","author":"Li Yangyan","year":"2018","unstructured":"Yangyan Li, Rui Bu, Mingchao Sun, Wei Wu, Xinhan Di, and Baoquan Chen. 2018. Pointcnn: Convolution on x-transformed points. In NeurIPS."},{"key":"e_1_3_3_2_32_1","unstructured":"Yangguang Li Zi-Xin Zou Zexiang Liu Dehu Wang Yuan Liang Zhipeng Yu Xingchao Liu Yuan-Chen Guo Ding Liang Wanli Ouyang et\u00a0al. 2025. TripoSG: High-Fidelity 3D Shape Synthesis using Large-Scale Rectified Flow Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2502.06608 (2025)."},{"key":"e_1_3_3_2_33_1","unstructured":"Yuchen Lin Chenguo Lin Panwang Pan Honglei Yan Yiqiang Feng Yadong Mu and Katerina Fragkiadaki. 2025. PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion Transformers. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2506.05573 (2025)."},{"key":"e_1_3_3_2_34_1","volume-title":"NeurIPS","author":"Lipman Yaron","year":"2024","unstructured":"Yaron Lipman, Ricky\u00a0TQ Chen, Heli Ben-Hamu, Maximilian Nickel, and Matt Le. 2024. Flow matching for generative modeling. In NeurIPS."},{"key":"e_1_3_3_2_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/3641519.3657482"},{"key":"e_1_3_3_2_36_1","unstructured":"Jian Liu Haohan Weng Biwen Lei Xianghui Yang Zibo Zhao Zhuo Chen Song Guo Tao Han and Chunchao Guo. 2025b. FreeMesh: Boosting Mesh Generation with Coordinates Merging. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2505.13573 (2025)."},{"key":"e_1_3_3_2_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.00960"},{"key":"e_1_3_3_2_38_1","unstructured":"Minghua Liu Mikaela\u00a0Angelina Uy Donglai Xiang Hao Su Sanja Fidler Nicholas Sharp and Jun Gao. 2025a. PARTFIELD: Learning 3D Feature Fields for Part Segmentation and Beyond. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2504.11451 (2025)."},{"key":"e_1_3_3_2_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.02082"},{"key":"e_1_3_3_2_40_1","volume-title":"ICLR","author":"Liu Yuan","year":"2024","unstructured":"Yuan Liu, Cheng Lin, Zijiao Zeng, Xiaoxiao Long, Lingjie Liu, Taku Komura, and Wenping Wang. 2024b. SyncDreamer: Learning to Generate Multiview-consistent Images from a Single-view Image. In ICLR."},{"key":"e_1_3_3_2_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.00951"},{"key":"e_1_3_3_2_42_1","doi-asserted-by":"crossref","unstructured":"Ben Mildenhall Pratul\u00a0P Srinivasan Matthew Tancik Jonathan\u00a0T Barron Ravi Ramamoorthi and Ren Ng. 2021. Nerf: Representing scenes as neural radiance fields for view synthesis. Commun. ACM 65 1 (2021) 99\u2013106.","DOI":"10.1145\/3503250"},{"key":"e_1_3_3_2_43_1","unstructured":"Maxime Oquab Timoth\u00e9e Darcet Th\u00e9o Moutakanni Huy Vo Marc Szafraniec Vasil Khalidov Pierre Fernandez Daniel Haziza Francisco Massa Alaaeldin El-Nouby et\u00a0al. 2023. Dinov2: Learning robust visual features without supervision. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2304.07193 (2023)."},{"key":"e_1_3_3_2_44_1","volume-title":"ICLR","author":"Poole Ben","year":"2023","unstructured":"Ben Poole, Ajay Jain, Jonathan\u00a0T Barron, and Ben Mildenhall. 2023. DreamFusion: Text-to-3D using 2D Diffusion. In ICLR."},{"key":"e_1_3_3_2_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52734.2025.00745"},{"key":"e_1_3_3_2_46_1","volume-title":"CVPR","author":"Qi Charles\u00a0R","year":"2017","unstructured":"Charles\u00a0R Qi, Hao Su, Kaichun Mo, and Leonidas\u00a0J Guibas. 2017a. Pointnet: Deep learning on point sets for 3d classification and segmentation. In CVPR."},{"key":"e_1_3_3_2_47_1","volume-title":"NeurIPS","author":"Qi Charles\u00a0Ruizhongtai","year":"2017","unstructured":"Charles\u00a0Ruizhongtai Qi, Li Yi, Hao Su, and Leonidas\u00a0J Guibas. 2017b. Pointnet++: Deep hierarchical feature learning on point sets in a metric space. In NeurIPS."},{"key":"e_1_3_3_2_48_1","unstructured":"Zhangyang Qi Yunhan Yang Mengchen Zhang Long Xing Xiaoyang Wu Tong Wu Dahua Lin Xihui Liu Jiaqi Wang and Hengshuang Zhao. 2024. Tailor3d: Customized 3d assets editing and generation with dual-side images. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2407.06191 (2024)."},{"key":"e_1_3_3_2_49_1","volume-title":"NeurIPS","author":"Qian Guocheng","year":"2022","unstructured":"Guocheng Qian, Yuchen Li, Houwen Peng, Jinjie Mai, Hasan Hammoud, Mohamed Elhoseiny, and Bernard Ghanem. 2022. Pointnext: Revisiting pointnet++ with improved training and scaling strategies. In NeurIPS."},{"key":"e_1_3_3_2_50_1","volume-title":"ICML","author":"Radford Alec","year":"2021","unstructured":"Alec Radford, Jong\u00a0Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et\u00a0al. 2021. Learning transferable visual models from natural language supervision. In ICML."},{"key":"e_1_3_3_2_51_1","unstructured":"Nikhila Ravi Valentin Gabeur Yuan-Ting Hu Ronghang Hu Chaitanya Ryali Tengyu Ma Haitham Khedr Roman R\u00e4dle Chloe Rolland Laura Gustafson et\u00a0al. 2024. Sam 2: Segment anything in images and videos. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.00714 (2024)."},{"key":"e_1_3_3_2_52_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01042"},{"key":"e_1_3_3_2_53_1","unstructured":"Ruoxi Shi Hansheng Chen Zhuoyang Zhang Minghua Liu Chao Xu Xinyue Wei Linghao Chen Chong Zeng and Hao Su. 2023. Zero123++: a single image to consistent multi-view diffusion base model. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2310.15110 (2023)."},{"key":"e_1_3_3_2_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.02001"},{"key":"e_1_3_3_2_55_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.01855"},{"key":"e_1_3_3_2_56_1","unstructured":"George Tang William Zhao Logan Ford David Benhaim and Paul Zhang. 2024. Segment Any Mesh: Zero-shot Mesh Part Segmentation via Lifting Segment Anything 2 to 3D. arXiv:https:\/\/arXiv.org\/abs\/2408.13679 (2024)."},{"key":"e_1_3_3_2_57_1","volume-title":"ICLR","author":"Tang Jiaxiang","year":"2025","unstructured":"Jiaxiang Tang, Zhaoshuo Li, Zekun Hao, Xian Liu, Gang Zeng, Ming-Yu Liu, and Qinsheng Zhang. 2025a. Edgerunner: Auto-regressive auto-encoder for artistic mesh generation. In ICLR."},{"key":"e_1_3_3_2_58_1","unstructured":"Jiaxiang Tang Ruijie Lu Zhaoshuo Li Zekun Hao Xuan Li Fangyin Wei Shuran Song Gang Zeng Ming-Yu Liu and Tsung-Yi Lin. 2025b. Efficient Part-level 3D Object Generation via Dual Volume Packing. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2506.09980 (2025)."},{"key":"e_1_3_3_2_59_1","volume-title":"ECCV","author":"Thai Anh","year":"2024","unstructured":"Anh Thai, Weiyao Wang, Hao Tang, Stefan Stojanov, Matt Feiszli, and James\u00a0M Rehg. 2024. 3x2: 3D Object Part Segmentation by 2D Semantic Correspondences. In ECCV."},{"key":"e_1_3_3_2_60_1","unstructured":"Yuxuan Wang Xuanyu Yi Haohan Weng Qingshan Xu Xiaokang Wei Xianghui Yang Chunchao Guo Long Chen and Hanwang Zhang. 2025. Nautilus: Locality-aware Autoencoder for Scalable Mesh Generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2501.14317 (2025)."},{"key":"e_1_3_3_2_61_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52734.2025.02015"},{"key":"e_1_3_3_2_62_1","unstructured":"Haohan Weng Zibo Zhao Biwen Lei Xianghui Yang Jian Liu Zeqiang Lai Zhuo Chen Yuhong Liu Jie Jiang Chunchao Guo et\u00a0al. 2024. Scaling mesh generation via compressive tokenization. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2411.07025 (2024)."},{"key":"e_1_3_3_2_63_1","volume-title":"NeurIPS","author":"Wu Shuang","year":"2024","unstructured":"Shuang Wu, Youtian Lin, Feihu Zhang, Yifei Zeng, Jingxi Xu, Philip Torr, Xun Cao, and Yao Yao. 2024. Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer. In NeurIPS."},{"key":"e_1_3_3_2_64_1","doi-asserted-by":"crossref","unstructured":"Jianfeng Xiang Zelong Lv Sicheng Xu Yu Deng Ruicheng Wang Bowen Zhang Dong Chen Xin Tong and Jiaolong Yang. 2024. Structured 3d latents for scalable and versatile 3d generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2412.01506 (2024).","DOI":"10.1109\/CVPR52734.2025.02000"},{"key":"e_1_3_3_2_65_1","unstructured":"Jiale Xu Weihao Cheng Yiming Gao Xintao Wang Shenghua Gao and Ying Shan. 2024. Instantmesh: Efficient 3d mesh generation from a single image with sparse-view large reconstruction models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2404.07191 (2024)."},{"key":"e_1_3_3_2_66_1","unstructured":"Yuheng Xue Nenglun Chen Jun Liu and Wenyun Sun. 2023. ZeroPS: High-quality Cross-modal Knowledge Transfer for Zero-Shot 3D Part Segmentation. arXiv:https:\/\/arXiv.org\/abs\/2311.14262 (2023)."},{"key":"e_1_3_3_2_67_1","unstructured":"Han Yan Mingrui Zhang Yang Li Chao Ma and Pan Ji. 2024. PhyCAGE: Physically Plausible Compositional 3D Asset Generation from a Single Image. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2411.18548 (2024)."},{"key":"e_1_3_3_2_68_1","doi-asserted-by":"crossref","unstructured":"Yunhan Yang Shuo Chen Yukun Huang Xiaoyang Wu Yuan-Chen Guo Edmund\u00a0Y Lam Hengshuang Zhao Tong He and Xihui Liu. 2025a. DreamComposer++: Empowering Diffusion Models with Multi-View Conditions for 3D Content Generation. IEEE Transactions on Pattern Analysis and Machine Intelligence (2025).","DOI":"10.1109\/TPAMI.2025.3568190"},{"key":"e_1_3_3_2_69_1","unstructured":"Yunhan Yang Yuan-Chen Guo Yukun Huang Zi-Xin Zou Zhipeng Yu Yangguang Li Yan-Pei Cao and Xihui Liu. 2025b. HoloPart: Generative 3D Part Amodal Segmentation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2504.07943 (2025)."},{"key":"e_1_3_3_2_70_1","unstructured":"Yunhan Yang Yukun Huang Yuan-Chen Guo Liangjun Lu Xiaoyang Wu Edmund\u00a0Y Lam Yan-Pei Cao and Xihui Liu. 2024a. Sampart3d: Segment any part in 3d objects. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2411.07184 (2024)."},{"key":"e_1_3_3_2_71_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.00775"},{"key":"e_1_3_3_2_72_1","unstructured":"Yunhan Yang Xiaoyang Wu Tong He Hengshuang Zhao and Xihui Liu. 2023. Sam3d: Segment anything in 3d scenes. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2306.03908 (2023)."},{"key":"e_1_3_3_2_73_1","doi-asserted-by":"crossref","unstructured":"Biao Zhang Jiapeng Tang Matthias Niessner and Peter Wonka. 2023. 3dshape2vecset: A 3d shape representation for neural fields and generative diffusion models. ACM Transactions On Graphics (TOG) 42 4 (2023) 1\u201316.","DOI":"10.1145\/3592442"},{"key":"e_1_3_3_2_74_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-72649-1_16"},{"key":"e_1_3_3_2_75_1","unstructured":"Gaoyang Zhang Bingtao Fu Qingnan Fan Qi Zhang Runxing Liu Hong Gu Huaqi Zhang and Xinguo Liu. 2024a. CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2412.13195 (2024)."},{"key":"e_1_3_3_2_76_1","doi-asserted-by":"crossref","unstructured":"Lvmin Zhang and Maneesh Agrawala. 2024. Transparent Image Layer Diffusion using Latent Transparency. ACM Transactions on Graphics (TOG) 43 4 (2024) 1\u201315.","DOI":"10.1145\/3658150"},{"key":"e_1_3_3_2_77_1","doi-asserted-by":"crossref","unstructured":"Longwen Zhang Ziyu Wang Qixuan Zhang Qiwei Qiu Anqi Pang Haoran Jiang Wei Yang Lan Xu and Jingyi Yu. 2024b. CLAY: A Controllable Large-scale Generative Model for Creating High-quality 3D Assets. ACM Transactions on Graphics (TOG) 43 4 (2024) 1\u201320.","DOI":"10.1145\/3658146"},{"key":"e_1_3_3_2_78_1","unstructured":"Susan Zhang Stephen Roller Naman Goyal Mikel Artetxe Moya Chen Shuohui Chen Christopher Dewan Mona Diab Xian Li Xi\u00a0Victoria Lin et\u00a0al. 2022. Opt: Open pre-trained transformer language models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2205.01068 (2022)."},{"key":"e_1_3_3_2_79_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01595"},{"key":"e_1_3_3_2_80_1","unstructured":"Ruowen Zhao Junliang Ye Zhengyi Wang Guangce Liu Yiwen Chen Yikai Wang and Jun Zhu. 2025. DeepMesh: Auto-Regressive Artist-mesh Creation with Reinforcement Learning. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2503.15265 (2025)."},{"key":"e_1_3_3_2_81_1","volume-title":"NeurIPS","author":"Zhao Zibo","year":"2024","unstructured":"Zibo Zhao, Wen Liu, Xin Chen, Xianfang Zeng, Rui Wang, Pei Cheng, Bin Fu, Tao Chen, Gang Yu, and Shenghua Gao. 2024. Michelangelo: Conditional 3d shape generation based on shape-image-text aligned latent representation. In NeurIPS."},{"key":"e_1_3_3_2_82_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-72980-5_11"},{"key":"e_1_3_3_2_83_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.00983"}],"event":{"name":"SA Conference Papers '25: SIGGRAPH Asia 2025 Conference Papers","location":"Hong Kong Hong Kong","acronym":"SA Conference Papers '25","sponsor":["SIGGRAPH ACM Special Interest Group on Computer Graphics and Interactive Techniques"]},"container-title":["Proceedings of the SIGGRAPH Asia 2025 Conference Papers"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3757377.3763872","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,12,9]],"date-time":"2025-12-09T03:32:06Z","timestamp":1765251126000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3757377.3763872"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,14]]},"references-count":82,"alternative-id":["10.1145\/3757377.3763872","10.1145\/3757377"],"URL":"https:\/\/doi.org\/10.1145\/3757377.3763872","relation":{},"subject":[],"published":{"date-parts":[[2025,12,14]]},"assertion":[{"value":"2025-12-14","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}