{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,23]],"date-time":"2026-04-23T07:59:32Z","timestamp":1776931172975,"version":"3.51.2"},"publisher-location":"New York, NY, USA","reference-count":89,"publisher":"ACM","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,12,15]]},"DOI":"10.1145\/3757377.3763857","type":"proceedings-article","created":{"date-parts":[[2025,12,8]],"date-time":"2025-12-08T16:27:29Z","timestamp":1765211249000},"page":"1-12","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["LLM-Primitives: Large Language Model for 3D Reconstruction with Primitives"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-9409-5842","authenticated-orcid":false,"given":"Kuan","family":"Tian","sequence":"first","affiliation":[{"name":"Tencent AIPD, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0430-5875","authenticated-orcid":false,"given":"Zhihao","family":"Hu","sequence":"additional","affiliation":[{"name":"Tencent AIPD, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-0643-3187","authenticated-orcid":false,"given":"Yonghang","family":"Guan","sequence":"additional","affiliation":[{"name":"Tencent AIPD, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5579-7094","authenticated-orcid":false,"given":"Jun","family":"Zhang","sequence":"additional","affiliation":[{"name":"Tencent AIPD, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,12,14]]},"reference":[{"key":"e_1_3_3_3_2_1","unstructured":"Antonio Alliegro Yawar Siddiqui Tatiana Tommasi and Matthias Nie\u00dfner. 2023. Polydiff: Generating 3d polygonal meshes with diffusion models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2312.11417 (2023)."},{"key":"e_1_3_3_3_3_1","unstructured":"Jinze Bai Shuai Bai Shusheng Yang Shijie Wang Sinan Tan Peng Wang Junyang Lin Chang Zhou and Jingren Zhou. 2023. Qwen-vl: A versatile vision-language model for understanding localization text reading and beyond. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2308.12966 1 2 (2023) 3."},{"key":"e_1_3_3_3_4_1","unstructured":"Tom Brown Benjamin Mann Nick Ryder Melanie Subbiah Jared\u00a0D Kaplan Prafulla Dhariwal Arvind Neelakantan Pranav Shyam Girish Sastry Amanda Askell et\u00a0al. 2020. Language models are few-shot learners. Advances in neural information processing systems 33 (2020) 1877\u20131901."},{"key":"e_1_3_3_3_5_1","volume-title":"ShapeNet: An Information-Rich 3D Model Repository","author":"Chang Angel\u00a0X.","year":"2015","unstructured":"Angel\u00a0X. Chang, Thomas Funkhouser, Leonidas Guibas, Pat Hanrahan, Qixing Huang, Zimo Li, Silvio Savarese, Manolis Savva, Shuran Song, Hao Su, Jianxiong Xiao, Li Yi, and Fisher Yu. 2015. ShapeNet: An Information-Rich 3D Model Repository. Technical Report arXiv:https:\/\/arXiv.org\/abs\/1512.03012 [cs.GR]. Stanford University \u2014 Princeton University \u2014 Toyota Technological Institute at Chicago."},{"key":"e_1_3_3_3_6_1","unstructured":"Yiwen Chen Yikai Wang Yihao Luo Zhengyi Wang Zilong Chen Jun Zhu Chi Zhang and Guosheng Lin. 2024a. Meshanything v2: Artist-created mesh generation with adjacent mesh tokenization. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.02555 (2024)."},{"key":"e_1_3_3_3_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.02283"},{"key":"e_1_3_3_3_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.00433"},{"key":"e_1_3_3_3_9_1","unstructured":"Matt Deitke Ruoshi Liu Matthew Wallingford Huong Ngo Oscar Michel Aditya Kusupati Alan Fan Christian Laforte Vikram Voleti Samir\u00a0Yitzhak Gadre et\u00a0al. 2024. Objaverse-xl: A universe of 10m+ 3d objects. Advances in Neural Information Processing Systems 36 (2024)."},{"key":"e_1_3_3_3_10_1","doi-asserted-by":"crossref","unstructured":"Haoxiang Guo Shilin Liu Hao Pan Yang Liu Xin Tong and Baining Guo. 2022. Complexgen: Cad reconstruction by b-rep chain complex generation. ACM Transactions on Graphics (TOG) 41 4 (2022) 1\u201318.","DOI":"10.1145\/3528223.3530078"},{"key":"e_1_3_3_3_11_1","unstructured":"Anchit Gupta Wenhan Xiong Yixin Nie Ian Jones and Barlas O\u011fuz. 2023. 3dgen: Triplane latent diffusion for textured mesh generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2303.05371 (2023)."},{"key":"e_1_3_3_3_12_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-73242-3_26"},{"key":"e_1_3_3_3_13_1","unstructured":"Jonathan Ho Ajay Jain and Pieter Abbeel. 2020. Denoising diffusion probabilistic models. Advances in neural information processing systems 33 (2020) 6840\u20136851."},{"key":"e_1_3_3_3_14_1","unstructured":"Yicong Hong Kai Zhang Jiuxiang Gu Sai Bi Yang Zhou Difan Liu Feng Liu Kalyan Sunkavalli Trung Bui and Hao Tan. 2023. Lrm: Large reconstruction model for single image to 3d. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2311.04400 (2023)."},{"key":"e_1_3_3_3_15_1","unstructured":"Edward\u00a0J Hu Yelong Shen Phillip Wallis Zeyuan Allen-Zhu Yuanzhi Li Shean Wang Lu Wang Weizhu Chen et\u00a0al. 2022. Lora: Low-rank adaptation of large language models. ICLR 1 2 (2022) 3."},{"key":"e_1_3_3_3_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/3550469.3555394"},{"key":"e_1_3_3_3_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00964"},{"key":"e_1_3_3_3_18_1","unstructured":"Kacper Kania Maciej Zieba and Tomasz Kajdanowicz. 2020. UCSG-NET-unsupervised discovering of constructive solid geometry tree. Advances in neural information processing systems 33 (2020) 8776\u20138786."},{"key":"e_1_3_3_3_19_1","doi-asserted-by":"crossref","unstructured":"Bernhard Kerbl Georgios Kopanas Thomas Leimk\u00fchler and George Drettakis. 2023. 3d gaussian splatting for real-time radiance field rendering. ACM Trans. Graph. 42 4 (2023) 139\u20131.","DOI":"10.1145\/3592433"},{"key":"e_1_3_3_3_20_1","doi-asserted-by":"crossref","unstructured":"Mohammad\u00a0Sadil Khan Sankalp Sinha Talha Uddin Didier Stricker Sk\u00a0Aziz Ali and Muhammad\u00a0Zeshan Afzal. 2024. Text2cad: Generating sequential cad designs from beginner-to-expert level text prompts. Advances in Neural Information Processing Systems 37 (2024) 7552\u20137579.","DOI":"10.52202\/079017-0242"},{"key":"e_1_3_3_3_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/3550469.3555424"},{"key":"e_1_3_3_3_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01258"},{"key":"e_1_3_3_3_23_1","first-page":"19730","volume-title":"International conference on machine learning","author":"Li Junnan","year":"2023","unstructured":"Junnan Li, Dongxu Li, Silvio Savarese, and Steven Hoi. 2023b. Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models. In International conference on machine learning. PMLR, 19730\u201319742."},{"key":"e_1_3_3_3_24_1","first-page":"12888","volume-title":"International conference on machine learning","author":"Li Junnan","year":"2022","unstructured":"Junnan Li, Dongxu Li, Caiming Xiong, and Steven Hoi. 2022. Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation. In International conference on machine learning. PMLR, 12888\u201312900."},{"key":"e_1_3_3_3_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00276"},{"key":"e_1_3_3_3_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.01610"},{"key":"e_1_3_3_3_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/3588432.3591522"},{"key":"e_1_3_3_3_28_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58607-2_32"},{"key":"e_1_3_3_3_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.02484"},{"key":"e_1_3_3_3_30_1","unstructured":"Haotian Liu Chunyuan Li Qingyang Wu and Yong\u00a0Jae Lee. 2024b. Visual instruction tuning. Advances in neural information processing systems 36 (2024)."},{"key":"e_1_3_3_3_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.00853"},{"key":"e_1_3_3_3_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00270"},{"key":"e_1_3_3_3_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.00847"},{"key":"e_1_3_3_3_34_1","unstructured":"Zhen Liu Yao Feng Michael\u00a0J Black Derek Nowrouzezahrai Liam Paull and Weiyang Liu. 2023a. Meshdiffusion: Score-based generative 3d mesh modeling. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2303.08133 (2023)."},{"key":"e_1_3_3_3_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00286"},{"key":"e_1_3_3_3_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.01242"},{"key":"e_1_3_3_3_37_1","doi-asserted-by":"crossref","unstructured":"Ben Mildenhall Pratul\u00a0P Srinivasan Matthew Tancik Jonathan\u00a0T Barron Ravi Ramamoorthi and Ren Ng. 2021. Nerf: Representing scenes as neural radiance fields for view synthesis. Commun. ACM 65 1 (2021) 99\u2013106.","DOI":"10.1145\/3503250"},{"key":"e_1_3_3_3_38_1","unstructured":"Tom Monnier Jake Austin Angjoo Kanazawa Alexei Efros and Mathieu Aubry. 2023. Differentiable blocks world: Qualitative 3d decomposition by rendering primitives. Advances in Neural Information Processing Systems 36 (2023) 5791\u20135807."},{"key":"e_1_3_3_3_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.00421"},{"key":"e_1_3_3_3_40_1","first-page":"7220","volume-title":"International conference on machine learning","author":"Nash Charlie","year":"2020","unstructured":"Charlie Nash, Yaroslav Ganin, SM\u00a0Ali Eslami, and Peter Battaglia. 2020. Polygen: An autoregressive generative model of 3d meshes. In International conference on machine learning. PMLR, 7220\u20137229."},{"key":"e_1_3_3_3_41_1","unstructured":"Alex Nichol Heewoo Jun Prafulla Dhariwal Pamela Mishkin and Mark Chen. 2022. Point-e: A system for generating 3d point clouds from complex prompts. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2212.08751 (2022)."},{"key":"e_1_3_3_3_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.01059"},{"key":"e_1_3_3_3_43_1","unstructured":"Ben Poole Ajay Jain Jonathan\u00a0T Barron and Ben Mildenhall. 2022. Dreamfusion: Text-to-3d using 2d diffusion. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2209.14988 (2022)."},{"key":"e_1_3_3_3_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01225"},{"key":"e_1_3_3_3_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.00403"},{"key":"e_1_3_3_3_46_1","doi-asserted-by":"publisher","DOI":"10.1111\/cgf.14601"},{"key":"e_1_3_3_3_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00578"},{"key":"e_1_3_3_3_48_1","unstructured":"Gopal Sharma Rishabh Goyal Difan Liu Evangelos Kalogerakis and Subhransu Maji. 2019. Neural shape parsers for constructive solid geometry. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1912.11393 (2019)."},{"key":"e_1_3_3_3_49_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58571-6_16"},{"key":"e_1_3_3_3_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.02000"},{"key":"e_1_3_3_3_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.01855"},{"key":"e_1_3_3_3_52_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00064"},{"key":"e_1_3_3_3_53_1","unstructured":"Jiaming Song Chenlin Meng and Stefano Ermon. 2020. Denoising diffusion implicit models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2010.02502 (2020)."},{"key":"e_1_3_3_3_54_1","doi-asserted-by":"crossref","unstructured":"Chun-Yu Sun Qian-Fang Zou Xin Tong and Yang Liu. 2019. Learning adaptive hierarchical cuboid abstractions of 3d shape collections. ACM Transactions on Graphics (TOG) 38 6 (2019) 1\u201313.","DOI":"10.1145\/3355089.3356529"},{"key":"e_1_3_3_3_55_1","unstructured":"Zhicong Tang Shuyang Gu Chunyu Wang Ting Zhang Jianmin Bao Dong Chen and Baining Guo. 2023. Volumediffusion: Flexible text-to-3d generation with efficient volumetric encoder. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2312.11459 (2023)."},{"key":"e_1_3_3_3_56_1","unstructured":"Dmitry Tochilkin David Pankratz Zexiang Liu Zixuan Huang Adam Letts Yangguang Li Ding Liang Christian Laforte Varun Jampani and Yan-Pei Cao. 2024. Triposr: Fast 3d object reconstruction from a single image. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2403.02151 (2024)."},{"key":"e_1_3_3_3_57_1","unstructured":"Hugo Touvron Louis Martin Kevin Stone Peter Albert Amjad Almahairi Yasmine Babaei Nikolay Bashlykov Soumya Batra Prajjwal Bhargava Shruti Bhosale et\u00a0al. 2023. Llama 2: Open foundation and fine-tuned chat models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2307.09288 (2023)."},{"key":"e_1_3_3_3_58_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.160"},{"key":"e_1_3_3_3_59_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.00203"},{"key":"e_1_3_3_3_60_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01155"},{"key":"e_1_3_3_3_61_1","doi-asserted-by":"crossref","unstructured":"Narunas Vaskevicius and Andreas Birk. 2017. Revisiting superquadric fitting: A numerically stable formulation. IEEE transactions on pattern analysis and machine intelligence 41 1 (2017) 220\u2013233.","DOI":"10.1109\/TPAMI.2017.2779493"},{"key":"e_1_3_3_3_62_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.00443"},{"key":"e_1_3_3_3_63_1","unstructured":"Xinlong Wang Xiaosong Zhang Zhengxiong Luo Quan Sun Yufeng Cui Jinsheng Wang Fan Zhang Yueze Wang Zhen Li Qiying Yu et\u00a0al. 2024b. Emu3: Next-token prediction is all you need. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2409.18869 (2024)."},{"key":"e_1_3_3_3_64_1","unstructured":"Zhengyi Wang Jonathan Lorraine Yikai Wang Hang Su Jun Zhu Sanja Fidler and Xiaohui Zeng. 2024a. LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2411.09595 (2024)."},{"key":"e_1_3_3_3_65_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01539"},{"key":"e_1_3_3_3_66_1","doi-asserted-by":"crossref","unstructured":"Karl\u00a0DD Willis Yewen Pu Jieliang Luo Hang Chu Tao Du Joseph\u00a0G Lambourne Armando Solar-Lezama and Wojciech Matusik. 2021. Fusion 360 gallery: A dataset and environment for programmatic cad construction from human design sequences. ACM Transactions on Graphics (TOG) 40 4 (2021) 1\u201324.","DOI":"10.1145\/3450626.3459818"},{"key":"e_1_3_3_3_67_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00670"},{"key":"e_1_3_3_3_68_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-19812-0_28"},{"key":"e_1_3_3_3_69_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.00820"},{"key":"e_1_3_3_3_70_1","doi-asserted-by":"crossref","unstructured":"Jianfeng Xiang Zelong Lv Sicheng Xu Yu Deng Ruicheng Wang Bowen Zhang Dong Chen Xin Tong and Jiaolong Yang. 2024. Structured 3D Latents for Scalable and Versatile 3D Generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2412.01506 (2024).","DOI":"10.1109\/CVPR52734.2025.02000"},{"key":"e_1_3_3_3_71_1","doi-asserted-by":"crossref","unstructured":"Bojun Xiong Si-Tong Wei Xin-Yang Zheng Yan-Pei Cao Zhouhui Lian and Peng-Shuai Wang. 2024. OctFusion: Octree-based Diffusion Models for 3D Shape Generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.14732 (2024).","DOI":"10.1111\/cgf.70198"},{"key":"e_1_3_3_3_72_1","unstructured":"Jiale Xu Weihao Cheng Yiming Gao Xintao Wang Shenghua Gao and Ying Shan. 2024a. Instantmesh: Efficient 3d mesh generation from a single image with sparse-view large reconstruction models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2404.07191 (2024)."},{"key":"e_1_3_3_3_73_1","unstructured":"Jingwei Xu Chenyu Wang Zibo Zhao Wen Liu Yi Ma and Shenghua Gao. 2024b. CAD-MLLM: Unifying Multimodality-Conditioned CAD Generation With MLLM. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2411.04954 (2024)."},{"key":"e_1_3_3_3_74_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.00827"},{"key":"e_1_3_3_3_75_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00600"},{"key":"e_1_3_3_3_76_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00275"},{"key":"e_1_3_3_3_77_1","doi-asserted-by":"crossref","unstructured":"Kaizhi Yang and Xuejin Chen. 2021. Unsupervised learning for cuboid shape abstraction via joint segmentation from point clouds. ACM Transactions On Graphics (TOG) 40 4 (2021) 1\u201311.","DOI":"10.1145\/3450626.3459873"},{"key":"e_1_3_3_3_78_1","unstructured":"Lei Yang Yongqing Liang Xin Li Congyi Zhang Guying Lin Alla Sheffer Scott Schaefer John Keyser and Wenping Wang. 2023. Neural parametric surfaces for shape modeling. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2309.09911 (2023)."},{"key":"e_1_3_3_3_79_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01147"},{"key":"e_1_3_3_3_80_1","first-page":"454","volume-title":"European Conference on Computer Vision","author":"Yu Fenggen","year":"2024","unstructured":"Fenggen Yu, Yiming Qian, Xu Zhang, Francisca Gil-Ureta, Brian Jackson, Eric Bennett, and Hao Zhang. 2024. Dpa-net: Structured 3d abstraction from sparse views via differentiable primitive assembly. In European Conference on Computer Vision. Springer, 454\u2013471."},{"key":"e_1_3_3_3_81_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-72630-9_27"},{"key":"e_1_3_3_3_82_1","unstructured":"Bowen Zhang Yiji Cheng Jiaolong Yang Chunyu Wang Feng Zhao Yansong Tang Dong Chen and Baining Guo. 2024a. GaussianCube: Structuring Gaussian Splatting using Optimal Transport for 3D Generative Modeling. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2403.19655 (2024)."},{"key":"e_1_3_3_3_83_1","doi-asserted-by":"crossref","unstructured":"Biao Zhang Jiapeng Tang Matthias Niessner and Peter Wonka. 2023. 3dshape2vecset: A 3d shape representation for neural fields and generative diffusion models. ACM Transactions on Graphics (TOG) 42 4 (2023) 1\u201316.","DOI":"10.1145\/3592442"},{"key":"e_1_3_3_3_84_1","doi-asserted-by":"crossref","unstructured":"Longwen Zhang Ziyu Wang Qixuan Zhang Qiwei Qiu Anqi Pang Haoran Jiang Wei Yang Lan Xu and Jingyi Yu. 2024c. CLAY: A Controllable Large-scale Generative Model for Creating High-quality 3D Assets. ACM Transactions on Graphics (TOG) 43 4 (2024) 1\u201320.","DOI":"10.1145\/3658146"},{"key":"e_1_3_3_3_85_1","unstructured":"Yunzhi Zhang Zizhang Li Matt Zhou Shangzhe Wu and Jiajun Wu. 2024b. The scene language: Representing scenes with programs words and embeddings. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2410.16770 (2024)."},{"key":"e_1_3_3_3_86_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-72913-3_17"},{"key":"e_1_3_3_3_87_1","unstructured":"Zibo Zhao Zeqiang Lai Qingxiang Lin Yunfei Zhao Haolin Liu Shuhui Yang Yifei Feng Mingxin Yang Sheng Zhang Xianghui Yang et\u00a0al. 2025a. Hunyuan3d 2.0: Scaling diffusion models for high resolution textured 3d assets generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2501.12202 (2025)."},{"key":"e_1_3_3_3_88_1","unstructured":"Zibo Zhao Wen Liu Xin Chen Xianfang Zeng Rui Wang Pei Cheng Bin Fu Tao Chen Gang Yu and Shenghua Gao. 2024. Michelangelo: Conditional 3d shape generation based on shape-image-text aligned latent representation. Advances in Neural Information Processing Systems 36 (2024)."},{"key":"e_1_3_3_3_89_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00577"},{"key":"e_1_3_3_3_90_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.103"}],"event":{"name":"SA Conference Papers '25: SIGGRAPH Asia 2025 Conference Papers","location":"Hong Kong Hong Kong","acronym":"SA Conference Papers '25","sponsor":["SIGGRAPH ACM Special Interest Group on Computer Graphics and Interactive Techniques"]},"container-title":["Proceedings of the SIGGRAPH Asia 2025 Conference Papers"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3757377.3763857","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,12,9]],"date-time":"2025-12-09T03:28:58Z","timestamp":1765250938000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3757377.3763857"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,14]]},"references-count":89,"alternative-id":["10.1145\/3757377.3763857","10.1145\/3757377"],"URL":"https:\/\/doi.org\/10.1145\/3757377.3763857","relation":{},"subject":[],"published":{"date-parts":[[2025,12,14]]},"assertion":[{"value":"2025-12-14","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}