{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,26]],"date-time":"2026-02-26T07:20:43Z","timestamp":1772090443341,"version":"3.50.1"},"reference-count":39,"publisher":"Association for Computing Machinery (ACM)","issue":"6","license":[{"start":{"date-parts":[[2012,11,1]],"date-time":"2012-11-01T00:00:00Z","timestamp":1351728000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100002855","name":"Ministry of Science and Technology of the People's Republic of China","doi-asserted-by":"publisher","award":["2009CB320801"],"award-info":[{"award-number":["2009CB320801"]}],"id":[{"id":"10.13039\/501100002855","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["6.13E+15"],"award-info":[{"award-number":["6.13E+15"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2012,11]]},"abstract":"<jats:p>We present an interactive approach to semantic modeling of indoor scenes with a consumer-level RGBD camera. Using our approach, the user first takes an RGBD image of an indoor scene, which is automatically segmented into a set of regions with semantic labels. If the segmentation is not satisfactory, the user can draw some strokes to guide the algorithm to achieve better results. After the segmentation is finished, the depth data of each semantic region is used to retrieve a matching 3D model from a database. Each model is then transformed according to the image depth to yield the scene. For large scenes where a single image can only cover one part of the scene, the user can take multiple images to construct other parts of the scene. The 3D models built for all images are then transformed and unified into a complete scene. We demonstrate the efficiency and robustness of our approach by modeling several real-world scenes.<\/jats:p>","DOI":"10.1145\/2366145.2366155","type":"journal-article","created":{"date-parts":[[2012,11,14]],"date-time":"2012-11-14T20:36:17Z","timestamp":1352925377000},"page":"1-11","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":176,"title":["An interactive approach to semantic modeling of indoor scenes with an RGBD camera"],"prefix":"10.1145","volume":"31","author":[{"given":"Tianjia","family":"Shao","sequence":"first","affiliation":[{"name":"Tsinghua University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Weiwei","family":"Xu","sequence":"additional","affiliation":[{"name":"Hangzhou Normal University and Microsoft Research Asia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Kun","family":"Zhou","sequence":"additional","affiliation":[{"name":"Zhejiang University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jingdong","family":"Wang","sequence":"additional","affiliation":[{"name":"Microsoft Research Asia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Dongping","family":"Li","sequence":"additional","affiliation":[{"name":"Zhejiang University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Baining","family":"Guo","sequence":"additional","affiliation":[{"name":"Microsoft Research Asia"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2012,11]]},"reference":[{"key":"e_1_2_1_1_1","unstructured":"Anand A. Koppula H. S. Joachims T. and Saxena A. 2011. Contextually guided semantic labeling and search for 3d point clouds. CoRR abs\/1111.5358.  Anand A. Koppula H. S. Joachims T. and Saxena A. 2011. Contextually guided semantic labeling and search for 3d point clouds. CoRR abs\/1111.5358 ."},{"key":"e_1_2_1_2_1","volume-title":"NIPS Workshop on Deep Learning and Unsupervised Feature Learning.","author":"Blum M.","unstructured":"Blum , M. , Springenberg , J. T. , Wulfing , J. , and Riedmiller , M . 2011. On the applicability of unsupervised feature learning for object recognition in rgb-d data . In NIPS Workshop on Deep Learning and Unsupervised Feature Learning. Blum, M., Springenberg, J. T., Wulfing, J., and Riedmiller, M. 2011. On the applicability of unsupervised feature learning for object recognition in rgb-d data. In NIPS Workshop on Deep Learning and Unsupervised Feature Learning."},{"key":"e_1_2_1_3_1","volume-title":"IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS).","author":"Bo L.","unstructured":"Bo , L. , Ren , X. , and Fox , D . 2011. Depth kernel descriptors for object recognition . In IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS). Bo, L., Ren, X., and Fox, D. 2011. Depth kernel descriptors for object recognition. In IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS)."},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.969114"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1010933404324"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/1899404.1899405"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/MC.2011.320"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1007607513941"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/2030112.2030123"},{"key":"e_1_2_1_10_1","volume-title":"Proceedings of the 33rd international conference on Pattern recognition, Springer-Verlag, Berlin, Heidelberg, DAGM'11, 101--110","author":"Fanelli G.","unstructured":"Fanelli , G. , Weise , T. , Gall , J. , and Gool , L. V . 2011. Real time head pose estimation from consumer depth cameras . In Proceedings of the 33rd international conference on Pattern recognition, Springer-Verlag, Berlin, Heidelberg, DAGM'11, 101--110 . Fanelli, G., Weise, T., Gall, J., and Gool, L. V. 2011. Real time head pose estimation from consumer depth cameras. In Proceedings of the 33rd international conference on Pattern recognition, Springer-Verlag, Berlin, Heidelberg, DAGM'11, 101--110."},{"key":"e_1_2_1_11_1","volume-title":"Proceedings of AAAI'99\/IAAI'99","author":"Fox D.","unstructured":"Fox , D. , Burgard , W. , Dellaert , F. , and Thrun , S . 1999. Monte carlo localization: efficient position estimation for mobile robots . In Proceedings of AAAI'99\/IAAI'99 , American Association for Artificial Intelligence, Menlo Park, CA, USA, AAAI '99\/IAAI '99, 343--349. Fox, D., Burgard, W., Dellaert, F., and Thrun, S. 1999. Monte carlo localization: efficient position estimation for mobile robots. In Proceedings of AAAI'99\/IAAI'99, American Association for Artificial Intelligence, Menlo Park, CA, USA, AAAI '99\/IAAI '99, 343--349."},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/588272.588279"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2009.161"},{"key":"e_1_2_1_14_1","doi-asserted-by":"crossref","unstructured":"Furukawa Y. Curless B. Seitz S. M. and Szeliski R. 2009. Manhattan-world stereo. CVPR 1422--1429.  Furukawa Y. Curless B. Seitz S. M. and Szeliski R. 2009. Manhattan-world stereo. CVPR 1422--1429.","DOI":"10.1109\/CVPR.2009.5206867"},{"key":"e_1_2_1_15_1","doi-asserted-by":"crossref","unstructured":"Furukawa Y. Curless B. Seitz S. M. and Szeliski R. 2009. Reconstructing Building Interiors from Images. In ICCV.  Furukawa Y. Curless B. Seitz S. M. and Szeliski R. 2009. Reconstructing Building Interiors from Images. In ICCV .","DOI":"10.1109\/ICCV.2009.5459145"},{"key":"e_1_2_1_16_1","doi-asserted-by":"crossref","unstructured":"Golovinskiy A. Kim V. and Funkhouser T. 2009. Shape-based recognition of 3d point clouds in urban environments. In ICCV.  Golovinskiy A. Kim V. and Funkhouser T. 2009. Shape-based recognition of 3d point clouds in urban environments. In ICCV .","DOI":"10.1109\/ICCV.2009.5459471"},{"key":"e_1_2_1_17_1","doi-asserted-by":"crossref","unstructured":"Hartley R. I. and Zisserman A. 2004. Multiple View Geometry in Computer Vision second ed. Cambridge University Press.   Hartley R. I. and Zisserman A. 2004. Multiple View Geometry in Computer Vision second ed. Cambridge University Press.","DOI":"10.1017\/CBO9780511811685"},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1177\/0278364911434148"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/2047196.2047270"},{"key":"e_1_2_1_20_1","volume-title":"ICCV Workshop on Consumer Depth Cameras for Computer Vision, 1168--1174","author":"Janoch A.","unstructured":"Janoch , A. , Karayev , S. , Jia , Y. , Barron , J. T. , Fritz , M. , Saenko , K. , and Darrell , T . 2011. A category-level 3-d object dataset: Putting the kinect to work . In ICCV Workshop on Consumer Depth Cameras for Computer Vision, 1168--1174 . Janoch, A., Karayev, S., Jia, Y., Barron, J. T., Fritz, M., Saenko, K., and Darrell, T. 2011. A category-level 3-d object dataset: Putting the kinect to work. In ICCV Workshop on Consumer Depth Cameras for Computer Vision, 1168--1174."},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.765655"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1145\/2366145.2366157"},{"key":"e_1_2_1_23_1","unstructured":"Koppula H. Anand A. Joachims T. and Saxena A. 2011. Semantic labeling of 3d point clouds for indoor scenes. In NIPS.  Koppula H. Anand A. Joachims T. and Saxena A. 2011. Semantic labeling of 3d point clouds for indoor scenes. In NIPS ."},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1177\/0278364910369190"},{"key":"e_1_2_1_25_1","volume-title":"Proc. of IEEE International Conference on Robotics and Automation.","author":"Lai K.","unstructured":"Lai , K. , Bo , L. , Ren , X. , and Fox , D . 2011. A large-scale hierarchical multi-view rgb-d object dataset . In Proc. of IEEE International Conference on Robotics and Automation. Lai, K., Bo, L., Ren, X., and Fox, D. 2011. A large-scale hierarchical multi-view rgb-d object dataset. In Proc. of IEEE International Conference on Robotics and Automation."},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/1015706.1015719"},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/2010324.1964947"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1023\/B:VISI.0000029664.99615.94"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/2010324.1964982"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/2366145.2366156"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.1467-8659.2007.01016.x"},{"key":"e_1_2_1_32_1","volume-title":"Proceedings of the International Conference on Computer Vision - Workshop on 3D Representation and Recognition.","author":"Silberman N.","unstructured":"Silberman , N. , and Fergus , R . 2011. Indoor scene segmentation using a structured light sensor . In Proceedings of the International Conference on Computer Vision - Workshop on 3D Representation and Recognition. Silberman, N., and Fergus, R. 2011. Indoor scene segmentation using a structured light sensor. In Proceedings of the International Conference on Computer Vision - Workshop on 3D Representation and Recognition."},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33715-4_54"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/1141911.1141964"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11042-007-0181-0"},{"key":"e_1_2_1_36_1","doi-asserted-by":"crossref","unstructured":"Triggs B. Mclauchlan P. Hartley R. and Fitzgibbon A. 2000. Bundle adjustment a modern synthesis. In Vision Algorithms: Theory and Practice LNCS Springer Verlag 298--375.   Triggs B. Mclauchlan P. Hartley R. and Fitzgibbon A. 2000. Bundle adjustment a modern synthesis. In Vision Algorithms: Theory and Practice LNCS Springer Verlag 298--375.","DOI":"10.1007\/3-540-44480-7_21"},{"key":"e_1_2_1_37_1","volume-title":"Proceedings of the 2nd international conference on 3-D digital imaging and modeling, 3DIM'99","author":"Whitaker R. T.","unstructured":"Whitaker , R. T. , Gregor , J. , and Chen , P. F . 1999. Indoor scene reconstruction from sets of noisy range images . In Proceedings of the 2nd international conference on 3-D digital imaging and modeling, 3DIM'99 , 348--357. Whitaker, R. T., Gregor, J., and Chen, P. F. 1999. Indoor scene reconstruction from sets of noisy range images. In Proceedings of the 2nd international conference on 3-D digital imaging and modeling, 3DIM'99, 348--357."},{"key":"e_1_2_1_38_1","doi-asserted-by":"crossref","unstructured":"Xiong X. and Huber D. 2010. Using context to create semantic 3d models of indoor environments. In BMVC 1--11.  Xiong X. and Huber D. 2010. Using context to create semantic 3d models of indoor environments. In BMVC 1--11.","DOI":"10.5244\/C.24.45"},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/2010324.1964981"}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2366145.2366155","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2366145.2366155","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T09:34:05Z","timestamp":1750239245000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2366145.2366155"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2012,11]]},"references-count":39,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2012,11]]}},"alternative-id":["10.1145\/2366145.2366155"],"URL":"https:\/\/doi.org\/10.1145\/2366145.2366155","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"value":"0730-0301","type":"print"},{"value":"1557-7368","type":"electronic"}],"subject":[],"published":{"date-parts":[[2012,11]]},"assertion":[{"value":"2012-11-01","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}