{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,8]],"date-time":"2026-05-08T16:37:42Z","timestamp":1778258262283,"version":"3.51.4"},"reference-count":56,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2014,5,1]],"date-time":"2014-05-01T00:00:00Z","timestamp":1398902400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000006","name":"Office of Naval Research","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100000006","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100000001","name":"National Science Foundation","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100000145","name":"Division of Information and Intelligent Systems","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100000145","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2014,5]]},"abstract":"<jats:p>We present a user-friendly image editing system that supports a drag-and-drop object insertion (where the user merely drags objects into the image, and the system automatically places them in 3D and relights them appropriately), postprocess illumination editing, and depth-of-field manipulation. Underlying our system is a fully automatic technique for recovering a comprehensive 3D scene model (geometry, illumination, diffuse albedo, and camera parameters) from a single, low dynamic range photograph. This is made possible by two novel contributions: an illumination inference algorithm that recovers a full lighting model of the scene (including light sources that are not directly visible in the photograph), and a depth estimation algorithm that combines data-driven depth transfer with geometric reasoning about the scene layout. A user study shows that our system produces perceptually convincing results, and achieves the same level of realism as techniques that require significant user interaction.<\/jats:p>","DOI":"10.1145\/2602146","type":"journal-article","created":{"date-parts":[[2014,6,10]],"date-time":"2014-06-10T12:50:17Z","timestamp":1402404617000},"page":"1-15","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":134,"title":["Automatic Scene Inference for 3D Object Compositing"],"prefix":"10.1145","volume":"33","author":[{"given":"Kevin","family":"Karsch","sequence":"first","affiliation":[{"name":"University of Illinois"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Kalyan","family":"Sunkavalli","sequence":"additional","affiliation":[{"name":"Adobe Research"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Sunil","family":"Hadap","sequence":"additional","affiliation":[{"name":"Adobe Research"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Nathan","family":"Carr","sequence":"additional","affiliation":[{"name":"Adobe Research"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hailin","family":"Jin","sequence":"additional","affiliation":[{"name":"Adobe Research"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Rafael","family":"Fonte","sequence":"additional","affiliation":[{"name":"University of Illinois"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Michael","family":"Sittig","sequence":"additional","affiliation":[{"name":"University of Illinois"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"David","family":"Forsyth","sequence":"additional","affiliation":[{"name":"University of Illinois"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2014,6,2]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2012.120"},{"key":"e_1_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.10"},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/383259.383270"},{"key":"e_1_2_2_4_1","volume-title":"Proceedings of the Annual ACM SIGGRAPH Conference on Computer Graphics and Interactive Techniques.","author":"Boyadzhiev I.","unstructured":"I. Boyadzhiev , S. Paris , and K. Bala . 2013. Example-based synthesis of 3d object arrangements . In Proceedings of the Annual ACM SIGGRAPH Conference on Computer Graphics and Interactive Techniques. I. Boyadzhiev, S. Paris, and K. Bala. 2013. Example-based synthesis of 3d object arrangements. In Proceedings of the Annual ACM SIGGRAPH Conference on Computer Graphics and Interactive Techniques."},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1026598000963"},{"key":"e_1_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/280814.280864"},{"key":"e_1_2_2_7_1","volume-title":"Proceedings of the International Symposium on Virtual Reality, Archaeology, and Culturage Heritage.","author":"Debevec P.","year":"2005","unstructured":"P. Debevec . 2005 . Making \u201cthe parthenon \u201d. In Proceedings of the International Symposium on Virtual Reality, Archaeology, and Culturage Heritage. P. Debevec. 2005. Making \u201cthe parthenon\u201d. In Proceedings of the International Symposium on Virtual Reality, Archaeology, and Culturage Heritage."},{"key":"e_1_2_2_8_1","volume-title":"Proceedings of the International Symposium on Robotics Research (ISRR'05)","author":"Delage E.","unstructured":"E. Delage , H. Lee , and A. Y. Ng . 2005. Automatic single-image 3d reconstructions of indoor manhattan world scenes . In Proceedings of the International Symposium on Robotics Research (ISRR'05) . 305--321. E. Delage, H. Lee, and A. Y. Ng. 2005. Automatic single-image 3d reconstructions of indoor manhattan world scenes. In Proceedings of the International Symposium on Robotics Research (ISRR'05). 305--321."},{"key":"e_1_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.27"},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10851-013-0442-7"},{"key":"e_1_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1167\/4.9.11"},{"key":"e_1_2_2_12_1","volume-title":"Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'09)","author":"Furukawa Y.","unstructured":"Y. Furukawa , B. Curless , S. M. Seitz , and R. Szeliski . 2009. Manhattan-world stereo . In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'09) . IEEE, 1422--1429. Y. Furukawa, B. Curless, S. M. Seitz, and R. Szeliski. 2009. Manhattan-world stereo. In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'09). IEEE, 1422--1429."},{"key":"e_1_2_2_13_1","volume-title":"Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'10)","author":"Gallup D.","unstructured":"D. Gallup , J.-M. Frahm , and M. Pollefeys . 2010. Piecewise planar and non-planar stereo for urban scene reconstruction . In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'10) . D. Gallup, J.-M. Frahm, and M. Pollefeys. 2010. Piecewise planar and non-planar stereo for urban scene reconstruction. In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'10)."},{"key":"e_1_2_2_14_1","volume-title":"Proceedings of the Eurographics Symposium on Rendering (EGSR'00)","author":"Gibson S.","unstructured":"S. Gibson and A. Murta . 2000. Interactive rendering with real-world illumination . In Proceedings of the Eurographics Symposium on Rendering (EGSR'00) . Springer, 365--376. S. Gibson and A. Murta. 2000. Interactive rendering with real-world illumination. In Proceedings of the Eurographics Symposium on Rendering (EGSR'00). Springer, 365--376."},{"key":"e_1_2_2_15_1","volume-title":"Proceedings of the International Conference on Computer Vision (ICCV'09)","author":"Grosse R.","unstructured":"R. Grosse , M. K. Johnson , E. H. Adelson , and W. Freeman . 2009. Ground truth dataset and baseline evaluations for intrinsic image algorithms . In Proceedings of the International Conference on Computer Vision (ICCV'09) . R. Grosse, M. K. Johnson, E. H. Adelson, and W. Freeman. 2009. Ground truth dataset and baseline evaluations for intrinsic image algorithms. In Proceedings of the International Conference on Computer Vision (ICCV'09)."},{"key":"e_1_2_2_16_1","doi-asserted-by":"crossref","unstructured":"R. Hartley and A. Zisserman. 2003. Multiple View Geometry in Computer Vision. Cambridge University Press. R. Hartley and A. Zisserman. 2003. Multiple View Geometry in Computer Vision. Cambridge University Press.","DOI":"10.1017\/CBO9780511811685"},{"key":"e_1_2_2_17_1","volume-title":"Proceedings of the International Conference on Computer Vision (ICCV'09)","author":"Hedau V.","unstructured":"V. Hedau , D. Hoiem , and D. Forsyth . 2009. Recovering the spatial layout of cluttered rooms . In Proceedings of the International Conference on Computer Vision (ICCV'09) . V. Hedau, D. Hoiem, and D. Forsyth. 2009. Recovering the spatial layout of cluttered rooms. In Proceedings of the International Conference on Computer Vision (ICCV'09)."},{"key":"e_1_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2005.107"},{"key":"e_1_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/1073204.1073232"},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/258734.258854"},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1037\/0278-7393.15.2.179"},{"key":"e_1_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1145\/1150402.1150429"},{"key":"e_1_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/1073170.1073171"},{"key":"e_1_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIFS.2007.903848"},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/2024156.2024191"},{"key":"e_1_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33715-4_56"},{"key":"e_1_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/1179352.1141937"},{"key":"e_1_2_2_28_1","volume-title":"Proceedings of the International Conference on Computer Vision (ICCV'09)","author":"Lalonde J.","unstructured":"J. Lalonde , A. A. Efros , and S. Narasimhan . 2009. Estimating natural illumination from a single outdoor image . In Proceedings of the International Conference on Computer Vision (ICCV'09) . J. Lalonde, A. A. Efros, and S. Narasimhan. 2009. Estimating natural illumination from a single outdoor image. In Proceedings of the International Conference on Computer Vision (ICCV'09)."},{"key":"e_1_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/1275808.1276381"},{"key":"e_1_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2006.68"},{"key":"e_1_2_2_31_1","volume-title":"Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'09)","author":"Lee D. C.","unstructured":"D. C. Lee , M. Hebert , and T. Kanade . 2009. Geometric reasoning for single image structure recovery . In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'09) . 2136--2143. D. C. Lee, M. Hebert, and T. Kanade. 2009. Geometric reasoning for single image structure recovery. In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'09). 2136--2143."},{"key":"e_1_2_2_32_1","volume-title":"Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'10)","author":"Liu B.","unstructured":"B. Liu , S. Gould , and D. Koller . 2010. Single image depth estimation from predicted semantic labels . In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'10) . 1253--1260. B. Liu, S. Gould, and D. Koller. 2010. Single image depth estimation from predicted semantic labels. In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'10). 1253--1260."},{"key":"e_1_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33783-3_42"},{"key":"e_1_2_2_34_1","volume-title":"Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'12)","author":"Lombardi S.","unstructured":"S. Lombardi and K. Nishino . 2012b. Single image multimaterial estimation . In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'12) . S. Lombardi and K. Nishino. 2012b. Single image multimaterial estimation. In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'12)."},{"key":"e_1_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cag.2010.08.004"},{"key":"e_1_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.5555\/2383815.2383845"},{"key":"e_1_2_2_37_1","volume-title":"Proceedings of the Eurographics Symposium on Rendering (EGSR). 359--373","author":"Nimeroff J. S.","unstructured":"J. S. Nimeroff , E. Simoncelli , and J. Dorsey . 1994. Efficient rerendering of naturally illuminated environments . In Proceedings of the Eurographics Symposium on Rendering (EGSR). 359--373 . J. S. Nimeroff, E. Simoncelli, and J. Dorsey. 1994. Efficient rerendering of naturally illuminated environments. In Proceedings of the Eurographics Symposium on Rendering (EGSR). 359--373."},{"key":"e_1_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1145\/1015706.1015783"},{"key":"e_1_2_2_39_1","unstructured":"J. Nocedal and S. J. Wright. 2006. Numerical Optimization 2nd Ed. Springer. J. Nocedal and S. J. Wright. 2006. Numerical Optimization 2 nd Ed. Springer."},{"key":"e_1_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/383259.383310"},{"key":"e_1_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1011139631724"},{"key":"e_1_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2011.5995585"},{"key":"e_1_2_2_43_1","unstructured":"M. Pharr and G. Humphreys. 2010. Physically Based Rendering: From Theory to Implementation 2nd Ed. Morgan Kaufmann San Fransisco. M. Pharr and G. Humphreys. 2010. Physically Based Rendering: From Theory to Implementation 2 nd Ed. Morgan Kaufmann San Fransisco."},{"key":"e_1_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1145\/1027411.1027416"},{"key":"e_1_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1145\/1276377.1276472"},{"key":"e_1_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-88693-8_63"},{"key":"e_1_2_2_47_1","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV'10)","author":"Romeiro F.","unstructured":"F. Romeiro and T. Zickler . 2010. Blind reflectometry . In Proceedings of the European Conference on Computer Vision (ECCV'10) . F. Romeiro and T. Zickler. 2010. Blind reflectometry. In Proceedings of the European Conference on Computer Vision (ECCV'10)."},{"key":"e_1_2_2_48_1","volume-title":"Proceedings of the 2nd British Machine Vision Conference.","author":"Satkin S.","unstructured":"S. Satkin , J. Lin , and M. Hebert . 2012. Data-driven scene understanding from 3d models . In Proceedings of the 2nd British Machine Vision Conference. S. Satkin, J. Lin, and M. Hebert. 2012. Data-driven scene understanding from 3d models. In Proceedings of the 2nd British Machine Vision Conference."},{"key":"e_1_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2008.132"},{"key":"e_1_2_2_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/166117.166135"},{"key":"e_1_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33783-3_22"},{"key":"e_1_2_2_52_1","doi-asserted-by":"crossref","first-page":"267","DOI":"10.1111\/j.2517-6161.1996.tb02080.x","article-title":"Regression shrinkage and selection via the lasso","volume":"58","author":"Tibshirani R.","year":"1996","unstructured":"R. Tibshirani . 1996 . Regression shrinkage and selection via the lasso . J. Roy. Statist. Soc. B58 , 1, 267 -- 288 . R. Tibshirani. 1996. Regression shrinkage and selection via the lasso. J. Roy. Statist. Soc. B58, 1, 267--288.","journal-title":"J. Roy. Statist. Soc."},{"key":"e_1_2_2_53_1","volume-title":"Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'12)","author":"Xiao J.","unstructured":"J. Xiao , K. A. Ehinger , A. Oliva , and A. Torralba . 2012. Recognizing scene viewpoint using panoramic place representation . In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'12) . J. Xiao, K. A. Ehinger, A. Oliva, and A. Torralba. 2012. Recognizing scene viewpoint using panoramic place representation. In Proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR'12)."},{"key":"e_1_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1145\/311535.311559"},{"key":"e_1_2_2_55_1","doi-asserted-by":"publisher","DOI":"10.1145\/2407156.2407188"},{"key":"e_1_2_2_56_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.155"}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2602146","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2602146","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T07:00:47Z","timestamp":1750230047000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2602146"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2014,5]]},"references-count":56,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2014,5]]}},"alternative-id":["10.1145\/2602146"],"URL":"https:\/\/doi.org\/10.1145\/2602146","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"value":"0730-0301","type":"print"},{"value":"1557-7368","type":"electronic"}],"subject":[],"published":{"date-parts":[[2014,5]]},"assertion":[{"value":"2013-07-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2014-02-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2014-06-02","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}