{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,12]],"date-time":"2026-07-12T15:06:47Z","timestamp":1783868807522,"version":"3.55.0"},"reference-count":60,"publisher":"Springer Science and Business Media LLC","issue":"9","license":[{"start":{"date-parts":[[2024,4,24]],"date-time":"2024-04-24T00:00:00Z","timestamp":1713916800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2024,4,24]],"date-time":"2024-04-24T00:00:00Z","timestamp":1713916800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"EPFL Lausanne"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Int J Comput Vis"],"published-print":{"date-parts":[[2024,9]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>\n                    Wildlife observation with camera traps has great potential for ethology and ecology, as it gathers data non-invasively in an automated way. However, camera traps produce large amounts of uncurated data, which is time-consuming to annotate. Existing methods to label these data automatically commonly use a fixed pre-defined set of distinctive classes and require many labeled examples per class to be trained. Moreover, the attributes of interest are sometimes rare and difficult to find in large data collections. Large pretrained vision-language models, such as contrastive language image pretraining (CLIP), offer great promises to facilitate the annotation process of camera-trap data. Images can be described with greater detail, the set of classes is not fixed and can be extensible on demand and pretrained models can help to retrieve rare samples. In this work, we explore the potential of CLIP to retrieve images according to environmental and ecological attributes. We create WildCLIP by fine-tuning CLIP on wildlife camera-trap images and to further increase its flexibility, we add an adapter module to better expand to novel attributes in a few-shot manner. We quantify WildCLIP\u2019s performance and show that it can retrieve novel attributes in the Snapshot Serengeti dataset. Our findings outline new opportunities to facilitate annotation processes with complex and multi-attribute captions. The code is available at\n                    <jats:ext-link xmlns:xlink=\"http:\/\/www.w3.org\/1999\/xlink\" ext-link-type=\"uri\" xlink:href=\"https:\/\/github.com\/amathislab\/wildclip\">https:\/\/github.com\/amathislab\/wildclip<\/jats:ext-link>\n                    .\n                  <\/jats:p>","DOI":"10.1007\/s11263-024-02026-6","type":"journal-article","created":{"date-parts":[[2024,4,24]],"date-time":"2024-04-24T05:02:23Z","timestamp":1713934943000},"page":"3770-3786","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":24,"title":["WildCLIP: Scene and Animal Attribute Retrieval from Camera Trap Data with Domain-Adapted Vision-Language Models"],"prefix":"10.1007","volume":"132","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-3352-1825","authenticated-orcid":false,"given":"Valentin","family":"Gabeff","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6612-5744","authenticated-orcid":false,"given":"Marc","family":"Ru\u00dfwurm","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0374-2459","authenticated-orcid":false,"given":"Devis","family":"Tuia","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3777-2202","authenticated-orcid":false,"given":"Alexander","family":"Mathis","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2024,4,24]]},"reference":[{"key":"2026_CR1","first-page":"23716","volume":"35","author":"JB Alayrac","year":"2022","unstructured":"Alayrac, J. B., Donahue, J., Luc, P., Miech, A., Barr, I., Hasson, Y., Lenc, K., Mensch, A., Millican, K., Reynolds, M., et al. (2022). Flamingo: A visual language model for few-shot learning. Advances in Neural Information Processing Systems, 35, 23716\u201323736.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"2026_CR2","unstructured":"Beery, S., Morris, D., & Yang, S. (2019). Efficient pipeline for camera trap image review. arXiv preprint arXiv:1907.06772"},{"key":"2026_CR3","unstructured":"Beery, S., Van\u00a0Horn, G., & Perona, P. (2018). In Proceedings of the European conference on computer vision (ECCV)(pp. 456\u2013473)."},{"key":"2026_CR4","doi-asserted-by":"crossref","unstructured":"Brookes, O., Mirmehdi, M., K\u00fchl, H., & Burghardt, T. (2023). Triple-stream deep metric learning of great ape behavioural actions. arXiv preprint arXiv:2301.02642","DOI":"10.5220\/0011798400003417"},{"key":"2026_CR5","first-page":"1877","volume":"33","author":"T Brown","year":"2020","unstructured":"Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al. (2020). Language models are few-shot learners. Advances in Neural Information Processing Systems, 33, 1877\u20131901.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"2026_CR6","unstructured":"Burghardt, T., & Calic, J. (2006). In 2006 8th seminar on neural network applications in electrical engineering (pp. 27\u201332). IEEE."},{"issue":"3","key":"2026_CR7","doi-asserted-by":"publisher","first-page":"675","DOI":"10.1111\/1365-2664.12432","volume":"52","author":"AC Burton","year":"2015","unstructured":"Burton, A. C., Neilson, E., Moreira, D., Ladle, A., Steenweg, R., Fisher, J. T., Bayne, E., & Boutin, S. (2015). Wildlife camera trapping: A review and recommendations for linking surveys to ecological processes. Journal of Applied Ecology, 52(3), 675\u2013685.","journal-title":"Journal of Applied Ecology"},{"issue":"6521","key":"2026_CR8","doi-asserted-by":"publisher","first-page":"1219","DOI":"10.1126\/science.abc7791","volume":"370","author":"ER Bush","year":"2020","unstructured":"Bush, E. R., Whytock, R. C., Bahaa-El-Din, L., Bourgeois, S., Bunnefeld, N., Cardoso, A. W., Dikangadissi, J. T., Dimbonda, P., Dimoto, E., Edzang Ndong, J., et al. (2020). Long-term collapse in fruit availability threatens central African forest megafauna. Science, 370(6521), 1219\u20131222.","journal-title":"Science"},{"issue":"3","key":"2026_CR9","doi-asserted-by":"publisher","first-page":"109","DOI":"10.1002\/rse2.48","volume":"3","author":"A Caravaggi","year":"2017","unstructured":"Caravaggi, A., Banks, P. B., Burton, A. C., Finlay, C. M., Haswell, P. M., Hayward, M. W., Rowcliffe, M. J., & Wood, M. D. (2017). A review of camera trapping for conservation behaviour research. Remote Sensing in Ecology and Conservation, 3(3), 109\u2013122.","journal-title":"Remote Sensing in Ecology and Conservation"},{"key":"2026_CR10","unstructured":"Chen, G., Han, T. X., He, Z., Kays, R., &\u00a0Forrester, R. (2014). In 2014 IEEE international conference on image processing (ICIP) (pp. 858\u2013862). IEEE."},{"key":"2026_CR11","doi-asserted-by":"publisher","first-page":"617,996","DOI":"10.3389\/fevo.2021.617996","volume":"9","author":"ZJ Delisle","year":"2021","unstructured":"Delisle, Z. J., Flaherty, E. A., Nobbe, M. R., Wzientek, C. M., & Swihart, R. K. (2021). Next-generation camera trapping: systematic review of historic trends suggests keys to expanded research applications in ecology and conservation. Frontiers in Ecology and Evolution, 9, 617,996.","journal-title":"Frontiers in Ecology and Evolution"},{"key":"2026_CR12","unstructured":"Devlin, J., Chang, M. W., Lee, K., & Toutanova, K. (2018). Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805"},{"key":"2026_CR13","unstructured":"Ding, Y., Liu, L., Tian, C., Yang, J., & Ding, H. (2022). Don\u2019t stop learning: Towards continual learning for the clip model. arXiv preprint arXiv:2207.09248"},{"issue":"2","key":"2026_CR14","doi-asserted-by":"publisher","first-page":"581","DOI":"10.1007\/s11263-023-01891-x","volume":"132","author":"P Gao","year":"2021","unstructured":"Gao, P., Geng, S., Zhang, R., Ma, T., Fang, R., Zhang, Y., Li, H., & Qiao, Y. (2021). Clip-adapter: Better vision-language models with feature adapters. International Journal of Computer Vision, 132(2), 581\u2013595.","journal-title":"International Journal of Computer Vision"},{"key":"2026_CR15","unstructured":"Ilharco, G., Wortsman, M., Wightman, R., Gordon, C., Carlini, N., Taori, R., Dave, A., Shankar, V., Namkoong, H., Miller, J., Hajishirzi, H., Farhadi, A., & Schmidt, L. (2021). Openclip."},{"key":"2026_CR16","unstructured":"Jia, C., Yang, Y., Xia, Y., Chen, Y. T., Parekh, Z., Pham, H., Le, Q., Sung, Y. H., Li, Z., & Duerig, T. (2021) In International conference on machine learning (PMLR, 2021) (pp. 4904\u20134916)."},{"key":"2026_CR17","doi-asserted-by":"publisher","first-page":"139","DOI":"10.1016\/j.rse.2018.06.028","volume":"216","author":"B Kellenberger","year":"2018","unstructured":"Kellenberger, B., Marcos, D., & Tuia, D. (2018). Detecting mammals in UAV images: Best practices to address a substantially imbalanced dataset with deep learning. Remote Sensing of Environment, 216, 139\u2013153.","journal-title":"Remote Sensing of Environment"},{"issue":"12","key":"2026_CR18","doi-asserted-by":"publisher","first-page":"1716","DOI":"10.1111\/2041-210X.13489","volume":"11","author":"B Kellenberger","year":"2020","unstructured":"Kellenberger, B., Tuia, D., & Morris, D. (2020). Aide: Accelerating image-based ecological surveys with interactive machine learning. Methods in Ecology and Evolution, 11(12), 1716\u20131727.","journal-title":"Methods in Ecology and Evolution"},{"key":"2026_CR19","unstructured":"Kinney, R., Anastasiades, C., Authur, R., Beltagy, I., Bragg, J., Buraczynski, A., Cachola, I., Candra, S., Chandrasekhar, Y., & Cohan, A. et\u00a0al. (2023). The semantic scholar open data platform. arXiv preprint arXiv:2301.10140"},{"issue":"13","key":"2026_CR20","doi-asserted-by":"publisher","first-page":"3521","DOI":"10.1073\/pnas.1611835114","volume":"114","author":"J Kirkpatrick","year":"2017","unstructured":"Kirkpatrick, J., Pascanu, R., Rabinowitz, N., Veness, J., Desjardins, G., Rusu, A. A., Milan, K., Quan, J., Ramalho, T., Grabska-Barwinska, A., et al. (2017). Overcoming catastrophic forgetting in neural networks. Proceedings of the National Academy of Sciences, 114(13), 3521\u20133526.","journal-title":"Proceedings of the National Academy of Sciences"},{"issue":"12","key":"2026_CR21","doi-asserted-by":"publisher","first-page":"2935","DOI":"10.1109\/TPAMI.2017.2773081","volume":"40","author":"Z Li","year":"2017","unstructured":"Li, Z., & Hoiem, D. (2017). Learning without forgetting. IEEE Transactions on Pattern Analysis and Machine Intelligence, 40(12), 2935\u20132947.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2026_CR22","unstructured":"LILA BC (Labeled Image Library of Alexandria: Biology and Conservation). https:\/\/lila.science\/"},{"key":"2026_CR23","unstructured":"Liu, D., Hou, J., Huang, S., Liu, J., He, Y., Zheng, B., Ning, J., & Zhang, J. (2023). In Proceedings of the IEEE\/CVF international conference on computer vision (pp. 20064\u201320075)."},{"key":"2026_CR24","unstructured":"Loshchilov, I., & Hutter, F. (2016). SGDR: Stochastic gradient descent with warm restarts. arXiv preprint arXiv:1608.03983"},{"key":"2026_CR25","unstructured":"Loshchilov, I., & Hutter, F. (2017). Decoupled weight decay regularization. arXiv preprint arXiv:1711.05101."},{"key":"2026_CR26","unstructured":"Lu, J., Batra, D., Parikh, D., & Lee, S. (2019). Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks. Advances in neural information processing systems 32."},{"key":"2026_CR27","unstructured":"Miguel, A., Beery, S., Flores, E., Klemesrud, L., & Bayrakcismith, R. (2016). In 2016 IEEE international conference on image processing (ICIP) (pp. 1334\u20131338). IEEE."},{"issue":"8","key":"2026_CR28","doi-asserted-by":"publisher","first-page":"1973","DOI":"10.1111\/1365-2656.13515","volume":"90","author":"MH Murray","year":"2021","unstructured":"Murray, M. H., Fidino, M., Lehrer, E. W., Simonis, J. L., & Magle, S. B. (2021). A multi-state occupancy model to non-invasively monitor visible signs of wildlife health with camera traps that accounts for image quality. Journal of Animal Ecology, 90(8), 1973\u20131984.","journal-title":"Journal of Animal Ecology"},{"issue":"7","key":"2026_CR29","doi-asserted-by":"publisher","first-page":"2152","DOI":"10.1038\/s41596-019-0176-0","volume":"14","author":"T Nath","year":"2019","unstructured":"Nath, T., Mathis, A., Chen, A. C., Patel, A., Bethge, M., & Mathis, M. W. (2019). Using deeplabcut for 3d markerless pose estimation across species and behaviors. Nature Protocols, 14(7), 2152\u20132176.","journal-title":"Nature Protocols"},{"issue":"1","key":"2026_CR30","doi-asserted-by":"publisher","first-page":"150","DOI":"10.1111\/2041-210X.13504","volume":"12","author":"MS Norouzzadeh","year":"2021","unstructured":"Norouzzadeh, M. S., Morris, D., Beery, S., Joshi, N., Jojic, N., & Clune, J. (2021). A deep active learning system for species identification and counting in camera trap images. Methods in Ecology and Evolution, 12(1), 150\u2013161.","journal-title":"Methods in Ecology and Evolution"},{"issue":"25","key":"2026_CR31","doi-asserted-by":"publisher","first-page":"E5716","DOI":"10.1073\/pnas.1719367115","volume":"115","author":"MS Norouzzadeh","year":"2018","unstructured":"Norouzzadeh, M. S., Nguyen, A., Kosmala, M., Swanson, A., Palmer, M. S., Packer, C., & Clune, J. (2018). Automatically identifying, counting, and describing wild animals in camera-trap images with deep learning. Proceedings of the National Academy of Sciences, 115(25), E5716\u2013E5725.","journal-title":"Proceedings of the National Academy of Sciences"},{"key":"2026_CR32","doi-asserted-by":"publisher","DOI":"10.1007\/978-4-431-99495-4","volume-title":"Camera traps in animal ecology: Methods and analyses","author":"AF O\u2019Connell","year":"2011","unstructured":"O\u2019Connell, A. F., Nichols, J. D., & Karanth, K. U. (2011). Camera traps in animal ecology: Methods and analyses (Vol. 271). Springer."},{"key":"2026_CR33","first-page":"27730","volume":"35","author":"L Ouyang","year":"2022","unstructured":"Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., et al. (2022). Training language models to follow instructions with human feedback. Advances in Neural Information Processing Systems, 35, 27730\u201327744.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"2026_CR34","unstructured":"Pantazis, O., Brostow, G., Jones, K., & Mac\u00a0Aodha, O. (2022). Svl-adapter: Self-supervised adapter for vision-language pretrained models. In Proceedings of The 33rd British Machine Vision Conference. The British Machine Vision Association (BMVA)."},{"key":"2026_CR35","unstructured":"Pantazis, O., Brostow, G. J., Jones, K. E., Mac\u00a0Aodha, O. (2021). In Proceedings of the IEEE\/CVF international conference on computer vision (pp. 10583\u201310592)."},{"key":"2026_CR36","unstructured":"Pennington, J., Socher, R., & Manning, C.D. (2014). In Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP) (pp. 1532\u20131543)."},{"key":"2026_CR37","unstructured":"Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., & Clark, J. et\u00a0al. (2021). In International conference on machine learning (PMLR, 2021) (pp. 8748\u20138763)."},{"issue":"1","key":"2026_CR38","first-page":"5485","volume":"21","author":"C Raffel","year":"2020","unstructured":"Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., & Liu, P. J. (2020). Exploring the limits of transfer learning with a unified text-to-text transformer. The Journal of Machine Learning Research, 21(1), 5485\u20135551.","journal-title":"The Journal of Machine Learning Research"},{"key":"2026_CR39","doi-asserted-by":"publisher","first-page":"105","DOI":"10.1007\/s00442-020-04803-9","volume":"195","author":"CD Reddell","year":"2021","unstructured":"Reddell, C. D., Abadi, F., Delaney, D. K., Cain, J. W., & Roemer, G. W. (2021). Urbanization\u2019s influence on the distribution of mange in a carnivore revealed with multistate occupancy models. Oecologia, 195, 105\u2013116.","journal-title":"Oecologia"},{"key":"2026_CR40","doi-asserted-by":"crossref","unstructured":"Rigoudy, N., Dussert, G., Benyoub, A., Besnard, A., Birck, C., Boyer, J., Bollet, Y., Bunz, Y., Caussimont, G., & Chetouane, E. et\u00a0al. (2022). The deepfaune initiative: a collaborative effort towards the automatic identification of the French fauna in camera-trap images. bioRxiv (pp. 2022\u201303).","DOI":"10.1101\/2022.03.15.484324"},{"key":"2026_CR41","doi-asserted-by":"crossref","unstructured":"Rose, S., Engel, D., Cramer, N., & Cowley, W. (2010). Automatic keyword extraction from individual documents. Text mining: applications and theory (pp. 1\u201320).","DOI":"10.1002\/9780470689646.ch1"},{"key":"2026_CR42","unstructured":"Schneider, S., Taylor, G. W., & Kremer, S. (2018). In 2018 15th conference on computer and robot vision (CRV) (pp. 321\u2013328). IEEE."},{"issue":"7","key":"2026_CR43","doi-asserted-by":"publisher","first-page":"3503","DOI":"10.1002\/ece3.6147","volume":"10","author":"S Schneider","year":"2020","unstructured":"Schneider, S., Greenberg, S., Taylor, G. W., & Kremer, S. C. (2020). Three critical factors affecting automated image species recognition performance for camera traps. Ecology and Evolution, 10(7), 3503\u20133517.","journal-title":"Ecology and Evolution"},{"key":"2026_CR44","unstructured":"Shen, Y., Song, K., Tan, X., Li, D., Lu, W., & Zhuang, Y. (2023). HuggingGPT: Solving AI tasks with chatGPT and its friends in huggingface. Advances in Neural Information Processing Systems, 36."},{"key":"2026_CR45","unstructured":"Singh, P., Lindshield, S. M., Zhu, F., & Reibman, A. R. (2020). In 2020 IEEE southwest symposium on image analysis and interpretation (SSIAI) (pp. 66\u201369). IEEE."},{"key":"2026_CR46","unstructured":"Snapshot Serengeti labeled information, library of Alexandria: Biology and conservation website. https:\/\/lila.science\/datasets\/snapshot-serengeti"},{"issue":"1","key":"2026_CR47","doi-asserted-by":"publisher","first-page":"26","DOI":"10.1002\/fee.1448","volume":"15","author":"R Steenweg","year":"2017","unstructured":"Steenweg, R., Hebblewhite, M., Kays, R., Ahumada, J., Fisher, J. T., Burton, C., Townsend, S. E., Carbone, C., Rowcliffe, J. M., Whittington, J., et al. (2017). Scaling-up camera traps: Monitoring the planet\u2019s biodiversity with networks of remote sensors. Frontiers in Ecology and the Environment, 15(1), 26\u201334.","journal-title":"Frontiers in Ecology and the Environment"},{"key":"2026_CR48","doi-asserted-by":"crossref","unstructured":"Sur\u00eds, D., Menon, S., & Vondrick, C. (2023). Vipergpt: Visual inference via python execution for reasoning. arXiv preprint arXiv:2303.08128","DOI":"10.1109\/ICCV51070.2023.01092"},{"key":"2026_CR49","doi-asserted-by":"crossref","unstructured":"Swanson, A., Kosmala, M., Lintott, C., Simpson, R., Smith, A., & Packer, C. (2015). Snapshot serengeti, high-frequency annotated camera trap images of 40 mammalian species in an african savanna. Scientific Data, 2(1), 1\u201314.","DOI":"10.1038\/sdata.2015.26"},{"key":"2026_CR50","doi-asserted-by":"crossref","unstructured":"Tabak, M. A., Falbel, D., Hamzeh, T., Brook, R. K., Goolsby, J. A., Zoromski, L. D., Boughton, R. K., Snow, N. P., VerCauteren, K. C., & Miller, R. S. (2022). Cameratrapdetector: Automatically detect, classify, and count animals in camera trap images using artificial intelligence. bioRxiv (pp. 2022\u201302).","DOI":"10.1101\/2022.02.07.479461"},{"issue":"4","key":"2026_CR51","doi-asserted-by":"publisher","first-page":"585","DOI":"10.1111\/2041-210X.13120","volume":"10","author":"MA Tabak","year":"2019","unstructured":"Tabak, M. A., Norouzzadeh, M. S., Wolfson, D. W., Sweeney, S. J., VerCauteren, K. C., Snow, N. P., Halseth, J. M., Di Salvo, P. A., Lewis, J. S., White, M. D., et al. (2019). Machine learning to classify animal species in camera trap images: Applications in ecology. Methods in Ecology and Evolution, 10(4), 585\u2013590.","journal-title":"Methods in Ecology and Evolution"},{"issue":"1","key":"2026_CR52","doi-asserted-by":"publisher","first-page":"792","DOI":"10.1038\/s41467-022-27980-y","volume":"13","author":"D Tuia","year":"2022","unstructured":"Tuia, D., Kellenberger, B., Beery, S., Costelloe, B. R., Zuffi, S., Risse, B., Mathis, A., Mathis, M. W., van Langevelde, F., Burghardt, T., et al. (2022). Perspectives in machine learning for wildlife conservation. Nature Communications, 13(1), 792.","journal-title":"Nature Communications"},{"key":"2026_CR53","unstructured":"Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, \u0141., & Polosukhin, I. (2017). Attention is all you need. Advances in neural information processing systems 30."},{"key":"2026_CR54","doi-asserted-by":"crossref","unstructured":"Wang, Z., Wu, Z., Agarwal, D., & Sun J. (2022). Medclip: Contrastive learning from unpaired medical images and text. arXiv preprint arXiv:2210.10163","DOI":"10.18653\/v1\/2022.emnlp-main.256"},{"issue":"6","key":"2026_CR55","doi-asserted-by":"publisher","first-page":"1080","DOI":"10.1111\/2041-210X.13576","volume":"12","author":"RC Whytock","year":"2021","unstructured":"Whytock, R. C., \u015awie\u017cewski, J., Zwerts, J. A., Bara-S\u0142upski, T., Koumba Pambo, A. F., Rogala, M., Bahaa-el din, L., Boekee, K., Brittain, S., Cardoso, A. W., et al. (2021). Robust ecological analysis of camera trap data labelled by a machine learning model. Methods in Ecology and Evolution, 12(6), 1080\u20131092.","journal-title":"Methods in Ecology and Evolution"},{"key":"2026_CR56","unstructured":"Wilber, M. J., Scheirer, W. J., Leitner, P., Heflin, B., Zott, J., Reinke, D., Delaney, D. K., Boult, T. E. (2013). In 2013 IEEE workshop on applications of computer vision (WACV) (pp. 206\u2013213). IEEE."},{"issue":"1","key":"2026_CR57","doi-asserted-by":"publisher","first-page":"80","DOI":"10.1111\/2041-210X.13099","volume":"10","author":"M Willi","year":"2019","unstructured":"Willi, M., Pitman, R. T., Cardoso, A. W., Locke, C., Swanson, A., Boyer, A., Veldthuis, M., & Fortson, L. (2019). Identifying animal species in camera trap images using deep learning and citizen science. Methods in Ecology and Evolution, 10(1), 80\u201391.","journal-title":"Methods in Ecology and Evolution"},{"key":"2026_CR58","unstructured":"Ye, S., Filippova, A., Lauer, J., Vidal, M., Schneider, S., Qiu, T., Mathis, A. & Mathis, M. W. (2022). Superanimal models pretrained for plug-and-play analysis of animal behavior. arXiv preprint arXiv:2203.07436"},{"key":"2026_CR59","doi-asserted-by":"publisher","unstructured":"Ye, S., Lauer, J., Zhou, M., Mathis, A., Mathis, M. W. (2023). AmadeusGPT: A natural language interface for interactive animal behavioral analysis. Advances in neural information processing systems, 1. https:\/\/doi.org\/10.48550\/arXiv.2307.04858","DOI":"10.48550\/arXiv.2307.04858"},{"key":"2026_CR60","first-page":"1","volume":"1","author":"X Yu","year":"2013","unstructured":"Yu, X., Wang, J., Kays, R., Jansen, P. A., Wang, T., & Huang, T. (2013). Automated identification of animal species in camera trap images. EURASIP Journal on Image and Video Processing, 1, 1\u201310.","journal-title":"EURASIP Journal on Image and Video Processing"}],"container-title":["International Journal of Computer Vision"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-024-02026-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11263-024-02026-6\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-024-02026-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,8,27]],"date-time":"2024-08-27T03:38:26Z","timestamp":1724729906000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11263-024-02026-6"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,4,24]]},"references-count":60,"journal-issue":{"issue":"9","published-print":{"date-parts":[[2024,9]]}},"alternative-id":["2026"],"URL":"https:\/\/doi.org\/10.1007\/s11263-024-02026-6","relation":{"has-preprint":[{"id-type":"doi","id":"10.1101\/2023.12.22.572990","asserted-by":"object"}]},"ISSN":["0920-5691","1573-1405"],"issn-type":[{"value":"0920-5691","type":"print"},{"value":"1573-1405","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,4,24]]},"assertion":[{"value":"15 September 2023","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"29 January 2024","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"24 April 2024","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}