{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:39:23Z","timestamp":1750307963642,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":17,"publisher":"ACM","license":[{"start":{"date-parts":[[2007,7,9]],"date-time":"2007-07-09T00:00:00Z","timestamp":1183939200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2007,7,9]]},"DOI":"10.1145\/1282280.1282298","type":"proceedings-article","created":{"date-parts":[[2012,10,11]],"date-time":"2012-10-11T15:35:23Z","timestamp":1349969723000},"page":"109-112","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["ACADI showcase - automatic character indexing in audiovisual document"],"prefix":"10.1145","author":[{"given":"Fr\u00e9d\u00e9ric","family":"Gianni","sequence":"first","affiliation":[{"name":"IRIT - Universit\u00e9 Paul Sabatier, Toulouse Cedex, France"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Julien","family":"Pinquier","sequence":"additional","affiliation":[{"name":"IRIT - Universit\u00e9 Paul Sabatier, Toulouse Cedex, France"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ewa Kijak","family":"Irisa","sequence":"additional","affiliation":[{"name":"Campus de Beaulieu, Rennes Cedex, France"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2007,7,9]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"Proceedings of the 5th European Workshop on Image Analysis for Multimedia Interactive Services","author":"Albiol A.","year":"2004","unstructured":"A. Albiol , L. Torres , and E. Delp . Two are better than one: when audio comes to the rescue of video . In Proceedings of the 5th European Workshop on Image Analysis for Multimedia Interactive Services , Lisbon, Portugal , Apr. 2004 . A. Albiol, L. Torres, and E. Delp. Two are better than one: when audio comes to the rescue of video. In Proceedings of the 5th European Workshop on Image Analysis for Multimedia Interactive Services, Lisbon, Portugal, Apr. 2004."},{"key":"e_1_3_2_1_2_1","volume-title":"DARPA Speech Recognition Workshop","author":"Chen S.","year":"1998","unstructured":"S. Chen and P. Gopalakrishnan . Speaker, environment and channel change detection and clustering via the Bayesian Information Criterion . In DARPA Speech Recognition Workshop , 1998 . S. Chen and P. Gopalakrishnan. Speaker, environment and channel change detection and clustering via the Bayesian Information Criterion. In DARPA Speech Recognition Workshop, 1998."},{"key":"e_1_3_2_1_3_1","first-page":"739","volume-title":"Proceedings of the IEEE International Workshop of Neural Networks for Signal Processing","author":"Delakis M.","year":"1998","unstructured":"M. Delakis and C. Garcia . Training Convolutional Filters for Robust Face Detection . In Proceedings of the IEEE International Workshop of Neural Networks for Signal Processing , pages 739 -- 748 , Toulouse, France , Sept. 1998 . M. Delakis and C. Garcia. Training Convolutional Filters for Robust Face Detection. In Proceedings of the IEEE International Workshop of Neural Networks for Signal Processing, pages 739--748, Toulouse, France, Sept. 1998."},{"volume-title":"Association of Audio and Video Segmentations for Automatic Person Indexing. In International Workshop on Content-Based Multimedia Indexing","author":"El Khoury E.","key":"e_1_3_2_1_4_1","unstructured":"E. El Khoury , G. Jaffr\u00e9 , J. Pinquier , and C. S\u00e9nac . Association of Audio and Video Segmentations for Automatic Person Indexing. In International Workshop on Content-Based Multimedia Indexing , page to be published, Bordeaux, France, June 2007. E. El Khoury, G. Jaffr\u00e9, J. Pinquier, and C. S\u00e9nac. Association of Audio and Video Segmentations for Automatic Person Indexing. In International Workshop on Content-Based Multimedia Indexing, page to be published, Bordeaux, France, June 2007."},{"volume-title":"IEEE International Conference on Acoustics, Speech and Signal Processing","author":"El Khoury E.","key":"e_1_3_2_1_5_1","unstructured":"E. El Khoury , C. S\u00e9nac , and R. Andr\u00e9-obrecht . Speaker Diarization: Towards a More Robust and Portable System . In IEEE International Conference on Acoustics, Speech and Signal Processing , page to appear, Honolulu, USA, Apr. 2007. E. El Khoury, C. S\u00e9nac, and R. Andr\u00e9-obrecht. Speaker Diarization: Towards a More Robust and Portable System. In IEEE International Conference on Acoustics, Speech and Signal Processing, page to appear, Honolulu, USA, Apr. 2007."},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPR.2002.1048232"},{"key":"e_1_3_2_1_8_1","first-page":"421","volume-title":"Audiovisual Integration for Tennis Broadcast Structuring. In International Workshop on Content-Based Multimedia Indexing","author":"Kijak E.","year":"2003","unstructured":"E. Kijak , G. Gravier , L. Oisel , and P. Gros . Audiovisual Integration for Tennis Broadcast Structuring. In International Workshop on Content-Based Multimedia Indexing , pages 421 -- 428 , Rennes, France , Sept. 2003 . GDR-PRC ISIS. E. Kijak, G. Gravier, L. Oisel, and P. Gros. Audiovisual Integration for Tennis Broadcast Structuring. In International Workshop on Content-Based Multimedia Indexing, pages 421--428, Rennes, France, Sept. 2003. GDR-PRC ISIS."},{"key":"e_1_3_2_1_9_1","first-page":"180","volume-title":"Fusion of visual and audio features for person identification in real video","author":"Li D.","year":"2001","unstructured":"D. Li , G. Wei , I. K. Sethi , and N. Dimitrova . Fusion of visual and audio features for person identification in real video . In SPIE , the International Society for Optical Engineering, volume 4315 , pages 180 -- 187 , San Diego, USA, Aug. 2001 . D. Li, G. Wei, I. K. Sethi, and N. Dimitrova. Fusion of visual and audio features for person identification in real video. In SPIE, the International Society for Optical Engineering, volume 4315, pages 180--187, San Diego, USA, Aug. 2001."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.patrec.2004.01.004"},{"key":"e_1_3_2_1_11_1","unstructured":"Multimedia understanding through semantics computation and learning. http:\/\/www.muscle-noe.org\/.  Multimedia understanding through semantics computation and learning. http:\/\/www.muscle-noe.org\/."},{"key":"e_1_3_2_1_12_1","unstructured":"OpenCV. http:\/\/www.intel.com\/research\/mrl\/research\/opencv\/.  OpenCV. http:\/\/www.intel.com\/research\/mrl\/research\/opencv\/."},{"key":"e_1_3_2_1_13_1","volume-title":"Issues in Visual and Audio-Visual Speech Processing","author":"Potamianos G.","year":"2004","unstructured":"G. Potamianos , C. Neti , J. Luettin , and I. Matthews . Audio-Visual Automatic Speech Recognition: An Overview . In G. Bailly, E. Vatikiotis-Bateson, and P. E. Perrier, editors, Issues in Visual and Audio-Visual Speech Processing . MIT Press , 2004 . G. Potamianos, C. Neti, J. Luettin, and I. Matthews. Audio-Visual Automatic Speech Recognition: An Overview. In G. Bailly, E. Vatikiotis-Bateson, and P. E. Perrier, editors, Issues in Visual and Audio-Visual Speech Processing. MIT Press, 2004."},{"key":"e_1_3_2_1_14_1","first-page":"873","volume-title":"IEEE International Conference on Acoustics, Speech and Signal Processing","author":"Siu M.","year":"1991","unstructured":"M. Siu , H. Gish , and R. Rohlicek . Segregation of speaker for speech recognition and speaker identification . In IEEE International Conference on Acoustics, Speech and Signal Processing , pages 873 -- 876 , Toronto, Canada , May 1991 . M. Siu, H. Gish, and R. Rohlicek. Segregation of speaker for speech recognition and speaker identification. In IEEE International Conference on Acoustics, Speech and Signal Processing, pages 873--876, Toronto, Canada, May 1991."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1023\/B:MTAP.0000046380.27575.a5"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP.2004.1418850"},{"key":"e_1_3_2_1_17_1","first-page":"725","volume-title":"IEEE International Conference on Acoustics, Speech and Signal Processing","author":"Tsai W.","year":"2005","unstructured":"W. Tsai , S. Cheng , Y. Chao , and H. Wang . Clustering speech utterances by the speaker using EigenVoice-Motivated vector space models . In IEEE International Conference on Acoustics, Speech and Signal Processing , pages 725 -- 728 , Philadelphia, USA , Mar. 2005 . W. Tsai, S. Cheng, Y. Chao, and H. Wang. Clustering speech utterances by the speaker using EigenVoice-Motivated vector space models. In IEEE International Conference on Acoustics, Speech and Signal Processing, pages 725--728, Philadelphia, USA, Mar. 2005."},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.982883"}],"event":{"name":"CIVR07: International Conference on Image and Video Retrieval 2007","sponsor":["SIGMM ACM Special Interest Group on Multimedia","SIGIR ACM Special Interest Group on Information Retrieval"],"location":"Amsterdam The Netherlands","acronym":"CIVR07"},"container-title":["Proceedings of the 6th ACM international conference on Image and video retrieval"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1282280.1282298","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/1282280.1282298","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T14:58:07Z","timestamp":1750258687000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1282280.1282298"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2007,7,9]]},"references-count":17,"alternative-id":["10.1145\/1282280.1282298","10.1145\/1282280"],"URL":"https:\/\/doi.org\/10.1145\/1282280.1282298","relation":{},"subject":[],"published":{"date-parts":[[2007,7,9]]},"assertion":[{"value":"2007-07-09","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}