{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,14]],"date-time":"2026-02-14T07:35:33Z","timestamp":1771054533761,"version":"3.50.1"},"reference-count":42,"publisher":"Springer Science and Business Media LLC","issue":"2","license":[{"start":{"date-parts":[[2011,10,12]],"date-time":"2011-10-12T00:00:00Z","timestamp":1318377600000},"content-version":"tdm","delay-in-days":0,"URL":"http:\/\/www.springer.com\/tdm"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Lang Resources &amp; Evaluation"],"published-print":{"date-parts":[[2012,6]]},"DOI":"10.1007\/s10579-011-9165-9","type":"journal-article","created":{"date-parts":[[2011,10,11]],"date-time":"2011-10-11T05:54:16Z","timestamp":1318312456000},"page":"155-176","source":"Crossref","is-referenced-by-count":33,"title":["A survey of methods to ease the development of highly multilingual text mining applications"],"prefix":"10.1007","volume":"46","author":[{"given":"Ralf","family":"Steinberger","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2011,10,12]]},"reference":[{"key":"9165_CR1","unstructured":"Balahur-Dobrescu, A., Steinberger, R., Kabadjov, M., Zavarella, V., van der Goot, E., Halkia, M., et al. (2010). Sentiment analysis in the news. In Proceedings of LREC. Valletta, Malta."},{"key":"9165_CR2","unstructured":"Bender, E., & Flickinger, D. (2005). Rapid prototyping of scalable grammars: Towards modularity in extensions to a language-independent core. In Proceedings of IJCNLP. Jeju Island, Korea."},{"key":"9165_CR3","first-page":"364","volume-title":"Evaluating cross-lingual annotation transfer in the MultiSemCor corpus","author":"L Bentivogli","year":"2004","unstructured":"Bentivogli, L., Forner, P., & Pianta, E. (2004). Evaluating cross-lingual annotation transfer in the MultiSemCor corpus (pp. 364\u2013370). Geneva, Switzerland: CoLing."},{"key":"9165_CR4","first-page":"42","volume-title":"Corpora and evaluation tools for multilingual named entity grammar development. Proceedings of the multilingual corpora workshop at corpus linguistics","author":"C Bering","year":"2003","unstructured":"Bering, C., Dro\u017cd\u017cy\u0144ski, W., Erbach, G., Guasch, L., Homola, P., Lehmann, S., et al. (2003). Corpora and evaluation tools for multilingual named entity grammar development. Proceedings of the multilingual corpora workshop at corpus linguistics (pp. 42\u201352). UK: Lancaster."},{"issue":"1","key":"9165_CR5","doi-asserted-by":"crossref","first-page":"20","DOI":"10.1109\/MIS.2007.11","volume":"22","author":"M Carenini","year":"2007","unstructured":"Carenini, M., Whyte, A., Bertorello, L., & Vanocchi, M. (2007). Improving communication in E-democracy using natural language processing. IEEE Intelligent Systems, 22(1), 20\u201327.","journal-title":"IEEE Intelligent Systems"},{"key":"9165_CR6","unstructured":"Ehrmann, M., & Turchi, M. (2010). Building multilingual named entity-annotated corpora exploiting parallel corpora. In Proceedings of the workshop on annotation and exploitation of parallel corpora (AEPC) (pp. 24\u201333). Tartu, Estonia."},{"key":"9165_CR7","unstructured":"Gamon, M., Lozano, C., Pinkham, J., & Reutter, T. (1997). Practical experience with grammar sharing in multilingual NLP. In Proceedings of ACL\/EACL, Madrid, Spain."},{"key":"9165_CR8","unstructured":"Gong, Y., & Liu, X. (2002). Generic text summarization using relevance measure and latent semantic analysis. In Proceedings of ACM SIGIR. New Orleans, USA."},{"key":"9165_CR9","volume-title":"Learning machine translation","author":"C Goutte","year":"2009","unstructured":"Goutte, C., Cancedda, N., Dymetman, M., & Foster, G. (2009). Learning machine translation. Cambridge, USA: MIT Press."},{"key":"9165_CR10","unstructured":"Grefenstette, G. (2010). Proposition for a web 2.0 version of linguistic resource creation. Presentation at FLaReNet Forum 2010 in Barcelona on 12.02.2010."},{"key":"9165_CR11","unstructured":"Ignat C., Pouliquen, B., Ribeiro, A., & Steinberger, R. (2003). Extending an information extraction tool set to central and eastern European languages. In Proceedings of the workshop information extraction for Slavonic and other central and eastern European languages (IESL), held at RANLP. Borovets, Bulgaria, September 8\u20139, 2003."},{"key":"9165_CR12","unstructured":"Kabadjov, M., Atkinson, M., Steinberger, J., Steinberger, R., & van der Goot, E. (2010). NewsGist: A multilingual statistical news summarizer. In J. L. Balc\u00e1zar, F. Bonchi, A. Gionis, & M. Sebag (Eds.), Proceedings of the European conference on machine learning and principles and practice of knowledge discovery in databases (ECML-PKDD). Barcelona, Spain, September 20\u201324, 2010. Lecture Notes in Computer Science (Vol. 6323, pp. 591\u2013594). Berlin: Springer."},{"key":"9165_CR13","doi-asserted-by":"crossref","unstructured":"Larkey, L., Feng, F., Connell, M., & Lavrenko, V. (2004). Language-specific models in multilingual topic tracking. In Proceedings of the 27th annual international ACM SIGIR conference on research and development in information retrieval (pp. 402\u2013409).","DOI":"10.1145\/1008992.1009061"},{"issue":"1","key":"9165_CR14","doi-asserted-by":"crossref","first-page":"67","DOI":"10.1016\/j.ins.2004.10.006","volume":"176","author":"C-J Lee","year":"2006","unstructured":"Lee, C.-J., Chang, J. S., & Jang, J.-S. R. (2006). Extraction of transliteration pairs from parallel corpora using a statistical transliteration model. Information Sciences, 176(1), 67\u201390.","journal-title":"Information Sciences"},{"key":"9165_CR15","doi-asserted-by":"crossref","unstructured":"Linge, J., Steinberger, R., Weber, T., Yangarber, R., van der Goot, E., Al Khudhairy, D., et al. (2009). Internet surveillance systems for early alerting of health threats. Euro Surveillance, 14(13). Stockholm, April 2, 2009.","DOI":"10.2807\/ese.14.13.19162-en"},{"key":"9165_CR16","unstructured":"Maynard, D., Tablan, V., & Cunningham, H. (2003). NE Recognition without training data on a language you don\u2019t speak. In Proceedings of the ACL workshop on multilingual and mixed-language NER: Combining statistical and symbolic methods. Sapporo, Japan."},{"issue":"3","key":"9165_CR17","doi-asserted-by":"crossref","first-page":"257","DOI":"10.1017\/S1351324902002930","volume":"8","author":"D Maynard","year":"2002","unstructured":"Maynard, D., Tablan, V., Cunningham, H., Ursu, C., Saggion, H., Bontcheva, K., et al. (2002). Architectural elements of language engineering robustness. Journal of Natural Language Engineering, 8(3), 257\u2013274. Special issue on robust methods in analysis of natural language data.","journal-title":"Journal of Natural Language Engineering"},{"key":"9165_CR18","volume-title":"Named entities\u2014recognition, classification and use","author":"D Nadeau","year":"2009","unstructured":"Nadeau, D., & Sekine, S. (2009). A survey of entity recognition and classification. In S. Sekine & E. Ranchhod (Eds.), Named entities\u2014recognition, classification and use. Amsterdam\/Philadelphia: John Benjamins Publishing Company."},{"key":"9165_CR19","unstructured":"Nivre, J., Hall, J., K\u00fcbler, S., McDonald, R., Nilsson, J., Riedel, S., et al. (2007). The CoNLL 2007 shared task on dependency parsing. In Proceedings of the CoNLL shared task session of EMNLP-CoNLL (pp. 915\u2013932). Prague, Czech Republic."},{"issue":"1\u20132","key":"9165_CR20","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1561\/1500000011","volume":"2","author":"B Pang","year":"2008","unstructured":"Pang, B., & Lee, L. (2008). Opinion mining and sentiment analysis. Foundations and Trends in Information Retrieval, 2(1\u20132), 1\u2013135.","journal-title":"Foundations and Trends in Information Retrieval"},{"key":"9165_CR21","unstructured":"Pastra, K., Maynard, D., Hamza, O., Cunningham, H., & Wilks, Y. (2002). How feasible is the reuse of grammars for named entity recognition? In Proceedings of LREC, Las Palmas, Spain."},{"key":"9165_CR22","unstructured":"Pouliquen, B., Steinberger, R., & Best, C. (2007a). Automatic detection of quotations in multilingual news. In Proceedings of the international conference recent advances in natural language processing (RANLP) (pp. 487\u2013492). Borovets, Bulgaria, September 27\u201329, 2007."},{"key":"9165_CR23","unstructured":"Pouliquen, B., Steinberger, R., & Belyaeva, J. (2007b). Multilingual multi-document continuously updated social networks. In Proceedings of the workshop multi-source multilingual information extraction and summarization (MMIES) held at RANLP (pp. 25\u201332). Borovets, Bulgaria, September 26, 2007."},{"key":"9165_CR24","unstructured":"Ranta, A. (2009). The GF resource grammar library. In Linguistic issues in language technology LiLT 2:2. December 2009."},{"key":"9165_CR25","unstructured":"Rayner, M., & Bouillon, P. (1996). Adapting the core language engine to French and Spanish. In Proceedings of the international conference NLP+IA (pp. 224\u2013232), Mouncton, Canada."},{"key":"9165_CR26","unstructured":"Rosen, A. (2010). Mediating between incompatible tagsets. In Proceedings of the workshop on annotation and exploitation of parallel corpora (pp. 53\u201362), Tartu, Estonia."},{"key":"9165_CR27","doi-asserted-by":"crossref","unstructured":"Shinyama, Y., & Sekine, S. (2004). Named entity discovery using comparable news articles. In Proceedings of the 20th international conference on computational linguistics (CoLing) (pp. 848\u2013853). Geneva, Switzerland.","DOI":"10.3115\/1220355.1220477"},{"key":"9165_CR28","unstructured":"Spreyer, K., & Frank, A. (2008). Projection-based acquisition of a temporal labeller. In Proceedings of the 3rd international joint conference on natural language processing (IJCNLP) (pp. 489\u2013496). Hyderabad, India."},{"key":"9165_CR29","unstructured":"Steinberger, J., Kabadjov, M., Pouliquen, B., Steinberger, R., & Poesio, M., (2009). WB-JRC-UT\u2019s participation in TAC 2009: Update summarization and AESOP tasks. In Proceedings of the text analysis conference 2009 (TAC\u20192009). National Institute of Standards and Technology, Gaithersburg, Maryland USA, November 16\u201317, 2009."},{"key":"9165_CR50","unstructured":"Steinberger, J., Lenkova, P., Ebrahim, M., Ehrmann, M., V\u00e1zquez, S., H\u00fcrriyeto\u011flu, A., Kabadjov, M., Steinberger, R., Tanev, H., & Zavarella, V. (2011). Creating sentiment dictionaries via triangulation. In Proceedings of the 2nd workshop on computational approaches to subjectivity and sentiment analysis, WASSA, held at the ACL-HLT conference (pp. 28\u201336). Portland, Oregon, USA, 24 June 2011."},{"key":"9165_CR30","unstructured":"Steinberger, R., & Pouliquen, B. (2007). Cross-lingual named entity recognition. In S. Sekine & E. Ranchhod (Eds.), Journal Linguisticae Investigationes, Special issue on named entity recognition and categorisation. LI, 30(1), 135\u2013162. Amsterdam: John Benjamins Publishing Company."},{"key":"9165_CR31","first-page":"217","volume-title":"Mining massive data sets for security","author":"R Steinberger","year":"2008","unstructured":"Steinberger, R., Pouliquen, B., & Ignat, C. (2008). Using language-independent rules to achieve high multilinguality in text mining. In F.-S. Fran\u00e7oise, D. Perrotta, J. Piskorski, & R. Steinberger (Eds.), Mining massive data sets for security (pp. 217\u2013240). Amsterdam, The Netherlands: IOS Press."},{"key":"9165_CR32","unstructured":"Steinberger, R., Pouliquen, B., & van der Goot, E. (2009). An introduction to the europe media monitor family of applications. In F. Gey, N. Kando, & J. Karlgren (Eds.), Information access in a multilingual world\u2014Proceedings of the SIGIR 2009 Workshop (SIGIR-CLIR) (pp. 1\u20138). Boston, USA, July 23, 2009."},{"key":"9165_CR33","unstructured":"Steinberger, R., Pouliquen, B., Widiger, A., Ignat, C., Erjavec, T., Tufi\u015f, D., et al. (2006). The JRC-acquis: A multilingual aligned parallel corpus with 20+ languages. In Proceedings of the 5th international conference on language resources and evaluation (LREC) (pp. 2142\u20132147). Genoa, Italy, May 24\u201326, 2006."},{"key":"9165_CR34","doi-asserted-by":"crossref","unstructured":"Steinberger, R., Ombuya, S., Kabadjov, M., Pouliquen, B., Della Rocca, L., Belyaeva, J., De Paola, M., & van der Goot, E. (2011). Expanding a multilingual media monitoring and information extraction tool to a new language: Swahili. Language Resources and Evaluation Journal, 45(3), 311\u2013330.","DOI":"10.1007\/s10579-011-9155-y"},{"key":"9165_CR35","unstructured":"Tanev, H., Zavarella, V., Linge, J., Kabadjov,M., Piskorski, J., Atkinson, M., et al. (2009). Exploiting machine learning techniques to build an event extraction system for Portuguese and Spanish. In linguaM\u00c1TICA\u2014Revista para o Processamento Autom\u00e1tico das L\u00ednguas Ib\u00e9ricas (Vol. 2, pp. 55\u201367)."},{"key":"9165_CR36","doi-asserted-by":"crossref","unstructured":"Turchi, M., Steinberger, J., Kabadjov, M., & Steinberger, R. (2010). Using parallel corpora for multilingual (multi-document) summarisation evaluation. In Conference on multilingual and multimodal information access evaluation (CLEF). Padua, Italy, September 20\u201323, 2010. Springer Lecture Notes for Computer Science LNCS.","DOI":"10.1007\/978-3-642-15998-5_7"},{"key":"9165_CR37","unstructured":"Vergne, J. (2002). Une m\u00e9thode pour l\u2019analyse descendante et calculatoire de corpus multilingues: Application au calcul des relations sujet-verbe. In Proceedings of TALN. Nancy, France."},{"key":"9165_CR38","unstructured":"Vergne, J. (2009). Defining the chunk as the period of the functions length and frequency of words on the syntagmatic axis. In Proceedings of the language technology conference LTC. Poznan, Poland."},{"key":"9165_CR39","doi-asserted-by":"crossref","unstructured":"Wehrli, E. (2007). Fips, a \u201cDeep\u201d linguistic multilingual parser. In Proceedings of the ACL workshop on deep linguistic processing (pp. 120\u2013127). Prague, Czech Republic.","DOI":"10.3115\/1608912.1608931"},{"key":"9165_CR40","doi-asserted-by":"crossref","unstructured":"Yarowski, D., Ngai, G., & Wicentowski, R. (2001). Inducing multilingual text analysis tools via robust projection across aligned corpora. In Proceedings of the 1st international conference on Human Language Technology research (HLT) (pp. 1\u20138). Stroudsburg, PA, USA.","DOI":"10.3115\/1072133.1072187"},{"key":"9165_CR41","unstructured":"Zaghouani, W., Pouliquen, B., Ibrahim, M., & Steinberger, R., (2010). Adapting a resource-light highly multilingual Named Entity Recognition system to Arabic. In Proceedings of LREC, Valletta, Malta."}],"container-title":["Language Resources and Evaluation"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s10579-011-9165-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s10579-011-9165-9\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s10579-011-9165-9","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2021,12,10]],"date-time":"2021-12-10T00:27:24Z","timestamp":1639096044000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s10579-011-9165-9"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2011,10,12]]},"references-count":42,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2012,6]]}},"alternative-id":["9165"],"URL":"https:\/\/doi.org\/10.1007\/s10579-011-9165-9","relation":{},"ISSN":["1574-020X","1574-0218"],"issn-type":[{"value":"1574-020X","type":"print"},{"value":"1574-0218","type":"electronic"}],"subject":[],"published":{"date-parts":[[2011,10,12]]}}}