{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,7]],"date-time":"2026-07-07T12:25:56Z","timestamp":1783427156155,"version":"3.54.6"},"reference-count":139,"publisher":"Informa UK Limited","issue":"4","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Human\u2013Computer Interaction"],"published-print":{"date-parts":[[2000,12]]},"DOI":"10.1207\/s15327051hci1504_1","type":"journal-article","created":{"date-parts":[[2004,1,14]],"date-time":"2004-01-14T14:16:50Z","timestamp":1074089810000},"page":"263-322","source":"Crossref","is-referenced-by-count":211,"title":["Designing the User Interface for Multimodal Speech and Pen-Based Gesture Applications: State-of-the-Art Systems and Future Research Directions"],"prefix":"10.1080","volume":"15","author":[{"given":"Sharon","family":"Oviatt","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Phil","family":"Cohen","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Lizhong","family":"Wu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Lisbeth","family":"Duncan","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Bernhard","family":"Suhm","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Josh","family":"Bers","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Thomas","family":"Holzman","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Terry","family":"Winograd","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"James","family":"Landay","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jim","family":"Larson","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"David","family":"Ferro","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"301","published-online":{"date-parts":[[2009,12,9]]},"reference":[{"key":"p_1_1","doi-asserted-by":"crossref","unstructured":"Adjoudani, A. & Benoit, C. (1996). On the integration of auditory and visual parameters in an HMM-based ASR. In D. Stork & M. Hennecke (Eds.), Speechreading by humans and machines, NATO ASI series, series F: Computer and systems science (pp. 461-473). Berlin: Springer.","DOI":"10.1007\/978-3-662-13015-5_35"},{"key":"p_2_2","doi-asserted-by":"publisher","DOI":"10.1016\/0004-3702(84)90008-0"},{"key":"p_3_3","doi-asserted-by":"publisher","DOI":"10.1016\/0004-3702(80)90042-9"},{"key":"p_4_4","doi-asserted-by":"publisher","DOI":"10.1080\/088395199117333"},{"key":"p_5_5","doi-asserted-by":"crossref","unstructured":"Appelt, D. (1985). Planning English sentences. Cambridge, England: Cambridge University Press.","DOI":"10.1017\/CBO9780511624575"},{"key":"p_6_6","unstructured":"Benoit, C., Martin, J.C., Pelachaud, C., Schomaker, L. & Suhm, B. (2000). Audio-visual and multimodal speech systems. In D. Gibbon, I. Mertins, & R. Moore (Eds.), Handbook of multimodal and spoken dialogue systems: Resources, terminology and product evaluation. Norwell, MA: Kluwer Academic."},{"key":"p_7_7","unstructured":"Bers, J., Miller, S. & Makhoul, J. (1998). Designing conversational interfaces with multimodal interaction. DARPA Workshop on Broadcast News Understanding Systems, 319-321."},{"key":"p_8_8","doi-asserted-by":"crossref","first-page":"262","DOI":"10.1145\/965105.807503","volume":"14","author":"Bolt R. A.","year":"1980","journal-title":"Computer Graphics"},{"key":"p_9_9","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.1993.319179"},{"key":"p_10_10","doi-asserted-by":"crossref","unstructured":"Bretier, P. & Sadek, D. (1997). A rational agent as the kernel of a cooperative spoken dialogue system: Implementing a logical theory of interaction.Intelligent Agents III: Proceedings of the Third International Workshop on Agent Theories, Architectures and Languages (ATAL-96), 189-204. Heidelberg, Germany: Springer-Verlag.","DOI":"10.1007\/BFb0013586"},{"key":"p_11_11","unstructured":"Calder, J. (1987). Typed unification for natural language processing. In E. Klein & J. van Benthem (Eds.), Categories, polymorphisms, and unification (pp. 65-72). Edinburgh, Scotland: Center for Cognitive Science, University of Edinburgh."},{"key":"p_12_12","unstructured":"Carberry, S. (1990). Plan recognition in natural language dialogue. Cambridge, MA: ACL-MIT Press Series in Natural Language Processing, MIT Press."},{"key":"p_13_13","unstructured":"Carpenter, R. (1990). Typed feature structures: Inheritance, (in)equality, and extensionality. Proceedings of the ITK Workshop: Inheritance in Natural Language Processing, 9-18. Tilburg, The Netherlands: Institute for Language Technology and Artificial Intelligence, Tilburg University."},{"key":"p_14_14","doi-asserted-by":"crossref","unstructured":"Carpenter, R. (1992). The logic of typed feature structures. Cambridge, England: Cambridge University Press.","DOI":"10.1017\/CBO9780511530098"},{"key":"p_15_15","unstructured":"Cassell, J. & Stone, M. (1999). Living hand to mouth: Psychological theories about speech and gesture in interactive dialogue systems. Working notes of the AAAI'99 Fall Symposium Series on Psychological Models of Communication in Collaborative Systems, 34-42. Menlo Park, CA: AAAI Press."},{"key":"p_16_16","doi-asserted-by":"crossref","unstructured":"Cassell, J., Sullivan, J., Prevost, S. & Churchill, E. (Eds.). (2000). Embodied conversational agents. Cambridge, MA: MIT Press.","DOI":"10.7551\/mitpress\/2697.001.0001"},{"key":"p_17_17","unstructured":"Cheyer, A. & Julia, L. (1995). Multimodal maps: An agent-based approach. International Conference on Cooperative Multimodal Communication (CMC'95), 103-113. Eindhoven, The Netherlands: Eindhoven University of Technology."},{"key":"p_18_18","unstructured":"Cheyer, A., Julia, L. & Martin, J. C. (1998). A unified framework for constructing multimodal experiments and applications. Conference on Cooperative Multimodal Communication (CMC'98), 63-69. Tilburg, The Netherlands: Institute for Language Technology and Artificial Intelligence, Tilburg University."},{"key":"p_19_19","first-page":"277","volume":"2","author":"Clow J.","year":"1998","journal-title":"Proceedings of the International Conference on Spoken Language Processing"},{"key":"p_20_20","doi-asserted-by":"crossref","unstructured":"Codella, C., Jalili, R., Koved, L., Lewis, J., Ling, D., Lipscomb, J., Rabenhorst, D., Wang, C., Norton, A., Sweeney, P. & Turk, C. (1992). Interactive simulation in a multi-person virtual world. Proceedings of the Conference on Human Factors in Computing Systems (CHI'92), 329-334. New York: ACM Press.","DOI":"10.1145\/142750.142825"},{"key":"p_21_21","unstructured":"Coen, M. (1998). Design principles for intelligent environments. Proceedings of the Fifteenth National Conference on Artificial Intelligence (AAAI'98), 547-554. Menlo Park, CA: AAAI Press."},{"key":"p_22_22","unstructured":"Cohen, P. R., Cheyer, A., Wang, M. & Baeg, S. C. (1997). An open agent architecture. In M. Huhns & M. Singh (Eds.), Readings in agents (pp. 197-204). San Francisco: Kaufmann. (Reprinted from AAAI '94 Spring Symposium Series on Software Agents, 1-8, 1994, Menlo Park, CA: AAAI Press)"},{"key":"p_23_23","doi-asserted-by":"crossref","unstructured":"Cohen, P. R., Dalrymple, M., Moran, D. B., Pereira, F. C. N., Sullivan, J. W., Gargan, R. A., Schlossberg, J. L. & Tyler, S. W. (1998). Synergistic use of direct manipulation and natural language. In M. Maybury & W. Wahlster (Eds.), Readings in intelligent user interfaces (pp. 29-37). San Francisco: Kaufmann. (Reprinted from Proceedings of the Conference on Human Factors in Computing Systems (CHI'89), 227-234, 1989, New York: ACM Press)","DOI":"10.1145\/67449.67494"},{"key":"p_24_24","first-page":"249","volume":"2","author":"Cohen P. R.","year":"1998","journal-title":"Proceedings of the International Conference on Spoken Language Processing"},{"key":"p_25_25","doi-asserted-by":"crossref","unstructured":"Cohen, P. R., Johnston, M., McGee, D., Oviatt, S., Pittman, J., Smith, I., Chen, L. & Clow, J. (1997). Quickset: Multimodal interaction for distributed applications. Proceedings of the Fifth ACM International Multimedia Conference, 31-40. New York: ACM Press.","DOI":"10.1145\/266180.266328"},{"key":"p_26_26","unstructured":"Cohen, P. R. & Levesque, H. J. (1990). Rational interaction as the basis for communication. In P. R. Cohen, J. Morgan, & M. E. Pollack (Eds.), Intentions in communication (pp. 221-255). Cambridge, MA: SDF Benchmark Series, MIT Press."},{"key":"p_27_27","doi-asserted-by":"publisher","DOI":"10.1109\/38.773958"},{"key":"p_28_28","doi-asserted-by":"crossref","first-page":"9921","DOI":"10.1073\/pnas.92.22.9921","volume":"92","author":"Cohen P. R.","year":"1995","journal-title":"Proceedings of the National Academy of Sciences"},{"issue":"3","key":"p_29_29","doi-asserted-by":"crossref","first-page":"177","DOI":"10.1207\/s15516709cog0303_1","volume":"3","author":"Cohen P. R.","year":"1979","journal-title":"Cognitive Science"},{"key":"p_30_30","doi-asserted-by":"publisher","DOI":"10.1109\/89.365385"},{"key":"p_31_31","unstructured":"Cole, R., Mariani, J., Uszkoreit, H., Varile, G., Zaenen, A., Zue, V. & Zampolli, A. (Eds.). (1997). Survey of the state of the art in human language technology. Cambridge, MA: Cambridge University Press."},{"key":"p_32_32","unstructured":"Cypher, A. (Ed.). (1993). Watch what I do: Programming by demonstration. Cambridge, MA: MIT Press."},{"key":"p_33_33","unstructured":"Duncan, L., Holmback, H., Kau, A., Poteet, S., Miller, S., Harrison, P., Powell, J. & Jenkins, T. (1994). Message processing and data extraction for the Real-Time Information Extraction System (RTIMS) project (Boeing Technical Report BCSTECH-94-057)."},{"key":"p_34_34","unstructured":"Duncan, L., Brown, W., Esposito, C., Holmback, H. & Xue, P. (1999). Enhancing virtual maintenance environments with speech understanding. Boeing M&CT TechNet."},{"key":"p_35_35","doi-asserted-by":"publisher","DOI":"10.1016\/0364-0213(90)90002-E"},{"key":"p_36_36","doi-asserted-by":"crossref","unstructured":"Erman, L. D. & Lesser, V. R. (1975). A multi-level organization for problem solving using many diverse sources of knowledge. Proceedings of the 4th International Joint Conference on Artificial Intelligence, 483-490. Georgia, USSR: IJCAI Press.","DOI":"10.21236\/ADA012916"},{"key":"p_37_37","doi-asserted-by":"crossref","unstructured":"Feiner, S., MacIntyre, B., Hollerer, T. & Webster, A. (1997). A touring machine: Prototyping 3D mobile augmented reality systems for exploring the urban environment. International Symposium on Wearable Computing (ISWC'97), 208-217. Cambridge, MA: CS Press.","DOI":"10.1109\/ISWC.1997.629922"},{"key":"p_38_38","doi-asserted-by":"crossref","unstructured":"Feiner, S. K. & McKeown, K. R. (1991). COMET: Generating coordinated multimedia explanations. Proceedings of the Conference on Human Factors in Computing Systems (CHI'91), 449-450. New York: ACM Press.","DOI":"10.1145\/108844.108997"},{"key":"p_39_39","doi-asserted-by":"crossref","unstructured":"Fell, H., Delta, H., Peterson, R., Ferrier, L., Mooraj, Z. & Valleau, M. (1994). Using the baby-babble-blanket for infants with motor problems. Proceedings of the Conference on Assistive Technologies (ASSETS'94), 77-84. Marina del Rey, CA. Retrieved f r o m t h e Wo r l d Wi d e We b : h t t p : \/ \/ w w w. a c m . o r g \/ s i g c a p h \/ a ssets\/assets98\/assets98index.html","DOI":"10.1145\/191028.191049"},{"key":"p_40_40","first-page":"58","volume":"73","author":"Flanagan J.","year":"1991","journal-title":"Acustica"},{"key":"p_41_41","doi-asserted-by":"publisher","DOI":"10.1016\/0097-8493(94)90157-0"},{"key":"p_42_42","unstructured":"Goldschen, A. & Loehr, D. (1999). The Role of the DARPA communicator architecture as a human computer interface for distributed simulations. Spring Simulation Interoperability Workshop. Orlando, FL: Simulation Interoperability Standards Organization."},{"key":"p_43_43","doi-asserted-by":"publisher","DOI":"10.1016\/0167-6393(94)00059-J"},{"key":"p_44_44","doi-asserted-by":"crossref","unstructured":"Hauptmann, A. G. (1989). Speech and gestures for graphic image manipulation. Proceedings of the International Conference on Human-Computer Interaction (CHI'89), 1, 241-245. New York: ACM Press.","DOI":"10.1145\/67449.67496"},{"key":"p_45_45","unstructured":"Holmback, H., Duncan, L. & Harrison, P. (2000). A word sense checking application for Simplified English. Proceedings of the Third International Workshop on Controlled Language Applications, Seattle."},{"key":"p_46_46","doi-asserted-by":"crossref","unstructured":"Holmback, H., Greaves, M. & Bradshaw, J. (1999). A pragmatic principle for agent communication. In J. Bradshaw, O. Etzioni, & J. Mueller (Eds.), Proceedings of Autonomous Agents'99. New York: ACM Press.","DOI":"10.1145\/301136.301238"},{"key":"p_47_47","doi-asserted-by":"publisher","DOI":"10.1145\/301153.301160"},{"key":"p_48_48","doi-asserted-by":"crossref","unstructured":"Horvitz, E. (1999). Principles of mixed-initiative user interfaces. Proceedings of the Conference on Human Factors in Computing Systems (CHI'99), 159-166. New York: ACM Press.","DOI":"10.1145\/302979.303030"},{"key":"p_49_49","doi-asserted-by":"crossref","unstructured":"Huang, X., Acero, A., Alleva, F., Hwang, M.Y., Jiang, L. & Mahajan, M. (1995). Microsoft Windows highly intelligent speech recognizer: Whisper. Proceedings of 1995 International Conference on Acoustics, Speech, and Signal Processing, 1, 93-96. Detroit, MI: IEEE.","DOI":"10.1109\/ICASSP.1995.479281"},{"key":"p_50_50","doi-asserted-by":"crossref","unstructured":"Iverson, P., Bernstein, L. & Auer, E. (1998). Modeling the interaction of phonemic intelligibility and lexical structure in audiovisual word recogniton. Speech Communication, 26(1&2), 45-63.","DOI":"10.1016\/S0167-6393(98)00049-1"},{"key":"p_51_51","unstructured":"Johnston, M. (1998). Unification-based multimodal parsing. Proceedings of the International Joint Conference of the Association for Computational Linguistics and the International Committee on Computational Linguistics (COLING-ACL'98), 624-630. Montreal, Canada: University of Montreal Press."},{"key":"p_52_52","doi-asserted-by":"crossref","unstructured":"Johnston, M., Cohen, P. R., McGee, D., Oviatt, S. L., Pittman, J. A. & Smith, I. (1997). Unification-based multimodal integration. Proceedings of the 35th Annual Meeting of the Association for Computational Linguistics, 281-288. San Francisco: Kaufmann.","DOI":"10.3115\/976909.979653"},{"key":"p_53_53","doi-asserted-by":"publisher","DOI":"10.1109\/78.175747"},{"key":"p_54_54","doi-asserted-by":"crossref","unstructured":"Karat, C.M., Halverson, C., Horn, D. & Karat, J. (1999). Patterns of entry and correction in large vocabulary continuous speech recognition systems. Proceedings of the International Conference for Computer-Human Interaction (CHI'99), 568-575. New York: ACM Press.","DOI":"10.1145\/302979.303160"},{"key":"p_55_55","unstructured":"Karshmer, A. I. & Blattner, M. (Organizers). (1998). Proceedings of the Third International ACM Proceedings of the Conference on Assistive Technologies (ASSETS'98). Marina del Rey, CA. http:\/\/www.acm.org\/sigcaph\/assets\/assets98\/assets98index.html"},{"key":"p_57_56","doi-asserted-by":"crossref","unstructured":"Kay, M. (1979). Functional grammar. Proceedings of the Fifth Annual Meeting of the Berkeley Linguistics Society, 142-158. Berkeley, CA: Berkeley Linguistics Society.","DOI":"10.3765\/bls.v5i0.3262"},{"key":"p_58_57","doi-asserted-by":"publisher","DOI":"10.1145\/357417.357420"},{"key":"p_59_58","unstructured":"Kempf, J. (1994). Preliminary handwriting recognition engine application program interface for Solaris 2. Mountain View, CA: Sun Microsystems."},{"key":"p_60_59","unstructured":"Kendon, A. (1980). Gesticulation and speech: Two aspects of the process of utterance. In M. Key (Ed.), The relationship of verbal and nonverbal communication (pp. 207-227). The Hague, The Netherlands: Mouton."},{"key":"p_61_60","first-page":"1","volume":"2000","author":"Klemmer S.","year":"2000","journal-title":"Proceedings of User Interface Software and Technology"},{"key":"p_62_61","doi-asserted-by":"crossref","first-page":"459","DOI":"10.1145\/336595.337570","volume":"2000","author":"Kumar S.","year":"2000","journal-title":"Fourth International Conference on Autonomous Agents"},{"key":"p_63_62","doi-asserted-by":"crossref","unstructured":"Lai, J. & Vergo, J. (1997). MedSpeak: Report creation with continuous speech recognition. Proceedings of the Conference on Human Factors in Computing (CHI'97), 431-438. New York: ACM Press.","DOI":"10.1145\/258549.258829"},{"key":"p_64_63","unstructured":"Landay, J. A. (1996). Interactive sketching for the early stages of user interface design. Unpublished doctoral dissertation, Carnegie Mellon University, Pittsburgh, PA."},{"key":"p_65_64","doi-asserted-by":"crossref","unstructured":"Landay, J. A. & Myers, B. A. (1995). Interactive sketching for the early stages of user interface design. Proceedings of the Conference on Human Factors in Computing Systems (CHI'95), 43-50. New York: ACM Press.","DOI":"10.1145\/223904.223910"},{"key":"p_66_65","doi-asserted-by":"crossref","unstructured":"Landay, J. A. & Myers, B. A. (1996). Sketching storyboards to illustrate interface behavior. Human Factors in Computing Systems (CHI'96), Conference Companion, 193-194. New York: ACM Press.","DOI":"10.1145\/257089.257257"},{"key":"p_67_66","doi-asserted-by":"crossref","unstructured":"Larson, J. A., Oviatt, S. L. & Ferro, D. (1999). Designing the user interface for pen and speech applications. CHI'99 Workshop, Conference on Human Factors in Computing Systems (CHI'99). New York: ACM Press.","DOI":"10.1145\/632716.632826"},{"key":"p_68_67","doi-asserted-by":"publisher","DOI":"10.1080\/088395199117324"},{"key":"p_69_68","doi-asserted-by":"crossref","first-page":"133","DOI":"10.1016\/0749-596X(85)90021-X","volume":"24","author":"Levelt W.","year":"1985","journal-title":"Journal of Memory and Language"},{"key":"p_70_69","unstructured":"Martin, A., Fiscus, J., Fisher, B., Pallett, D. & Przybocki, M. (1997). System descriptions and performance summary. Proceedings of the Conversational Speech Recognition Workshop\/DARPA Hub-5E Evaluation. San Francisco: Kaufmann."},{"key":"p_71_70","doi-asserted-by":"publisher","DOI":"10.1080\/088395199117504"},{"key":"p_72_71","unstructured":"McGee, D. & Cohen, P. R. (2001). Creating tangible interfaces by augmenting physical objects with mutilmodal language. Proceedings of the International Conference on Intelligent User Interfaces, 113-119. Santa Fe, NM: ACM Press."},{"key":"p_73_72","unstructured":"McGee, D., Cohen, P. R. & Oviatt, S. L. (1998). Confirmation in multimodal systems. Proceedings of the International Joint Conference of the Association for Computational Linguistics and the International Committee on Computational Linguistics (COLING-ACL'98), 823-829. Montreal, Canada: University of Montreal Press."},{"key":"p_74_73","doi-asserted-by":"crossref","first-page":"71","DOI":"10.1145\/354666.354674","volume":"2000","author":"McGee D. R.","year":"2000","journal-title":"Proceedings of the Designing Augmented Reality Environments Conference"},{"key":"p_75_74","doi-asserted-by":"crossref","first-page":"746","DOI":"10.1038\/264746a0","volume":"264","author":"McGurk H.","year":"1976","journal-title":"Nature"},{"key":"p_76_75","unstructured":"McNeill, D. (1992). Hand and mind: What gestures reveal about thought. Chicago: University of Chicago Press."},{"key":"p_77_76","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.1996.543250"},{"key":"p_78_77","unstructured":"Minsky, M. (1975). A framework for representing knowledge. In P. Winston (Ed.), The psychology of computer vision (pp. 211-277). New York: McGraw-Hill."},{"key":"p_79_78","unstructured":"Neal, J. G. & Shapiro, S. C. (1991). Intelligent multimedia interface technology. In J. Sullivan & S. Tyler (Eds.), Intelligent user interfaces (pp. 11-43). New York: ACM Press."},{"key":"p_80_79","unstructured":"Nicolino, T. (1994). A natural language processing based situation display. Proceedings of the 1994 Symposium on Command and Control Research and Decision Aids, 575-580. Monterey, CA: Naval Postgraduate School."},{"key":"p_81_80","unstructured":"Oviatt, S. L. (1992). Pen\/voice: Complementary multimodal communication. Proceedings of Speech Tech'92, New York."},{"key":"p_82_81","doi-asserted-by":"crossref","unstructured":"Oviatt, S. L. (1996). User-centered design of spoken language and multimodal interfaces. In M. Maybury & W. Wahlster (Eds.), Readings on intelligent user interfaces (pp. 620-629). San Francisco: Kaufmann. (Reprinted from IEEE Multimedia, 3(4), 26-35, 1996)","DOI":"10.1109\/93.556458"},{"key":"p_83_82","doi-asserted-by":"crossref","unstructured":"Oviatt, S. L. (1997). Multimodal interactive maps: Designing for human performance. Human-Computer Interaction, 12, 93-129.","DOI":"10.1207\/s15327051hci1201&2_4"},{"key":"p_84_83","doi-asserted-by":"crossref","unstructured":"Oviatt, S. L. (1999a). Mutual disambiguation of recognition errors in a multimodal architecture. Proceedings of the Conference on Human Factors in Computing Systems (CHI'99), 576-583. New York: ACM Press.","DOI":"10.1145\/302979.303163"},{"key":"p_85_84","doi-asserted-by":"publisher","DOI":"10.1145\/319382.319398"},{"key":"p_86_85","doi-asserted-by":"crossref","unstructured":"Oviatt, S. L. (2000a). Multimodal system processing in mobile environments. Proceedings of the 13th annual ACM Symposium on User Interface Software and Technology (UIST 2000), 21-30. New York: ACM Press.","DOI":"10.1145\/354401.354408"},{"key":"p_87_86","doi-asserted-by":"publisher","DOI":"10.1145\/348941.348979"},{"key":"p_88_87","unstructured":"Oviatt, S. L., Bernard, J. & Levow, G. (1999). Linguistic adaptation during error resolution with spoken and multimodal systems. Language and Speech, 41(3&4), 415-438."},{"issue":"4","key":"p_89_88","doi-asserted-by":"crossref","first-page":"297","DOI":"10.1016\/0885-2308(91)90001-7","volume":"5","author":"Oviatt S. L.","year":"1991","journal-title":"Computer Speech and Language"},{"key":"p_90_89","doi-asserted-by":"publisher","DOI":"10.1145\/330534.330538"},{"key":"p_91_90","first-page":"1351","volume":"2","author":"Oviatt S. L.","year":"1992","journal-title":"Proceedings of the International Conference on Spoken Language Processing"},{"key":"p_92_91","doi-asserted-by":"publisher","DOI":"10.1016\/0167-6393(94)90079-5"},{"key":"p_93_92","doi-asserted-by":"crossref","unstructured":"Oviatt, S. L., DeAngeli, A. & Kuhn, K. (1997). Integration and synchronization of input modes during multimodal human-computer interaction. Proceedings of Conference on Human Factors in Computing Systems (CHI'97), 415-422. New York: ACM Press.","DOI":"10.1145\/258549.258821"},{"key":"p_94_93","first-page":"2339","volume":"6","author":"Oviatt S. L.","year":"1998","journal-title":"Proceedings of the International Conference on Spoken Language Processing"},{"key":"p_95_94","unstructured":"Oviatt, S. L. & Olsen, E. (1994). Integration themes in multimodal human-computer interaction. In K. Shirai, S. Furui, & K. Kakehi (Eds.), Proceedings of the International Conference on Spoken Language Processing (Vol. 2, pp. 551-554). Yokohama, Japan: Acoustical Society of Japan."},{"key":"p_96_95","unstructured":"Oviatt, S. L. & Pothering, J. (1998). Interacting with animated characters: Research infrastructure and next-generation interface design. Proceedings of the First Workshop on Embodied Conversational Characters, 159-165, Tahoe City, CA."},{"key":"p_97_96","first-page":"204","volume":"2","author":"Oviatt S. L.","year":"1996","journal-title":"Proceedings of the International Conference on Spoken Language Processing"},{"key":"p_98_97","doi-asserted-by":"crossref","unstructured":"Pankanti, S., Bolle, R. M. & Jain, A. (Eds.). (2000). Biometrics: The future of identification [Special issue]. Computer, 33(2).","DOI":"10.1109\/2.820038"},{"key":"p_99_98","unstructured":"Papineni, K. A., Roukos, S. & Ward, R. T. (1997). Feature-based language understanding. Proceedings of the 5th European Conference On Speech Communication and Technology, 3, 1435-1438. Rhodes, Greece: European Speech Communication Association."},{"key":"p_100_99","doi-asserted-by":"crossref","unstructured":"Pavlovic, V., Berry, G. & Huang, T. S. (1997). Integration of audio\/visual information for use in human-computer intelligent interaction. Proceedings of IEEE International Conference on Image Processing, 121-124. Washington, DC: IEEE Press.","DOI":"10.1109\/ICIP.1997.647399"},{"key":"p_101_100","unstructured":"Pavlovic, V. & Huang, T. S. (1998). Multimodal prediction and classification on audio-visual features. AAAI'98 Workshop on Representations for Multi-Modal Human-Computer Interaction, 55-59. Menlo Park, CA: AAAI Press."},{"key":"p_102_101","doi-asserted-by":"crossref","unstructured":"Pentland, A. (1996, April). Smart rooms. Scientific American, 68-76.","DOI":"10.1038\/scientificamerican0496-68"},{"key":"p_103_102","unstructured":"Perrault, C. R. (1990). An application of default logic to speech act theory. In P. R. Cohen, J. Morgan, & M. E. Pollack (Eds.), Intentions in communication (pp. 161-186). Cambridge, MA: MIT Press."},{"key":"p_104_103","doi-asserted-by":"crossref","unstructured":"Rabiner, L. R. (1989). A tutorial on hidden Markov models and selected applications in speech recognition. IEEE Proceedings, 267-296.","DOI":"10.1109\/5.18626"},{"key":"p_105_104","unstructured":"Reithinger, N. & Klesen, M. (1997). Dialogue act classification using language models. Proceedings of Eurospeech'97, Vol. 4, 2235-2238. Grenoble: European Speech Communication Association."},{"key":"p_106_105","unstructured":"Rhyne, J. R. & Wolf, C. G. (1993). Recognition-based user interfaces. In H. R. Hartson & D. Hix (Eds.), Advances in human-computer interaction (Vol. 4, pp. 191-250). Norwood, NJ: Ablex."},{"key":"p_107_106","unstructured":"Roe, D. B. & Wilpon, J. G. (Eds.). (1994). Voice communication between humans and machines. Washington, DC: National Academy Press."},{"key":"p_108_107","doi-asserted-by":"crossref","unstructured":"Rogozan, A. & Deleglise, P. (1998). Adaptive fusion of acoustic and visual sources for automatic speech recognition. Speech Communication, 26 (1&2), 149-161.","DOI":"10.1016\/S0167-6393(98)00056-9"},{"key":"p_109_108","unstructured":"Roth, S. F., Mattis, J. & Mesnard, X. (1991). Graphics and natural language as components of automatic explanation. In J. W. Sullivan & S. W. Tyler (Eds.), Intelligent user interfaces (pp. 207-239). New York: ACM Press."},{"key":"p_110_109","unstructured":"Rubin, P., Vatikiotis-Bateson, E. & Benoit, C. (1998). Special issue on audio-visual speech processing. Speech Communication, 26(1&2)."},{"key":"p_111_110","unstructured":"Rudnicky, A. & Hauptmann, A. (1992). Multimodal interactions in speech systems. In M. Blattner & R. Dannenberg (Eds.), Multimedia interface design (pp. 147-172). New York: ACM Press."},{"key":"p_112_111","unstructured":"Sadek, D. (1991). Dialogue acts are rational plans. Proceedings of the ESCA\/ETRW Workshop on the Structure of Multimodal Dialogue, 1-29. Maratea, Italy: European Speech Communication Association."},{"key":"p_113_112","unstructured":"Sag, I. A. & Wasow, T. (1999). Syntactic theory: A formal introduction. Stanford, CA: CSLI Publications."},{"key":"p_114_113","unstructured":"Schwartz, D. G. (1993). Cooperating heterogeneous systems: A blackboard-based meta approach. Unpublished doctoral thesis, Case Western Reserve University."},{"key":"p_115_114","doi-asserted-by":"crossref","unstructured":"Searle, J. R. (1969). Speech acts: An essay in the philosophy of language. Cambridge, England: Cambridge University Press.","DOI":"10.1017\/CBO9781139173438"},{"key":"p_116_115","unstructured":"Sejnowski, T., Yuhas, B., Goldstein, M. & Jenkins, R. (1990). Combining visual and acoustic speech signal with a neural network improves intelligibility. In D. Touretzky (Ed.), Advances in neural information processing systems (pp. 232-239). Denver."},{"key":"p_117_116","unstructured":"Seneff, S., Hurley, E., Lau, R., Pao, C., Schmid, P. & Zue, V. (1998). Galaxy-II: A reference architecture for conversational system development. Proceedings of the International Conference on Spoken Language Processing, 1153. Sydney, Australia: ASSTA, Inc."},{"key":"p_118_117","unstructured":"Shaikh, A., Juth, S., Medl, A., Marsic, I., Kulikowski, C. & Flanagan, J. (1997). An architecture for multimodal information fusion. Proceedings of the Workshop on Perceptual User Interfaces (PUI'97), 91-93. Banff, Canada."},{"key":"p_119_118","doi-asserted-by":"crossref","unstructured":"Spiegel, M. & Kamm, C. (Eds.). (1997). Interactive voice technology for telecommunications applications (IVTTA'96) [Special issue]. Speech Communication, 23(1&2).","DOI":"10.1016\/S0167-6393(97)00051-4"},{"key":"p_120_119","unstructured":"Stork, D. G. & Hennecke, M. E. (Eds.). (1995). Speechreading by humans and machines. New York: Springer-Verlag."},{"key":"p_121_120","unstructured":"Suhm, B. (1998). Multimodal interactive error recovery for non-conversational speech user interfaces. Unpublished doctoral thesis, Fredericiana University, Germany."},{"key":"p_122_121","first-page":"861","volume":"2","author":"Suhm B.","year":"1996","journal-title":"Proceedings of the International Conference on Spoken Language Processing"},{"key":"p_124_122","doi-asserted-by":"crossref","unstructured":"Sutton, S. & Cole, R. (1997). The CSLU Toolkit: Rapid prototyping of spoken language systems. Proceedings of UIST'97: The ACM Symposium on User Interface Software and Technology, 85-86. New York: ACM Press.","DOI":"10.1145\/263407.263517"},{"key":"p_125_123","doi-asserted-by":"crossref","unstructured":"Turk, M. & Robertson, G. (Eds.). (2000). Perceptual user interfaces [Special issue]. Communications of the ACM, 43(3).","DOI":"10.1145\/330534.330535"},{"key":"p_126_124","unstructured":"Vergo, J. (1998). A statistical approach to multimodal natural language interaction. Proceedings of the AAAI'98 Workshop on Representations for Multimodal Human-Computer Interaction, 81-85. Madison, WI: AAAI Press."},{"key":"p_127_125","unstructured":"Vo, M. T. (1998). A framework and toolkit for the construction of multimodal learning interfaces. Unpublished doctoral thesis, Carnegie Mellon University, Pittsburgh, PA."},{"key":"p_128_126","unstructured":"Vo, M. T., Houghton, R., Yang, J., Bub, U., Meier, U., Waibel, A. & Duchnowski, P. (1995). Multimodal learning interfaces. Proceedings of the DARPA Spoken Language Technology Workshop, 233-237. Barton Creek: Kaufmann. Retrieved from the World Wide Web: http:\/\/werner.ira.uka.de\/ISL.publications.html#1995"},{"key":"p_129_127","first-page":"3545","volume":"6","author":"Vo M. T.","year":"1996","journal-title":"Speech and Signal Processing"},{"key":"p_130_128","unstructured":"Wagner, A. (1990). Prototyping: A day in the life of an interface designer. In B. Laurel (Ed.), The Art of human-computer interface design (pp. 79-84). Reading, MA: Addison-Wesley."},{"key":"p_131_129","doi-asserted-by":"publisher","DOI":"10.1016\/0004-3702(93)90022-4"},{"key":"p_132_130","doi-asserted-by":"crossref","unstructured":"Waibel, A., Hanazawa, T., Hinton, G., Shikano, K. & Lang, K. J. (1989). Phenome recognition using time-delay neural networks. IEEE Transactions on Acoustic, Speech, and Signal Processing, 37, 328-339.","DOI":"10.1109\/29.21701"},{"key":"p_133_131","unstructured":"Waibel, A., Vo, M. T., Duchnowski, P. & Manke, S. (1995). Multimodal interfaces [Special issue]. Artificial Intelligence Review, 10(3&4)."},{"key":"p_134_132","unstructured":"Wang, J. (1995). Integration of eye-gaze, voice and manual response in multimodal user interfaces. Proceedings of IEEE International Conference on Systems, Man and Cybernetics, 3938-3942. IEEE Press."},{"key":"p_135_133","doi-asserted-by":"crossref","first-page":"227","DOI":"10.1177\/001872088302500209","volume":"25","author":"Wickens C. D.","year":"1983","journal-title":"Human Factors"},{"key":"p_136_134","doi-asserted-by":"crossref","first-page":"533","DOI":"10.1177\/001872088402600505","volume":"26","author":"Wickens C. D.","year":"1984","journal-title":"Human Factors"},{"key":"p_137_135","unstructured":"Wojcik, R. & Holmback, H. (1996). Getting a controlled language off the ground at Boeing. Proceedings of the First International Workshop on Controlled Language Applications, 22-31. Seattle, WA."},{"key":"p_138_136","doi-asserted-by":"crossref","unstructured":"Wolf, C. G. & Morrel-Samuels, P. (1987). The use of hand-drawn gestures for text editing. International Journal of Man-Machine Studies, 27, 91-102.","DOI":"10.1016\/S0020-7373(87)80045-7"},{"key":"p_139_137","unstructured":"Wu, L., Oviatt, S. & Cohen, P. (1998). From members to teams to committee: A robust approach to gestural and multimodal recognition. Manuscript submitted for publication."},{"key":"p_140_138","doi-asserted-by":"publisher","DOI":"10.1109\/6046.807953"},{"key":"p_141_139","doi-asserted-by":"crossref","unstructured":"Zhai, S., Morimoto, C. & Ihde, S. (1999). Manual and gaze input cascaded (MAGIC) pointing. Proceedings of the Conference on Human Factors in Computing Systems (CHI'99), 246-253. New York: ACM Press.","DOI":"10.1145\/302979.303053"}],"container-title":["Human\u2013Computer Interaction"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.tandfonline.com\/doi\/pdf\/10.1207\/S15327051HCI1504_1","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,1,11]],"date-time":"2024-01-11T22:30:11Z","timestamp":1705012211000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.tandfonline.com\/doi\/full\/10.1207\/S15327051HCI1504_1"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2000,12]]},"references-count":139,"journal-issue":{"issue":"4","published-online":{"date-parts":[[2009,12,9]]},"published-print":{"date-parts":[[2000,12]]}},"alternative-id":["10.1207\/S15327051HCI1504_1"],"URL":"https:\/\/doi.org\/10.1207\/s15327051hci1504_1","relation":{},"ISSN":["0737-0024","1532-7051"],"issn-type":[{"value":"0737-0024","type":"print"},{"value":"1532-7051","type":"electronic"}],"subject":[],"published":{"date-parts":[[2000,12]]}}}