Pular para o conteúdo principal

ETHNOS_APP

Início • Busca • Periódicos • Lista 0

Designing the User Interface for Multimodal Speech and Pen-Based Gesture Applications

State-of-the-Art Systems and Future Research Directions

Dados Bibliográficos

ID15216144
AutoresSharon Oviatt (0000-0003-4664-1412, Oregon Health & Science University, autor correspondente), Philip R Cohen (0000-0003-1303-064X, Oregon Research Institute), Phil Cohen, Lizhong Wu (Software (Germany)), Lisbeth Duncan, Bernhard Suhm (Boeing (Australia)), Josh Bers, Thomas F Holzman, Thomas Holzman, Terry Winograd (0000-0003-4435-7215), James A Landay (0000-0003-1520-8894, Stanford University), James Landay, Jim Larson (0000-0002-3814-2659, University of California, Berkeley), David Ferro (Intel (United States))
Ano2000
Volume15
Fascículo4
Páginas263-322
Data de publicação2000-12-01
Peer ReviewedSim
Open AccessNão
TipoARTICLE
PeriódicoHuman-Computer Interaction (JOURNAL)
Identificadores do periódicoISSN: 0737-0024 • E-ISSN: 1532-7051
EditoraTaylor & Francis (PUBLISHER • GB)
DOI10.1207/s15327051hci1504_1
OpenAlexW2009803366
IdiomaEN
Citações recebidas3
Referências citadas78

The growing interest in multimodal interface design is inspired in large part by the goals of supporting more transparent, flexible, efficient, and powerfully expressive means of humancomputer interaction than in the past. Multimodal interfaces are expected to support a wider range of diverse applications, to be usable by a broader spectrum of the average population, and to function more reliably under realistic and challenging usage conditions. In this paper, we summarize the emerging architectural approaches for interpreting speech and pen-based gestural input in a robust manner--- including early and late fusion approaches, and the new hybrid symbolic/statistical approach. We also describe a diverse collection of state-of-the-art multimodal systems that process users' spoken and gestural input. These applications range from map-based and virtual reality systems for engaging in simulations and training, to field medic systems for mobile use in noisy environments, to web-based transactions and standard text-editing applications that will reshape daily computing and have a significant commercial impact. To realize successful multimodal systems of the future, many key research challenges remain to be addressed. Among these challenges are the development of cognitive theories to guide multimodal system design, and the development of effective natural language processing, dialogue processing, and error handling techniques. In addition, new multimodal systems will be needed that can function more robustly and adaptively, and with support for collaborative multi-person use. Before this new class of systems can proliferate, toolkits also will be needed to promote software development for both simulated and functioning systems. Multimodal Speech and Gesture Interfaces 3 CONT

Field (mathematics · Function (biology · Gesture · Human–computer interaction · Interface (matter · Multimedia · Multimodal interaction · Multimodality · Process (computing · USable · User interface · World Wide Web · Computer Science · Multi-Agent Systems and Negotiation · Natural Language Processing Techniques · Speech and dialogue systems · Artificial Intelligence

  • Immersive virtual reality in the age of the Metaverse

    Open Access•Ersin Dincelli, Alper Yayla•The Journal of Strategic…•2022

  • From Siri to Bixby

    Ritika Chopra, Seema Bhardwaj et al.•Journal of Information…•2026

  • Exiting the Cleanroom

    Scott Carter, Scott L Carter et al.•Human-Computer Interaction•2008

  • Multimedia interface design

    Blessed Numetu, Meera M Blattner et al.•Multimedia interface design•1992

  • Embodied Conversational Agents

    Justine Cassell, Joseph Sullivan et al.•Embodied Conversational Agents•2000

  • Hearing lips and seeing voices

    Open Access•Harry Mcgurk, John Macdonald•Nature•1976

  • Gesticulation and Speech

    Adam Kendon•Relationship of Verbal and…•1980

  • A tutorial on hidden Markov models and selected applications in speech recognition

    Open Access•L R Rabiner•Proceedings of the IEEE•1989

  • Finding Structure in Time

    Open Access•Jeffrey L Elman•Cognitive Science•1990

  • Functional Grammar

    Open Access•Martin Kay•Proceedings of the Annual Meeting…•1979

  • Computer-human interface solutions for emergency medical care

    Open Access•Thomas G Holzman•interactions•1999

  • Mulitmodal Interactive Maps

    Sharon Oviatt•Human-Computer Interaction•1997

  • Syntactic Theory

    Open Access•Ivan A Sag, Thomas Wasow•Computational Linguistics•2000

  • Survey of the State of the Art in Human Language Technology

    K Bretonnel Cohen, Kay Cohen et al.•Language•2000

  • Linguistic Adaptations During Spoken and Multimodal Error Resolution

    Open Access•Sharon Oviatt, Jon Bernard et al.•Language and Speech•1998

  • Speech Acts

    Open Access•John R Searle•Speech acts•1969

Obras citantes distintas3
Citações por ano0,17
Intervalo de citações2008 - 2026 (19)
Velocidade de citaçãocurrent
Altamente citadoNão
Tipos de citaçãoNeutras: 3
Ethnos_APP • Projeto Open Source • Licença MIT • Frontend v2.0.0 • Privacidade e Cookies • Documentação da API: api.ethnos.app/docs • Código da API: GitHub • DOI: 10.5281/zenodo.17049435 • Código do Frontend: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae