Multimodal Interaction Sessions All-at-Once
Tutorial: 14. Designing Speech User Interfaces
Sunday 09:00 - 17:30.
Nicole Yankelovich, Sun Microsystems Lab;
Jennifer Lai, IBM T. J. Watson Research Center
Benefits
You will explore techniques for designing speech user interfaces by examining
existing products and research prototypes.
Origins
This is a new tutorial for CHI 98.
Features
- learn about basic concepts of speech input (recognition) and
output (synthesis)
- understand the breadth of products and types of applications
where speech is used
- discover design issues that affect both multimodal and speech-only
systems including techniques for
providing user feedback, strategies for designing effective prompts,
methods for handling user and system
errors
- examine a range of user studies that are appropriate at different
stages of a speech application's life cycle
Audience
User interface designers and application developers who are interested in
understanding the issues involved in designing effective speech interfaces.
Participants are not expected to have prior knowledge of the basic underlying
speech input and output concepts.
Presentation
Lecture and small-group exercises.
Instructors
Nicole Yankelovich is a Principal Investigator of the Speech Applications project
at Sun Microsystems Laboratories. She specializes in designing speech user
interfaces and has published numerous papers on user interface design, speech
applications and hypertext. Jennifer Lai has been working with the Speech
Recognition effort at IBM Research for 9 years. She has conducted requirements
gathering, design, development and usability testing for speech systems,
including the IBM product MedSpeak/Radiology. She has published papers on the use
of speech in multimedia systems, the development of statistical language models
and has the patent for a Method and System for Natural Language Translation.
Papers: About Faces
Tuesday 16:00 - 17:30.
Session Chair: Robert J. K. Jacob, Tufts University
- Visual Tracking for Multimodal Human Computer Interaction
Jie Yang, Rainer Stifelhagen, Uwe Meier, Alex Waibel, Carnegie Mellon University
- When My Face is the Interface: An Experimental Comparison of
Interacting With One's Own Face or
Someone Else's Face
Clifford Nass, Eun-Young Kim, Eun-Ju Lee, Stanford University
- Digital Smart Kiosk Project
Andrew D. Christian, Brian L. Avery, Digital
Late-Breaking Results: The Raw and The Cooked: Experiments and Applications of Speech Interaction
Tuesday 16:00 - 17:30.
Session Chair: Debby Hindus, Interval Research Corporation
- The Sound of Your Stuff: Designing a Complex Auditory Display for
an Interactive Museum Exhibit
Maribeth Back, Xerox PARC; Jonathan Cohen, Interval Research
- Synchronization of Speech and Hand Gestures during Multimodal
Human-Computer Interaction
Marie-Luce Bourguet, Arkio Ando, NHK (Japan Broadcasting Corp.)
- "Just Speak Naturally": Designing for Naturalness in Automated
Spoken Dialogues
David Williams, Vocalis; Christine Cheepen, University of Surrey
- Speech Recognition, Children and Reading
Don Nix, Peter Fairweather, Bill Adams, IBM T. J. Watson Research Center
- Play It Again: A Study of the Factors Underlying Speech Browsing Behavior
Steve Whittaker, Julia Hirschberg, Christine H. Nakatani, ATT Labs Research
- All Talk and All Action: Strategies for Managing Voicemail Messages
Steve Whittaker, Julia Hirschberg, Christine H. Nakatani, ATT Labs Research
Late-Breaking Results: Great E-Scapes: Electronic Landscapes and Soundscapes
Wednesday 16:00 - 17:30.
Session Chair: Marilyn Salzman, George Mason University
- Experiments in Inhabited TV
Steve Benford, Chris Greenhalgh, Chris Brown, Nottingham University;
Graham Walker, Tim Regan, Paul Rea, Jason Morphett, BT Laboratories;
John Wyver, Illuminations Television
- The Lightwork Performance: Algorithmically Mediated Interaction
for Virtual Environments
John Bowers, Royal Institute of Technology & University of Manchester;
Sten-Olof Hellstrom, Kai-Mikael Jaa-Aro, Royal Institute of Technology
- Look Who's Talking: the GAZE Groupware System
Roel Vertegaal, Twente University; Harro Vons, Baan Development;
Robert Slagter, Twente University
- Looking for Sound? Selling Perceptual Space in Hierarchically Nested Boxes
Roel Vertegaal, University of Twente; Barry Eaglestone, University of Bradford
- Using Earcons to Improve the Usability of Tool Palettes
Stephen A. Brewster, University of Glasgow
- Multi-Parameter Controllers for Audio Mixing
Craig Wisneski, Ed Hammond, MIT Media Laboratory
Papers: Hear Here!
Thursday 11:00 - 12:30.
Session Chair: Sharon Oviatt, Oregon Graduate Institute of Science & Technology
- Designing Audio Aura
Elizabeth D. Mynatt, Maribeth Back, Roy Want, Michael Baer,
Jason B. Ellis, Xerox PARC
- Communicating Graphical Information to Blind Users Using Music:
The Role of Context in Design
James L. Alty, Loughborough University;
Dimitrios I. Rigas, Liverpool John Moores University
- What Can I Say?: Evaluating a Spoken Language Interface to Email
Marilyn A. Walker, AT&T Labs Research;
Jeanne Fromer, MIT AI Lab;
Giuseppe Di Fabbrizio, Craig Mestel, Don Hindle, AT&T Labs Research