A design space for multimodal systems

Laurence Nigay, Joëlle Coutaz

1993 · 301 citations · 12 references

TL;DR

Multimodal interaction lets users communicate with a computer using voice, gesture, and typing. The paper analyzes how multiple modalities are integrated, arguing that concurrency of processing and data fusion are the two essential features, and proposes a design space and classification method for multimodal systems. Using a software‑engineering perspective, the authors clarify the concept of a multimodal system and present a software architecture model that implements concurrency and data fusion. The architecture is illustrated by two team‑developed systems, VoicePaint and NoteBook, demonstrating the proposed design space.

Abstract

Multimodal interaction enables the user to employ different modalities such as voice, gesture and typing for communicating with a computer. This paper presents an analysis of the integration of multiple communication modalities within an interactive system. To do so, a software engineering perspective is adopted. First, the notion of "multimodal system" is clarified. We aim at proving that two main features of a multimodal system are the concurrency of processing and the fusion of input/output data. On the basis of these two features, we then propose a design space and a method for classifying multimodal systems. In the last section, we present a software architecture model of multimodal systems which supports these two salient properties: concurrency of processing and data fusion. Two multimodal systems developed in our team, VoicePaint and NoteBook, are used to illustrate the discussion.

References

12