The method which is called the “tandem approach ” in speech recog-nition has been shown to increase performance by using classifier posterior probabilities as observations in a hidden Markov model. We study the effect of using visual tandem features in audio-visual speech recognition using a novel setup which uses multiple classi-fiers to obtain multiple visual tandem features. We adopt the ap-proach of multi-stream hidden Markov models where visual tandem features from two different classifiers are considered as additional streams in the model. It is shown in our experiments that using multiple visual tandem features improve the recognition accuracy in various noise conditions. In addition, in order to handle asynchrony between audio and v...
Abstract — Visual speech information from the speaker’s mouth region has been successfully shown to ...
We investigate the use of local, frame-dependent reliability indicators of the audio and visual moda...
87 p.Thesis (Ph.D.)--University of Illinois at Urbana-Champaign, 2003.Differences in the characteris...
The method which is called the “tandem approach ” in speech recog-nition has been shown to increase ...
This paper presents a novel Hidden Markov Model architecture to model the joint probability of pair...
With the increase in the computational complexity of recent computers, audio-visual speech recogniti...
This paper presents a novel Hidden Markov Model architecture to model the joint probability of pairs...
Extending automatic speech recognition (ASR) to the vi sual modality has been shown to greatly incre...
Abstract—This paper presents the design and evaluation of a speaker-independent audio-visual speech ...
Speech recognition can be improved by using visual information in the form of lip movements of the s...
The Multi-Stream automatic speech recognition approach was investigated in this work as a framework ...
Extending automatic speech recognition (ASR) to the visual modality has been shown to greatly increa...
Extending automatic speech recognition (ASR) to the visual modality has been shown to greatly increa...
This paper describes a complete system for audio-visual recognition of continuous speech including r...
We address the problem of robust lip tracking, visual speech feature extraction, and sensor integrat...
Abstract — Visual speech information from the speaker’s mouth region has been successfully shown to ...
We investigate the use of local, frame-dependent reliability indicators of the audio and visual moda...
87 p.Thesis (Ph.D.)--University of Illinois at Urbana-Champaign, 2003.Differences in the characteris...
The method which is called the “tandem approach ” in speech recog-nition has been shown to increase ...
This paper presents a novel Hidden Markov Model architecture to model the joint probability of pair...
With the increase in the computational complexity of recent computers, audio-visual speech recogniti...
This paper presents a novel Hidden Markov Model architecture to model the joint probability of pairs...
Extending automatic speech recognition (ASR) to the vi sual modality has been shown to greatly incre...
Abstract—This paper presents the design and evaluation of a speaker-independent audio-visual speech ...
Speech recognition can be improved by using visual information in the form of lip movements of the s...
The Multi-Stream automatic speech recognition approach was investigated in this work as a framework ...
Extending automatic speech recognition (ASR) to the visual modality has been shown to greatly increa...
Extending automatic speech recognition (ASR) to the visual modality has been shown to greatly increa...
This paper describes a complete system for audio-visual recognition of continuous speech including r...
We address the problem of robust lip tracking, visual speech feature extraction, and sensor integrat...
Abstract — Visual speech information from the speaker’s mouth region has been successfully shown to ...
We investigate the use of local, frame-dependent reliability indicators of the audio and visual moda...
87 p.Thesis (Ph.D.)--University of Illinois at Urbana-Champaign, 2003.Differences in the characteris...