Skip to search form
Skip to main content
Skip to account menu
Semantic Scholar
Semantic Scholar's Logo
Search 238,266,953 papers from all fields of science
Search
Sign In
Create Free Account
Audio-visual speech recognition
Known as:
AVSR
, Audio visual speech recognition
Audio visual speech recognition (AVSR) is a technique that uses image processing capabilities in lip reading to aid speech recognition systems in…
Expand
Wikipedia
(opens in a new tab)
Create Alert
Alert
Related topics
Related topics
3 relations
Broader (1)
Computational linguistics
Image processing
Speech recognition
Papers overview
Semantic Scholar uses AI to extract papers important to this topic.
2008
2008
The application of manifold based visual speech units for visual speech recognition
Dahai Yu
2008
Corpus ID: 59846388
This dissertation presents a new learning-based representation that is referred to as a Visual Speech Unit for visual speech…
Expand
2007
2007
BUILDING A DATA CORPUS FOR AUDIO-VISUAL SPEECH RECOGNITION
A. G. ChiŃu
,
L. Rothkrantz
,
A. Chitu
,
L. Rothkrantz
2007
Corpus ID: 15718600
Data corpora are an important part of any audio-visual research. However, the time and effort needed to build a good dataset are…
Expand
2006
2006
Audio-visual speech recognition in the presence of a competing speaker
Xu Shao
,
J. Barker
Interspeech
2006
Corpus ID: 9210189
This paper examines the problem of estimating stream weights for a multistream audio-visual speech recogniser in the context of a…
Expand
2005
2005
A system for audio-visual speech recognition
I. Shdaifat
,
R. Grigat
Interspeech
2005
Corpus ID: 18171518
In this work, a system of audio visual speech recognition will be presented. A new hybrid visual feature combination, which is…
Expand
2002
2002
Large Vocabulary Audio-Visual Speech Recognition
C. Neti
,
G. Potamianos
2002
Corpus ID: 60187104
Abstract : This document is a report of the Large Vocabulary Audio-Visual Speech Recognition.
2002
2002
Audio-visual speech recognition for difficult environments
J. Gowdy
,
E. Patterson
2002
Corpus ID: 195932569
The work presented in this dissertation focuses on audio-visual speech recognition for difficult environments where background…
Expand
2002
2002
Medium vocabulary continuous audio-visual speech recognition
P. Wiggers
,
J. Wojdel
,
L. Rothkrantz
Interspeech
2002
Corpus ID: 3237568
This paper presents our experiments on continuous audio-visual speech recognition. A number of bimodal systems using feature…
Expand
2001
2001
Comparing audio- and a-posteriori-probability-based stream confidence measures for audio-visual speech recognition
Martin Heckmann
,
T. Wild
,
F. Berthommier
,
K. Kroschel
Interspeech
2001
Corpus ID: 155781
During the fusion of audio and video information for speech recognition, the estimation of the reliability of the noise affected…
Expand
2001
2001
A HYBRID ANN / HMM AUDIO-VISUAL SPEE SYSTEM
Martin Heckmann
,
F. Berthommier
,
Kristian Kroschel
2001
Corpus ID: 16828762
In this paper we present a system for audio-visual speech recognition based on a hybrid Artificial Neural Network/Hidden Markov…
Expand
1997
1997
Lip Tracking for Audio-Visual Speech Recognition.
R. Kaucic
1997
Corpus ID: 59984480
Abstract : Human speech is conveyed through both acoustic and visual channels and is therefore inherently multi-modal. Further…
Expand