Skip to search form
Skip to main content
Skip to account menu
Semantic Scholar
Semantic Scholar's Logo
Search 237,047,380 papers from all fields of science
Search
Sign In
Create Free Account
Speech recognition
Known as:
Speech recognizer
, Voice dialing
, Voice to text
Expand
Speech recognition (SR) is the inter-disciplinary sub-field of computational linguistics which incorporates knowledge and research in the linguistics…
Expand
Wikipedia
(opens in a new tab)
Create Alert
Alert
Related topics
Related topics
50 relations
AI winter
Articulatory speech recognition
Audio-visual speech recognition
BINA48
Expand
Papers overview
Semantic Scholar uses AI to extract papers important to this topic.
Highly Cited
2006
Highly Cited
2006
Automatic Synchronization between Lyrics and Music CD Recordings Based on Viterbi Alignment of Segregated Vocal Signals
Hiromasa Fujihara
,
Masataka Goto
,
J. Ogata
,
Kazunori Komatani
,
T. Ogata
,
Hiroshi G. Okuno
IEEE International Symposium on Multimedia
2006
Corpus ID: 17445483
This paper describes a system that can automatically synchronize between polyphonic musical audio signals and corresponding…
Expand
2006
2006
Towards recognizing emotion with affective dimensions through body gestures
P. Ravindra
,
De Silva
,
M. Osano
,
A. Marasinghe
,
A. Madurapperuma
International Conference on Automatic Face and…
2006
Corpus ID: 18238252
Due to the ever-increasing importance of computers in many areas of today's society such as e-learning, telehome-health care, and…
Expand
2006
2006
A Comparative Study of Discriminative Methods for Reranking LVCSR N-Best Hypotheses in Domain Adaptation and Generalization
Zhengyu Zhou
,
Jianfeng Gao
,
F. Soong
,
H. Meng
IEEE International Conference on Acoustics Speech…
2006
Corpus ID: 7421452
This paper is an empirical study on the performance of different discriminative approaches to reranking the N-best hypotheses…
Expand
2005
2005
Improved MO-LRT VAD based on bispectra Gaussian model
J. Górriz
,
J. Ramírez
,
J. C. Segura
,
C. Puntonet
2005
Corpus ID: 18646069
A robust algorithm for voice activity detection (VAD) is presented. It defines a likelihood ratio test (LRT) involving multiple…
Expand
2004
2004
Spontaneous handwriting recognition and classification
A. Rossi
,
Alfons Juan-Císcar
,
E. Vidal
Proceedings of the 17th International Conference…
2004
Corpus ID: 13386034
Finite-state models are used to implement a handwritten text recognition and classification system for a real application…
Expand
1997
1997
The GlobalPhone Project: Multilingual LVCSR with JANUS-3
Tanja Schultz
,
M. Westphal
,
A. Waibel
1997
Corpus ID: 12815324
. This paper describes our recent effort in developing the Global Phone database for multilingual large vocabulary continuous…
Expand
Highly Cited
1996
Highly Cited
1996
A probabilistic framework for feature-based speech recognition
James R. Glass
,
Jane W. Chang
,
M. McCandless
Proceeding of Fourth International Conference on…
1996
Corpus ID: 1207578
Most current speech recognizers use an observation space which is based on a temporal sequence of "frames" (e.g. Mel-cepstra…
Expand
1996
1996
Robust distant-talking speech recognition
J. Pearson
,
Q. Lin
,
+4 authors
J. Flanagan
IEEE International Conference on Acoustics…
1996
Corpus ID: 11574578
Most contemporary speech recognizers are designed to operate with close-talking speech and they work best in a quiet laboratory…
Expand
Highly Cited
1989
Highly Cited
1989
The Lincoln robust continuous speech recognizer
D. Paul
IEEE International Conference on Acoustics…
1989
Corpus ID: 61239657
The Lincoln stress-resistant HMM (hidden Markov model) CSR has been extended to large-vocabulary continuous speech for both…
Expand
1989
1989
Token Passing : a Simple Conceptual Model for ConnectedSpeech Recognition
SystemsS
,
J. ..
,
+4 authors
ThorntonCambridge
1989
Corpus ID: 18288351
This paper describes a simple but powerful abstract model in which connected word recognition is viewed as a process of passing…
Expand