articleNov 1, 2012Closed access
Deep Neural Networks for Acoustic Modeling in Speech Recognition
Abstract
Most current speech recognition systems use hidden Markov models (HMMs) to deal with the temporal variability of speech and Gaussian mixture models to determine how well each state of each HMM fits a frame or a short window of frames of coefficients that represents the acoustic input. An alternative way to evaluate the fit is to use a feedforward neural network that takes several frames of coefficients as input and produces posterior probabilities over HMM states as output. Deep neural networks with many hidden layers, that are trained using new methods have been shown to outperform Gaussian mixture models on a variety of speech recognition benchmarks, sometimes by a large margin. This paper provides an…
Citation impact
1,903
total citations
- FWCI
- 182.61
- Percentile
- 100%
- References
- 64
Citations per year
Authors
11Topics & keywords
Topics
Keywords
- Hidden Markov model
- Computer science
- Speech recognition
- Mixture model
- Artificial neural network
- Margin (machine learning)
- Deep neural networks
- Pattern recognition (psychology)
No related works found for this paper.