Back to results
Bibliographic record · Consultation and access
Artículo

Deep Neural Networks for Acoustic Modeling in Speech Recognition: The Shared Views of Four Research Groups

Geoffrey E. Hinton; Li Deng; Dong Yu; George E. Dahl; Abdelrahman Mohamed; Navdeep Jaitly; Andrew Senior; Vincent Vanhoucke · IEEE Signal Processing Magazine · 2012

Resource page
Quick overview. Review the resource’s basic details, then access the content using the main button. This page shows only the information needed to identify, cite, and open the work.

Resource access

Open the content from the main option or choose another available source.

OpenAlex OpenAlex Works
Entrar por OpenAlex
Main access

Resource page

Resource reference page. Full text availability has not been automatically confirmed.
Open resource

Summary

Descripción general del contenido del recurso.

Most current speech recognition systems use hidden Markov models (HMMs) to deal with the temporal variability of speech and Gaussian mixture models (GMMs) to determine how well each state of each HMM fits a frame or a short window of frames of coefficients that represents the acoustic input. An alternative way to evaluate the fit is to use a feed-forward neural network that takes several frames of coefficients as input and produces posterior probabilities over HMM states as output. Deep neural networks (DNNs) that have many hidden layers and are trained using new methods have been shown to outperform GMMs on a variety of speech recognition benchmarks, sometimes by a large margin. This article provides an overview of this progress and represents the shared views of four research groups that have had recent successes in using DNNs for acoustic modeling in speech recognition.

How to cite

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

Hinton, G. E, Deng, L, Yu, D, Dahl, G. E, Mohamed, A, Jaitly, N, Senior, A, & Vanhoucke, V. (2012). Deep Neural Networks for Acoustic Modeling in Speech Recognition: The Shared Views of Four Research Groups. https://doi.org/10.1109/msp.2012.2205597

MLA

Hinton, Geoffrey E, et al. "Deep Neural Networks for Acoustic Modeling in Speech Recognition: The Shared Views of Four Research Groups." 2012. https://doi.org/10.1109/msp.2012.2205597.

Chicago

Hinton, Geoffrey E, Li Deng, Dong Yu, George E. Dahl, Abdelrahman Mohamed, Navdeep Jaitly, Andrew Senior, and Vincent Vanhoucke. 2012. "Deep Neural Networks for Acoustic Modeling in Speech Recognition: The Shared Views of Four Research Groups.". https://doi.org/10.1109/msp.2012.2205597.

Harvard

Hinton, G. E. et al. 2012, Deep Neural Networks for Acoustic Modeling in Speech Recognition: The Shared Views of Four Research Groups, IEEE Signal Processing Magazine, available at: https://doi.org/10.1109/msp.2012.2205597 [Accessed 7 Aug. 2026].

Share and print

Save the record, copy its permanent link, or print it as a PDF.

Export reference

You can export the record in common formats for use in a reference manager.

Resource details

Bibliographic information to help confirm that this is the correct material.

Title
Deep Neural Networks for Acoustic Modeling in Speech Recognition: The Shared Views of Four Research Groups
Author / contributors
Geoffrey E. Hinton; Li Deng; Dong Yu; George E. Dahl; Abdelrahman Mohamed; Navdeep Jaitly; Andrew Senior; Vincent Vanhoucke
Publisher
IEEE Signal Processing Magazine
Publication year
2012
Language
English

Subjects

Explore related resources through these subjects.

Copied