On the similarities of representations in artificial and brain neural networks for speech recognition.

Wingfield, Cai; Zhang, Chao; Devereux, Barry; Fonteneau, Elisabeth; Thwaites, Andrew; Liu, Xunying; Woodland, Phil; Marslen-Wilson, William; Su, Li

doi:10.3389/fncom.2022.1057439

On the similarities of representations in artificial and brain neural networks for speech recognition.

Published version

Peer-reviewed

Repository URI

https://www.repository.cam.ac.uk/handle/1810/346328

Repository DOI

https://doi.org/10.17863/CAM.93749

Files

Published version (1.47 MB)

Type

Article

Authors

Wingfield, Cai

Zhang, Chao

Devereux, Barry

Fonteneau, Elisabeth

Thwaites, Andrew

https://orcid.org/0000-0002-6237-7140

Show 4 more

Abstract

INTRODUCTION: In recent years, machines powered by deep learning have achieved near-human levels of performance in speech recognition. The fields of artificial intelligence and cognitive neuroscience have finally reached a similar level of performance, despite their huge differences in implementation, and so deep learning models can-in principle-serve as candidates for mechanistic models of the human auditory system. METHODS: Utilizing high-performance automatic speech recognition systems, and advanced non-invasive human neuroimaging technology such as magnetoencephalography and multivariate pattern-information analysis, the current study aimed to relate machine-learned representations of speech to recorded human brain representations of the same speech. RESULTS: In one direction, we found a quasi-hierarchical functional organization in human auditory cortex qualitatively matched with the hidden layers of deep artificial neural networks trained as part of an automatic speech recognizer. In the reverse direction, we modified the hidden layer organization of the artificial neural network based on neural activation patterns in human brains. The result was a substantial improvement in word recognition accuracy and learned speech representations. DISCUSSION: We have demonstrated that artificial and brain neural networks can be mutually informative in the domain of speech recognition.

Keywords

auditory cortex, automatic speech recognition, deep neural network, representational similarity analysis, speech recognition

Journal Title

Front Comput Neurosci

Journal ISSN

1662-5188
1662-5188

Volume Title

16

Publisher

Frontiers Media SA

Publisher DOI

https://doi.org/10.3389/fncom.2022.1057439

Rights

Attribution 4.0 International

Sponsorship

European Research Council (230570)
European Research Council (669820)

Collections

Jisc Publications Router

On the similarities of representations in artificial and brain neural networks for speech recognition.

Published version

Peer-reviewed

Repository URI

Repository DOI

Files

Type

Change log

Authors

Abstract

Description

Keywords

Journal Title

Conference Name

Journal ISSN

Volume Title

Publisher

Publisher DOI

Rights

Sponsorship

Collections