bioRxiv · 10.1101/2025.01.30.635775
Parallel hierarchical encoding of linguistic representations in the human auditory cortex and recurrent automatic speech recognition systems
Abstract
Transforming continuous acoustic speech signals into discrete linguistic meaning is a remarkable computational feat accomplished by both the human brain and modern artificial intelligence. A key scientific question is whether these biological and artificial systems, despite their different architectures, converge on similar strategies to solve this challenge. While ASR systems now achieve human-level performance, research on their parallels with the brain has been limited by biologically implausible, non-causal models and comparisons that stop at predicting brain activity without detailing the alignment of the underlying representations. Furthermore, studies using text-based models overlook the crucial acoustic stages of speech processing. Here, using high-resolution intracranial recordings and a causal, recurrent ASR model, we bridge these gaps by uncovering a striking correspondence between the brains processing hierarchy and the models internal representations. Specifically, we demonstrate a deep alignment in their algorithmic approach: neural activity in distinct cortical regions maps topographically to corresponding model layers, and critically, the representational content at each stage follows a parallel progression from acoustic to phonetic, lexical, and semantic information. This work thus moves beyond demonstrating simple model-brain alignment to specifying the shared underlying representations at each stage of processing, providing direct evidence that both systems converge on a similar computational strategy for transforming sound into meaning.
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Keshishian, M., Mischler, G., Thomas, S., Kingsbury, B., Bickel, S., Mehta, A. D., Mesgarani, N.. 2025-02-01. Parallel hierarchical encoding of linguistic representations in the human auditory cortex and recurrent automatic speech recognition systems. https://doi.org/10.1101/2025.01.30.635775
Cite the original work for its findings. Save a collection to share your selection of sources.