Electrophysiological study demonstrates that human superior temporal gyrus resets activity at word boundaries, indicating a dynamic mechanism for segmenting and encoding auditory word forms.
We perceive continuous speech as a series of discrete words, despite the lack of clear acoustic boundaries. The superior temporal gyrus (STG) encodes phonetic elements like consonants and vowels, but it is unclear how whole words are encoded. Using high-density cortical recordings and spoken narratives, we investigated how the human brain represents auditory word forms. STG activity exhibits a distinctive reset at word boundaries, marked by a sharp drop in cortical activity. Between resets, STG encodes acoustic-phonetic, prosodic, and lexical features, supporting integration of phonological features into coherent word forms. This process tracks the relative elapsed time within words, independent of absolute duration, providing a flexible encoding of variable word lengths. Similar dynamics were found in deeper layers of a self-supervised artificial speech network. Finally, a bistable word perception task revealed trial-by-trial STG responses to perceived word boundaries. Together, these findings support a new dynamical model of auditory word forms.
No takes yet. Share an insight, caveat, or question.
Zhang et al. (2025) studied this question.
Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context: