Total: 1
Speech conveys not only linguistic information but also critical cues to speaker identity. Previous studies have identified bilateral superior temporal cortex (STG/STS) as central to human voice perception. However, speaker identity, comprising multiple-layer information, has largely been treated as a unitary construct. The present fMRI study investigated whether distinct speaker traits, such as gender, age, and accent, are processed by shared or dissociable neural mechanisms. Results revealed a common voice-processing core centered on bilateral STG/STS, which was engaged across all conditions. However, neural responses within this core were not uniform, with accent processing eliciting stronger activation than gender and age. Beyond this shared region, each trait recruited partially distinct cortical regions. These findings revealed a "core-plus-extension" model of speaker identification, in which shared auditory mechanisms are complemented by trait-specific cortical regions.