Data-driven machine learning models for decoding speech categorization from evoked brain responses.
Data-driven machine learning models for decoding speech categorization from evoked brain responses.
复制标题
数据驱动的机器学习模型,用于从诱发的大脑反应中解码语音分类。
DOI:
10.1088/1741-2552/abecf0
复制
发表时间:
2021-03-23
影响因子:
4
通讯作者:
Bidelman GM
中科院分区:
文献类型:
--
作者:
Mahmud MS;Yeasin M;Bidelman GM
Categorical perception (CP) of audio is critical to understand how the human brain perceives speech sounds despite widespread variability in acoustic properties. Here, we investigated the spatiotemporal characteristics of auditory neural activity that reflects CP for speech (i.e. differentiates phonetic prototypes from ambiguous speech sounds). We recorded 64-channel electroencephalograms as listeners rapidly classified vowel sounds along an acoustic-phonetic continuum. We used support vector machine classifiers and stability selection to determine when and where in the brain CP was best decoded across space and time via source-level analysis of the event-related potentials. We found that early (120 ms) whole-brain data decoded speech categories (i.e. prototypical vs. ambiguous tokens) with 95.16% accuracy (area under the curve 95.14%; F1-score 95.00%). Separate analyses on left hemisphere (LH) and right hemisphere (RH) responses showed that LH decoding was more accurate and earlier than RH (89.03% vs. 86.45% accuracy; 140 ms vs. 200 ms). Stability (feature) selection identified 13 regions of interest (ROIs) out of 68 brain regions [including auditory cortex, supramarginal gyrus, and inferior frontal gyrus (IFG)] that showed categorical representation during stimulus encoding (0–260 ms). In contrast, 15 ROIs (including fronto-parietal regions, IFG, motor cortex) were necessary to describe later decision stages (later 300–800 ms) of categorization but these areas were highly associated with the strength of listeners’ categorical hearing (i.e. slope of behavioral identification functions). Our data-driven multivariate models demonstrate that abstract categories emerge surprisingly early (~120 ms) in the time course of speech processing and are dominated by engagement of a relatively compact fronto-temporal-parietal brain network.
登录
查看更多内容
影响因子:
5.7
作者:
Alho J;Green BM;May PJC;Sams M;Tiitinen H;Rauschecker JP;Jääskeläinen IP
通讯作者:
Jääskeläinen IP
影响因子:
2.6
作者:
Deschamps, Isabelle;Baum, Shari R.;Gracco, Vincent L.
通讯作者:
Gracco, Vincent L.
影响因子:
2.9
作者:
Carter JA;Bidelman GM
通讯作者:
Bidelman GM
影响因子:
16.6
作者:
Du Y;Buchsbaum BR;Grady CL;Alain C
通讯作者:
Alain C
影响因子:
4.3
作者:
Bidelman, Gavin M.;Bush, Lauren C.;Boudreaux, Alex M.
通讯作者:
Boudreaux, Alex M.