Speech Privacy Leakage from Shared Gradients in Distributed Learning
Speech Privacy Leakage from Shared Gradients in Distributed Learning
复制标题
DOI:
10.1109/icassp49357.2023.10095443
复制
发表时间:
2023-02
期刊:
影响因子:
--
通讯作者:
Zhuohang Li;Jiaxin Zhang;Jian Liu
中科院分区:
文献类型:
--
作者:
Zhuohang Li;Jiaxin Zhang;Jian Liu
Distributed machine learning paradigms, such as federated learning, have been recently adopted in many privacy-critical applications for speech analysis. However, such frameworks are vulnerable to privacy leakage attacks from shared gradients. Despite extensive efforts in the image domain, the exploration of speech privacy leakage from gradients is quite limited. In this paper, we explore methods for recovering private speech/speaker information from the shared gradients in distributed learning settings. We conduct experiments on a keyword spotting model with two different types of speech features to quantify the amount of leaked information by measuring the similarity between the original and recovered speech signals. We further demonstrate the feasibility of inferring various levels of side-channel information, including speech content and speaker identity, under the distributed learning framework without accessing the user’s data.