Solving Google's Continuous Audio CAPTCHA with HMM-Based Automatic Speech Recognition
Solving Google's Continuous Audio CAPTCHA with HMM-Based Automatic Speech Recognition
复制标题
DOI:
10.1007/978-3-642-41383-4_3
复制
发表时间:
2013-11
期刊:
影响因子:
--
通讯作者:
Shotaro Sano;Takuma Otsuka;HIroshi G. Okuno
中科院分区:
文献类型:
--
作者:
Shotaro Sano;Takuma Otsuka;HIroshi G. Okuno
CAPTCHAs play critical roles in maintaining the security of various Web services by distinguishing humans from automated programs and preventing Web services from being abused. CAPTCHAs are designed to block automated programs by presenting questions that are easy for humans but difficult for computers, e.g., recognition of visual digits or audio utterances. Recent audio CAPTCHAs, such as Google’s audio reCAPTCHA, have presented overlapping and distorted target voices with stationary background noise. We investigate the security of overlapping audio CAPTCHAs by developing an audio reCAPTCHA solver. Our solver is constructed based on speech recognition techniques using hidden Markov models (HMMs). It is implemented by using an off-the-shelf library HMM Toolkit. Our experiments revealed vulnerabilities in the current version of audio reCAPTCHA with the solver cracking 52% of the questions. We further explain that background stationary noise did not contribute to enhance security against our solver.