Artificial speech detection using image-based features and random forest classifier

Tan, Choon Beng and Mohd Hanafi Ahmad Hijazi and Frazier Kok and Mohd Saberi Mohamad and Puteri Nor Ellyza Nohuddin (2022) Artificial speech detection using image-based features and random forest classifier. IAES International Journal of Artificial Intelligence (IJ-AI), 11 (1). pp. 161-172. ISSN 2252-8938

[img] Text
Artificial speech detection using image.pdf

Download (40kB)
[img] Text
Artificial speech detection using image1.pdf
Restricted to Registered users only

Download (411kB) | Request a copy


The ASV spoof 2015 Challenge was one of the efforts of the research community in the field of speech processing to foster the development of generalized countermeasures against spoofing attacks. However, most countermeasures submitted to the ASV spoof 2015 Challenge failed to detect the S10 attack effectively, the only attack that was generated using the waveform concatenation approach. Hence, more informative features are needed to detect previously unseen spoofing attacks. This paper presents an approach that uses data transformation techniques to engineer image-based features together with random forest classifier to detect artificial speech. The objectives are two-fold: (i) to extract image-based features from the Mel frequency cepstral coefficients representation of the speech signal and (ii) to compare the performance of using the extracted features and Random Forest to determine the authenticity of voices with the existing approaches. An audio-to-image transformation technique was used to engineer new features in classifying genuine and spoof voices. An experiment was conducted to find the appropriate combination of the engineered features and classifier. Experimental results showed that the proposed approach was able to detect speech synthesis and voice conversion attacks effectively, with an equal error rate of 0.10% and accuracy of 99.93%.

Item Type: Article
Uncontrolled Keywords: Anti-spoofing voice recognition , Artificial speech detection , Speaker recognition , Speaker verification , Voice presentation attack detection
Subjects: Q Science > QA Mathematics > QA1-939 Mathematics > QA71-90 Instruments and machines > QA75.5-76.95 Electronic computers. Computer science
Divisions: FACULTY > Faculty of Computing and Informatics
Date Deposited: 09 Jun 2022 08:41
Last Modified: 09 Jun 2022 08:41

Actions (login required)

View Item View Item