Overlapped Frequency-Distributed Network: Frequency-Aware Voice Spoofing Countermeasure

Citations

WEB OF SCIENCE

7
Citations

SCOPUS

12

초록

Numerous IT companies around the world are developing and deploying artificial voice assistants via their products, but they are still vulnerable to spoofing attacks. Since 2015, the competition “Automatic Speaker Verification Spoofing and Countermeasures Challenge (ASVspoof)” has been held every two years to encourage people to design systems that can detect spoofing attacks. In this paper, we focused on developing spoofing countermeasure systems mainly based on Convolutional Neural Networks (CNNs). However, CNNs have translation invariant property, which may cause loss of frequency information when a spectrogram is used as input. Hence, we propose models which split inputs along the frequency axis: 1) Overlapped Frequency-Distributed (OFD) model and 2) Non-overlapped Frequency-Distributed (Non-OFD) model. Using ASVspoof 2019 dataset, we measured their performances with two different activations; ReLU and Max feature map (MFM). The best performing model on LA dataset is the Non-OFD model with ReLU which achieved an equal error rate (EER) of 1.35%, and the best performing model on PA dataset is the OFD model with MFM which achieved an EER of 0.35%. Copyright © 2022 ISCA.

키워드

audio deep synthesis; countermeasure; Deep learning; fake audio detection; spoofing
제목
Overlapped Frequency-Distributed Network: Frequency-Aware Voice Spoofing Countermeasure
저자
Choi, S.; Kwak, I.-Y.; Oh, S.
DOI
10.21437/Interspeech.2022-657
발행일
2022-09
유형
Proceedings Paper
저널명
Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH
권
2022-September
페이지
3558 ~ 3562