AWaveFormer: Audio Wavelet Transformer Network for Generalized Audio Deepfake Detection

Rapid advancements in speech synthesis technology have made it easier to produce realistic synthetic speech, which poses serious threats to public privacy and security. Recent studies have investigated pre-trained models for feature extraction and adopted advanced architectures, including convolu…