Open-Set Source Tracing of Audio Deepfake Systems
Authors: Nicholas Klein, Hemlata Tak, Elie Khoury
Published: 2025-07-09 01:03:36+00:00
AI Summary
This paper addresses the challenge of open-set source tracing in audio deepfakes. It introduces a novel softmax energy (SME) score for out-of-distribution (OOD) detection, significantly improving open-set source tracing performance compared to existing energy-based methods. The authors achieve an FPR95 of 8.3% by combining SME with data augmentation techniques.
Abstract
Existing research on source tracing of audio deepfake systems has focused primarily on the closed-set scenario, while studies that evaluate open-set performance are limited to a small number of unseen systems. Due to the large number of emerging audio deepfake systems, robust open-set source tracing is critical. We leverage the protocol of the Interspeech 2025 special session on source tracing to evaluate methods for improving open-set source tracing performance. We introduce a novel adaptation to the energy score for out-of-distribution (OOD) detection, softmax energy (SME). We find that replacing the typical temperature-scaled energy score with SME provides a relative average improvement of 31% in the standard FPR95 (false positive rate at true positive rate of 95%) measure. We further explore SME-guided training as well as copy synthesis, codec, and reverberation augmentations, yielding an FPR95 of 8.3%.