The EU AI Act voice watermarking rules took effect on August 2, 2026. Every AI system that generates synthetic audio, image, video or text must now mark its output in a machine-readable format that can be detected as artificially generated, and the fines for missing it run to EUR 15 million or 3% of worldwide annual turnover. If you build on a TTS API or run a voice agent for EU users, the marking duty sits with your provider, but the compliance risk lands on your product.

What does the EU AI Act require for AI audio?

Article 50 of the EU AI Act covers transparency for generative AI, and the part that matters for voice is Article 50(2). Providers of AI systems that generate synthetic audio must ensure the output is marked in a machine-readable format and detectable as artificially generated or manipulated. That covers text-to-speech output, voice clones, AI dubbing, and general-purpose AI models with audio generation.

The obligation is outcome based. The law does not name a specific watermarking technology. It requires the marking to be effective, interoperable, robust and reliable, as far as this is technically feasible. The phrasing matters, because audio is harder than images. Most providers are settling on a layered approach: an inaudible watermark embedded in the signal plus signed provenance metadata.