Audio Watermark

Audio Watermark explained in the WithFeeling sonic branding glossary

Audio Watermark.

What is an audio watermark?

An audio watermark is a signal hidden inside a piece of audio so that machines can identify, track or authenticate it, while people hear nothing unusual. It carries a small payload, typically an identifier and a timestamp, and it is designed to survive whatever happens to the audio on its way to a listener: compression, broadcast, streaming, even being played through a speaker and picked up again by a microphone. A sonic logo is designed to be heard. An audio watermark is designed not to be.

How an audio watermark works

Human hearing has blind spots, and watermarking lives in them. When a loud sound is present, quieter sounds close to it in frequency and time become inaudible, an effect called masking. The same effect lets MP3 and AAC throw away most of a recording without anyone noticing. A watermarking encoder analyses the audio moment by moment, finds the parts that are masked, and writes its data there, spread across many frequencies and repeated many times so that no single damaged fragment loses the message.

Several techniques exist and they are often combined.

  • Spread spectrum. The payload is turned into a faint noise pattern shaped to sit under the music. A detector that knows the pattern can pull it back out even when the music is louder by a factor of thousands.
  • Echo hiding. Very short echoes, too close to the original to be heard as echoes, encode ones and zeros by their delay.
  • Phase coding. The timing relationship between frequency components is nudged in ways the ear does not register.
  • Frequency-band tones. Low-level tones are placed in bands that the surrounding audio masks, a method used for broadcast audience measurement because it survives the acoustic path from a television to a meter in the room.

A robust watermark is built to survive everything, including re-recording. A fragile watermark is built to break, so that a broken mark proves the audio was altered. Both have their uses.

Watermark or fingerprint?

These are the two ways a machine can recognise audio, and they are opposites. A watermark adds something to the audio before it is released. A fingerprint takes something from the audio after the fact: a compact description of its spectral shape that can be matched against a database. Shazam, YouTube’s Content ID and most music recognition work on fingerprints, which is why they can identify a recording nobody marked. Watermarks win when you need to tell two identical copies apart, for example which broadcaster aired a spot or which reviewer leaked a screener. Fingerprints win when you need to recognise audio you never had the chance to mark.

Where audio watermarks are used

Audience measurement. The most famous application. Nielsen’s broadcast watermarking encodes a station identifier and time into every programme and advert on the air, and portable meters carried by panel members decode what they were exposed to, whether the set was in the living room or a bar. Ratings, and the advertising rates built on them, rest on marks nobody can hear.

Advertising verification. Brands and agencies mark their spots so that monitoring services can prove where and when each one aired, which matters when a media plan promises a thousand airings across fifty stations.

Anti-piracy. Film screeners and cinema prints carry forensic marks that identify the recipient, so a leaked copy points back to a source. Cinavia, embedded in film soundtracks, tells consumer players to mute copies that have left the authorised chain.

Rights and royalties. Production music libraries and broadcasters use marks to log usage automatically, replacing the cue sheets that used to be filled in by hand.

AI provenance. The newest use, and the one that will matter most for brands. Generated speech and music are now marked at the point of creation so that a detector can later say whether a clip came from a synthesis model. Google DeepMind’s SynthID marks audio produced by its models, and Meta’s AudioSeal does the same for speech, with the aim of localising which seconds of a recording are synthetic. As cloned voices become indistinguishable from real ones, this is the only scalable way to tell them apart.

Why this matters in sonic branding

Three reasons, and they are getting more important every year.

  • Proof of airplay. A sonic logo is an asset that earns its value through exposure. Watermarking the master versions lets a brand measure, rather than assume, how often the signature was actually heard, and it settles disputes with media owners with data instead of affidavits.
  • Protecting a brand voice. Once a brand commissions a synthetic voice or licenses a real one for cloning, every authorised output should be marked. A marked voice can be told apart from an impersonation, and an unmarked clip claiming to be the brand can be dismissed.
  • Rights hygiene. A watermark does not replace a contract, but it makes enforcement possible. If your sonic logo turns up in someone else’s advert, a mark shows whose master they used.

Our own practice is to deliver brand voice and sonic identity masters unmarked to the client, and to recommend marking at the distribution stage, where the platform or measurement partner can choose the scheme that fits their detectors. Marking twice with incompatible systems is a waste, and marking a master that will be re-edited a hundred times is fragile by definition.

What a watermark cannot do

It cannot stop copying; it can only make copying traceable. It cannot survive every attack: heavy pitch shifting, time stretching and re-synthesis can strip robust marks, which is why detection systems combine watermarks with fingerprints. And it is not free of audible cost. Aggressive marks on sparse, quiet material such as solo piano or spoken word can be heard by a trained ear, so a good encoder is tuned per programme, not applied blindly.

Frequently asked questions

Can you hear an audio watermark?
A well-made one, no. Marks are placed where the audio itself masks them, and encoders test their own output against psychoacoustic models. Badly tuned marks on quiet material can be audible as faint noise or a slight thickening of the sound.

Does an audio watermark survive MP3 compression or a phone recording?
Robust marks are designed to survive both, and broadcast measurement depends on surviving the journey through a loudspeaker and a room. Fragile marks are designed to fail, which is the point of them.

Is a watermark the same as metadata?
No. Metadata such as ID3 tags or a file’s embedded rights fields sits beside the audio and is lost the moment the file is converted, trimmed or re-recorded. A watermark lives inside the audio itself.

Should a brand watermark its sonic logo?
Mark the distributed versions if you intend to measure airplay or need to trace misuse, and keep an unmarked master. For brand voices built with AI, mark every output.

Related terms: audio fingerprinting, synthetic voice, sound trademark, master rights. For the discipline this protects, see sonic branding.

Your Bazaar sonic identity
From our work · Sonic BrandingYour Bazaar

Gold at the Transform Awards MEA 2025. A full sonic identity built around a sonic logo.

View case study

WithFeeling
Privacy Overview

This website uses cookies so that we can provide you with the best user experience possible. Cookie information is stored in your browser and performs functions such as recognising you when you return to our website and helping our team to understand which sections of the website you find most interesting and useful.