Updated Aug 20, 2026

Voice Cloning

Recreating a specific person's voice from a short sample, then making it say anything.

Share

What it means

Voice cloning reproduces an individual's vocal characteristics from a sample — historically hours of studio recording, now seconds of ordinary audio. Anyone who has spoken publicly, appeared on a podcast, or posted a video has supplied enough material.

The legitimate uses are real: narrators scaling their own voice across languages, accessibility tools preserving the voice of someone losing speech, and localization that keeps an actor's voice in a language they don't speak.

The illegitimate use is fraud, and it is the dominant real-world harm. A cloned voice authorizing a payment or requesting an urgent transfer defeats the informal authentication most organizations actually rely on — recognizing who is speaking.

Legal protection is uneven. Some jurisdictions recognize voice as a protected personal attribute; many do not clearly, and enforcement lags the capability badly.

Why it matters

This breaks a trust assumption almost everyone still holds: that a familiar voice identifies the speaker. Any process where hearing someone constitutes authorization needs redesigning, and that is a procedural fix available today rather than a technical one to wait for.

In practice

Establish a verification channel for consequential requests — a callback to a known number, a shared code word — and make it policy rather than judgment. Reputable vendors require consent verification for cloning; the ones that matter for fraud do not.

Where this shows up

Tools and models in our catalog.

Related terms