Voice cloning AI can replicate a person's voice from as little as 3 seconds of audio. This technology is increasingly combined with intimate video deepfakes to create more convincing content, which can be used as leverage in various schemes. Understanding your exposure can help you act on any findings related to voice cloning.

How voice cloning is used in schemes

Voice cloning is combined with deepfake video to create more convincing fake intimate content. Synthetic audio of individuals is used in schemes, falsely claiming to be recordings of intimate conversations. Voice deepfakes are sent to contacts, employers, or family members as a separate harassment vector from visual deepfakes.

Detection and prevention

Voice deepfakes can often be detected through spectrogram analysis and AI detection tools. Commercial voice detection services identify synthetic audio with improving accuracy. However, detection lags generation quality as with visual deepfakes.

Understanding voice deepfakes

Recognizing the implications of voice deepfakes involves understanding that some states have separate voice rights protections. Misuse of voice cloning may involve various fraud statutes.

Frequently asked questions

Can I find deepfake audio of myself on the internet?

Audio-visual deepfake content combining voice cloning with intimate imagery should be reported to relevant platform trust and safety teams. ScanErase's biometric scanning covers content identifiable through facial likeness.

Is voice cloning without consent illegal?

Creating and distributing synthetic voice content for harmful purposes may violate various statutes depending on context. The most clear-cut illegality is when voice cloning is combined with intimate visual imagery.